paper-with-me

Papers

Mojito: LLM-Aided Motion Instructor with Jitter-Reduced Inertial Tokens

2025-02-22 · Ziwei Shan, Yaoyu He, Chengfeng Zhao, Jiashen Du, Jingyan Zhang, Qixuan Zhang, Jingyi Yu, Lan Xu

Human bodily movements convey critical insights into action intentions and cognitive processes, yet existing multimodal systems primarily focused on understanding human motion via language, vision, and audio, which struggle to capture the dynamic forces and torques inherent in 3D motion. Inertial measurement units (IMUs) present a promising alternative, offering lightweight, wearable, and privacy-conscious motion sensing. However, processing of streaming IMU data faces challenges such as wireless transmission instability, sensor noise, and drift, limiting their utility for long-term real-time motion capture (MoCap), and more importantly, online motion analysis. To address these challenges, we introduce Mojito, an intelligent motion agent that integrates inertial sensing with large language models (LLMs) for interactive motion capture and behavioral analysis.

📄 PDF Abstract BibTeX arXiv:2502.16175

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Mojito: Motion Trajectory and Intensity Control for Video Generation

2024-12-12 · Xuehai He, Shuohang Wang, Jianwei Yang, Xiaoxia Wu 외

Recent advancements in diffusion models have shown great promise in producing high-quality video content. However, efficiently training diffusion models capable of integrating directional guidance and controllable motion…

Computational EfficiencyOptical Flow EstimationText-to-Video GenerationVideo Generation

FlowMotion: Target-Predictive Conditional Flow Matching for Jitter-Reduced Text-Driven Human Motion Generation

2025-04-02 · Manolo Canales Cuba, Vinícius do Carmo Melício, João Paulo Gois

Achieving high-fidelity and temporally smooth 3D human motion generation remains a challenge, particularly within resource-constrained environments. We introduce FlowMotion, a novel method leveraging Conditional Flow Mat…

Computational EfficiencyMotion GenerationMotion Synthesis

MOJITO: Modal Joint Learning for Unified End-to-End Autonomous Driving

2026-07-26 · Zhijing Cheng, Xuancheng Zhang, Donglin Di, Lei Fan 외 arxiv

End-to-end autonomous driving systems commonly follow a cascaded two-stage pipeline where a perception stage compresses multi-modal sensor inputs into a compact context and a downstream planner predicts trajectories cond…

Instruction FollowingAutonomous Driving

StableFace: Analyzing and Improving Motion Stability for Talking Face Generation

2022-08-29 · Jun Ling, Xu Tan, Liyang Chen, Runnan Li 외

While previous speech-driven talking face generation methods have made significant progress in improving the visual quality and lip-sync quality of the synthesized videos, they pay less attention to lip motion jitters wh…

Face GenerationTalking Face GenerationVideo Generation

Attention Mixtures for Time-Aware Sequential Recommendation

2023-04-17 · Viet-Anh Tran, Guillaume Salha-Galvan, Bruno Sguerra, Romain Hennequin

Transformers emerged as powerful methods for sequential recommendation. However, existing architectures often overlook the complex dependencies between user preferences and the temporal context. In this short paper, we i…

Recommendation SystemsSequential Recommendation