arXiv · 2502.16175
Mojito: LLM-Aided Motion Instructor with Jitter-Reduced Inertial Tokens
Abstract
Human bodily movements convey critical insights into action intentions and cognitive processes, yet existing multimodal systems primarily focused on understanding human motion via language, vision, and audio, which struggle to capture the dynamic forces and torques inherent in 3D motion. Inertial measurement units (IMUs) present a promising alternative, offering lightweight, wearable, and privacy-conscious motion sensing. However, processing of streaming IMU data faces challenges such as wireless transmission instability, sensor noise, and drift, limiting their utility for long-term real-time motion capture (MoCap), and more importantly, online motion analysis. To address these challenges, we introduce Mojito, an intelligent motion agent that integrates inertial sensing with large language models (LLMs) for interactive motion capture and behavioral analysis.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Ziwei Shan, Yaoyu He, Chengfeng Zhao, Jiashen Du, Jingyan Zhang, Qixuan Zhang, Jingyi Yu, Lan Xu. 2025-02-22. Mojito: LLM-Aided Motion Instructor with Jitter-Reduced Inertial Tokens. https://arxiv.org/abs/2502.16175
Cite the original work for its findings. Save a collection to share your selection of sources.