From 3D Pose to Prose: Biomechanics-Grounded Vision-Language Coaching
Yuyang Ji, Yixuan Shen, Shengjie Zhu, Yu Kong, Feng Liu
摘要
We present BioCoach, a biomechanics-grounded visionlanguage framework for fitness coaching from streaming video. BioCoach fuses visual appearance and 3D skeletal kinematics, through a novel three-stage pipeline: an exercise-specific degree-of-freedom selector that focuses analysis on salient joints; a structured biomechanical context that pairs individualized morphometrics with cycle and constraint analysis; and a vision-biomechanics conditioned feedback module that applies cross-attention to generate precise, actionable text. Using parameter-efficient training that freezes the vision and language backbones, BioCoach yields transparent, personalized reasoning rather than pattern matching. To enable learning and fair evaluation, we augment QEVD-fit-coach with biomechanicsoriented feedback to create QEVD-bio-fit-coach, and we introduce a biomechanics-aware LLM judge metric. Bio-Coach delivers clear gains on QEVD-bio-fit-coach across lexical and judgment metrics while maintaining temporal triggering; on the original QEVD-fit-coach, it improves text quality and correctness with near-parity timing, demonstrating that explicit kinematics and constraints are key to accurate, phase-aware coaching. Project
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper19
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 被引用 11,349 次
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- Flamingo: a Visual Language Model for Few-Shot LearningJean-Baptiste Alayrac, Jeff Donahue, Pauline Luc, Antoine Miech 等NeurIPS 2022 · 被引用 6,707 次
- InstructBLIP: Towards General-purpose Vision-Language Models with Instruction TuningWenliang Dai, Junnan Li, Dongxu Li, Anthony Meng Huat Tiong 等NeurIPS 2023 · 被引用 4,013 次
- MotionGPT: Human Motion as a Foreign LanguageBiao Jiang, Xin Chen, Wen Liu, Jingyi Yu 等NeurIPS 2023 · 被引用 698 次
相关 Paper
- AgentCoach: LLM-Based Adaptive Coaching Feedback for Motor Skill LearningDizhi Ma, Jiakun Yu, Xinyi Wang, Xiyun Hu 等CHI 2026 · 被引用 1 次
- TechCoach: Towards Technical-Point-Aware Descriptive Action CoachingYuan-Ming Li, An-Lan Wang, Ling-An Zeng, Kun-Yu Lin 等AAAI 2026
- ExpertAF: Expert Actionable Feedback from VideoKumar Ashutosh, Tushar Nagarajan, Georgios Pavlakos, Kris Kitani 等CVPR 2025
- ViSTAR: Virtual Skill Training with Augmented Reality with 3D Avatars and LLM coaching agentChunggi Lee, Hayato Saiki, Tica Lin, Eiji Ikeda 等CHI 2026 · 被引用 1 次
- VisMimic: Integrating Motion Chain in Feedback Video Generation for Motor CoachingLiqi Cheng, Xiao Xie, Yiwei Peng, Minghao Feng 等UIST 2025 · 被引用 2 次
