Lune

CVPR2026Top-tier venue

From 3D Pose to Prose: Biomechanics-Grounded Vision-Language Coaching

Yuyang Ji, Yixuan Shen, Shengjie Zhu, Yu Kong, Feng Liu

2026Year
4Citations

Abstract

We present BioCoach, a biomechanics-grounded visionlanguage framework for fitness coaching from streaming video. BioCoach fuses visual appearance and 3D skeletal kinematics, through a novel three-stage pipeline: an exercise-specific degree-of-freedom selector that focuses analysis on salient joints; a structured biomechanical context that pairs individualized morphometrics with cycle and constraint analysis; and a vision-biomechanics conditioned feedback module that applies cross-attention to generate precise, actionable text. Using parameter-efficient training that freezes the vision and language backbones, BioCoach yields transparent, personalized reasoning rather than pattern matching. To enable learning and fair evaluation, we augment QEVD-fit-coach with biomechanicsoriented feedback to create QEVD-bio-fit-coach, and we introduce a biomechanics-aware LLM judge metric. Bio-Coach delivers clear gains on QEVD-bio-fit-coach across lexical and judgment metrics while maintaining temporal triggering; on the original QEVD-fit-coach, it improves text quality and correctness with near-parity timing, demonstrating that explicit kinematics and constraints are key to accurate, phase-aware coaching. Project

Ask about this paper

Your agent reads all of it.

Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.

Questions to start from

Your agent calls

Luneget_paper_fulltext

Ask in Lune

Free to start. No credit card required.

lune papers fulltext 8d4fe99f-544e-438c-9dd8-31cf791b7f5d

Builds on19

Related papers

Dusk over the sea between two cliffs drawn in fine vertical lines