Categorical Codebook Matching for Embodied Character Controllers
Sebastian Starke, Paul Starke, Nicky He, Taku Komura, Yuting Ye
Abstract
Translating motions from a real user onto a virtual embodied avatar is a key challenge for character animation in the metaverse. In this work, we present a novel generative framework that enables mapping from a set of sparse sensor signals to a full body avatar motion in real-time while faithfully preserving the motion context of the user. In contrast to existing techniques that require training a motion prior and its mapping from control to motion separately, our framework is able to learn the motion manifold as well as how to sample from it at the same time in an end-to-end manner. To achieve that, we introduce a technique called codebook matching which matches the probability distribution between two categorical codebooks for the inputs and outputs for synthesizing the character motions. We demonstrate this technique can successfully handle ambiguity in motion generation and produce high quality character controllers from unstructured motion capture data. Our method is especially useful for interactive applications like virtual reality or video games where high accuracy and responsiveness are needed.
Ask about this paper
Ask your agent about it.
Lune has read the top-tier papers around this one, so every answer names the papers it rests on.
Your agent calls
Lunesearch_papers
Free to start. No credit card required.
Terminal
Install the CLIlune papers get dddf3dd5-b163-4e99-be60-cb191d444005Cited by top-tier papers15
- EgoPoseFormer v2: Accurate Egocentric Human Motion Estimation for AR/VRZhenyu Li, Sai Kumar Dwivedi, Filip Maric, Carlos Chacón et al.CVPR 2026 · 3 citations
- SMGDiff: Soccer Motion Generation using Diffusion Probabilistic ModelsHongdi Yang, Chengyang Li, Zhenxuan Wu, Gaozheng Li et al.ICCV 2025 · 2 citations
- DartControl: A Diffusion-Based Autoregressive Motion Model for Real-Time Text-Driven Motion ControlKaifeng Zhao, Gen Li, Siyu TangICLR 2025 · 1 citation
- Group Inertial Poser: Multi-Person Pose and Global Translationfrom Sparse Inertial Sensors and Ultra-Wideband RangingYing Xue, Jiaxi Jiang, Rayan Armani, Dominik Hollidt et al.ICCV 2025 · 1 citation
- PuppetChat: Fostering Intimate Communication through Bidirectional Actions and MicronarrativesEmma Jiren Wang, Siying Hu, Zhicong LuCHI 2026 · 1 citation
Related papers
- Realistic Full-Body Tracking from Sparse Observations via Joint-Level ModelingXiaozheng Zheng, Zhuo Su, Chao Wen, Zhou Xue et al.ICCV 2023 · 57 citations
- CoolMoves: User Motion Accentuation in Virtual RealityKaran Ahuja, Eyal Ofek, Mar González-Franco, Christian Holz et al.UbiComp 2021 · 69 citations
- FLAG: Flow-based 3D Avatar Generation from Sparse ObservationsSadegh Aliakbarian, Pashmina Cameron, Federica Bogo, Andrew W. Fitzgibbon et al.CVPR 2022
- Driving-signal aware full-body avatarsTimur M. Bagautdinov, Chenglei Wu, Tomas Simon, Fabián Prada et al.SIGGRAPH 2021 · 71 citations
- Continuous Intermediate Token Learning with Implicit Motion Manifold for Keyframe Based Motion InterpolationClinton Ansun Mo, Kun Hu, Chengjiang Long, Zhiyong WangCVPR 2023
