EgoLM: Multi-Modal Language Model of Egocentric Motions
Fangzhou Hong, Vladimir Guzov, Hyo Jin Kim, Yuting Ye, Richard A. Newcombe, Ziwei Liu, Lingni Ma
2025Year
15Top-tier citations
Abstract
The person is standing straight as she puts the piece of clothing on the hanger." "The person turns around then walks out of the bedroom.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers15
- Whole-Body Conditioned Egocentric Video PredictionYutong Bai, Danny Tran, Amir Bar, Yann LeCun et al.NeurIPS 2025 · 33 citations
- Leveraging the Power of MLLMs for Gloss-Free Sign Language TranslationJungeun Kim, Hyeongwoo Jeon, Jongseong Bae, Ha Young KimICCV 2025 · 10 citations
- Interaction-aware Representation Modeling With Co-Occurrence Consistency for Egocentric Hand-Object ParsingYUEJIAO SU, Yi Wang, Lei Yao, Yawen Cui et al.ICLR 2026 · 5 citations
- EgoPoseFormer v2: Accurate Egocentric Human Motion Estimation for AR/VRZhenyu Li, Sai Kumar Dwivedi, Filip Maric, Carlos Chacón et al.CVPR 2026 · 3 citations
- UniEgoMotion: A Unified Model for Egocentric Motion Reconstruction, Forecasting, and GenerationChaitanya Patel, Hiroki Nakamura, Yuta Kyuragi, Kazuki Kozuka et al.ICCV 2025 · 3 citations
Builds on21
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah et al.NeurIPS 2020 · 64,255 citations
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
- MotionGPT: Human Motion as a Foreign LanguageBiao Jiang, Xin Chen, Wen Liu, Jingyi Yu et al.NeurIPS 2023 · 698 citations
- Action-Conditioned 3D Human Motion Synthesis with Transformer VAEMathis Petrovich, Michael J. Black, Gül VarolICCV 2021 · 672 citations
Related papers
- Guiding Human-Object Interactions with Rich Geometry and RelationsMengqing Xue, Yifei Liu, Ling Guo, Shaoli Huang et al.CVPR 2025
- Ego4o: Egocentric Human Motion Capture and Understanding from Multi-Modal InputJian Wang, Rishabh Dabral, Diogo C. Luvizon, Zhe Cao et al.CVPR 2025
- Clothe and PoseNakul Sharma, Aayush Bansal, Minh VoCVPR 2026
- Learning Efficient Robotic Garment Manipulation with StandardizationChangshi Zhou, Feng Luan, Jiarui Hu, Shaoqiang Meng et al.ICML 2025
- MoMask: Generative Masked Modeling of 3D Human MotionsChuan Guo, Yuxuan Mu, Muhammad Gohar Javed, Sen Wang et al.CVPR 2024
