UniPose: A Unified Multimodal Framework for Human Pose Comprehension, Generation and Editing
Yiheng Li, Ruibing Hou, Hong Chang, Shiguang Shan, Xilin Chen
2025Year
9Top-tier citations
Abstract
Please generate the SMPL pose from the description: This person supports his body with the right leg, stretches his left leg backward, and extends his hands to both sides.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 01ca2538-ff30-4aa9-91fa-14e701294253Cited by top-tier papers9
- RealisMotion: Decomposed Human Motion Control and Video Generation in the World SpaceJingyun Liang, Jingkai Zhou, Shikai Li, Chenjie Cao et al.ICML 2026 · 9 citations
- HIS-GPT: Towards 3D Human-In-Scene Multimodal UnderstandingJiahe Zhao, Ruibing Hou, Zejie Tian, Hong Chang et al.ICCV 2025 · 6 citations
- Beyond Global Alignment: Fine-Grained Motion-Language Retrieval via Pyramidal Shapley-Taylor LearningHanmo Chen, Guangtao Lyu, Chenghao Xu, Jiexi Yan et al.ICML 2026 · 3 citations
- HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric ScenesKeliang Li, Hongze Shen, Hao Shi, Ruibing Hou et al.ICLR 2026 · 2 citations
- MotionCtrl: A Real-Time Controllable Vision-Language-Motion ModelBin Cao, Sipeng Zheng, Ye Wang, Lujie Xia et al.ICCV 2025 · 1 citation
Builds on35
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
- MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language ModelsDeyao Zhu, Jun Chen, Xiaoqian Shen, Xiang Li et al.ICLR 2024 · 3,079 citations
Related papers
- ChatPose: Chatting about 3D Human PoseYao Feng, Jing Lin, Sai Kumar Dwivedi, Yu Sun et al.CVPR 2024
- ScaMo: Exploring the Scaling Law in Autoregressive Motion Generation ModelShunlin Lu, Jingbo Wang, Zeyu Lu, Ling-Hao Chen et al.CVPR 2025
- Prepose: Privacy, Security, and Reliability for Gesture-Based ProgrammingLucas Silva Figueiredo, Benjamin Livshits, David Molnar, Margus VeanesS&P 2016 · 30 citations
- Learning to Dress 3D People in Generative ClothingQianli Ma, Jinlong Yang, Anurag Ranjan, Sergi Pujades et al.CVPR 2020
- 3D Multi-bodies: Fitting Sets of Plausible 3D Human Models to Ambiguous Image DataBenjamin Biggs, David Novotný, Sébastien Ehrhardt, Hanbyul Joo et al.NeurIPS 2020 · 79 citations
