ChatHuman: Chatting about 3D Humans with Tools
Jing Lin, Yao Feng, Weiyang Liu, Michael J. Black
摘要
Tool-Conditioned Transform Tool Docs The man in the dark blue shirt is preparing to sit down. Please find him out and estimate his 3d human pose. ⋆ Equal contribution. This work was done while YF and JL were at Meshcapade. specialized 3D outputs into comprehensible formats. Experiments demonstrate that ChatHuman surpasses existing models in both tool selection accuracy and overall performance across various 3D human tasks, and it supports interactive chatting with users. ChatHuman represents a significant step toward consolidating diverse analytical methods into a unified, robust system for 3D human tasks. Code and data are available at chathuman.github.io. What are the height, chest, hip, waist circumferences (in meter), and weight (in kg) of the person in the image? Height: 1.69 m Chest: 0.91 m Hip: 1.0 m Waist: 0.75 m Weight: 59.9 kg The woman is about 1.68 m tall, with a chest circumference of 0.90 m, a hip circumference of 0.93 m, a waist circumference of 0.78 m, weighs approximately 69 kg Height: 1.44 m Chest: 0.83 m Hip: 0.91 m Waist: 0.7 m Weight: 45.5 kg Tool ChatHuman GT Label What are the height, chest, hip, waist circumferences (in meter), and weight (in kg) of the person in the image?
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- EchoMotion: Unified Human Video and Motion Generation via Dual-Modality Diffusion TransformerYuxiao Yang, Hualian Sheng, Sijia Cai, Jing Lin 等ICLR 2026 · 被引用 12 次
- FrankenMotion: Part-level Human Motion Generation and CompositionChuqiao Li, Xianghui Xie, Yong Cao, Andreas Geiger 等CVPR 2026 · 被引用 10 次
- HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric ScenesKeliang Li, Hongze Shen, Hao Shi, Ruibing Hou 等ICLR 2026 · 被引用 2 次
它引用的顶会 Paper31
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Retrieval-Augmented Generation for Knowledge-Intensive NLP TasksPatrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni 等NeurIPS 2020 · 被引用 19,162 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 被引用 11,349 次
相关 Paper
- ChatGarment: Garment Estimation, Generation and Editing via Large Language ModelsSiyuan Bian, Chenghao Xu, Yuliang Xiu, Artur Grigorev 等CVPR 2025
- ChainHOI: Joint-based Kinematic Chain Modeling for Human-Object Interaction GenerationLing-An Zeng, Guohong Huang, Yi-Lin Wei, Shengbo Gu 等CVPR 2025
- CapHuman: Capture Your Moments in Parallel UniversesChao Liang, Fan Ma, Linchao Zhu, Yingying Deng 等CVPR 2024
- Per Garment Capture and Synthesis for Real-time Virtual Try-onToby Long Hin Chong, I-Chao Shen, Nobuyuki Umetani, Takeo IgarashiUIST 2021 · 被引用 9 次
- A Focused Human Body Model for Accurate Anthropometric Measurements ExtractionShuhang Chen, Xianliang Huang, Zhizhou Zhong, Juhong Guan 等CVPR 2025
