FLAG3D: A 3D Fitness Activity Dataset with Language Instruction
Yansong Tang, Jinpeng Liu, Aoyang Liu, Bin Yang, Wenxun Dai, Yongming Rao, Jiwen Lu, Jie Zhou, Xiu Li
Abstract
Figure 1. An overview of the proposed FLAG3D dataset, which contains 180K videos of 60 daily fitness activities. Our dataset is comprised of (a) 3D activity sequences captured from advanced MoCap system, (b) rendered videos of different people with their SMPL parameters, and (c) real-world videos obtained by cost-effective phones from both indoor and outdoor natural environments. FLAG3D also provides a series of detailed and professional sentence-level language instructions for each fitness activity. All figures are best viewed in color.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1fa206cf-219d-41f7-aabd-6dca1c509d66Cited by top-tier papers9
- Video Action DifferencingJames Burgess, Xiaohan Wang, Yuhui Zhang, Anita Rau et al.ICLR 2025 · 1,149 citations
- FineDance: A Fine-grained Choreography Dataset for 3D Full Body Dance GenerationRonghui Li, Junfan Zhao, Yachao Zhang, Mingyang Su et al.ICCV 2023 · 110 citations
- MotionCraft: Crafting Whole-Body Motion with Plug-and-Play Multimodal ControlsYuxuan Bian, Ailing Zeng, Xuan Ju, Xian Liu et al.AAAI 2025 · 22 citations
- Inter-X: Towards Versatile Human-Human Interaction AnalysisLiang Xu, Xintao Lv, Yichao Yan, Xin Jin et al.CVPR 2024 · 18 citations
- LDPose: Towards Inclusive Human Pose Estimation for Limb-Deficient Individuals in the WildJiaying Ying, Heming Du, Kaihao Zhang, Lincheng Li et al.ICCV 2025 · 3 citations
Builds on41
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 35,902 citations
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 11,743 citations
- AMASS: Archive of Motion Capture As Surface ShapesNaureen Mahmood, Nima Ghorbani, Nikolaus F. Troje, Gerard Pons-Moll et al.ICCV 2019 · 1,784 citations
- Learning to Reconstruct 3D Human Pose and Shape via Model-Fitting in the LoopNikos Kolotouros, Georgios Pavlakos, Michael J. Black, Kostas DaniilidisICCV 2019 · 1,139 citations
Related papers
- M3GYM: A Large-Scale Multimodal Multi-view Multi-person Pose Dataset for Fitness Activity Understanding in Real-world SettingsQingzheng Xu, Ru Cao, Xin Shen, Heming Du et al.CVPR 2025
- AIFit: Automatic 3D Human-Interpretable Feedback Models for Fitness TrainingMihai Fieraru, Mihai Zanfir, Silviu Cristian Pirlea, Vlad Olaru et al.CVPR 2021
- HmPEAR: A Dataset for Human Pose Estimation and Action RecognitionYitai Lin, Zhijie Wei, Wanfa Zhang, Xiping Lin et al.ACM MM 2024 · 5 citations
- BABEL: Bodies, Action and Behavior With English LabelsAbhinanda R. Punnakkal, Arjun Chandrasekaran, Nikos Athanasiou, Alejandra Quiros-Ramirez et al.CVPR 2021
- Muscles in ActionMia Chiquier, Carl VondrickICCV 2023 · 1 citation
