ZARA: Training-Free Motion Time-Series Reasoning via Evidence-Grounded LLM Agents
Zechen Li, Baiyu Chen, Hao Xue, Flora D. Salim
Abstract
Motion sensor time-series are central to Human Activity Recognition (HAR), yet conventional approaches are constrained to fixed activity sets and typically require costly parameter retraining to adapt to new behaviors. While Large Language Models (LLMs) offer promising open-set reasoning capabilities, applying them directly to numerical time-series often leads to hallucinations and weak grounding. To address this challenge, we propose ZARA (Zero-training Activity Reasoning Agents), a knowledge-and retrieval-augmented agentic framework for motion time-series reasoning in a training-free inference setting. Rather than relying on black-box projections, ZARA distills reference data into a statistically grounded textual knowledge base that transforms implicit signal patterns into verifiable naturallanguage priors. Guided by retrieved evidence, ZARA iteratively selects discriminative cues and performs grounded reasoning over candidate activities. Extensive experiments on eight benchmarks show that ZARA generalizes robustly to unseen subjects and across datasets, demonstrating strong transferability across heterogeneous sensor domains. These results mark a step toward trustworthy, plugand-play motion understanding beyond datasetspecific artifacts. Our code is available at https://github.com/zechenli03/ZARA .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on12
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh et al.ICML 2021 · 47,906 citations
- MOMENT: A Family of Open Time-series Foundation ModelsMononito Goswami, Konrad Szafer, Arjun Choudhry, Yifu Cai et al.ICML 2024 · 442 citations
- K-LITE: Learning Transferable Visual Models with External KnowledgeSheng Shen, Chunyuan Li, Xiaowei Hu, Yujia Xie et al.NeurIPS 2022 · 111 citations
- Attend and Discriminate: Beyond the State-of-the-Art for Human Activity Recognition Using Wearable SensorsAlireza Abedin, Mahsa Ehsanpour, Qinfeng Shi, Hamid Rezatofighi et al.UbiComp 2021 · 104 citations
- UniMTS: Unified Pre-training for Motion Time SeriesXiyuan Zhang, Diyan Teng, Ranak Roy Chowdhury, Shuheng Li et al.NeurIPS 2024 · 49 citations
Related papers
- SensorLLM: Aligning Large Language Models with Motion Sensors for Human Activity RecognitionZechen Li, Shohreh Deldari, Linyao Chen, Hao Xue et al.EMNLP 2025 · 9 citations
- IMUZero: Zero-Shot Human Activity Recognition by Language-Based Cross Modality FusionJie Su, Fengtong Ge, Zhenyu Wen, Taotao Li et al.UbiComp 2026 · 2 citations
- Open-ended Human Activity Understanding via LLM-assisted Motion Decomposition and Semantic FusionQingxin Wei, Kai Hu, Jiaming Huang, Cheng Guo et al.UbiComp 2026
- Boosting Skeleton-based Zero-Shot Action Recognition with Training-Free Test-Time AdaptationJingmin Zhu, Anqi Zhu, Hossein Rahmani, Jun Liu et al.NeurIPS 2025 · 3 citations
- Zero-Shot Open-Vocabulary Human Motion Grounding with Test-Time TrainingYunjiao Zhou, Xinyan Chen, Junlang Qian, Lihua Xie et al.AAAI 2026 · 2 citations
