RoVaR: Robust Multi-agent Tracking through Dual-layer Diversity in Visual and RF Sensing
Mallesham Dasari, Ramanujan K. Sheshadri, Karthikeyan Sundaresan, Samir R. Das
Abstract
The plethora of sensors in our commodity devices provides a rich substrate for sensor-fused tracking. Yet, today's solutions are unable to deliver robust and high tracking accuracies across multiple agents in practical, everyday environments -a feature central to the future of immersive and collaborative applications. This can be attributed to the limited scope of diversity leveraged by these fusion solutions, preventing them from catering to the multiple dimensions of accuracy, robustness (diverse environmental conditions) and scalability (multiple agents) simultaneously.
In this work, we take an important step towards this goal by introducing the notion of dual-layer diversity to the problem of sensor fusion in multi-agent tracking. We demonstrate that the fusion of complementary tracking modalities, -passive/relative (e.g. visual odometry) and active/absolute tracking (e.g.infrastructure-assisted RF localization) offer a key first layer of diversity that brings scalability while the second layer of diversity lies in the methodology of fusion, where we bring together the complementary strengths of algorithmic (for robustness) and data-driven (for accuracy) approaches. RoVaR is an embodiment of such a dual-layer diversity approach that intelligently attends to cross-modal information using algorithmic and data-driven techniques that jointly share the burden of accurately tracking multiple agents in the wild. Extensive evaluations reveal RoVaR's multi-dimensional benefits in terms of tracking accuracy (median of ≈15cm), robustness (in unseen environments), light weight (runs in real-time on mobile platforms such as Jetson Nano/TX2), to enable practical multi-agent immersive applications in everyday environments.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers2
- Malicious Attacks against Multi-Sensor Fusion in Autonomous DrivingYi Zhu, Chenglin Miao, Hongfei Xue, Yunnan Yu et al.MobiCom 2024 · 28 citations
- SHARE: Towards Head-Mounted AR with User-Centric SLAM in Shared Human-Robot WorkspacesTianyuan Du, Tianyi Hu, Hanting Ye, Maria GorlatovaUbiComp 2026
Builds on1
Related papers
- MIROS: Elusive Unauthorized AAV Positioning by Multi-View Radar-Vision Cognitive FusionGuangyu Wu, Yuxin Zhao, Haibo Zhou, Yuben Qu et al.INFOCOM 2026
- CoRA: A Collaborative Robust Architecture with Hybrid Fusion for Efficient PerceptionGong Chen, Chaokun Zhang, Pengcheng Lv, Xiaohui XieAAAI 2026 · 4 citations
- ULoc: Low-Power, Scalable and cm-Accurate UWB-Tag Localization and Tracking for Indoor ApplicationsMinghui Zhao, Tyler Chang, Aditya Arun, Roshan Sai Ayyalasomayajula et al.UbiComp 2021 · 78 citations
- Weakly Misalignment-Free Adaptive Feature Alignment for UAVs-Based Multimodal Object DetectionChen Chen, Jiahao Qi, Xingyue Liu, Kangcheng Bin et al.CVPR 2024
- Enabling RFID-Based Tracking for Multi-Objects with Visual Aids: A Calibration-Free SolutionChunhui Duan, Wenlei Shi, Fan Dang, Xuan DingINFOCOM 2020 · 20 citations
