BlinkBud: Detecting Hazards from Behind via Sampled Monocular 3D Detection on a Single Earbud
Yunzhe Li, Jiajun Yan, Yuzhou Wei, Kechen Liu, Yize Zhao, Chong Zhang, Hongzi Zhu, Li Lu, Shan Chang, Minyi Guo
Abstract
Failing to be aware of speeding vehicles approaching from behind poses a huge threat to the road safety of pedestrians and cyclists. In this paper, we propose BlinkBud, which utilizes a single earbud and a paired phone to online detect hazardous objects approaching from behind of a user. The core idea is to accurately track visually identified objects utilizing a small number of sampled camera images taken from the earbud. To minimize the power consumption of the earbud and the phone while guaranteeing the best tracking accuracy, a novel 3D object tracking algorithm is devised, integrating both a Kalman filter based trajectory estimation scheme and an optimal image sampling strategy based on reinforcement learning. Moreover, the impact of constant user head movements on the tracking accuracy is significantly eliminated by leveraging the estimated pitch and yaw angles to correct the object depth estimation and align the camera coordinate system to the user's body coordinate system, respectively. We implement a prototype BlinkBud system and conduct extensive real-world experiments. Results show that BlinkBud is lightweight with ultra-low mean power consumptions of 29.8 mW and 702.6 mW on the earbud and smartphone, respectively, and can accurately detect hazards with a low average false positive ratio (FPR) and false negative ratio (FNR) of 4.90% and 1.47%, respectively.
CCS Concepts: • Human-centered computing → Ubiquitous and mobile computing.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 1bc305ca-982e-4c9d-98d0-bed776a45d09Builds on12
- Digging Into Self-Supervised Monocular Depth EstimationClément Godard, Oisin Mac Aodha, Michael Firman, Gabriel J. BrostowICCV 2019 · 2,416 citations
- TrackFormer: Multi-Object Tracking with TransformersTim Meinhardt, Alexander Kirillov, Laura Leal-Taixé, Christoph FeichtenhoferCVPR 2022 · 927 citations
- Global Tracking TransformersXingyi Zhou, Tianwei Yin, Vladlen Koltun, Philipp KrähenbühlCVPR 2022 · 180 citations
- Hybrid-SORT: Weak Cues Matter for Online Multi-Object TrackingMingzhan Yang, Guangxin Han, Bin Yan, Wenhua Zhang et al.AAAI 2024 · 171 citations
- VIPS: real-time perception fusion for infrastructure-assisted autonomous drivingShuyao Shi, Jiahe Cui, Zhehao Jiang, Zhenyu Yan et al.MobiCom 2022 · 126 citations
Related papers
- BlinkListener: "Listen" to Your Eye Blink Using Your SmartphoneJialin Liu, Dong Li, Lei Wang, Jie XiongUbiComp 2021 · 41 citations
- BioFace-3D: continuous 3d facial reconstruction through lightweight single-ear biosensorsYi Wu, Vimal Kakaraparthi, Zhuohang Li, Tien Pham et al.MobiCom 2021 · 25 citations
- VueBuds: Visual Intelligence with Wireless EarbudsMaruchi Kim, Rasya Fawwaz, Zhi Yang Lim, Brinda Moudgalya et al.CHI 2026 · 1 citation
- EarMeter: Continuous Respiration Volume Monitoring with EarablesYang Liu, Qiang Yang, Kayla-Jade Butkow, Jake Stuchbury-Wass et al.UbiComp 2026 · 6 citations
- MR Object Identification and Interaction: Fusing Object Situation Information from Heterogeneous SourcesJannis Strecker, Khakim Akhunov, Federico Carbone, Kimberly García et al.UbiComp 2023 · 15 citations
