VideoLoc: Video-based Indoor Localization with Text Information
Shusheng Li, Wenbo He
摘要
Indoor localization serves an important role in various scenarios such as navigation in shopping malls or hospitals. However, the existing technology is usually based on additional deployment and the signals suffer from strong environmental interference in the complex indoor environment. In this paper, we propose video-based indoor localization with text information (i.e. "VideoLoc") without the deployment of additional equipment. Videos taken by the phone carriers cover more critical information (e.g. logos in malls), while a single photo may fail to capture it. To reduce redundant information in the video, we propose key-frame selection based on deep learning model and clustering algorithm. Video frames are characterized with deep visual descriptors and the clustering algorithm efficiently clusters these descriptors into a set of non-overlapping snippets. We select keyframes from these non-overlapping snippets in terms of the cluster centroid that represents each snippet. Then, we propose text detection and recognition with the perspective transformation to make full use of stable and discriminative text information (e.g. logos or room numbers) in keyframes for localization. Finally, we obtain the location of the phone carrier via the triangulation algorithm. The experimental results show that VideoLoc achieves high precision of localization and is robust to dynamic environments.
问问这篇 Paper
问问你的智能体。
Lune 读过与它相关的顶会 Paper,每个回答都会注明依据哪几篇。
相关 Paper
- TextPlace: Visual Place Recognition and Topological Localization Through Reading Scene TextsZiyang Hong, Yvan R. Petillot, David Lane, Yishu Miao 等ICCV 2019 · 被引用 60 次
- Large-Scale Localization Datasets in Crowded Indoor SpacesDonghwan Lee, Soo-Hyun Ryu, Suyong Yeon, Yonghan Lee 等CVPR 2021
- Train Once, Locate Anytime for Anyone: Adversarial Learning based Wireless LocalizationDanyang Li, Jingao Xu, Zheng Yang, Yumeng Lu 等INFOCOM 2021 · 被引用 57 次
- DyLoc: Dynamic Localization for Massive MIMO Using Predictive Recurrent Neural NetworksFarzam Hejazi, Katarina Vuckovic, Nazanin RahnavardINFOCOM 2021 · 被引用 35 次
- Perspective from a Broader Context: Can Room Style Knowledge Help Visual Floorplan Localization?Bolei Chen, Shengsheng Yan, Yongzheng Cui, Jiaxu Kang 等AAAI 2026 · 被引用 1 次
