Understanding Spatiotemporal-Aware Multimodal Conversational Search in the Outdoor Urban Space
Jiangnan Xu, Suyeon Seo, Joni Salminen, Michael Saker, Joongi Shin, Alan Chamberlain, Konstantinos Papangelis, Dae Hyun Kim
摘要
Emerging multimodal conversational search (MCS) tools (e.g., Gemini Live) allow users to search for spatiotemporal information through natural language dialogues as they move through urban space. Despite the growing popularity of these tools, there is limited understanding of how people engage with this technology. To address this gap, we developed UrbanSearch, an MCS technology probe designed to capture the user’s current geolocation, time, and visual surroundings. A contextual inquiry (N=23) revealed that MCS tools provide two core values: requiring low effort in forming queries while offering highly relevant responses, and functioning as a central information gateway. As a promising technology, MCS supports environmental learning, in-situ decision making, and personalized navigation. Participants also revealed unmet needs for spatial reasoning and transparent integration of multi-source information, along with concerns related to peripheral awareness, social context, and personal space. Drawing from the findings, we discuss design implications for future MCS tools in urban spaces.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- Self-Consistency Improves Chain of Thought Reasoning in Language ModelsXuezhi Wang, Jason Wei, Dale Schuurmans, Quoc V. Le 等ICLR 2023 · 被引用 681 次
- What Is Wrong With Scene Text Recognition Model Comparisons? Dataset and Model AnalysisJeonghun Baek, Geewook Kim, Junyeop Lee, Sungrae Park 等ICCV 2019 · 被引用 551 次
- AutoDroid: LLM-powered Task Automation in AndroidHao Wen, Yuanchun Li, Guohong Liu, Shanhui Zhao 等MobiCom 2024 · 被引用 94 次
- NeRF-LiDAR: Generating Realistic LiDAR Point Clouds with Neural Radiance FieldsJunge Zhang, Feihu Zhang, Shaochen Kuang, Li ZhangAAAI 2024 · 被引用 76 次
- AiGet: Transforming Everyday Moments into Hidden Knowledge Discovery with AI Assistance on Smart GlassesRunze Cai, Nuwan Janaka, Hyeongcheol Kim, Yang Chen 等CHI 2025 · 被引用 26 次
相关 Paper
- Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-SpeechRui Liu, Shuwei He, Yifan Hu, Haizhou LiAAAI 2025 · 被引用 8 次
- NAVIGATE: Evaluating Visual-Guided Search Decision-Making on the Open WebYaoQi Fan, Zhe Chen, Zhu Wei, Kangxin Yin 等ICML 2026
- Tesseract: Querying Spatial Design Recordings by Manipulating Worlds in MiniatureKarthik Mahadevan, Qian Zhou, George W. Fitzmaurice, Tovi Grossman 等CHI 2023 · 被引用 16 次
- UrbanFeel:A Comprehensive Benchmark for Temporal and Perceptual Understanding of City Scenes through Human PerspectiveJun He, Yi Lin, Zilong Huang, Jiacong Yin 等ICLR 2026 · 被引用 7 次
- Textual Supervision Enhances Geospatial Representations in Vision-Language ModelsMarcelo Sartori Locatelli, Fernando Tonucci, Jea Kwon, Luiz Felipe Vecchietti 等ICML 2026
