StreetViewAI: Making Street View Accessible Using Context-Aware Multimodal AI
Jon E. Froehlich, Alexander J. Fiannaca, Nimer Jaber, Victor Tsaran, Shaun K. Kane
Abstract
Interactive streetscape mapping tools such as Google Street View (GSV) and Meta Mapillary enable users to virtually navigate and experience real-world environments via immersive 360° imagery but remain fundamentally inaccessible to blind users. We introduce StreetReaderAI, the first-ever accessible street view tool, which combines context-aware, multimodal AI, accessible navigation controls, and conversational speech. With StreetReaderAI, blind users can virtually examine destinations, engage in open-world exploration, or virtually tour any of the over 220 billion images and 100+ countries where GSV is deployed. We iteratively designed StreetReaderAI with a mixed-visual ability team and performed an evaluation with eleven blind users. Our findings demonstrate the value of an accessible street view in supporting POI investigations and remote route planning. We close by enumerating key guidelines for future work.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext f13e9c40-7054-4ed9-be62-15b2f936480dCited by top-tier papers4
- ADCanvas: Accessible and Conversational Audio Description Authoring for Blind and Low Vision CreatorsFranklin Mingzhe Li, Michael Xieyang Liu, Cynthia L. Bennett, Shaun K. KaneCHI 2026 · 2 citations
- NaviNote: Enabling In-situ Spatial Annotation Authoring to Support Exploration and Navigation for Blind and Low Vision PeopleRuijia Chen, Yuheng Wu, Charlie Houseago, Filipe Gaspar et al.CHI 2026 · 1 citation
- Towards LLM-powered Assistive Drone for Blind and Low Vision UsersYize Wei, Ibnu Taimiyyah Bin Adam, Hanjun Wu, Moritz Messerschmidt et al.CHI 2026 · 1 citation
- RAVEN: Realtime Accessibility in Virtual ENvironments for Blind and Low-Vision PeopleXinyun Cao, Kexin Phyllis Ju, Chenglin Li, Venkatesh Potluri et al.CHI 2026 · 1 citation
Builds on13
- Visual Instruction TuningHaotian Liu, Chunyuan Li, Qingyang Wu, Yong Jae LeeNeurIPS 2023 · 11,349 citations
- Depth Anything V2Lihe Yang, Bingyi Kang, Zilong Huang, Zhen Zhao et al.NeurIPS 2024 · 2,305 citations
- Virtual Reality Without Vision: A Haptic and Auditory White Cane to Navigate Complex Virtual WorldsAlexa F. Siu, Mike Sinclair, Robert Kovacs, Eyal Ofek et al.CHI 2020 · 115 citations
- ImageExplorer: Multi-Layered Touch Exploration to Encourage Skepticism Towards Imperfect AI-Generated Image CaptionsJaewook Lee, Jaylin Herskovitz, Yi-Hao Peng, Anhong GuoCHI 2022 · 55 citations
- WorldScribe: Towards Context-Aware Live Visual DescriptionsRuei-Che Chang, Yuxuan Liu, Anhong GuoUIST 2024 · 54 citations
Related papers
- SceneScout: Towards AI-Driven Access to Street Level Imagery for Blind UsersGaurav Jain, Leah Findlater, Cole GleasonCHI 2026 · 2 citations
- GeoVisA11y: An AI-based Geovisualization Question-Answering System for Screen-Reader UsersChu Li, Rock Yuren Pang, Arnavi Chheda-Kothary, Ather Sharif et al.CHI 2026 · 1 citation
- Understanding the Use of a Large Language Model-Powered Guide to Make Virtual Reality Accessible for Blind and Low Vision PeopleJazmin Collins, Sharon Y. Lin, Tianqi Liu, Andrea Stevenson Won et al.CHI 2026 · 3 citations
- From Tactile to NavTile: Opportunities and Challenges with Multi-Modal Feedback for Guiding Surfaces during Non-Visual NavigationSaiganesh Swaminathan, Yellina Yim, Scott E. Hudson, Cynthia L. Bennett et al.CHI 2021 · 13 citations
- How Multimodal Large Language Models Support Access to Visual Information: A Diary Study With Blind and Low Vision PeopleRicardo E. Gonzalez Penuela, Crescentia Jung, Sharon Y. Lin, Ruiying Hu et al.CHI 2026 · 1 citation
