Around the World in 80 Timesteps: A Generative Approach to Global Visual Geolocation
Nicolas Dufour, Vicky Kalogeiton, David Picard, Loïc Landrieu
2025年份
12顶会引用
摘要
Figure 1. Geolocation as a Generative Process. We explore diffusion and flow matching for visual geolocation by sampling and denoising random locations. This process generates trajectories onto the Earth's surface, whose endpoints provide location estimates. Our models also provide probability densities for every possible image locations. We illustrate these trajectories and the log-densities for three images from different datasets: an Andean condor from iNat21 [74], an African open-air market from YFCC-100M [1], and a dashcam snapshot from OSV-5M [2]. The predicted image locations are indicated by and the true ones by .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper12
- Recognition through Reasoning: Reinforcing Image Geo-localization with Large Vision-Language ModelsLing Li, Yao Zhou, Yuxuan Liang, Fugee Tsung 等NeurIPS 2025 · 被引用 30 次
- GeoRanker: Distance-Aware Ranking for Worldwide Image GeolocalizationPengyue Jia, Seongheon Park, Song Gao, Xiangyu Zhao 等NeurIPS 2025 · 被引用 22 次
- GRE Suite: Geo-localization Inference via Fine-Tuned Vision-Language Models and Enhanced Reasoning ChainsChun Wang, Xiaojun Ye, Xiaoran Pan, Zihao Pan 等NeurIPS 2025 · 被引用 18 次
- Scaling Image Geo-Localization to Continent LevelPhilipp Lindenberger, Paul-Edouard Sarlin, Jan Hosang, Marc Pollefeys 等NeurIPS 2025 · 被引用 11 次
- GeoAgent: Learning to Geolocate Everywhere with Reinforced Geographic CharacteristicsModi Jin, Yiming Zhang, Boyuan Sun, Dingwen Zhang 等CVPR 2026 · 被引用 7 次
它引用的顶会 Paper23
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Denoising Diffusion Implicit ModelsJiaming Song, Chenlin Meng, Stefano ErmonICLR 2021 · 被引用 11,743 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
相关 Paper
- LocDiff: Identifying Locations on Earth by Diffusing in the Hilbert SpaceZhangyu Wang, Zeping Liu, Jielu Zhang, Zhongliang Zhou 等NeurIPS 2025 · 被引用 7 次
- TopicGeo: An Efficient Unified Framework for GeolocationXin Wang, Xinlin Wang, Shuiping GouICCV 2025 · 被引用 2 次
- DiffTraj: Generating GPS Trajectory with Diffusion Probabilistic ModelYuanshao Zhu, Yongchao Ye, Shiyao Zhang, Xiangyu Zhao 等NeurIPS 2023 · 被引用 134 次
- Probability Density Geodesics in Image Diffusion Latent SpaceQingtao Yu, Jaskirat Singh, Zhaoyuan Yang, Peter Henry Tu 等CVPR 2025
- Low-Rank Prior-Induced Consistency Flow Matching for Efficient Traffic ImputationXiaowei Mao, Tingrui Wu, Yawen Yang, Shengnan Guo 等KDD 2026
