Presence-Only Geographical Priors for Fine-Grained Image Classification
Oisin Mac Aodha, Elijah Cole, Pietro Perona
Abstract
Appearance information alone is often not sufficient to accurately differentiate between fine-grained visual categories. Human experts make use of additional cues such as where, and when, a given image was taken in order to inform their final decision. This contextual information is readily available in many online image collections but has been underutilized by existing image classifiers that focus solely on making predictions based on the image contents. We propose an efficient spatio-temporal prior, that when conditioned on a geographical location and time, estimates the probability that a given object category occurs at that location. Our prior is trained from presence-only observation data and jointly models object categories, their spatio-temporal distributions, and photographer biases. Experiments performed on multiple challenging image classification datasets show that combining our prior with the predictions from image classifiers results in a large improvement in final classification performance.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers33
- Geography-Aware Self-Supervised LearningKumar Ayush, Burak Uzkent, Chenlin Meng, Kumar Tanmay et al.ICCV 2021 · 304 citations
- SatCLIP: Global, General-Purpose Location Embeddings with Satellite ImageryKonstantin Klemmer, Esther Rolf, Caleb Robinson, Lester Mackey et al.AAAI 2025 · 173 citations
- Multi-Scale Representation Learning for Spatial Feature Distributions using Grid CellsGengchen Mai, Krzysztof Janowicz, Bo Yan, Rui Zhu et al.ICLR 2020 · 161 citations
- CSP: Self-Supervised Contrastive Spatial Pre-Training for Geospatial-Visual RepresentationsGengchen Mai, Ni Lao, Yutong He, Jiaming Song et al.ICML 2023 · 103 citations
- Spatial Implicit Neural Representations for Global-Scale Species MappingElijah Cole, Grant Van Horn, Christian Lange, Alexander Shepard et al.ICML 2023 · 70 citations
Related papers
- GT-Loc: Unifying When and Where in Images Through a Joint Embedding SpaceDavid G. Shatwell, Ishan Rajendrakumar Dave, Sirnam Swetha, Mubarak ShahICCV 2025 · 1 citation
- TIGER: A Unified Framework for Time, Images and Geo-location RetrievalDavid G. Shatwell, Sirnam Swetha, Mubarak ShahCVPR 2026 · 2 citations
- On Guiding Visual Attention with Language SpecificationSuzanne Petryk, Lisa Dunlap, Keyan Nasseri, Joseph Gonzalez et al.CVPR 2022 · 18 citations
- RANGE: Retrieval Augmented Neural Fields for Multi-Resolution Geo-EmbeddingsAayush Dhakal, Srikumar Sastry, Subash Khanal, Adeel Ahmad et al.CVPR 2025
- Learning a Dynamic Map of Visual AppearanceTawfiq Salem, Scott Workman, Nathan JacobsCVPR 2020
