Unsupervised Object Detection with Theoretical Guarantees
Marian Longa, João F. Henriques
Abstract
Unsupervised object detection using deep neural networks is typically a difficult problem with few to no guarantees about the learned representation. In this work we present the first unsupervised object detection method that is theoretically guaranteed to recover the true object positions up to quantifiable small shifts. We develop an unsupervised object detection architecture and prove that the learned variables correspond to the true object positions up to small shifts related to the encoder and decoder receptive field sizes, the object sizes, and the widths of the Gaussians used in the rendering process. We perform detailed analysis of how the error depends on each of these variables and perform synthetic experiments validating our theoretical predictions up to a precision of individual pixels. We also perform experiments on CLEVR-based data and show that, unlike current SOTA object detection methods (SAM, CutLER), our method's prediction errors always lie within our theoretical bounds. We hope that this work helps open up an avenue of research into object detection methods with theoretical guarantees.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 7fd2c70b-36ba-4a50-b762-5be6e993005fBuilds on8
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao et al.ICCV 2023 · 13,211 citations
- Object-Centric Learning with Slot AttentionFrancesco Locatello, Dirk Weissenborn, Thomas Unterthiner, Aravindh Mahendran et al.NeurIPS 2020 · 1,275 citations
- Weakly-Supervised Disentanglement Without CompromisesFrancesco Locatello, Ben Poole, Gunnar Rätsch, Bernhard Schölkopf et al.ICML 2020 · 361 citations
- Weakly supervised causal representation learningJohann Brehmer, Pim de Haan, Phillip Lippe, Taco S. CohenNeurIPS 2022 · 196 citations
- CITRIS: Causal Identifiability from Temporal Intervened SequencesPhillip Lippe, Sara Magliacane, Sindy Löwe, Yuki M. Asano et al.ICML 2022 · 136 citations
Related papers
- DyStaB: Unsupervised Object Segmentation via Dynamic-Static BootstrappingYanchao Yang, Brian Lai, Stefano SoattoCVPR 2021
- A Causal Debiasing Framework for Unsupervised Salient Object DetectionXiangru Lin, Ziyi Wu, Guanqi Chen, Guanbin Li et al.AAAI 2022 · 34 citations
- MOVE: Unsupervised Movable Object Segmentation and DetectionAdam Bielski, Paolo FavaroNeurIPS 2022 · 30 citations
- SdalsNet: Self-Distilled Attention Localization and Shift Network for Unsupervised Camouflaged Object DetectionPeiyao Shou, Yixiu Liu, Wei Wang, Yaoqi Sun et al.AAAI 2025 · 4 citations
- Unsupervised Object Representation Learning using Translation and Rotation Group Equivariant VAEAlireza Nasiri, Tristan BeplerNeurIPS 2022 · 18 citations
