Simple Multi-dataset Detection
Xingyi Zhou, Vladlen Koltun, Philipp Krähenbühl
Abstract
How do we build a general and broad object detection system? We use all labels of all concepts ever annotated. These labels span diverse datasets with potentially inconsistent taxonomies. In this paper, we present a simple method for training a unified detector on multiple large-scale datasets. We use dataset-specific training protocols and losses, but share a common detection architecture with dataset-specific outputs. We show how to automatically integrate these dataset-specific outputs into a common semantic taxonomy. In contrast to prior work, our approach does not require manual taxonomy reconciliation. Experiments show our learned taxonomy outperforms a expert-designed taxonomy in all datasets. Our multi-dataset detector performs as well as dataset-specific models on each training domain, and can generalize to new unseen dataset without fine-tuning on them. Code is available at https://github.com/xingyizhou/UniDet.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext e67da3b4-4319-4c7d-80e1-ce70473e0b07Cited by top-tier papers56
- ReCLIP: A Strong Zero-Shot Baseline for Referring Expression ComprehensionSanjay Subramanian, William Merrill, Trevor Darrell, Matt Gardner et al.ACL 2022 · 172 citations
- HRS-Bench: Holistic, Reliable and Scalable Benchmark for Text-to-Image ModelsEslam Mohamed Bakr, Pengzhan Sun, Xiaoqian Shen, Faizan Farooq Khan et al.ICCV 2023 · 115 citations
- Cascade-DETR: Delving into High-Quality Universal Object DetectionMingqiao Ye, Lei Ke, Siyuan Li, Yu-Wing Tai et al.ICCV 2023 · 62 citations
- VideoTetris: Towards Compositional Text-to-Video GenerationYe Tian, Ling Yang, Haotian Yang, Yuan Gao et al.NeurIPS 2024 · 62 citations
- DAMEX: Dataset-aware Mixture-of-Experts for visual understanding of mixture-of-datasetsYash Jain, Harkirat S. Behl, Zsolt Kira, Vibhav VineetNeurIPS 2023 · 43 citations
Builds on12
- Objects365: A Large-Scale, High-Quality Dataset for Object DetectionShuai Shao, Zeming Li, Tianyuan Zhang, Chao Peng et al.ICCV 2019 · 1,018 citations
- Transductive Learning for Zero-Shot Object DetectionShafin Rahman, Salman H. Khan, Nick BarnesICCV 2019 · 82 citations
- Universal-RCNN: Universal Object Detector via Transferable Graph R-CNNHang Xu, Linpu Fang, Xiaodan Liang, Wenxiong Kang et al.AAAI 2020 · 26 citations
- VarifocalNet: An IoU-Aware Dense Object DetectorHaoyang Zhang, Ying Wang, Feras Dayoub, Niko SünderhaufCVPR 2021
- MSeg: A Composite Dataset for Multi-Domain Semantic SegmentationJohn Lambert, Zhuang Liu, Ozan Sener, James Hays et al.CVPR 2020
Related papers
- Detection Hub: Unifying Object Detection Datasets via Query Adaptation on Language EmbeddingLingchen Meng, Xiyang Dai, Yinpeng Chen, Pengchuan Zhang et al.CVPR 2023
- ScaleDet: A Scalable Multi-Dataset Object DetectorYanbei Chen, Manchen Wang, Abhay Mittal, Zhenlin Xu et al.CVPR 2023
- Detecting Everything in the Open World: Towards Universal Object DetectionZhenyu Wang, Yali Li, Xi Chen, Ser-Nam Lim et al.CVPR 2023
- LMSeg: Language-guided Multi-dataset SegmentationQiang Zhou, Yuang Liu, Chaohui Yu, Jingliang Li et al.ICLR 2023 · 2 citations
- Uni2Det: Unified and Universal Framework for Prompt-Guided Multi-dataset 3D DetectionYubin Wang, Zhikang Zou, Xiaoqing Ye, Xiao Tan et al.ICLR 2025
