Training-Time-Friendly Network for Real-Time Object Detection
Zili Liu, Tu Zheng, Guodong Xu, Zheng Yang, Haifeng Liu, Deng Cai
Abstract
Modern object detectors can rarely achieve short training time, fast inference speed, and high accuracy at the same time. To strike a balance among them, we propose the Training-Time-Friendly Network (TTFNet). In this work, we start with light-head, single-stage, and anchor-free designs, which enable fast inference speed. Then, we focus on shortening training time. We notice that encoding more training samples from annotated boxes plays a similar role as increasing batch size, which helps enlarge the learning rate and accelerate the training process. To this end, we introduce a novel approach using Gaussian kernels to encode training samples. Besides, we design the initiative sample weights for better information utilization. Experiments on MS COCO show that our TTFNet has great advantages in balancing training time, inference speed, and accuracy. It has reduced training time by more than seven times compared to previous real-time detectors while maintaining state-of-the-art performances. In addition, our super-fast version of TTFNet-18 and TTFNet-53 can outperform SSD300 and YOLOv3 by less than one-tenth of their training time, respectively. The code has been made available at https://github.com/ZJULearning/ttfnet .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a38c57c1-3985-46e1-b475-afc7ab2b75d1Cited by top-tier papers5
- MSG-Transformer: Exchanging Local Spatial Information by Manipulating Messenger TokensJiemin Fang, Lingxi Xie, Xinggang Wang, Xiaopeng Zhang et al.CVPR 2022 · 73 citations
- Fast Neural Network Adaptation via Parameter Remapping and Architecture SearchJiemin Fang, Yuzhu Sun, Kangjian Peng, Qian Zhang et al.ICLR 2020 · 36 citations
- DynamicISP: Dynamically Controlled Image Signal Processor for Image RecognitionMasakazu Yoshimura, Junji Otsuka, Atsushi Irie, Takeshi OhashiICCV 2023 · 28 citations
- Rawgment: Noise-Accounted RAW Augmentation Enables Recognition in a Wide Variety of EnvironmentsMasakazu Yoshimura, Junji Otsuka, Atsushi Irie, Takeshi OhashiCVPR 2023
- Rigidity-Aware Detection for 6D Object Pose EstimationYang Hai, Rui Song, Jiaojiao Li, Mathieu Salzmann et al.CVPR 2023
Builds on2
Related papers
- Bridging the Gap Between Anchor-Based and Anchor-Free Detection via Adaptive Training Sample SelectionShifeng Zhang, Cheng Chi, Yongqiang Yao, Zhen Lei et al.CVPR 2020
- ThunderNet: Towards Real-Time Generic Object Detection on Mobile DevicesZheng Qin, Zeming Li, Zhaoning Zhang, Yiping Bao et al.ICCV 2019 · 282 citations
- YOLO-ULM: Ultra-Lightweight Models for Real-Time Object DetectionShasha Han, Chong Li, Xinning Wang, Xuebo LiCVPR 2026
- POD: Practical Object Detection With Scale-Sensitive NetworkJunran Peng, Ming Sun, Zhaoxiang Zhang, Tieniu Tan et al.ICCV 2019 · 23 citations
- SCNet: Training Inference Sample Consistency for Instance SegmentationThang Vu, Haeyong Kang, Chang D. YooAAAI 2021 · 111 citations
