Training-Time-Friendly Network for Real-Time Object Detection
Zili Liu, Tu Zheng, Guodong Xu, Zheng Yang, Haifeng Liu, Deng Cai
摘要
Modern object detectors can rarely achieve short training time, fast inference speed, and high accuracy at the same time. To strike a balance among them, we propose the Training-Time-Friendly Network (TTFNet). In this work, we start with light-head, single-stage, and anchor-free designs, which enable fast inference speed. Then, we focus on shortening training time. We notice that encoding more training samples from annotated boxes plays a similar role as increasing batch size, which helps enlarge the learning rate and accelerate the training process. To this end, we introduce a novel approach using Gaussian kernels to encode training samples. Besides, we design the initiative sample weights for better information utilization. Experiments on MS COCO show that our TTFNet has great advantages in balancing training time, inference speed, and accuracy. It has reduced training time by more than seven times compared to previous real-time detectors while maintaining state-of-the-art performances. In addition, our super-fast version of TTFNet-18 and TTFNet-53 can outperform SSD300 and YOLOv3 by less than one-tenth of their training time, respectively. The code has been made available at https://github.com/ZJULearning/ttfnet .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- MSG-Transformer: Exchanging Local Spatial Information by Manipulating Messenger TokensJiemin Fang, Lingxi Xie, Xinggang Wang, Xiaopeng Zhang 等CVPR 2022 · 被引用 73 次
- Fast Neural Network Adaptation via Parameter Remapping and Architecture SearchJiemin Fang, Yuzhu Sun, Kangjian Peng, Qian Zhang 等ICLR 2020 · 被引用 36 次
- DynamicISP: Dynamically Controlled Image Signal Processor for Image RecognitionMasakazu Yoshimura, Junji Otsuka, Atsushi Irie, Takeshi OhashiICCV 2023 · 被引用 28 次
- Rawgment: Noise-Accounted RAW Augmentation Enables Recognition in a Wide Variety of EnvironmentsMasakazu Yoshimura, Junji Otsuka, Atsushi Irie, Takeshi OhashiCVPR 2023
- Rigidity-Aware Detection for 6D Object Pose EstimationYang Hai, Rui Song, Jiaojiao Li, Mathieu Salzmann 等CVPR 2023
它引用的顶会 Paper2
相关 Paper
- Bridging the Gap Between Anchor-Based and Anchor-Free Detection via Adaptive Training Sample SelectionShifeng Zhang, Cheng Chi, Yongqiang Yao, Zhen Lei 等CVPR 2020
- ThunderNet: Towards Real-Time Generic Object Detection on Mobile DevicesZheng Qin, Zeming Li, Zhaoning Zhang, Yiping Bao 等ICCV 2019 · 被引用 282 次
- YOLO-ULM: Ultra-Lightweight Models for Real-Time Object DetectionShasha Han, Chong Li, Xinning Wang, Xuebo LiCVPR 2026
- POD: Practical Object Detection With Scale-Sensitive NetworkJunran Peng, Ming Sun, Zhaoxiang Zhang, Tieniu Tan 等ICCV 2019 · 被引用 23 次
- SCNet: Training Inference Sample Consistency for Instance SegmentationThang Vu, Haeyong Kang, Chang D. YooAAAI 2021 · 被引用 111 次
