Scaled-YOLOv4: Scaling Cross Stage Partial Network
Chien-Yao Wang, Alexey Bochkovskiy, Hong-Yuan Mark Liao
摘要
We show that the YOLOv4 object detection neural network based on the CSP approach, scales both up and down and is applicable to small and large networks while maintaining optimal speed and accuracy. We propose a network scaling approach that modifies not only the depth, width, resolution, but also structure of the network. YOLOv4large model achieves state-of-the-art results: 55.5% AP (73.4% AP 50 ) for the MS COCO dataset at a speed of ∼16 FPS on Tesla V100, while with the test time augmentation, YOLOv4-large achieves 56.0% AP (73.3 AP 50 ). To the best of our knowledge, this is currently the highest accuracy on the COCO dataset among any published work. The YOLOv4-tiny model achieves 22.0% AP (42.0% AP 50 ) at a speed of ∼443 FPS on RTX 2080Ti, while by using Ten-sorRT, batch size = 4 and FP16-precision the YOLOv4-tiny achieves 1774 FPS.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper27
- YOLOv10: Real-Time End-to-End Object DetectionAo Wang, Hui Chen, Lihao Liu, Kai Chen 等NeurIPS 2024 · 被引用 6,113 次
- DETRs Beat YOLOs on Real-time Object DetectionYian Zhao, Wenyu Lv, Shangliang Xu, Jinman Wei 等CVPR 2024 · 被引用 3,046 次
- Non-deep NetworksAnkit Goyal, Alexey Bochkovskiy, Jia Deng, Vladlen KoltunNeurIPS 2022 · 被引用 112 次
- Simple Multi-dataset DetectionXingyi Zhou, Vladlen Koltun, Philipp KrähenbühlCVPR 2022 · 被引用 85 次
- MS-DETR: Efficient DETR Training with Mixed SupervisionChuyang Zhao, Yifan Sun, Wenhao Wang, Qiang Chen 等CVPR 2024 · 被引用 51 次
它引用的顶会 Paper15
- FCOS: Fully Convolutional One-Stage Object DetectionZhi Tian, Chunhua Shen, Hao Chen, Tong HeICCV 2019 · 被引用 6,042 次
- CenterNet: Keypoint Triplets for Object DetectionKaiwen Duan, Song Bai, Lingxi Xie, Honggang Qi 等ICCV 2019 · 被引用 3,348 次
- Once-for-All: Train One Network and Specialize it for Efficient DeploymentHan Cai, Chuang Gan, Tianzhe Wang, Zhekai Zhang 等ICLR 2020 · 被引用 1,522 次
- HarDNet: A Low Memory Traffic NetworkPing Chao, Chao-Yang Kao, Yu-Shan Ruan, Chien-Hsiang Huang 等ICCV 2019 · 被引用 303 次
- SM-NAS: Structural-to-Modular Neural Architecture Search for Object DetectionLewei Yao, Hang Xu, Wei Zhang, Xiaodan Liang 等AAAI 2020 · 被引用 83 次
相关 Paper
- YOLOv7: Trainable Bag-of-Freebies Sets New State-of-the-Art for Real-Time Object DetectorsChien-Yao Wang, Alexey Bochkovskiy, Hong-Yuan Mark LiaoCVPR 2023
- YOLOv12: Attention-Centric Real-Time Object DetectorsYunjie Tian, Qixiang Ye, David S. DoermannNeurIPS 2025 · 被引用 2,652 次
- YOLO-ULM: Ultra-Lightweight Models for Real-Time Object DetectionShasha Han, Chong Li, Xinning Wang, Xuebo LiCVPR 2026
- SpineNet: Learning Scale-Permuted Backbone for Recognition and LocalizationXianzhi Du, Tsung-Yi Lin, Pengchong Jin, Golnaz Ghiasi 等CVPR 2020
- POD: Practical Object Detection With Scale-Sensitive NetworkJunran Peng, Ming Sun, Zhaoxiang Zhang, Tieniu Tan 等ICCV 2019 · 被引用 23 次
