FoveaBox: Beyound Anchor-Based Object Detection

被引:662
|
作者
Kong, Tao [1 ]
Sun, Fuchun [2 ]
Liu, Huaping [2 ]
Jiang, Yuning [1 ]
Li, Lei [1 ]
Shi, Jianbo [3 ]
机构
[1] ByteDance AI Lab, Beijing 100098, Peoples R China
[2] Tsinghua Univ, Beijing Natl Res Ctr Informat Sci & Technol BNRis, Dept Comp Sci & Lochnol, Beijing 100084, Peoples R China
[3] Univ Penn, Grasp Lab, Philadelphia, PA 19104 USA
基金
美国国家科学基金会;
关键词
Object detection; anchor free; foveabox;
D O I
10.1109/TIP.2020.3002345
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
We present FoveaBox, an accurate, flexible, and completely anchor-free framework for object detection. While almost all state-of-the-art object detectors utilize predefined anchors to enumerate possible locations, scales and aspect ratios for the search of the objects, their performance and generalization ability are also limited to the design of anchors. Instead, FoveaBox directly learns the object existing possibility and the bounding box coordinates without anchor reference. This is achieved by: (a) predicting category-sensitive semantic maps for the object existing possibility, and (b) producing category-agnostic bounding box for each position that potentially contains an object. The scales of target boxes are naturally associated with feature pyramid representations. In FoveaBox, an instance is assigned to adjacent feature levels to make the model more accurate.We demonstrate its effectiveness on standard benchmarks and report extensive experimental analysis. Without bells and whistles, FoveaBox achieves state-of-the-art single model performance on the standard COCO and Pascal VOC object detection benchmark. More importantly, FoveaBox avoids all computation and hyper-parameters related to anchor boxes, which are often sensitive to the final detection performance. We believe the simple and effective approach will serve as a solid baseline and help ease future research for object detection. The code has been made publicly available at https://github.com/taokong/FoveaBox.
引用
收藏
页码:7389 / 7398
页数:10
相关论文
共 50 条
  • [31] Anchor-based optimization of energy density functionals
    Taninah, A.
    Afanasjev, A. V.
    PHYSICAL REVIEW C, 2023, 107 (04)
  • [32] Anchor-based fast spectral ensemble clustering
    Zhang, Runxin
    Hang, Shuaijun
    Sun, Zhensheng
    Nie, Feiping
    Wang, Rong
    Li, Xuelong
    INFORMATION FUSION, 2025, 113
  • [33] SP-Det: Anchor-based lane detection network with structural prior perception
    Sun, Libo
    Zhu, Hangyu
    Qin, Wenhu
    PATTERN RECOGNITION LETTERS, 2025, 188 : 60 - 66
  • [34] Dynamic adjustment of hyperparameters for anchor-based detection of objects with large image size differences
    Deng, Ying
    Hu, Xinliang
    Teng, Da
    Li, Bing
    Zhang, Congxuan
    Hu, Weiming
    PATTERN RECOGNITION LETTERS, 2023, 167 : 196 - 203
  • [35] An anchor-based convolutional network for the near-surface camouflaged personnel detection of UAVs
    Xu, Bin
    Wang, Congqing
    Liu, Yang
    Zhou, Yongjun
    VISUAL COMPUTER, 2024, 40 (03): : 1659 - 1671
  • [36] CCLane: Concise Curve Anchor-Based Lane Detection Model with MLP-Mixer
    Yang, Fan
    Zhao, Yanan
    Gao, Li
    Tan, Huachun
    Liu, Weijin
    Chen, Xue-mei
    Yang, Shijuan
    PATTERN RECOGNITION AND COMPUTER VISION, PRCV 2023, PT III, 2024, 14427 : 376 - 387
  • [37] An anchor-based convolutional network for the near-surface camouflaged personnel detection of UAVs
    Bin Xu
    Congqing Wang
    Yang Liu
    Yongjun Zhou
    The Visual Computer, 2024, 40 : 1659 - 1671
  • [38] Reliable Anchor-Based Sensor Localization in Irregular Areas
    Xiao, Bin
    Chen, Lin
    Xiao, Qingjun
    Li, Minglu
    IEEE TRANSACTIONS ON MOBILE COMPUTING, 2010, 9 (01) : 60 - 72
  • [39] Fast unsupervised embedding learning with anchor-based graph
    Zhang, Canyu
    Nie, Feiping
    Wang, Rong
    Li, Xuelong
    INFORMATION SCIENCES, 2022, 609 : 949 - 962
  • [40] SABR: Sparse, Anchor-Based Representation of the Speech Signal
    Liberatore, Christopher
    Aryal, Sandesh
    Wang, Zelun
    Polsley, Seth
    Gutierrez-Osuna, Ricardo
    16TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION (INTERSPEECH 2015), VOLS 1-5, 2015, : 608 - 612