FoveaBox: Beyond Anchor-based Object Detector
arXiv:1904.03797 · doi:10.1109/TIP.2020.3002345
Abstract
We present FoveaBox, an accurate, flexible, and completely anchor-free framework for object detection. While almost all state-of-the-art object detectors utilize predefined anchors to enumerate possible locations, scales and aspect ratios for the search of the objects, their performance and generalization ability are also limited to the design of anchors. Instead, FoveaBox directly learns the object existing possibility and the bounding box coordinates without anchor reference. This is achieved by: (a) predicting category-sensitive semantic maps for the object existing possibility, and (b) producing category-agnostic bounding box for each position that potentially contains an object. The scales of target boxes are naturally associated with feature pyramid representations. In FoveaBox, an instance is assigned to adjacent feature levels to make the model more accurate.We demonstrate its effectiveness on standard benchmarks and report extensive experimental analysis. Without bells and whistles, FoveaBox achieves state-of-the-art single model performance on the standard COCO and Pascal VOC object detection benchmark. More importantly, FoveaBox avoids all computation and hyper-parameters related to anchor boxes, which are often sensitive to the final detection performance. We believe the simple and effective approach will serve as a solid baseline and help ease future research for object detection. The code has been made publicly available at https://github.com/taokong/FoveaBox .
IEEE Transactions on Image Processing, code at: https://github.com/taokong/FoveaBox
References in corpus (1)
Cited by in corpus (17)
- Anchor-free Oriented Proposal Generator for Object Detection
- Detecting tiny objects in aerial images: A normalized Wasserstein distance and a new benchmark
- A Dataset And Benchmark Of Underwater Object Detection For Robot Picking
- CBNet: A Composite Backbone Network Architecture for Object Detection
- SSPNet: Scale Selection Pyramid Network for Tiny Person Detection from UAV Images
- A Survey on Approximate Edge AI for Energy Efficient Autonomous Driving Services
- PETDet: Proposal Enhancement for Two-Stage Fine-Grained Object Detection
- CBA: Contextual Background Attack against Optical Aerial Detection in the Physical World
- PDNet: Toward Better One-Stage Object Detection With Prediction Decoupling
- FoodLogoDet-1500: A Dataset for Large-Scale Food Logo Detection via Multi-Scale Feature Decoupling Network
- EARL: An Elliptical Distribution aided Adaptive Rotation Label Assignment for Oriented Object Detection in Remote Sensing Images
- Towards Balanced Learning for Instance Recognition
- RCNet: Reverse Feature Pyramid and Cross-scale Shift Network for Object Detection
- Local and Global Context-and-Object-part-Aware Superpixel-based Data Augmentation for Deep Visual Recognition
- hSDB-instrument: Instrument Localization Database for Laparoscopic and Robotic Surgeries
- Parallel Residual Bi-Fusion Feature Pyramid Network for Accurate Single-Shot Object Detection
- Real-Time Frame- and Event-based Object Detection with Spiking Neural Networks on Edge Neuromorphic Hardware: Design, Deployment and Benchmark