UA-DETRAC: A New Benchmark and Protocol for Multi-Object Detection and Tracking
arXiv:1511.04136
Abstract
In recent years, numerous effective multi-object tracking (MOT) methods are developed because of the wide range of applications. Existing performance evaluations of MOT methods usually separate the object tracking step from the object detection step by using the same fixed object detection results for comparisons. In this work, we perform a comprehensive quantitative study on the effects of object detection accuracy to the overall MOT performance, using the new large-scale University at Albany DETection and tRACking (UA-DETRAC) benchmark dataset. The UA-DETRAC benchmark dataset consists of 100 challenging video sequences captured from real-world traffic scenes (over 140,000 frames with rich annotations, including occlusion, weather, vehicle category, truncation, and vehicle bounding boxes) for object detection, object tracking and MOT system. We evaluate complete MOT systems constructed from combinations of state-of-the-art object detection and object tracking methods. Our analysis shows the complex effects of object detection accuracy on MOT system performance. Based on these observations, we propose new evaluation tools and metrics for MOT systems that consider both object detection and object tracking for comprehensive analysis.
18 pages, 11 figures, accepted by CVIU
References in corpus (3)
Cited by in corpus (10)
- MOT16: A Benchmark for Multi-Object Tracking
- Deep Learning in Video Multi-Object Tracking: A Survey
- Benchmarking Robustness in Object Detection: Autonomous Driving when Winter is Coming
- Joint Monocular 3D Vehicle Detection and Tracking
- An On-line Variational Bayesian Model for Multi-Person Tracking from Cluttered Scenes
- FAMNet: Joint Learning of Feature, Affinity and Multi-dimensional Assignment for Online Multiple Object Tracking
- Deep Affinity Network for Multiple Object Tracking
- Adversarial Image Composition with Auxiliary Illumination
- Focal Loss Dense Detector for Vehicle Surveillance
- Inserting Videos into Videos