DeepSportradar-v1: Computer Vision Dataset for Sports Understanding with High Quality Annotations
arXiv:2208.08190 · doi:10.1145/3552437.3555699
Abstract
With the recent development of Deep Learning applied to Computer Vision, sport video understanding has gained a lot of attention, providing much richer information for both sport consumers and leagues. This paper introduces DeepSportradar-v1, a suite of computer vision tasks, datasets and benchmarks for automated sport understanding. The main purpose of this framework is to close the gap between academic research and real world settings. To this end, the datasets provide high-resolution raw images, camera parameters and high quality annotations. DeepSportradar currently supports four challenging tasks related to basketball: ball 3D localization, camera calibration, player instance segmentation and player re-identification. For each of the four tasks, a detailed description of the dataset, objective, performance metrics, and the proposed baseline method are provided. To encourage further research on advanced methods for sport understanding, a competition is organized as part of the MMSports workshop from the ACM Multimedia 2022 conference, where participants have to develop state-of-the-art methods to solve the above tasks. The four datasets, development kits and baselines are publicly available.
References in corpus (7)
- Very Deep Convolutional Networks for Large-Scale Image Recognition
- Rethinking Atrous Convolution for Semantic Image Segmentation
- SoccerNet-Tracking: Multiple Object Tracking Dataset and Benchmark in Soccer Videos
- Torchreid: A Library for Deep Learning Person Re-Identification in Pytorch
- Ball 3D Localization From A Single Calibrated Image
- VIPriors 1: Visual Inductive Priors for Data-Efficient Deep Learning Challenges
- Accelerating the creation of instance segmentation training sets through bounding box annotation
Cited by in corpus (14)
- Body Part-Based Representation Learning for Occluded Person Re-Identification
- Fine-Grained Sports, Yoga, and Dance Postures Recognition: A Benchmark Analysis
- SoccerNet-Caption: Dense Video Captioning for Soccer Broadcasts Commentaries
- VARS: Video Assistant Referee System for Automated Soccer Decision Making from Multiple Views
- SoccerNet 2023 Challenges Results
- X-VARS: Introducing Explainability in Football Refereeing with Multi-Modal Large Language Model
- Keypoint Promptable Re-Identification
- Towards Active Learning for Action Spotting in Association Football Videos
- A Universal Protocol to Benchmark Camera Calibration for Sports
- Multi-task Learning for Joint Re-identification, Team Affiliation, and Role Classification for Sports Visual Tracking
- CLIP-ReIdent: Contrastive Training for Player Re-Identification
- KaliCalib: A Framework for Basketball Court Registration
- Context-Aware 3D Object Localization from Single Calibrated Images: A Study of Basketballs
- TrackID3x3: A Dataset and Algorithm for Multi-Player Tracking with Identification and Pose Estimation in 3x3 Basketball Full-court Videos