4 papers
MM-SEAL: A Large-scale Video Dataset of Multi-person Multi-grained Spatio-temporally Action Localization
Shimin Chen, Wei Li, Chen Chen +4
In this paper, we introduce a novel large-scale video dataset dubbed MM-SEAL for multi-person multi-grained spatio-temporal action localization among human daily life. We are the f…
Technical Report for ActivityNet Challenge 2022 -- Temporal Action Localization
Shimin Chen, Wei Li, Jianyang Gu +2
In the task of temporal action localization of ActivityNet-1.3 datasets, we propose to locate the temporal boundaries of each action and predict action class in untrimmed videos. W…
Technical Report for Soccernet 2023 -- Dense Video Captioning
Zheng Ruan, Ruixuan Liu, Shimin Chen +5
In the task of dense video captioning of Soccernet dataset, we propose to generate a video caption of each soccer action and locate the timestamp of the caption. Firstly, we apply…
Technical Report for SoccerNet Challenge 2022 -- Replay Grounding Task
Shimin Chen, Wei Li, Jiaming Chu +3
In order to make full use of video information, we transform the replay grounding problem into a video action location problem. We apply a unified network Faster-TAD proposed by us…