103 citations · 118 across the 3 of their papers we have counts for
3 papers
cs.CV2022★ 9 cited
Self-supervised Video Representation Learning with Motion-Aware Masked Autoencoders
Haosen Yang, Deng Huang, Bin Wen +5
Masked autoencoders (MAEs) have emerged recently as art self-supervised spatiotemporal representation learners. Inheriting from the image counterparts, however, existing video MAEs…
cs.CV2021★ 6 cited
Towards High-Quality Temporal Action Detection with Sparse Proposals
Jiannan Wu, Peize Sun, Shoufa Chen +4
Temporal Action Detection (TAD) is an essential and challenging topic in video understanding, aiming to localize the temporal segments containing human action instances and predict…
cs.MA2020★ 103 cited
SMARTS: Scalable Multi-Agent Reinforcement Learning Training School for Autonomous Driving
Ming Zhou, Jun Luo, Julian Villella +34
Multi-agent interaction is a fundamental aspect of autonomous driving in the real world. Despite more than a decade of research and development, the problem of how to competently i…