2 papers
cs.CV2024
Action Detection via an Image Diffusion Process
Lin Geng Foo, Tianjiao Li, Hossein Rahmani +1
Action detection aims to localize the starting and ending points of action instances in untrimmed videos, and predict the classes of those instances. In this paper, we make the obs…
cs.CV2023
Token Boosting for Robust Self-Supervised Visual Transformer Pre-training
Tianjiao Li, Lin Geng Foo, Ping Hu +4
Learning with large-scale unlabeled data has become a powerful tool for pre-training Visual Transformers (VTs). However, prior works tend to overlook that, in real-world scenarios,…