papers
Publications (2)
cs.CV2025
Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding
Ye Wang, Ziheng Wang, Boshen Xu +14
Temporal Video Grounding (TVG), the task of locating specific video segments based on language queries, is a core challenge in long-form video understanding. While recent Large Vis…
cs.CV2020
SMPR: Single-Stage Multi-Person Pose Regression
Junqi Lin, Huixin Miao, Junjie Cao +2
Existing multi-person pose estimators can be roughly divided into two-stage approaches (top-down and bottom-up approaches) and one-stage approaches. The two-stage methods either su…