Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Multi-Object Tracking Retrieval with LLaVA-Video: A Training-Free Solution to MOT25-StAG Challenge
Yi Yang, Yiming Xu, Timo Kaiser +3
In this report, we present our solution to the MOT25-Spatiotemporal Action Grounding (MOT25-StAG) Challenge. The aim of this challenge is to accurately localize and track multiple…
cs.CV2025
FDSG: Forecasting Dynamic Scene Graphs
Yi Yang, Yuren Cong, Hao Cheng +2
Dynamic scene graph generation extends scene graph generation from images to videos by modeling entity relationships and their temporal evolution. However, existing methods either…