3 papers
cs.CV2026
TennisExpert: Towards Expert-Level Analytical Sports Video Understanding
Zhaoyu Liu, Xi Weng, Lianyu Hu +4
Tennis is one of the most widely followed sports, generating extensive broadcast footage with strong potential for professional analysis, automated coaching, and real-time commenta…
cs.LG2025
Clustering Properties of Self-Supervised Learning
Xi Weng, Jianing An, Xudong Ma +5
Self-supervised learning (SSL) methods via joint embedding architectures have proven remarkably effective at capturing semantically rich representations with strong clustering prop…
cs.CV2025
TinyLLaVA-Video: Towards Smaller LMMs for Video Understanding with Group Resampler
Xingjian Zhang, Xi Weng, Yihao Yue +3
Video behavior recognition and scene understanding are fundamental tasks in multimodal intelligence, serving as critical building blocks for numerous real-world applications. Throu…