2 papers
cs.CV2026
VideoBrain: Learning Adaptive Frame Sampling for Long Video Understanding
Junbo Zou, Ziheng Huang, Shengjie Zhang +2
Long-form video understanding remains challenging for Vision-Language Models (VLMs) due to the inherent tension between computational constraints and the need to capture informatio…
cs.AI2025
LightAgent: Production-level Open-source Agentic AI Framework
Weige Cai, Tong Zhu, Jinyi Niu +6
With the rapid advancement of large language models (LLMs), Multi-agent Systems (MAS) have achieved significant progress in various application scenarios. However, substantial chal…