Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
ViSAGE: Constructing Self-Correcting Memories for Long-Form Video Understanding
Xinkui Zhao, Enbo Chen, Yifan Zhang +4
Multimodal agents operating in long-horizon environments must build and continually update multimedia memories to support entity-consistent, temporally grounded reasoning. However,…
cs.AI2025
LightRouter: Towards Efficient LLM Collaboration with Minimal Overhead
Yifan Zhang, Xinkui Zhao, Zuxin Wang +4
The rapid advancement of large language models has unlocked remarkable capabilities across a diverse array of natural language processing tasks. However, the considerable differenc…