1 citations · 1 across the 1 of their papers we have counts for
2 papers
cs.CV2025
TeleEgo: Benchmarking Egocentric AI Assistants in the Wild
Jiaqi Yan, Ruilong Ren, Jingren Liu +13
Egocentric AI assistants in real-world settings must process multi-modal inputs (video, audio, text), respond in real time, and retain evolving long-term memory. However, existing…
cs.CV2025★ 1 cited
Infinite Video Understanding
Dell Zhang, Xiangyu Chen, Jixiang Luo +6
The rapid advancements in Large Language Models (LLMs) and their multimodal extensions (MLLMs) have ushered in remarkable progress in video understanding. However, a fundamental ch…