3 papers
cs.CV2026
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
Shuailei Ma, Jiaqi Liao, Xinyang Wang +24
Despite the recent promise in robot control, video generative models suffer from a domain mismatch due to their primary focus on content creation. For example, their design inheren…
cs.IR2026
InterDeepResearch: Enabling Human-Agent Collaborative Information Seeking through Interactive Deep Research
Bo Pan, Lunke Pan, Yitao Zhou +4
Deep research systems powered by LLM agents have transformed complex information seeking by automating the iterative retrieval, filtering, and synthesis of insights from massive-sc…
cs.CV2025
VIS-Shepherd: Constructing Critic for LLM-based Data Visualization Generation
Bo Pan, Yixiao Fu, Ke Wang +15
Data visualization generation using Large Language Models (LLMs) has shown promising results but often produces suboptimal visualizations that require human intervention for improv…