2 papers
cs.AI2026
When Robots Do the Chores: A Benchmark and Agent for Long-Horizon Household Task Execution
Zilin Zhu, Longteng Guo, Yanghong Mei +5
Long-horizon household tasks demand robust high-level planning and sustained reasoning capabilities, which are largely overlooked by existing embodied AI benchmarks that emphasize…
cs.CV2026
Thinking in Streaming Video
Zikang Liu, Longteng Guo, Handong Li +7
Real-time understanding of continuous video streams is essential for interactive assistants and multimodal agents operating in dynamic environments. However, most existing video re…