Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
RescueBench: Can Embodied Agents Save Lives in the Wild ?
Kui Wu, Beiyu Guo, Hao Chen +6
Search-and-rescue (SAR) requires embodied agents to explore unfamiliar environments under multimodal uncertainty, perform multi-stage interactions, and retrieve spatial memory over…
cs.CV2025
VLM Can Be a Good Assistant: Enhancing Embodied Visual Tracking with Self-Improving Vision-Language Models
Kui Wu, Shuhang Xu, Hao Chen +4
We introduce a novel self-improving framework that enhances Embodied Visual Tracking (EVT) with Vision-Language Models (VLMs) to address the limitations of current active visual tr…
cs.CV2025
Hierarchical Instruction-aware Embodied Visual Tracking
Kui Wu, Hao Chen, Churan Wang +4
User-Centric Embodied Visual Tracking (UC-EVT) presents a novel challenge for reinforcement learning-based models due to the substantial gap between high-level user instructions an…