3 papers
cs.RO2026
VL-LN Bench: Towards Long-horizon Goal-oriented Navigation with Active Dialogs
Wensi Huang, Shaohao Zhu, Meng Wei +7
In most existing embodied navigation tasks, instructions are well-defined and unambiguous, such as instruction following and object searching. Under this idealized setting, agents…
cs.CV2025
MMSI-Video-Bench: A Holistic Benchmark for Video-Based Spatial Intelligence
Jingli Lin, Runsen Xu, Shaohao Zhu +11
Spatial understanding over continuous visual input is crucial for MLLMs to evolve into general-purpose assistants in physical environments. Yet there is still no comprehensive benc…
cs.RO2024
MAexp: A Generic Platform for RL-based Multi-Agent Exploration
Shaohao Zhu, Jiacheng Zhou, Anjun Chen +3
The sim-to-real gap poses a significant challenge in RL-based multi-agent exploration due to scene quantization and action discretization. Existing platforms suffer from the ineffi…