2 papers
cs.AI2026
MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror
Shengyu Guo, Tongrui Ye, Jianbo Zhang +3
Recent progress in Multimodal Large Language Models (MLLMs) has demonstrated remarkable advances in perception and reasoning, suggesting their potential for embodied intelligence.…
cs.CV2025
Static and Plugged: Make Embodied Evaluation Simple
Jiahao Xiao, Jianbo Zhang, BoWen Yan +9
Embodied intelligence is advancing rapidly, driving the need for efficient evaluation. Current benchmarks typically rely on interactive simulated environments or real-world setups,…