3 papers
cs.CV2026
LAVE: Latent Visual Evidence-Enhanced Planning for Video Tool-use Agents
Zijian Wang, Junnan Zhu, Rongzhen Li +7
Long-video understanding requires models to efficiently acquire and reuse sparse visual evidence from long and redundant video streams. Recent video tool-use agents address this ch…
cs.RO2026
ProtoAct: Turning Wet-Lab Protocols into Embodied Robotic Actions
Zhe Liu, Jiaming Gu, Zhaohui Du +7
Biological wet-lab protocols are written for trained researchers and often leave routine operations, state-dependent conditions, and contextual parameters implicit, making them dif…
cs.RO2026
BioVLN: A Simulation Platform for Visual Language Navigation in Biomedical Laboratories
Zhe Liu, Quan Lu, Zhaohui Du +7
Biomedical laboratory robots must navigate to instruments before performing experimental procedures. Existing embodied navigation platforms are designed for household environments…