From the 1 of 4 linked papers with an AI index.
4 papers
LAVE: Latent Visual Evidence-Enhanced Planning for Video Tool-use Agents
Zijian Wang, Junnan Zhu, Rongzhen Li +7
Long-video understanding requires models to efficiently acquire and reuse sparse visual evidence from long and redundant video streams. Recent video tool-use agents address this ch…
ProtoAct: Turning Wet-Lab Protocols into Embodied Robotic Actions
Zhe Liu, Jiaming Gu, Zhaohui Du +7
Biological wet-lab protocols are written for trained researchers and often leave routine operations, state-dependent conditions, and contextual parameters implicit, making them dif…
BioVLN: A Simulation Platform for Visual Language Navigation in Biomedical Laboratories
Zhe Liu, Quan Lu, Zhaohui Du +7
The paper presents BioVLN, a simulation platform that enables visual‑language navigation agents to safely approach biomedical laboratory instruments by modeling each instrument wit…
BioProVLA-Agent: An Affordable, Protocol-Driven, Vision-Enhanced VLA-Enabled Embodied Multi-Agent System with Closed-Loop-Capable Reasoning for Biological Laboratory Manipulation
Zhaohui Du, Zhe Wang, Dongzhan Zhou +9
Biological laboratory automation can reduce repetitive manual work and improve reproducibility, but reliable embodied execution in wet-lab environments remains challenging. Protoco…