4 papers
LabRobFail: A Benchmark for Robotic Failure Analysis in Chemical Self-driving Laboratory
Haobo Wang, Baoli Sun, Anqi Zou +8
The deployment of embodied agents in self-driving laboratories could accelerate scientific discovery, yet their reliability is constrained by the irreversible and safety-critical n…
LabOSBench: Benchmarking Computer Use Agents for Scientific Instrument Control
Anqi Zou, Han Deng, Chengyu Zhang +9
Current computer-use benchmarks primarily focus on software operation tasks in virtualized systems, whereas scientific instrumentation scenarios require coordinated control over co…
From Tokens to Regions: CUDA-Sensitive Instruction Tuning for GPU Kernel Generation
Wentao Chen, Jiace Zhu, Xing Zhe Chai +4
High-performance CUDA kernels are essential for scalable AI systems, while Large Language Models (LLMs) still struggle to generate correct kernels due to strict and implicit execut…
Owl-AuraID 1.0: An Intelligent System for Autonomous Scientific Instrumentation and Scientific Data Analysis
Han Deng, Anqi Zou, Hanling Zhang +14
Scientific discovery increasingly depends on high-throughput characterization, yet automation is hindered by proprietary GUIs and the limited generalizability of existing API-based…