2 papers
cs.AI2026
StepGuard: Guarding Web Navigation via Single-Step Calibration
Zhihao Cui, Yuchen Zhang, Xiyang Sun +6
Web navigation requires agents to follow natural language goals, interact with web pages, and produce accurate answers. While recent advances leverage vision-language models and re…
cs.RO2026
Safe Reinforcement Learning of Autonomous Highway Driving: A Unified Framework for Safety and Efficiency
Chufei Yan, Zhihao Cui, Yiyan Lv +3
Deep reinforcement learning (DRL) offers a compelling route to decision-making for advanced autonomous vehicles (AVs), yet its trial-and-error nature makes it difficult to guarante…