3 papers
cs.CL2026
Sandboxed Coding Agents are Competitive Omni-modal Task Solvers
Dongping Chen, Xuanao Huang, Zhihan Hu +3
As multimodal LLMs increasingly target video and audio, it is often assumed that such tasks require native omnimodal models. We show that this is not always the case: coding agents…
cs.RO2025
Modified-Emergency Index (MEI): A Criticality Metric for Autonomous Driving in Lateral Conflict
Hao Cheng, Yanbo Jiang, Qingyuan Shi +5
Effective, reliable, and efficient evaluation of autonomous driving safety is essential to demonstrate its trustworthiness. Criticality metrics provide an objective means of assess…
cs.AI2025
LinguaSim: Interactive Multi-Vehicle Testing Scenario Generation via Natural Language Instruction Based on Large Language Models
Qingyuan Shi, Qingwen Meng, Hao Cheng +2
The generation of testing and training scenarios for autonomous vehicles has drawn significant attention. While Large Language Models (LLMs) have enabled new scenario generation me…