Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
MobilityBench: A Benchmark for Evaluating Route-Planning Agents in Real-World Mobility Scenarios
Zhiheng Song, Jingshuai Zhang, Chuan Qin +6
Route-planning agents powered by large language models (LLMs) have emerged as a promising paradigm for supporting everyday human mobility through natural language interaction and t…
cs.AI2026
SciHorizon-DataEVA: An Agentic System for AI-Readiness Evaluation of Heterogeneous Scientific Data
Dianyu Liu, Chuan Qin, Xi Chen +6
AI-for-Science (AI4Science) is increasingly transforming scientific discovery by embedding machine learning models into prediction, simulation, and hypothesis generation workflows…