2 papers
cs.AI2025
Learned-Rule-Augmented Large Language Model Evaluators
Jie Meng, Jin Mao
Large language models (LLMs) are predominantly used as evaluators for natural language generation (NLG) tasks, but their application to broader evaluation scenarios remains limited…
cs.CV2025
NuScenes-SpatialQA: A Spatial Understanding and Reasoning Benchmark for Vision-Language Models in Autonomous Driving
Kexin Tian, Jingrui Mao, Yunlong Zhang +3
Recent advancements in Vision-Language Models (VLMs) have demonstrated strong potential for autonomous driving tasks. However, their spatial understanding and reasoning-key capabil…