3 papers
cs.CV2026
Thought Graph Traversal for Test-time Scaling in Chest X-ray VLLMs
Yue Yao, Zelin Wen, Yan Tong +5
Test-time scaling offers a promising way to improve the reasoning performance of vision-language large models (VLLMs) without additional training. In this paper, we explore a simpl…
cs.CL2026
SCAN: Structured Capability Assessment and Navigation for LLMs
Zongqi Wang, Tianle Gu, Chen Gong +3
Evaluating Large Language Models (LLMs) has become increasingly important, with automatic evaluation benchmarks gaining prominence as alternatives to human evaluation. While existi…
cs.CL2025
Enhancing Text-to-SQL Capabilities of Large Language Models via Domain Database Knowledge Injection
Xingyu Ma, Xin Tian, Lingxiang Wu +3
Text-to-SQL is a subtask in semantic parsing that has seen rapid progress with the evolution of Large Language Models (LLMs). However, LLMs face challenges due to hallucination iss…