bias amplification 1concept chaining 1implicit reasoning 1language model steering 1model robustness 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CL2026
Implicit Reasoning Steering via Concept Chaining
Xiao Ye, Sanika Chavan, Yuxi Huang +4
The paper introduces Concept Chaining, a method that creates short natural-language paragraphs linking question entities to a target answer via intermediate concepts, and uses cont…
cs.CL2025
Evaluating Medical LLMs by Levels of Autonomy: A Survey Moving from Benchmarks to Applications
Xiao Ye, Jacob Dineen, Zhaonan Li +11
Medical Large language models achieve strong scores on standard benchmarks; however, the transfer of those results to safe and reliable performance in clinical workflows remains a…
cs.CL2025
ArenaBencher: Automatic Benchmark Evolution via Multi-Model Competitive Evaluation
Qin Liu, Jacob Dineen, Yuxi Huang +4
Benchmarks are central to measuring the capabilities of large language models and guiding model development, yet widespread data leakage from pretraining corpora undermines their v…