chain-of-thought prompting 1information density 1large language models 1post-hoc compression 1reasoning efficiency 1
From the 1 of 2 linked papers with an AI index.
2 papers
cs.AI2026
Valid Necessary: Diagnosing Latent Inefficiency in Chain-of-Thought
Daeyeop Lee, Hwanjo Yu
The paper identifies and diagnoses inefficient reasoning steps in chain-of-thought prompting for large language models, introducing a benchmark and a training-free metric (CAID) to…
cs.CL2025
Everyday Physics in Korean Contexts: A Culturally Grounded Physical Reasoning Benchmark
Jihae Jeong, DaeYeop Lee, DongGeon Lee +1
Existing physical commonsense reasoning benchmarks predominantly focus on Western contexts, overlooking cultural variations in physical problem-solving. To address this gap, we int…