Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Test-Time Scaling for Scientific Equation Discovery
Haowei Lin, Hubert Lim, Xiangyu Wang +2
Test-time scaling (TTS) improves language model reasoning by allocating additional test-time compute, but prior work mainly studies closed-ended tasks such as math and coding. We s…
cs.CL2025
Generative Evaluation of Complex Reasoning in Large Language Models
Haowei Lin, Xiangyu Wang, Ruilin Yan +7
With powerful large language models (LLMs) demonstrating superhuman reasoning capabilities, a critical question arises: Do LLMs genuinely reason, or do they merely recall answers f…