2 papers
cs.CL2026
-Bench: LLMs Struggle with Resource-Rational Reasoning under Shared Budgets
Peisong Wang, Zhiwei Ma, Bowen Liu +6
In cognitive science, resource rationality asks how an agent should allocate limited computation to maximize expected value. Most reasoning and agent benchmarks use independent per…
cs.CL2026
CP-Agent: A Calibrated Risk-Controlled Agent for Feedback-Driven Competitive Programming
Peisong Wang, Bowen Liu, Zehua Li +4
Large language models still struggle with contest-level programming, while many agentic remedies rely on massive inference-time sampling or expensive multi-stage post-training. We…