2 papers
cs.LG2026
Trading Human Curation for Synthetic Augmentation in RLVR
Akshansh, Leonardo Rosa Rodrigues, Michael Korostelev +2
The supply of high-quality training tasks is a central bottleneck for reinforcement learning from verifiable rewards (RLVR) on agentic language models. Each task requires a sandbox…
cs.AI2026
SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science
Nithin Somasekharan, Youssef Hassan, Shiyao Lin +5
Large Language Models (LLMs) are increasingly deployed as scientific AI as- sistants, and a growing body of benchmarks evaluates their capabilities across knowledge retrieval, reas…