2 papers
cs.LG2026
Position: Quantum Program Generation Must Prioritize Validity Over Probabilistic Scaling
Junhao Song, Yu Zhou, William Knottenbelt +1
The scaling hypothesis assumes that increasing model parameters yields emergent reasoning capabilities. This position paper argues that applying this probabilistic paradigm to gene…
cs.CL2026
Beyond Function Calling: Benchmarking Tool-Using Agents under Tool-Environment Unreliability
Yang Tian, Zhengpeng Shi, Yu Zhou +1
Large language models are increasingly deployed as agents that solve tasks by interacting with external tool environments. Although recent tool-use benchmarks increasingly cover co…