2 papers
cs.LG2026
Continual Reasoning Gym: Diagnosing and Harnessing Shared Reasoning in Continual RLVR
Lirui Luo, Guoxi Zhang, Hongming Xu +3
Reinforcement learning with verifiable rewards (RLVR) commonly post-trains reasoning models on multiple tasks, while rerunning multitask RLVR (MTRL) as new tasks are added makes ca…
cs.CL2025
Do Theory of Mind Benchmarks Need Explicit Human-like Reasoning in Language Models?
Yi-Long Lu, Chunhui Zhang, Jiajun Song +2
Theory of Mind (ToM), the ability to attribute mental states to others, is fundamental for human social intelligence and a critical capability for advanced Artificial Intelligence.…