2 papers
cs.LG2026
Emotional Preferences as Goal-Priority Regulation
Shiqi Liu, Yihua Tan, Hu Fu +1
A core question in decision-making for agents is whether the relative priorities of competing lower-level objectives can be determined by emotional preferences autonomously generat…
cs.CL2026
The Chase Is the Curriculum, the Capture Anchors the Credit: Pursuit-Evasion Self-Play for Zero-Data LLM Reasoning
Jing Yu, Shengchao Chen, Yiyun Tan
Reinforcement learning with verifiable rewards has become the dominant recipe for improving large language model reasoning, yet it presumes large human-curated task collections. Ze…