2 papers
cs.CL2026
Writer-R1: Enhancing Generative Writing in LLMs via Memory-augmented Replay Policy Optimization
Jihao Zhao, Shuaishuai Zu, Zhiyuan Ji +2
As a typical open-ended generation task, creative writing lacks verifiable reference answers, which has long constrained reward modeling and automatic evaluation due to high human…
cs.CL2025
Invoke Interfaces Only When Needed: Adaptive Invocation for Large Language Models in Question Answering
Jihao Zhao, Chunlai Zhou, Daixuan Li +2
The collaborative paradigm of large and small language models (LMs) effectively balances performance and cost, yet its pivotal challenge lies in precisely pinpointing the moment of…