3 papers
cs.AI2026
TRCA: Transition-wise Rubric Credit Assignment for Long-horizon LLM Agents
Huan Zhang, Mingju Chen, Dongxu Zhou +5
Long-horizon large language model (LLM) agents are typically optimized with sparse terminal outcomes, making fine-grained credit assignment across multi-step interactions difficult…
cs.MA2026
A-MapReduce: Executing Wide Search via Agentic MapReduce
Mingju Chen, Guibin Zhang, Heng Chang +2
Contemporary large language model (LLM)-based multi-agent systems exhibit systematic advantages in deep research tasks, which emphasize iterative, vertically structured information…
cs.LG2025
Efficient Utility-Preserving Machine Unlearning with Implicit Gradient Surgery
Shiji Zhou, Tianbai Yu, Zhi Zhang +4
Machine unlearning (MU) aims to efficiently remove sensitive or harmful memory from a pre-trained model. The key challenge is to balance the potential tradeoff between unlearning e…