2 papers
cs.CL2026
LoopRPT: Reinforcement Pre-Training for Looped Language Models
Guo Tang, Shixin Jiang, Heng Chang +6
Looped language models (LoopLMs) perform iterative latent computation to refine internal representations, offering a promising alternative to explicit chain-of-thought (CoT) reason…
cs.LG2025
Efficient Utility-Preserving Machine Unlearning with Implicit Gradient Surgery
Shiji Zhou, Tianbai Yu, Zhi Zhang +4
Machine unlearning (MU) aims to efficiently remove sensitive or harmful memory from a pre-trained model. The key challenge is to balance the potential tradeoff between unlearning e…