From the 1 of 1 linked paper with an AI index.
1 paper
Mao-Lin Luo, Zhe-Xu Wang, Zi-Hao Zhou +4
The paper investigates catastrophic forgetting in continual post‑training of vision‑language models with reinforcement learning, introduces the MRCL benchmark, and proposes a repla…