1 paper
Yibei Liu, Jiajun Chen, Qianle Zhang +4
Reinforcement fine-tuning (RFT) is widely believed to inherently resist catastrophic forgetting in continual post-training of multimodal large language models. Under pronounced tas…