Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Towards Robust Endogenous Reasoning: Unifying Drift Adaptation in Non-Stationary Tuning
Xiaoyu Yang, En Yu, Wei Duan +1
Reinforcement Fine-Tuning (RFT) has established itself as a critical paradigm for the alignment of Multi-modal Large Language Models (MLLMs) with complex human values and domain-sp…
cs.LG2025
Resilient Contrastive Pre-training under Non-Stationary Drift
Xiaoyu Yang, Jie Lu, En Yu +1
The remarkable success of large-scale contrastive pre-training has been largely driven by by vast yet static datasets. However, as the scaling paradigm evolves, this paradigm encou…