3 papers
cs.CV2026
Turning Drift into Constraint: Robust Reasoning Alignment in Non-Stationary Multi-Stream Environments
Xiaoyu Yang, En Yu, Wei Duan +1
This paper identifies a critical yet underexplored challenge in reasoning alignment from multiple multi-modal large language models (MLLMs): In non-stationary environments, the div…
cs.LG2026
Towards Robust Endogenous Reasoning: Unifying Drift Adaptation in Non-Stationary Tuning
Xiaoyu Yang, En Yu, Wei Duan +1
Reinforcement Fine-Tuning (RFT) has established itself as a critical paradigm for the alignment of Multi-modal Large Language Models (MLLMs) with complex human values and domain-sp…
cs.LG2025
Resilient Contrastive Pre-training under Non-Stationary Drift
Xiaoyu Yang, Jie Lu, En Yu +1
The remarkable success of large-scale contrastive pre-training has been largely driven by by vast yet static datasets. However, as the scaling paradigm evolves, this paradigm encou…