2 papers
cs.LG2025
A Theoretical Perspective: How to Prevent Model Collapse in Self-consuming Training Loops
Shi Fu, Yingjie Wang, Yuzhu Chen +2
High-quality data is essential for training large generative models, yet the vast reservoir of real data available online has become nearly depleted. Consequently, models increasin…
cs.LG2025
HRP: High-Rank Preheating for Superior LoRA Initialization
Yuzhu Chen, Yingjie Wang, Shi Fu +4
This paper studies the crucial impact of initialization in Low-Rank Adaptation (LoRA). Through theoretical analysis, we demonstrate that the fine-tuned result of LoRA is highly sen…