Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Similar Models Learn Differently: Final-Window Pretraining Shapes Post-Training Beyond SFT
Cen Lu, Yung-Chen Tang, Andrea Cavallaro
Developers judge a model checkpoint by how it behaves. After supervised fine-tuning (SFT), two checkpoints that perform about the same across relevant benchmarks are treated as int…
cs.AI2026
Sparse Neuron Ablation Triggers Catastrophic Collapse of the Language Core in Large Vision-Language Models
Cen Lu, Yung-Chen Tang, Andrea Cavallaro
Large Vision-Language Models (LVLMs) have shown impressive multimodal understanding capabilities, yet the structures that sustain their functionality remain poorly understood from…