3 papers
cs.AI2026
Similar Models Learn Differently: Final-Window Pretraining Shapes Post-Training Beyond SFT
Cen Lu, Yung-Chen Tang, Andrea Cavallaro
Developers judge a model checkpoint by how it behaves. After supervised fine-tuning (SFT), two checkpoints that perform about the same across relevant benchmarks are treated as int…
cs.LG2026
Metric-Aware Hybrid Forecasting for the CTF4Science Lorenz Challenge
Cen Lu
We describe our approach to the CTF4Science Lorenz challenge, a benchmark that mixes short-horizon forecasting, long-time distribution matching, and trajectory reconstruction acros…
cs.CL2024
Analyzing Nobel Prize Literature with Large Language Models
Zhenyuan Yang, Zhengliang Liu, Jing Zhang +19
This study examines the capabilities of advanced Large Language Models (LLMs), particularly the o1 model, in the context of literary analysis. The outputs of these models are compa…