2 papers
cs.CL2026
Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models
Christine Zhang, Dan Jurafsky, Chen Shani
The next-token prediction (NTP) objective trains language models to predict a single token at each step, even though many continuations can express the same meaning. For example, i…
physics.chem-ph2025
MLIP Arena: Advancing Fairness and Transparency in Machine Learning Interatomic Potentials via an Open, Accessible Benchmark Platform
Yuan Chiang, Tobias Kreiman, Christine Zhang +11
Machine learning interatomic potentials (MLIPs) have revolutionized molecular and materials modeling, but existing benchmarks suffer from data leakage, limited transferability, and…