8 citations · 15 across the 7 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Mechanism-Driven Monitors for Preemptive Detection of LLM Training Instability
Ruixuan Huang, Hantao Huang, Yifan Huang +3
Frontier large language model training consumes massive accelerator fleets and long wall-clock computation, making stability failures costly when they occur. After a numerical or a…
cs.CL2024★ 5 cited
Generative Monoculture in Large Language Models
Fan Wu, Emily Black, Varun Chandrasekaran
We introduce {\em generative monoculture}, a behavior observed in large language models (LLMs) characterized by a significant narrowing of model output diversity relative to availa…