2 papers
cs.SE2026
ProgramBench: Can Language Models Rebuild Programs From Scratch?
John Yang, Kilian Lieret, Jeffrey Ma +9
Turning ideas into full software projects from scratch has become a popular use case for language models. Agents are being deployed to seed, maintain, and grow codebases over exten…
cs.LG2025
Understanding Silent Data Corruption in LLM Training
Jeffrey Ma, Hengzhi Pei, Leonard Lausen +1
As the scale of training large language models (LLMs) increases, one emergent failure is silent data corruption (SDC), where hardware produces incorrect computations without explic…