3 papers
cs.CY2026
Muse Spark Safety & Preparedness Report
Cristina Menghini, Peter Ney, Hamza Kwisaba +117
Muse Spark is the latest large language model developed by Meta. In this report, we first present evaluations for catastrophic risk domains under Meta's Advanced AI Scaling Framewo…
cs.LG2025
Remote Labor Index: Measuring AI Automation of Remote Work
Mantas Mazeika, Alice Gatti, Cristina Menghini +44
AIs have made rapid progress on research-oriented benchmarks of knowledge and reasoning, but it remains unclear how these gains translate into economic value and automation. To mea…
cs.LG2025
LLM Priors for ERM over Programs
Shivam Singhal, Priyadarsi Mishra, Eran Malach +1
We study program-learning methods that are efficient in both samples and computation. Classical learning theory suggests that when the target admits a short program description, fo…