2 papers
cs.LG2024
Safety vs. Performance: How Multi-Objective Learning Reduces Barriers to Market Entry
Meena Jagadeesan, Michael I. Jordan, Jacob Steinhardt
Emerging marketplaces for large language models and other large-scale machine learning (ML) models appear to exhibit market concentration, which has raised concerns about whether t…
cs.LG2024
Feedback Loops With Language Models Drive In-Context Reward Hacking
Alexander Pan, Erik Jones, Meena Jagadeesan +1
Language models influence the external world: they query APIs that read and write to web pages, generate content that shapes human behavior, and run system commands as autonomous a…