activity
20232025
most citedBenchmarking and Improving Generator-Validator Consistency of Language Models

2 citations · 3 across the 5 of their papers we have counts for

collaborators

13 papers

cs.LG2025

Pre-training under infinite compute

Konwoo Kim, Suhas Kotha, Percy Liang +1

Since compute grows much faster than web text available for language model pre-training, we ask how one should approach pre-training under fixed data and no compute constraints. We…

cs.CL2025

s1: Simple test-time scaling

Niklas Muennighoff, Zitong Yang, Weijia Shi +7

Test-time scaling is a promising new approach to language modeling that uses extra test-time compute to improve performance. Recently, OpenAI's o1 model showed this capability but…

cs.LG2025

Eliciting Language Model Behaviors with Investigator Agents

Xiang Lisa Li, Neil Chowdhury, Daniel D. Johnson +4

Language models exhibit complex, diverse behaviors when prompted with free-form text, making it difficult to characterize the space of possible outputs. We study the problem of beh…

cs.CL2025

Auditing Prompt Caching in Language Model APIs

Chenchen Gu, Xiang Lisa Li, Rohith Kuditipudi +2

Prompt caching in large language models (LLMs) results in data-dependent timing variations: cached prompts are processed faster than non-cached prompts. These timing differences in…

cs.CL2024

PERSONA: A Reproducible Testbed for Pluralistic Alignment

Louis Castricato, Nathan Lile, Rafael Rafailov +2

The rapid advancement of language models (LMs) necessitates robust alignment with diverse user values. However, current preference optimization approaches often fail to capture the…

cs.CL2024

AutoBencher: Towards Declarative Benchmark Construction

Xiang Lisa Li, Farzaan Kaiyom, Evan Zheran Liu +3

We present AutoBencher, a declarative framework for automatic benchmark construction, and use it to scalably discover novel insights and vulnerabilities of existing language models…