activity
20242026
most citedGemma 2: Improving Open Language Models at a Practical Size

145 citations · 152 across the 14 of their papers we have counts for

collaborators

17 papers

cs.CY2026

Can AI mediation improve democratic deliberation?

Michael Henry Tessler, Georgina Evans, Michiel A. Bakker +8

The strength of democracy lies in the free and equal exchange of diverse viewpoints. Living up to this ideal at scale faces inherent tensions: broad participation, meaningful delib…

cs.CL2025

The FACTS Leaderboard: A Comprehensive Benchmark for Large Language Model Factuality

Aileen Cheng, Alon Jacovi, Amir Globerson +62

We introduce The FACTS Leaderboard, an online leaderboard suite and associated set of benchmarks that comprehensively evaluates the ability of language models to generate factually…

cs.AI2025

Training LLM Agents to Empower Humans

Evan Ellis, Vivek Myers, Jens Tuyls +3

Assistive agents should not only take actions on behalf of a human, but also step out of the way and cede control when there are important decisions to be made. However, current me…

cs.AI2025

CTRL-Rec: Controlling Recommender Systems With Natural Language

Micah Carroll, Adeline Foote, Kevin Feng +4

When users are dissatisfied with recommendations from a recommender system, they often lack fine-grained controls for changing them. Large language models (LLMs) offer a solution b…

cs.LG2025

Evaluating Sparse Autoencoders for Monosemantic Representation

Moghis Fereidouni, Muhammad Umair Haider, Peizhong Ju +1

A key barrier to interpreting large language models is polysemanticity, where neurons activate for multiple unrelated concepts. Sparse autoencoders (SAEs) have been proposed to mit…

cs.AI2025

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Tomek Korbak, Mikita Balesni, Elizabeth Barnes +38

AI systems that "think" in human language offer a unique opportunity for AI safety: we can monitor their chains of thought (CoT) for the intent to misbehave. Like all other known A…