papers
Publications (4)
cs.AI2026
Human-AI Complementarity: A Goal for Amplified Oversight
Rishub Jain, Sophie Bridgers, Lili Janzer +3
cs.LG2022
Improving alignment of dialogue agents via targeted human judgements
Amelia Glaese, Nat McAleese, Maja TrÄbacz +31
cs.CL2025
Gemini: A Family of Highly Capable Multimodal Models
Gemini Team, Rohan Anil, Sebastian Borgeaud +1340
cs.AI2025
An Approach to Technical AGI Safety and Security
Rohin Shah, Alex Irpan, Alexander Matt Turner +27