papers

Publications (29)

cs.AI2023

Identifying and Mitigating the Security Risks of Generative AI

Clark Barrett, Brad Boyd, Elie Burzstein +20

cs.LG2020

Higher-Order Function Networks for Learning Composable 3D Object Representations

Eric Mitchell, Selim Engin, Volkan Isler +1

cs.LG2026

FSPO: Few-Shot Optimization of Synthetic Preferences Personalizes to Real Users

Anikait Singh, Sheryl Hsu, Kyle Hsu +5

cs.LG2024

A Critical Evaluation of AI Feedback for Aligning Large Language Models

Archit Sharma, Sedrick Keh, Eric Mitchell +3

cs.LG2021

Offline Meta-Reinforcement Learning with Advantage Weighting

Eric Mitchell, Rafael Rafailov, Xue Bin Peng +2

cs.RO2019

Pixels to Plans: Learning Non-Prehensile Manipulation by Imitating a Planner

Tarik Tosun, Eric Mitchell, Ben Eisner +6

cs.LG2022

Fast Model Editing at Scale

Eric Mitchell, Charles Lin, Antoine Bosselut +2

cs.CL2026

OpenAI GPT-5 System Card

Aaditya Singh, Adam Fry, Adam Perelman +483

cs.AI2022

Memory-Based Model Editing at Scale

Eric Mitchell, Charles Lin, Antoine Bosselut +2

cs.LG2024

RLVF: Learning from Verbal Feedback without Overgeneralization

Moritz Stephan, Alexander Khazatsky, Eric Mitchell +4

cs.LG2023

Self-Destructing Models: Increasing the Costs of Harmful Dual Uses of Foundation Models

Peter Henderson, Eric Mitchell, Christopher D. Manning +2

cs.AI2019

Q-Learning for Continuous Actions with Cross-Entropy Guided Policies

Riley Simmons-Edler, Ben Eisner, Eric Mitchell +2

cs.RO2019

Higher Order Function Networks for View Planning and Multi-View Reconstruction

Selim Engin, Eric Mitchell, Daewon Lee +2

cs.CL2023

An Emulator for Fine-Tuning Large Language Models using Small Language Models

Eric Mitchell, Rafael Rafailov, Archit Sharma +2

cs.CL2023

Just Ask for Calibration: Strategies for Eliciting Calibrated Confidence Scores from Language Models Fine-Tuned with Human Feedback

Katherine Tian, Eric Mitchell, Allan Zhou +5

cs.LG2026

Test-Time Alignment via Hypothesis Reweighting

Yoonho Lee, Jonathan Williams, Henrik Marklund +4

cs.LG2024

Online Adaptation of Language Models with a Memory of Amortized Contexts

Jihoon Tack, Jaehyung Kim, Eric Mitchell +3

cs.CL2023

Meta-Learning Online Adaptation of Language Models

Nathan Hu, Eric Mitchell, Christopher D. Manning +1

cs.CL2023

Fine-tuning Language Models for Factuality

Katherine Tian, Eric Mitchell, Huaxiu Yao +2

cs.CL2022

Enhancing Self-Consistency and Performance of Pre-Trained Language Models through Natural Language Inference

Eric Mitchell, Joseph J. Noh, Siyan Li +5

cs.LG2024

Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Rafael Rafailov, Archit Sharma, Eric Mitchell +3

cs.LG2024

Calibrating Language Models with Adaptive Temperature Scaling

Johnathan Xie, Annie S. Chen, Yoonho Lee +2

cs.LG2021

Reward Prediction Error as an Exploration Objective in Deep RL

Riley Simmons-Edler, Ben Eisner, Daniel Yang +4

cs.LG2022

On the Opportunities and Risks of Foundation Models

Rishi Bommasani, Drew A. Hudson, Ehsan Adeli +111

cs.CV2019

Siamese Encoding and Alignment by Multiscale Learning with Self-Supervision

Eric Mitchell, Stefan Keselj, Sergiy Popovych +2

cs.RO2021

Learning Language-Conditioned Robot Behavior from Offline Data and Crowd-Sourced Annotation

Suraj Nair, Eric Mitchell, Kevin Chen +3

cs.CL2023

DetectGPT: Zero-Shot Machine-Generated Text Detection using Probability Curvature

Eric Mitchell, Yoonho Lee, Alexander Khazatsky +2

cs.CL2023

RECKONING: Reasoning through Dynamic Knowledge Encoding

Zeming Chen, Gail Weiss, Eric Mitchell +2

cs.AI2026

OpenAI o1 System Card

OpenAI, :, Aaron Jaech +261