collaborators

6 papers

cs.AI2026

Can AI agents conduct open-ended AI research? Early evidence from two case studies

Peter Kirgis, Sayash Kapoor, Andrew Schwartz +21

The paper evaluates whether current AI agents can independently conduct open‑ended AI research by having them attempt to solve the central questions of two unpublished NeurIPS subm…

cs.CY2026

Are LLMs Bad at Moral Reasoning?

Menghang Zhu, Seth Lazar

For highly capable AI systems to operate safely in dynamic, open-ended environments, they must be able to identify, understand, and respond to moral reasons for action, and constra…

cs.CV2026

NoRA: Evaluating Grounded Reasonableness in Visual First-person Normative Action Reasoning

Sichao Li, Sai Ma, Daniel Kilov +3

LLMs and agentic systems are increasingly deployed in social environments, making normative competence critical for safe and appropriate behavior. However, existing approaches eith…

cs.CY2026

Legal Alignment for Safe and Ethical AI

Noam Kolt, Nicholas Caputo, Jack Boeglin +14

Alignment of artificial intelligence (AI) encompasses the normative problem of specifying how AI systems should act and the technical problem of ensuring AI systems comply with tho…

cs.AI2026

Open-World Evaluations for Measuring Frontier AI Capabilities

Sayash Kapoor, Peter Kirgis, Andrew Schwartz +15

Benchmark-based evaluation remains important for tracking frontier AI progress. But it can both overstate and understate deployed capability because it privileges tasks that can be…

cs.AI2025

Infrastructure for AI Agents

Alan Chan, Kevin Wei, Sihao Huang +5

AI agents plan and execute interactions in open-ended environments. For example, OpenAI's Operator can use a web browser to do product comparisons and buy online goods. Much resear…