◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

D. Duvenaud

13 papers hereh-index 5632.7k citations113 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author8
  • last author5

Across the 13 of 13 papers where every author was matched, so the position is known.

fields
  • cs.AI4
  • cs.CY3
  • cs.LG3
  • stat.ML2
  • cs.CL1

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedA Definition of AGI

3 citations · 7 across the 5 of their papers we have counts for

collaborators
Showing cs.LGShow all

3 papers · 1 filter

cs.LG2025

Full-Stack Alignment: Co-Aligning AI and Institutions with Thick Models of Value

Joe Edelman, Tan Zhi-Xuan, Ryan Lowe +30

Beneficial societal outcomes cannot be guaranteed by aligning individual AI systems with the intentions of their operators or users. Even an AI system that is perfectly aligned to…

cs.LG2024

Sabotage Evaluations for Frontier Models

Joe Benton, Misha Wagner, Eric Christiansen +13

Sufficiently capable models could subvert human oversight and decision-making in important contexts. For example, in the context of AI development, models could covertly sabotage e…

cs.LG2024

Experts Don't Cheat: Learning What You Don't Know By Predicting Pairs

Daniel D. Johnson, Daniel Tarlow, David Duvenaud +1

Identifying how much a model p​I^​¸(Y∣X) knows about the stochastic real-world process p(Y∣X) it was trained on is important to ensure it avoids producing incorrect o…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.