◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Tom Everitt

DeepMind

20 papers hereh-index 212.4k citations46 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author6
  • middle author10
  • last author3

Across the 19 of 20 papers where every author was matched, so the position is known.

fields
  • cs.AI14
  • cs.LG4
  • cs.GT1
  • cs.MA1
affiliations
  • DeepMind
Homepage
same name
  • Tom Everitt — 5 papers
  • Tom Everitt — 4 papers, h 7
  • Tom Everitt — 2 papers, h 3

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20152025
most citedAI Safety Gridworlds

117 citations · 184 across the 11 of their papers we have counts for

collaborators
Showing 2022Show all

2 papers · 1 filter

cs.AI2022★ 3 cited

Path-Specific Objectives for Safer Agent Incentives

Sebastian Farquhar, Ryan Carey, Tom Everitt

We present a general framework for training safe agents whose naive incentives are unsafe. As an example, manipulative or deceptive behaviour can improve rewards but should be avoi…

cs.AI2022

A Complete Criterion for Value of Information in Soluble Influence Diagrams

Chris van Merwijk, Ryan Carey, Tom Everitt

Influence diagrams have recently been used to analyse the safety and fairness properties of AI systems. A key building block for this analysis is a graphical criterion for value of…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.