◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Tom Everitt

DeepMind

20 papers hereh-index 212.4k citations46 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author6
  • middle author10
  • last author3

Across the 19 of 20 papers where every author was matched, so the position is known.

fields
  • cs.AI14
  • cs.LG4
  • cs.GT1
  • cs.MA1
affiliations
  • DeepMind
Homepage
same name
  • Tom Everitt — 5 papers
  • Tom Everitt — 4 papers, h 7
  • Tom Everitt — 2 papers, h 3

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20152025
most citedAI Safety Gridworlds

117 citations · 184 across the 11 of their papers we have counts for

collaborators
Showing 2019Show all

2 papers · 1 filter

cs.AI2019

Reward Tampering Problems and Solutions in Reinforcement Learning: A Causal Influence Diagram Perspective

Tom Everitt, Marcus Hutter, Ramana Kumar +1

Can humans get arbitrarily capable reinforcement learning (RL) agents to do their bidding? Or will sufficiently capable RL agents always find ways to bypass their intended objectiv…

cs.AI2019★ 9 cited

Modeling AGI Safety Frameworks with Causal Influence Diagrams

Tom Everitt, Ramana Kumar, Victoria Krakovna +1

Proposals for safe AGI systems are typically made at the level of frameworks, specifying how the components of the proposed system should be trained and interact with each other. I…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.