◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Jan Leike

34 papers hereh-index 3075k citations76 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author17
  • last author7

Across the 26 of 34 papers where every author was matched, so the position is known.

fields
  • cs.CL14
  • cs.LG13
  • cs.AI3
  • cs.CY2
  • cs.CR1
  • cs.SE1
same name
  • Jan Leike — 5 papers
  • Jan Leike — 5 papers, h 9
  • Jan Leike — 1 paper, h 1

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20172026
most citedTraining language models to follow instructions with human feedback

4.3k citations · 6.3k across the 28 of their papers we have counts for

collaborators
Showing cs.SEShow all

1 paper · 1 filter

cs.SE2024★ 8 cited

LLM Critics Help Catch LLM Bugs

Nat McAleese, Rai Michael Pokorny, Juan Felipe Ceron Uribe +3

Reinforcement learning from human feedback (RLHF) is fundamentally limited by the capacity of humans to correctly evaluate model output. To improve human evaluation ability and ove…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.