◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Charlie Griffin

0 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

same name
  • Charlie Griffin — 3 papers, h 3
  • Charlie Griffin — 3 papers, h 3
  • Charlie Griffin — 3 papers, h 3

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

2 papers

cs.LG2023

Reinforcement Learning Fine-tuning of Language Models is Biased Towards More Extractable Features

Diogo Cruz, Edoardo Pona, Alex Holness-Tofts +4

Many capable large language models (LLMs) are developed via self-supervised pre-training followed by a reinforcement-learning fine-tuning phase, often based on human or AI feedback…

cs.LG2023★ 1 cited

Goodhart's Law in Reinforcement Learning

Jacek Karwowski, Oliver Hayman, Xingjian Bai +3

Implementing a reward function that perfectly captures a complex task in the real world is impractical. As a result, it is often appropriate to think of the reward function as a pr…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.