◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Ryan J. Lowe

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author1
  • last author2

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CL2
  • cs.CY1
  • physics.flu-dyn1

identity via Semantic Scholar / OpenAlex

most citedTraining language models to follow instructions with human feedback

4.3k citations · 4.4k across the 3 of their papers we have counts for

collaborators
Showing cs.CLShow all

2 papers · 1 filter

cs.CL2022★ 4.3k cited

Training language models to follow instructions with human feedback

Long Ouyang, Jeff Wu, Xu Jiang +17

Making language models bigger does not inherently make them better at following a user's intent. For example, large language models can generate outputs that are untruthful, toxic,…

cs.CL2020

Learning an Unreferenced Metric for Online Dialogue Evaluation

Koustuv Sinha, Prasanna Parthasarathi, Jasmine Wang +3

Evaluating the quality of a dialogue interaction between two agents is a difficult task, especially in open-domain chit-chat style dialogue. There have been recent efforts to devel…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.