◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Percy Liang

15 papers hereh-index 14982 citations29 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author11
  • last author2

Across the 13 of 15 papers where every author was matched, so the position is known.

fields
  • cs.CY7
  • cs.CL4
  • cs.AI2
  • cs.LG2
same name
  • Percy Liang — 9 papers, h 7
  • Percy Liang — 8 papers, h 9
  • Percy Liang — 7 papers, h 7
  • Percy Liang — 6 papers, h 6
  • Percy Liang — 6 papers, h 6
  • Percy Liang — 5 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.CLShow all

4 papers · 1 filter

cs.CL2026

DelusionEval: Measuring Delusion-Linked Behaviors in AI Chatbots

Jared Moore, Andrea Mock, Yifan Mai +9

Mental health professionals have raised concerns about risks of psychological harm from interaction with large language models (LLMs), including "delusional spirals" in which conce…

cs.CL2025

LawInstruct: A Resource for Studying Language Model Adaptation to the Legal Domain

Joel Niklaus, Lucia Zheng, Arya D. McCarthy +7

Instruction tuning is an important step in making language models useful for direct user interaction. However, the legal domain is underrepresented in typical instruction datasets…

cs.CL2024

When Benchmarks are Targets: Revealing the Sensitivity of Large Language Model Leaderboards

Norah Alzahrani, Hisham Abdullah Alyahya, Yazeed Alnumay +9

Large Language Model (LLM) leaderboards based on benchmark rankings are regularly used to guide practitioners in model selection. Often, the published leaderboard rankings are take…

cs.CL2024

Introducing v0.5 of the AI Safety Benchmark from MLCommons

Bertie Vidgen, Adarsh Agrawal, Ahmed M. Ahmed +97

This paper introduces v0.5 of the AI Safety Benchmark, which has been created by the MLCommons AI Safety Working Group. The AI Safety Benchmark has been designed to assess the safe…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.