◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yang Liu

12 papers hereh-index 6249 citations19 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author5
  • last author5

Across the 10 of 12 papers where every author was matched, so the position is known.

fields
  • cs.AI4
  • cs.LG4
  • cs.SE2
  • cs.CL1
  • cs.CR1
same name
  • Yang Liu — 50 papers
  • Yang Liu — 22 papers, h 13
  • Yang Liu — 20 papers, h 2
  • Yang Liu — 18 papers, h 9
  • Yang Liu — 16 papers, h 3
  • Yang Liu — 15 papers, h 5

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedSPARK: Stepwise Process-Aware Rewards for Reference-Free Reinforcement Learning

1 citations · 1 across the 8 of their papers we have counts for

collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2025★ 1 cited

SPARK: Stepwise Process-Aware Rewards for Reference-Free Reinforcement Learning

Salman Rahman, Sruthi Gorantla, Arpit Gupta +3

Process reward models (PRMs) that provide dense, step-level feedback have shown promise for reinforcement learning, yet their adoption remains limited by the need for expensive ste…

cs.LG2025

SABER: Small Actions, Big Errors -- Safeguarding Mutating Steps in LLM Agents

Alejandro Cuadron, Pengfei Yu, Yang Liu +1

Despite rapid progress in LLM agents, performance on long-horizon, tool-using tasks remains fragile. To better understand this fragility, we ask a simple question: \emph{do all act…

cs.LG2025

MM-OPERA: Benchmarking Open-ended Association Reasoning for Large Vision-Language Models

Zimeng Huang, Jinxin Ke, Xiaoxuan Fan +9

Large Vision-Language Models (LVLMs) have exhibited remarkable progress. However, deficiencies remain compared to human intelligence, such as hallucination and shallow pattern matc…

cs.LG2025

KodCode: A Diverse, Challenging, and Verifiable Synthetic Dataset for Coding

Zhangchen Xu, Yang Liu, Yueqin Yin +2

We introduce KodCode, a synthetic dataset that addresses the persistent challenge of acquiring high-quality, verifiable training data across diverse difficulties and domains for tr…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.