◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yong Liu

6 papers hereh-index 458 citations21 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author5

Across the 5 of 6 papers where every author was matched, so the position is known.

fields
  • cs.AI3
  • cs.LG3
same name
  • Yong Liu — 20 papers, h 5
  • Yong Liu — 17 papers, h 12
  • Yong Liu — 14 papers, h 10
  • Yong Liu — 14 papers, h 10
  • Yong Liu — 13 papers, h 5
  • Yong Liu — 12 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.AIShow all

4 papers · 1 filter

cs.AI2026

Learn More with Less: Uncertainty Consistency Guided Query Selection for RLVR

Hao Yi, Yulan Hu, Xin Li +3

Large Language Models (LLMs) have recently improved mathematical reasoning through Reinforcement Learning with Verifiable Reward (RLVR). However, existing RLVR algorithms require l…

cs.AI2025

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning

Yulan Hu, Sheng Ouyang, Jinman Zhao +1

The Process Reward Model (PRM) plays a crucial role in mathematical reasoning tasks, requiring high-quality supervised process data. However, we observe that reasoning steps genera…

cs.AI2024

GUNDAM: Aligning Large Language Models with Graph Understanding

Sheng Ouyang, Yulan Hu, Ge Chen +1

Large Language Models (LLMs) have achieved impressive results in processing text data, which has sparked interest in applying these models beyond textual data, such as graphs. In t…

cs.AI2024

Towards Comprehensive Preference Data Collection for Reward Modeling

Yulan Hu, Qingyang Li, Sheng Ouyang +6

Reinforcement Learning from Human Feedback (RLHF) facilitates the alignment of large language models (LLMs) with human preferences, thereby enhancing the quality of responses gener…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.