◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Lin Yan

8 papers hereh-index 72.4k citations8 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3
  • last author5

Across the 8 of 8 papers where every author was matched, so the position is known.

fields
  • cs.LG4
  • cs.AI3
  • cs.CL1
same name
  • Lin Yan — 15 papers, h 15
  • Lin Yan — 8 papers, h 5
  • Lin Yan — 5 papers, h 2
  • Lin Yan — 3 papers, h 3
  • Lin Yan — 2 papers, h 1
  • Lin Yan — 2 papers, h 9

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2025

DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Qiying Yu, Zheng Zhang, Ruofei Zhu +32

Inference scaling empowers LLMs with unprecedented reasoning ability, with reinforcement learning as the core technique to elicit complex reasoning. However, key technical details…

cs.LG2025

A Unified Pairwise Framework for RLHF: Bridging Generative Reward Modeling and Policy Optimization

Wenyuan Xu, Xiaochen Zuo, Chao Xin +3

Reinforcement Learning from Human Feedback (RLHF) has emerged as a important paradigm for aligning large language models (LLMs) with human preferences during post-training. This fr…

cs.LG2025

Exploring Data Scaling Trends and Effects in Reinforcement Learning from Human Feedback

Wei Shen, Guanlin Liu, Zheng Wu +5

Reinforcement Learning from Human Feedback (RLHF) is crucial for aligning large language models with human preferences. While recent research has focused on algorithmic improvement…

cs.LG2025

What's Behind PPO's Collapse in Long-CoT? Value Optimization Holds the Secret

Yufeng Yuan, Yu Yue, Ruofei Zhu +2

Reinforcement learning (RL) is pivotal for enabling large language models (LLMs) to generate long chains of thought (CoT) for complex tasks like math and reasoning. However, Proxim…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.