◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yaodong Yang

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2
  • last author2

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.AI1
  • cs.CL1
  • cs.GT1
  • cs.LG1
same name
  • Yaodong Yang — 37 papers
  • Yaodong Yang — 27 papers
  • Yaodong Yang — 3 papers
  • Yaodong Yang — 1 paper
  • Yaodong Yang — 1 paper
  • Yaodong Yang — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedRedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?

1 citations · 1 across the 4 of their papers we have counts for

collaborators

4 papers

cs.LG2025★ 1 cited

RedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?

Haotian Xu, Xing Wu, Weinong Wang +11

Can scaling transform reasoning? In this work, we explore the untapped potential of scaling Long Chain-of-Thought (Long-CoT) data to 1000k samples, pioneering the development of a…

cs.CL2025

Stream Aligner: Efficient Sentence-Level Alignment via Distribution Induction

Hantao Lou, Jiaming Ji, Kaile Wang +1

The rapid advancement of large language models (LLMs) has led to significant improvements in their capabilities, but also to increased concerns about their alignment with human val…

cs.GT2025

Approximating N-Player Nash Equilibrium through Gradient Descent

Dongge Wang, Xiang Yan, Zehao Dou +3

Decoding how rational agents should behave in shared systems remains a critical challenge within theoretical computer science, artificial intelligence and economics studies. Centra…

cs.AI2024

Align Anything: Training All-Modality Models to Follow Instructions with Language Feedback

Jiaming Ji, Jiayi Zhou, Hantao Lou +16

Reinforcement learning from human feedback (RLHF) has proven effective in enhancing the instruction-following capabilities of large language models; however, it remains underexplor…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.