◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Dawei Yin

49 papers hereh-index 191.6k citations73 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author27
  • last author16

Across the 43 of 49 papers where every author was matched, so the position is known.

fields
  • cs.CL23
  • cs.IR9
  • cs.AI7
  • cs.LG6
  • cs.CV3
  • cs.DC1
same name
  • Dawei Yin — 12 papers, h 5
  • Dawei Yin — 11 papers, h 11
  • Dawei Yin — 8 papers, h 6
  • Dawei Yin — 8 papers, h 5
  • Dawei Yin — 7 papers, h 2
  • Dawei Yin — 7 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2026

CurES: From Gradient Analysis to Efficient Curriculum Learning for Reasoning LLMs

Yongcheng Zeng, Zexu Sun, Bokai Ji +7

Curriculum learning plays a crucial role in enhancing the training efficiency of large language models (LLMs) on reasoning tasks. However, existing methods often fail to adequately…

cs.LG2026

Solving the Granularity Mismatch: Hierarchical Preference Learning for Long-Horizon LLM Agents

Heyang Gao, Zexu Sun, Erxue Min +4

Large Language Models (LLMs) as autonomous agents are increasingly tasked with solving complex, long-horizon problems. Aligning these agents via preference-based offline methods li…

cs.LG2025

Staying in the Sweet Spot: Responsive Reasoning Evolution via Capability-Adaptive Hint Scaffolding

Ziheng Li, Zexu Sun, Jinman Zhao +8

Reinforcement learning with verifiable rewards (RLVR) has achieved remarkable success in enhancing the reasoning capabilities of large language models (LLMs). However, existing RLV…

cs.LG2025

Divide-Then-Aggregate: An Efficient Tool Learning Method via Parallel Tool Invocation

Dongsheng Zhu, Weixian Shi, Zhengliang Shi +4

Although current Large Language Models (LLMs) exhibit impressive capabilities, performing complex real-world tasks still requires tool learning. Mainstream methods, such as CoT/ReA…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.