◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yu Li

7 papers hereh-index 28 citations6 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author5

Across the 7 of 7 papers where every author was matched, so the position is known.

fields
  • cs.AI3
  • cs.CL2
  • cs.LG2
same name
  • Yu Li — 18 papers, h 10
  • Yu Li — 18 papers, h 7
  • Yu Li — 16 papers, h 8
  • Yu Li — 14 papers, h 5
  • Yu Li — 12 papers, h 9
  • Yu Li — 9 papers, h 16

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.AIShow all

3 papers · 1 filter

cs.AI2026

SRPO: Setwise Relative Policy Optimization for Multi-Agent LLMs

Shengtian Yang, Ziyu Xiong, Yu Li +3

Multi-agent large language models solve complex tasks by coordinating several policies in a shared environment. However, existing reinforcement learning methods usually optimize ea…

cs.AI2026

Phase-Aware Mixture of Experts for Agentic Reinforcement Learning

Shengtian Yang, Yu Li, Shuo He +4

Reinforcement learning (RL) has equipped LLM agents with a strong ability to solve complex tasks. However, existing RL methods normally use a \emph{single} policy network, causing…

cs.AI2026

Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning

Yu Li, Mingyang Yi, Xiuyu Li +6

Agentic Reinforcement Learning (ARL) trains large language models to interleave reasoning with external tool execution to solve complex tasks. Most existing ARL methods train a sin…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.