◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yiwei Li

24 papers hereh-index 9352 citations30 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author20

Across the 24 of 24 papers where every author was matched, so the position is known.

fields
  • cs.CL16
  • cs.LG4
  • cs.AI3
  • cs.CV1
same name
  • Yiwei Li — 21 papers, h 15
  • Yiwei Li — 7 papers, h 2
  • Yiwei Li — 4 papers, h 1
  • Yiwei Li — 4 papers, h 6
  • Yiwei Li — 4 papers, h 17
  • Yiwei Li — 4 papers, h 2

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2026

Breaking the Self-Confirming Loop: Diagnosing and Mitigating Systemic Reward Bias in Self-Rewarding RL

Chuyi Tan, Peiwen Yuan, Xinglin Wang +8

Reinforcement learning with verifiable rewards (RLVR) efficiently scales the reasoning ability of large language models (LLMs) but is bottlenecked by scarce labeled data. Reinforce…

cs.LG2026

PatternKV: Flattening KV Representation Expands Quantization Headroom

Ji Zhang, Yiwei Li, Shaoxiong Feng +8

KV cache in autoregressive LLMs eliminates redundant recomputation but has emerged as the dominant memory and bandwidth bottleneck during inference, notably with long contexts and…

cs.LG2025

Every Rollout Counts: Optimal Resource Allocation for Efficient Test-Time Scaling

Xinglin Wang, Yiwei Li, Shaoxiong Feng +7

Test-Time Scaling (TTS) improves the performance of Large Language Models (LLMs) by using additional inference-time computation to explore multiple reasoning paths through search.…

cs.LG2025

Beyond One-Size-Fits-All: Tailored Benchmarks for Efficient Evaluation

Peiwen Yuan, Yueqi Zhang, Shaoxiong Feng +7

Evaluating models on large benchmarks is very resource-intensive, especially during the period of rapid model evolution. Existing efficient evaluation methods estimate the performa…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.