◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Qiang Lin

15 papers hereh-index 469 citations20 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author10
  • last author4

Across the 15 of 15 papers where every author was matched, so the position is known.

fields
  • cs.AI6
  • cs.LG4
  • cs.SE2
  • cs.CL1
  • cs.CV1
  • cs.MA1
same name
  • Qiang Lin — 4 papers
  • Qiang Lin — 3 papers, h 4
  • Qiang Lin — 1 paper
  • Qiang Lin — 1 paper, h 7
  • Qiang Lin — 1 paper, h 3
  • Qiang Lin — 1 paper, h 8

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2026

Benchmarking the Residual: What Long-Horizon Evaluations Add Beyond Matched Short-Task Performance

Chao Peng, Zhiheng Lyu, Peijie Dong +2

Long-horizon benchmarks often show that agents fail more as tasks become longer. This observation is useful for deployment, but it does not by itself explain why failure occurs. Mo…

cs.LG2026

LEPO: Latent Reasoning Policy Optimization for Large Language Models

Yuyan Zhou, Jiarui Yu, Hande Dong +4

Recently, latent reasoning has been introduced into large language models (LLMs) to leverage rich information within a continuous space. However, without stochastic sampling, these…

cs.LG2025

GAPO: Robust Advantage Estimation for Real-World Code LLMs

Jianqing Zhang, Zhezheng Hao, Wei Xia +7

Reinforcement learning (RL) is widely used for post-training large language models (LLMs) in code editing, where group-relative methods, such as GRPO, are popular due to their crit…

cs.LG2025

Rethinking Entropy Interventions in RLVR: An Entropy Change Perspective

Zhezheng Hao, Hong Wang, Haoyang Liu +6

Reinforcement Learning with Verifiable Rewards (RLVR) serves as a cornerstone technique for enhancing the reasoning capabilities of Large Language Models (LLMs). However, its train…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.