◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Runze Liu

9 papers hereh-index 11620 citations24 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author6

Across the 9 of 9 papers where every author was matched, so the position is known.

fields
  • cs.LG5
  • cs.CL4
same name
  • Runze Liu — 3 papers
  • Runze Liu — 3 papers, h 6
  • Runze Liu — 3 papers
  • Runze Liu — 2 papers, h 6
  • Runze Liu — 2 papers
  • Runze Liu — 2 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedGenPRM: Scaling Test-Time Compute of Process Reward Models via Generative Reasoning

2 citations · 6 across the 8 of their papers we have counts for

collaborators
Showing cs.CLShow all

4 papers · 1 filter

cs.CL2025★ 2 cited

A Survey of Reinforcement Learning for Large Reasoning Models

Kaiyan Zhang, Yuxin Zuo, Bingxiang He +36

In this paper, we survey recent advances in Reinforcement Learning (RL) for reasoning with Large Language Models (LLMs). RL has achieved remarkable success in advancing the frontie…

cs.CL2025

ReviewRL: Towards Automated Scientific Review with RL

Sihang Zeng, Kai Tian, Kaiyan Zhang +9

Peer review is essential for scientific progress but faces growing challenges due to increasing submission volumes and reviewer fatigue. Existing automated review approaches strugg…

cs.CL2025★ 2 cited

GenPRM: Scaling Test-Time Compute of Process Reward Models via Generative Reasoning

Jian Zhao, Runze Liu, Kaiyan Zhang +8

Recent advancements in Large Language Models (LLMs) have shown that it is promising to utilize Process Reward Models (PRMs) as verifiers to enhance the performance of LLMs. However…

cs.CL2025★ 2 cited

Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling

Runze Liu, Junqi Gao, Jian Zhao +5

Test-Time Scaling (TTS) is an important method for improving the performance of Large Language Models (LLMs) by using additional computation during the inference phase. However, cu…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.