◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Jiaming Guo

33 papers hereh-index 10409 citations54 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author29

Across the 32 of 33 papers where every author was matched, so the position is known.

fields
  • cs.LG16
  • cs.AI5
  • cs.CL4
  • cs.CV4
  • cs.AR1
  • cs.DC1
same name
  • Jiaming Guo — 13 papers, h 17
  • Jiaming Guo — 3 papers, h 12
  • Jiaming Guo — 2 papers, h 1
  • Jiaming Guo — 2 papers, h 2
  • Jiaming Guo — 1 paper
  • Jiaming Guo — 1 paper, h 16

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20192026
most citedCausality-driven Hierarchical Structure Discovery for Reinforcement Learning

13 citations · 37 across the 31 of their papers we have counts for

collaborators
Showing 2026Show all

4 papers · 1 filter

cs.RO2026

PrimitiveVLA: Learning Reusable Motion Primitives for Efficient and Generalizable Robotic Manipulation

Yutai Li, Shaohui Peng, Jiaming Guo +8

Vision-Language-Action (VLA) models offer a promising paradigm for generalist robotic policies, yet their adaptation is hindered by data inefficiency and poor generalization. We ar…

cs.LG2026

LKV: End-to-End Learning of Head-wise Budgets and Token Selection for LLM KV Cache Eviction

Enshuai Zhou, Yifan Hao, Chao Wang +7

Long-context inference in Large Language Models (LLMs) is bottlenecked by the linear growth of Key-Value (KV) cache memory. Existing KV cache compression paradigms are fundamentall…

cs.SE2026

QiMeng-PRepair: Precise Code Repair via Edit-Aware Reward Optimization

Changxin Ke, Rui Zhang, Jiaming Guo +10

Large Language Models (LLMs) achieve strong program repair performance but often suffer from over-editing, where excessive modifications overwrite correct code and hinder bug local…

cs.LG2026

Segmental Advantage Estimation: Enhancing PPO for Long-Context LLM Training

Xue Gong, Qi Yi, Ziyuan Nan +8

Training Large Language Models (LLMs) for reasoning tasks is increasingly driven by Reinforcement Learning with Verifiable Rewards (RLVR), where Proximal Policy Optimization (PPO)…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.