◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Weiran He

14 papers hereh-index 122.9k citations16 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author6

Across the 6 of 14 papers where every author was matched, so the position is known.

fields
  • cs.CL4
  • cs.DC3
  • cs.LG3
  • astro-ph.IM1
  • cs.AI1
  • cs.CV1

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.DCShow all

3 papers · 1 filter

cs.DC2026

Prefill-as-a-Service: KVCache of Next-Generation Models Could Go Cross-Datacenter

Ruoyu Qin, Weiran He, Yaoyu Wang +5

Prefill-decode (PD) disaggregation has become the standard architecture for large-scale LLM serving, but in practice its deployment boundary is still determined by KVCache transfer…

cs.DC2026

Seer: Online Context Learning for Fast Synchronous LLM Reinforcement Learning

Ruoyu Qin, Weiran He, Weixiao Huang +7

Reinforcement Learning (RL) has emerged as a critical technique for advancing modern Large Language Models (LLMs), yet existing synchronous RL systems face severe performance bottl…

cs.DC2025

Mooncake: A KVCache-centric Disaggregated Architecture for LLM Serving

Ruoyu Qin, Zheming Li, Weiran He +4

Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI. It features a KVCache-centric disaggregated architecture that separates the prefill and de…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.