◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Xinjun Yang

3 papers hereh-index 9322 citations22 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author2

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.DC3

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

cs.DC2026

SAC: Disaggregated KV Cache System for Sparse Attention LLMs with CXL

Ruiyang Ma, Teng Ma, Junru Li +7

The scaling of LLMs toward long-context inference has shifted the primary serving system bottleneck from computation to memory capacity. Traditional solutions for dense attention m…

cs.DC2025

Beluga: A CXL-Based Memory Architecture for Scalable and Efficient LLM KVCache Management

Xinjun Yang, Qingda Hu, Junru Li +9

The rapid increase in LLM model sizes and the growing demand for long-context inference have made memory a critical bottleneck in GPU-accelerated serving systems. Although high-ban…

cs.DC2025

PolarStore: High-Performance Data Compression for Large-Scale Cloud-Native Databases

Qingda Hu, Xinjun Yang, Feifei Li +9

In recent years, resource elasticity and cost optimization have become essential for RDBMSs. While cloud-native RDBMSs provide elastic computing resources via disaggregated computi…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.