◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Zhongchun Zhou

4 papers hereh-index 283 citations6 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3

Across the 3 of 4 papers where every author was matched, so the position is known.

fields
  • cs.AR3
  • cs.PF1

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

cs.PF2026

PASCAL: A Phase-Aware Shared-Cache Model for Parallel Scans

Zhongchun Zhou, Chengtao Lai, Songtao Mao

In modern AI Accelerators and GPGPUs, many concurrent cores repeatedly access the same shared data. This pattern occurs in attention, where different query tiles share the same K/V…

cs.AR2026

Sim-FA: A GPGPU Simulator Framework for Fine-Grained Asynchronous Pipeline Analysis

Zhongchun Zhou, Yuhang Gu, Chengtao Lai +4

To efficiently support Large Language Models (LLMs), modern GPGPU architectures have introduced new features and programming paradigms, such as warp specialization. These features…

cs.AR2025

DCO: Dynamic Cache Orchestration for LLM Accelerators through Predictive Management

Zhongchun Zhou, Chengtao Lai, Yuhang Gu +1

The rapid adoption of large language models (LLMs) is pushing AI accelerators toward increasingly powerful and specialized designs. Instead of further complicating software develop…

cs.AR2025

LLaMCAT: Optimizing Large Language Model Inference with Cache Arbitration and Throttling

Zhongchun Zhou, Chengtao Lai, Wei Zhang

Large Language Models (LLMs) have achieved unprecedented success across various applications, but their substantial memory requirements pose significant challenges to current memor…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.