◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Alvin Cheung

10 papers hereh-index 6246 citations12 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author5
  • last author4

Across the 9 of 10 papers where every author was matched, so the position is known.

fields
  • cs.AI4
  • cs.CL1
  • cs.CV1
  • cs.DB1
  • cs.DC1
  • cs.LG1
same name
  • Alvin Cheung — 5 papers, h 3
  • Alvin Cheung — 5 papers, h 3
  • Alvin Cheung — 5 papers, h 4
  • Alvin Cheung — 3 papers, h 2
  • Alvin Cheung — 3 papers, h 4
  • Alvin Cheung — 1 paper, h 31

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.AIShow all

4 papers · 1 filter

cs.AI2026

Qrita: High-performance Top-k and Top-p using Pivot-based Truncation and Selection

Jongseok Park, Sunga Kim, Alvin Cheung +1

Despite their importance in model sampling, efficient implementation of Top-k and Top-p algorithms for large vocabularies remains a significant challenge. Existing approaches often…

cs.AI2026

Combee: Scaling Prompt Learning for Self-Improving Language Model Agents

Hanchen Li, Runyuan He, Qizheng Zhang +11

Recent advances in prompt learning allow large language model agents to acquire task-relevant knowledge from inference-time context without parameter changes. For example, existing…

cs.AI2025

TurboSpec: Closed-loop Speculation Control System for Optimizing LLM Serving Goodput

Xiaoxuan Liu, Jongseok Park, Langxiang Hu +10

Large Language Model (LLM) serving systems batch concurrent user requests to achieve efficient serving. However, in real-world deployments, such inter-request parallelism from batc…

cs.AI2024

Online Speculative Decoding

Xiaoxuan Liu, Lanxiang Hu, Peter Bailis +4

Speculative decoding is a pivotal technique to accelerate the inference of large language models (LLMs) by employing a smaller draft model to predict the target model's outputs. Ho…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.