◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Avinash Kumar

6 papers hereh-index 332 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author2

Across the 6 of 6 papers where every author was matched, so the position is known.

fields
  • cs.CL4
  • cs.CR1
  • cs.DC1
same name
  • Avinash Kumar — 3 papers, h 1
  • Avinash Kumar — 3 papers, h 3
  • Avinash Kumar — 2 papers, h 0
  • Avinash Kumar — 2 papers, h 1
  • Avinash Kumar — 1 paper, h 6
  • Avinash Kumar — 1 paper, h 0

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.CLShow all

4 papers · 1 filter

cs.CL2026

HiSpec: Hierarchical Speculative Decoding for LLMs

Avinash Kumar, Sujay Sanghavi, Poulami Das

Speculative decoding accelerates LLM inference by using a smaller draft model to speculate tokens that a larger target model verifies. Verification is often the bottleneck (e.g. ve…

cs.CL2026

Test-Time Speculation

Avinash Kumar, Sujay Sanghavi, Poulami Das

Speculative decoding accelerates LLM inference by using a fast draft model to generate tokens and a more accurate target model to verify them. Its performance depends on the $\text…

cs.CL2025

HELIOS: Adaptive Model And Early-Exit Selection for Efficient LLM Inference Serving

Avinash Kumar, Shashank Nag, Jason Clemons +2

Early-Exit Large Language Models (EE-LLMs) enable high throughput inference by allowing tokens to exit early at intermediate layers. However, their throughput is limited by the com…

cs.CL2025

Dialogue Without Limits: Constant-Sized KV Caches for Extended Responses in LLMs

Ravi Ghadia, Avinash Kumar, Gaurav Jain +2

Autoregressive Transformers rely on Key-Value (KV) caching to accelerate inference. However, the linear growth of the KV cache with context length leads to excessive memory consump…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.