◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

A. Benazir

6 papers hereh-index 425 citations11 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author1

Across the 5 of 6 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • eess.AS2
  • cs.OS1
  • cs.PF1

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.LGShow all

2 papers · 1 filter

cs.LG2026

Benchmarking Composable Compression Techniques in Mixture-of-Experts LLMs

Afsara Benazir, Chen Chen, Rongxiao Qu +3

Mixture-of-Experts (MoE) LLMs scale model capacity efficiently through sparse activation, but their large expert parameter footprint, routing imbalance, and long-context KV-cache g…

cs.LG2026

Efficient Mixture-of-Experts LLM Inference with Apple Silicon NPUs

Afsara Benazir, Felix Xiaozhu Lin

Apple Neural Engine (ANE) is a dedicated neural processing unit (NPU) present in every Apple Silicon chip. Mixture-of-Experts (MoE) LLMs improve inference efficiency via sparse act…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.