◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Quentin Anthony

11 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author4
  • middle author5
  • last author1

Across the 10 of 11 papers where every author was matched, so the position is known.

fields
  • cs.CL5
  • cs.DC4
  • cs.LG2
ORCID 0000-0002-6823-9080
same name
  • Quentin Anthony — 1 paper
  • Quentin Anthony — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedEmergent and Predictable Memorization in Large Language Models

26 citations · 71 across the 11 of their papers we have counts for

collaborators
Showing cs.LGShow all

2 papers · 1 filter

cs.LG2024★ 1 cited

The Zamba2 Suite: Technical Report

Paolo Glorioso, Quentin Anthony, Yury Tokpanov +5

In this technical report, we present the Zamba2 series -- a suite of 1.2B, 2.7B, and 7.4B parameter hybrid Mamba2-transformer models that achieve state of the art performance again…

cs.LG2024★ 1 cited

Exploiting Inter-Layer Expert Affinity for Accelerating Mixture-of-Experts Model Inference

Jinghan Yao, Quentin Anthony, Aamir Shafi +3

In large language models like the Generative Pre-trained Transformer, the Mixture of Experts paradigm has emerged as a powerful technique for enhancing model expressiveness and acc…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.