◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Avinash Maurya

9 papers hereh-index 8191 citations36 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author6

Across the 9 of 9 papers where every author was matched, so the position is known.

fields
  • cs.DC5
  • cs.LG2
  • cs.AI1
  • eess.IV1

identity via Semantic Scholar / OpenAlex

most citedDeep Optimizer States: Towards Scalable Training of Transformer Models Using Interleaved Offloading

9 citations · 9 across the 4 of their papers we have counts for

collaborators
Showing cs.LGShow all

2 papers · 1 filter

cs.LG2026

BOOST: BOttleneck-Optimized Scalable Training Framework for Low-Rank Large Language Models

Zhengyang Wang, Ziyue Liu, Ruijie Zhang +5

The scale of transformer model pre-training is constrained by the increasing computation and communication cost. Low-rank bottleneck architectures offer a promising solution to sig…

cs.LG2026★ 9 cited

Deep Optimizer States: Towards Scalable Training of Transformer Models Using Interleaved Offloading

Avinash Maurya, Jie Ye, M. Mustafa Rafique +2

Transformers and large language models~(LLMs) have seen rapid adoption in all domains. Their sizes have exploded to hundreds of billions of parameters and keep increasing. Under th…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.