◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Achintya Kundu

1 paper hereh-index 8164 citations15 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

fields
  • cs.LG1

identity via Semantic Scholar / OpenAlex

collaborators

1 paper

cs.LG2024

Enhancing Training Efficiency Using Packing with Flash Attention

Achintya Kundu, Rhui Dih Lee, Laura Wynter +2

Padding is often used in tuning LLM models by adding special tokens to shorter training examples to match the length of the longest sequence in each batch. While this ensures unifo…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.