◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Mayank Mishra

2 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.LG1
  • cs.SE1
ORCID 0000-0001-8964-4133
same name
  • Mayank Mishra — 3 papers, h 9
  • Mayank Mishra — 2 papers, h 4
  • Mayank Mishra — 2 papers
  • Mayank Mishra — 1 paper
  • Mayank Mishra — 1 paper
  • Mayank Mishra — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedSantaCoder: don't reach for the stars!

54 citations · 56 across the 2 of their papers we have counts for

collaborators

2 papers

cs.LG2024★ 2 cited

Dense Training, Sparse Inference: Rethinking Training of Mixture-of-Experts Language Models

Bowen Pan, Yikang Shen, Haokun Liu +5

Mixture-of-Experts (MoE) language models can reduce computational costs by 2-4× compared to dense models without sacrificing performance, making them more efficient in compu…

cs.SE2023★ 54 cited

SantaCoder: don't reach for the stars!

Loubna Ben Allal, Raymond Li, Denis Kocetkov +38

The BigCode project is an open-scientific collaboration working on the responsible development of large language models for code. This tech report describes the progress of the col…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.