◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Daniel Bershatsky

3 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author2

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG3
same name
  • Daniel Bershatsky — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedSurvey on Large Scale Neural Network Training

8 citations · 13 across the 2 of their papers we have counts for

collaborators

3 papers

cs.LG2022★ 8 cited

Survey on Large Scale Neural Network Training

Julia Gusak, Daria Cherniuk, Alena Shilova +8

Modern Deep Neural Networks (DNNs) require significant memory to store weight, activations, and other intermediate tensors during training. Hence, many models do not fit one GPU de…

cs.LG2022★ 5 cited

Few-Bit Backward: Quantized Gradients of Activation Functions for Memory Footprint Reduction

Georgii Novikov, Daniel Bershatsky, Julia Gusak +3

Memory footprint is one of the main limiting factors for large neural network training. In backpropagation, one needs to store the input to each operation in the computational grap…

cs.LG2022

Memory-Efficient Backpropagation through Large Linear Layers

Daniel Bershatsky, Aleksandr Mikhalev, Alexandr Katrutsa +3

In modern neural networks like Transformers, linear layers require significant memory to store activations during backward pass. This study proposes a memory reduction approach to…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.