◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Anuj Apte

9 papers hereh-index 436 citations10 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author1
  • first author6
  • middle author2

Across the 9 of 9 papers where every author was matched, so the position is known.

fields
  • quant-ph7
  • cs.LG2

identity via Semantic Scholar / OpenAlex

collaborators
Showing cs.LGShow all

2 papers · 1 filter

cs.LG2026

Scale Weight Decay and Train Better

Anuj Apte

The discovery of scaling laws has motivated training neural networks on ever increasing quantities of data. This is typically done with a constant decoupled weight decay which caus…

cs.LG2026

Anytime Training with Schedule-Free Spectral Optimization

Anuj Apte, Pranav Deshpande, Niraj Kumar +2

Standard neural network training relies on learning-rate schedules tied to a fixed horizon, leading to strong path dependence and costly re-tuning as data availability changes. Sch…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.