◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Chao Ma

6 papers hereh-index 447 citations11 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author5

Across the 6 of 6 papers where every author was matched, so the position is known.

fields
  • cs.LG4
  • cs.CR1
  • cs.IR1
same name
  • Chao Ma — 24 papers, h 25
  • Chao Ma — 15 papers, h 10
  • Chao Ma — 14 papers, h 27
  • Chao Ma — 13 papers, h 27
  • Chao Ma — 10 papers, h 5
  • Chao Ma — 8 papers, h 3

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedEdgeLeakage: Membership Information Leakage in Distributed Edge Intelligence Systems

3 citations · 3 across the 6 of their papers we have counts for

collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2025

Towards Efficient Optimizer Design for LLM via Structured Fisher Approximation with a Low-Rank Extension

Wenbo Gong, Meyer Scetbon, Chao Ma +1

Designing efficient optimizers for large language models (LLMs) with low-memory requirements and fast convergence is an important and challenging problem. This paper makes a step t…

cs.LG2025

Gradient Multi-Normalization for Stateless and Scalable LLM Training

Meyer Scetbon, Chao Ma, Wenbo Gong +1

Training large language models (LLMs) typically relies on adaptive optimizers like Adam (Kingma & Ba, 2015) which store additional state information to accelerate convergence but i…

cs.LG2024

SWAN: SGD with Normalization and Whitening Enables Stateless LLM Training

Chao Ma, Wenbo Gong, Meyer Scetbon +1

Adaptive optimizers such as Adam (Kingma & Ba, 2015) have been central to the success of large language models. However, they often require to maintain optimizer states throughout…

cs.LG2024

Understanding the Generalization Benefits of Late Learning Rate Decay

Yinuo Ren, Chao Ma, Lexing Ying

Why do neural networks trained with large learning rates for a longer time often lead to better generalization? In this paper, we delve into this question by examining the relation…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.