◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

S. S. Du

5 papers hereh-index 4118 citations11 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3
  • last author2

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.LG5

identity via Semantic Scholar / OpenAlex

collaborators

5 papers

cs.LG2024

Anytime Acceleration of Gradient Descent

Zihan Zhang, Jason D. Lee, Simon S. Du +1

This work investigates stepsize-based acceleration of gradient descent with {\em anytime} convergence guarantees. For smooth (non-strongly) convex optimization, we propose a stepsi…

cs.LG2024

Horizon-Free Regret for Linear Markov Decision Processes

Zihan Zhang, Jason D. Lee, Yuxin Chen +1

A recent line of works showed regret bounds in reinforcement learning (RL) can be (nearly) independent of planning horizon, a.k.a.~the horizon-free bounds. However, these regret bo…

cs.LG2023

Optimal Multi-Distribution Learning

Zihan Zhang, Wenhao Zhan, Yuxin Chen +2

Multi-distribution learning (MDL), which seeks to learn a shared model that minimizes the worst-case risk across k distinct data distributions, has emerged as a unified framework…

cs.LG2023

Dichotomy of Early and Late Phase Implicit Biases Can Provably Induce Grokking

Kaifeng Lyu, Jikai Jin, Zhiyuan Li +3

Recent work by Power et al. (2022) highlighted a surprising "grokking" phenomenon in learning arithmetic tasks: a neural net first "memorizes" the training set, resulting in perfec…

cs.LG2023

Robust Offline Reinforcement Learning -- Certify the Confidence Interval

Jiarui Yao, Simon Shaolei Du

Currently, reinforcement learning (RL), especially deep RL, has received more and more attention in the research area. However, the security of RL has been an obvious problem due t…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.