◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

D. Malik

7 papers hereh-index 6425 citations13 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author5
  • middle author2

Across the 7 of 7 papers where every author was matched, so the position is known.

fields
  • cs.LG5
  • cs.AI1
  • stat.ML1

identity via Semantic Scholar / OpenAlex

activity
20182024
most citedSample Efficient Reinforcement Learning In Continuous State Spaces: A Perspective Beyond Linearity

3 citations · 3 across the 4 of their papers we have counts for

collaborators
Showing 2022Show all

2 papers · 1 filter

cs.LG2022

How Does Adaptive Optimization Impact Local Neural Network Geometry?

Kaiqi Jiang, Dhruv Malik, Yuanzhi Li

Adaptive optimization methods are well known to achieve superior convergence relative to vanilla gradient methods. The traditional viewpoint in optimization, particularly in convex…

stat.ML2022

Complete Policy Regret Bounds for Tallying Bandits

Dhruv Malik, Yuanzhi Li, Aarti Singh

Policy regret is a well established notion of measuring the performance of an online learning algorithm against an adaptive adversary. We study restrictions on the adversary that e…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.