◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

M. Ghavamzadeh

50 papers hereh-index 5413.6k citations187 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author35
  • last author10

Across the 47 of 50 papers where every author was matched, so the position is known.

fields
  • cs.LG39
  • cs.AI4
  • stat.ML3
  • cs.CL1
  • cs.IR1
  • eess.SY1

identity via Semantic Scholar / OpenAlex

activity
20152026
most citedBenchmarking Batch Deep Reinforcement Learning Algorithms

159 citations · 619 across the 30 of their papers we have counts for

collaborators
Showing stat.MLShow all

3 papers · 1 filter

stat.ML2020

Active Model Estimation in Markov Decision Processes

Jean Tarbouriech, Shubhanshu Shekhar, Matteo Pirotta +2

We study the problem of efficient exploration in order to learn an accurate model of an environment, modeled as a Markov decision process (MDP). Efficient exploration in this probl…

stat.ML2019

Adaptive Sampling for Estimating Multiple Probability Distributions

Shubhanshu Shekhar, Tara Javidi, Mohammad Ghavamzadeh

We consider the problem of allocating samples to a finite set of discrete distributions in order to learn them uniformly well in terms of four common distance measures: ℓ22​,…

stat.ML2016★ 67 cited

Safe Policy Improvement by Minimizing Robust Baseline Regret

Marek Petrik, Yinlam Chow, Mohammad Ghavamzadeh

An important problem in sequential decision-making under uncertainty is to use limited data to compute a safe policy, i.e., a policy that is guaranteed to perform at least as well…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.