◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Miroslav Dudík

3 papers hereh-index 3536.1k citations79 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author1

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • q-bio.QM1

identity via Semantic Scholar / OpenAlex

most citedDoubly Robust Policy Evaluation and Learning

299 citations · 478 across the 3 of their papers we have counts for

collaborators
Showing cs.LGShow all

2 papers · 1 filter

cs.LG2011★ 119 cited

Efficient Optimal Learning for Contextual Bandits

Miroslav Dudik, Daniel Hsu, Satyen Kale +4

We address the problem of learning in an online setting where the learner repeatedly observes features, selects among a set of actions, and receives reward for the action taken. We…

cs.LG2011★ 299 cited

Doubly Robust Policy Evaluation and Learning

Miroslav Dudik, John Langford, Lihong Li

We study decision making in environments where the reward is only partially observed, but can be modeled as a function of an action and an observed context. This setting, known as…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.