◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Csaba Szepesvári

22 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author10
  • last author12

Across the 22 of 22 papers where every author was matched, so the position is known.

fields
  • cs.LG15
  • stat.ML4
  • cs.AI3
ORCID 0000-0002-9286-2892

identity via Semantic Scholar / OpenAlex

activity
20102024
most citedApprenticeship Learning using Inverse Reinforcement Learning and Gradient Methods

157 citations · 349 across the 22 of their papers we have counts for

collaborators
Showing cs.AIShow all

3 papers · 1 filter

cs.AI2014★ 13 cited

On Minimax Optimal Offline Policy Evaluation

Lihong Li, Remi Munos, Csaba Szepesvari

This paper studies the off-policy evaluation problem, where one aims to estimate the value of a target policy based on a sample of observations collected by another policy. We firs…

cs.AI2014★ 9 cited

Adaptive Monte Carlo via Bandit Allocation

James Neufeld, András György, Dale Schuurmans +1

We consider the problem of sequentially choosing between a set of unbiased Monte Carlo estimators to minimize the mean-squared-error (MSE) of a final combined estimate. By reducing…

cs.AI2012★ 3 cited

Speeding Up Planning in Markov Decision Processes via Automatically Constructed Abstractions

Alejandro Isaza, Csaba Szepesvari, Vadim Bulitko +1

In this paper, we consider planning in stochastic shortest path (SSP) problems, a subclass of Markov Decision Problems (MDP). We focus on medium-size problems whose state space can…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.