◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Csaba Szepesvári

23 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author10
  • last author13

Across the 23 of 23 papers where every author was matched, so the position is known.

fields
  • cs.LG16
  • stat.ML4
  • cs.AI3
ORCID 0000-0002-9286-2892
same name
  • Csaba Szepesvári — 13 papers, h 9

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20102024
most citedApprenticeship Learning using Inverse Reinforcement Learning and Gradient Methods

157 citations · 352 across the 23 of their papers we have counts for

collaborators
Showing 2014Show all

4 papers · 1 filter

cs.AI2014★ 13 cited

On Minimax Optimal Offline Policy Evaluation

Lihong Li, Remi Munos, Csaba Szepesvari

This paper studies the off-policy evaluation problem, where one aims to estimate the value of a target policy based on a sample of observations collected by another policy. We firs…

cs.LG2014★ 6 cited

Bayesian Optimal Control of Smoothly Parameterized Systems: The Lazy Posterior Sampling Algorithm

Yasin Abbasi-Yadkori, Csaba Szepesvari

We study Bayesian optimal control of a general class of smoothly parameterized Markov decision problems. Since computing the optimal control is computationally expensive, we design…

cs.LG2014★ 11 cited

Optimal Resource Allocation with Semi-Bandit Feedback

Tor Lattimore, Koby Crammer, Csaba Szepesvári

We study a sequential resource allocation problem involving a fixed number of recurring jobs. At each time-step the manager should distribute available resources among the jobs in…

cs.AI2014★ 9 cited

Adaptive Monte Carlo via Bandit Allocation

James Neufeld, András György, Dale Schuurmans +1

We consider the problem of sequentially choosing between a set of unbiased Monte Carlo estimators to minimize the mean-squared-error (MSE) of a final combined estimate. By reducing…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.