◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Pierre Ménard

7 papers hereh-index 14822 citations29 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author4
  • last author3

Across the 7 of 7 papers where every author was matched, so the position is known.

fields
  • stat.ML4
  • cs.LG3
same name
  • Pierre Ménard — 3 papers, h 4

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedA single algorithm for both restless and rested rotting bandits

6 citations · 10 across the 5 of their papers we have counts for

collaborators
Showing stat.MLShow all

4 papers · 1 filter

stat.ML2026★ 6 cited

A single algorithm for both restless and rested rotting bandits

Julien Seznec, Pierre Ménard, Alessandro Lazaric +1

In many application domains (e.g., recommender systems, intelligent tutoring systems), the rewards associated to the actions tend to decrease over time. This decay is either caused…

stat.ML2026

Proximal Point Nash Learning from Human Feedback

Daniil Tiapkin, Daniele Calandriello, Denis Belomestny +5

Traditional Reinforcement Learning from Human Feedback (RLHF) often relies on reward models, frequently assuming preference structures like the Bradley--Terry model, which may not…

stat.ML2025

Model-free Posterior Sampling via Learning Rate Randomization

Daniil Tiapkin, Denis Belomestny, Daniele Calandriello +6

In this paper, we introduce Randomized Q-learning (RandQL), a novel randomized model-free algorithm for regret minimization in episodic Markov Decision Processes (MDPs). To the bes…

stat.ML2024

Demonstration-Regularized RL

Daniil Tiapkin, Denis Belomestny, Daniele Calandriello +5

Incorporating expert demonstrations has empirically helped to improve the sample efficiency of reinforcement learning (RL). This paper quantifies theoretically to what extent this…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.