1 citations · 2 across the 13 of their papers we have counts for
Showing 2026 · cs.LGShow all
2 papers · 2 filters
cs.LG2026
Adaptive Bandit Algorithms for Contextual Matching Markets
Shiyun Lin, Simon Mauras, Vianney Perchet +1
We study bandit learning in matching markets, where players and arms constitute the two market sides, and the players' utilities are linear in the arm contexts. In each round, new…
cs.LG2026
Reinforcement Learning with Multi-Step Lookahead Information Via Adaptive Batching
Nadav Merlis
We study tabular reinforcement learning problems with multiple steps of lookahead information. Before acting, the learner observes steps of future transition and reward real…