3 citations · 3 across the 3 of their papers we have counts for
3 papers · 1 filter
Computing the Collection of Good Models for Rule Lists
Kota Mata, Kentaro Kanamori, Hiroki Arimura
Since the seminal paper by Breiman in 2001, who pointed out a potential harm of prediction multiplicities from the view of explainable AI, global analysis of a collection of all go…
BRPO: Batch Residual Policy Optimization
Sungryull Sohn, Yinlam Chow, Jayden Ooi +4
In batch reinforcement learning (RL), one often constrains a learned policy to be close to the behavior (data-generating) policy, e.g., by constraining the learned action distribut…
Enumeration of Distinct Support Vectors for Interactive Decision Making
Kentaro Kanamori, Satoshi Hara, Masakazu Ishihata +1
In conventional prediction tasks, a machine learning algorithm outputs a single best model that globally optimizes its objective function, which typically is accuracy. Therefore, u…