3 citations · 3 across the 2 of their papers we have counts for
4 papers
BRPO: Batch Residual Policy Optimization
Sungryull Sohn, Yinlam Chow, Jayden Ooi +4
In batch reinforcement learning (RL), one often constrains a learned policy to be close to the behavior (data-generating) policy, e.g., by constraining the learned action distribut…
Enumeration of Distinct Support Vectors for Interactive Decision Making
Kentaro Kanamori, Satoshi Hara, Masakazu Ishihata +1
In conventional prediction tasks, a machine learning algorithm outputs a single best model that globally optimizes its objective function, which typically is accuracy. Therefore, u…
An Efficient Algorithm for Enumerating Chordal Bipartite Induced Subgraphs in Sparse Graphs
Kazuhiro Kurita, Kunihiro Wasa, Hiroki Arimura +1
In this paper, we propose a characterization of chordal bipartite graphs and an efficient enumeration algorithm for chordal bipartite induced subgraphs. A chordal bipartite graph i…
On the Model Shrinkage Effect of Gamma Process Edge Partition Models
Iku Ohama, Issei Sato, Takuya Kida +1
The edge partition model (EPM) is a fundamental Bayesian nonparametric model for extracting an overlapping structure from binary matrix. The EPM adopts a gamma process (P) prior…