4 citations · 8 across the 2 of their papers we have counts for
2 papers
cs.AI2021★ 4 cited
PoBRL: Optimizing Multi-Document Summarization by Blending Reinforcement Learning Policies
Andy Su, Difei Su, John M. Mulvey +1
We propose a novel reinforcement learning based framework PoBRL for solving multi-document summarization. PoBRL jointly optimizes over the following three objectives necessary for…
cs.LG2021
MUSBO: Model-based Uncertainty Regularized and Sample Efficient Batch Optimization for Deployment Constrained Reinforcement Learning
DiJia Su, Jason D. Lee, John M. Mulvey +1
In many contemporary applications such as healthcare, finance, robotics, and recommendation systems, continuous deployment of new policies for data collection and online learning i…