22 citations · 27 across the 2 of their papers we have counts for
3 papers
ORL: Reinforcement Learning Benchmarks for Online Stochastic Optimization Problems
Bharathan Balaji, Jordan Bell-Masterson, Enes Bilgin +7
Reinforcement Learning (RL) has achieved state-of-the-art results in domains such as robotics and games. We build on this previous work by applying RL algorithms to a selection of…
Imitation-Regularized Offline Learning
Yifei Ma, Yu-Xiang Wang, Balakrishnan +1
We study the problem of offline learning in automated decision systems under the contextual bandits model. We are given logged historical data consisting of contexts, (randomized)…
Achieving Fluency and Coherency in Task-oriented Dialog
Rashmi Gangadharaiah, Balakrishnan Narayanaswamy, Charles Elkan
We consider real world task-oriented dialog settings, where agents need to generate both fluent natural language responses and correct external actions like database queries and up…