18 citations · 36 across the 10 of their papers we have counts for
4 papers · 2 filters
Neural Rate Control for Video Encoding using Imitation Learning
Hongzi Mao, Chenjie Gu, Miaosen Wang +9
In modern video encoders, rate control is a critical component and has been heavily engineered. It decides how many bits to spend to encode each frame, in order to optimize the rat…
A maximum-entropy approach to off-policy evaluation in average-reward MDPs
Nevena Lazic, Dong Yin, Mehrdad Farajtabar +4
This work focuses on off-policy evaluation (OPE) with function approximation in infinite-horizon undiscounted Markov decision processes (MDPs). For MDPs that are ergodic and linear…
Robotic Table Tennis with Model-Free Reinforcement Learning
Wenbo Gao, Laura Graesser, Krzysztof Choromanski +5
We propose a model-free algorithm for learning efficient policies capable of returning table tennis balls by controlling robot joints at a rate of 100Hz. We demonstrate that evolut…
Adaptive Approximate Policy Iteration
Botao Hao, Nevena Lazic, Yasin Abbasi-Yadkori +2
Model-free reinforcement learning algorithms combined with value function approximation have recently achieved impressive performance in a variety of application domains. However,…