42 citations · 46 across the 3 of their papers we have counts for
4 papers
MOORe: Model-based Offline-to-Online Reinforcement Learning
Yihuan Mao, Chao Wang, Bin Wang +1
With the success of offline reinforcement learning (RL), offline trained RL policies have the potential to be further improved when deployed online. A smooth transfer of the policy…
Towards robust and domain agnostic reinforcement learning competitions
William Hebgen Guss, Stephanie Milani, Nicholay Topin +26
Reinforcement learning competitions have formed the basis for standard research benchmarks, galvanized advances in the state-of-the-art, and shaped the direction of the field. Desp…
LadaBERT: Lightweight Adaptation of BERT through Hybrid Model Compression
Yihuan Mao, Yujing Wang, Chufan Wu +6
BERT is a cutting-edge language representation model pre-trained by a large corpus, which achieves superior performances on various natural language understanding tasks. However, a…
CrowdPose: Efficient Crowded Scenes Pose Estimation and A New Benchmark
Jiefeng Li, Can Wang, Hao Zhu +3
Multi-person pose estimation is fundamental to many computer vision tasks and has made significant progress in recent years. However, few previous methods explored the problem of p…