162 citations · 581 across the 59 of their papers we have counts for
Showing cs.LGShow all
3 papers · 1 filter
cs.LG2021
Equivalence Analysis between Counterfactual Regret Minimization and Online Mirror Descent
Weiming Liu, Huacong Jiang, Bin Li +1
Follow-the-Regularized-Lead (FTRL) and Online Mirror Descent (OMD) are regret minimization algorithms for Online Convex Optimization (OCO), they are mathematically elegant but less…
cs.LG2021★ 3 cited
Revisiting Knowledge Distillation: An Inheritance and Exploration Framework
Zhen Huang, Xu Shen, Jun Xing +6
Knowledge Distillation (KD) is a popular technique to transfer knowledge from a teacher model or ensemble to a student model. Its success is generally attributed to the privileged…
cs.LG2020
Masked Contrastive Representation Learning for Reinforcement Learning
Jinhua Zhu, Yingce Xia, Lijun Wu +4
Improving sample efficiency is a key research problem in reinforcement learning (RL), and CURL, which uses contrastive learning to extract high-level features from raw pixels of in…