activity
20192025
most citedRank the Episodes: A Simple Approach for Exploration in Procedurally-Generated Environments

5 citations · 5 across the 3 of their papers we have counts for

collaborators

5 papers

cs.LG2021

PASTO: Strategic Parameter Optimization in Recommendation Systems -- Probabilistic is Better than Deterministic

Weicong Ding, Hanlin Tang, Jingshuo Feng +17

Real-world recommendation systems often consist of two phases. In the first phase, multiple predictive models produce the probability of different immediate user actions. In the se…

cs.LG20215 cited

Rank the Episodes: A Simple Approach for Exploration in Procedurally-Generated Environments

Daochen Zha, Wenye Ma, Lei Yuan +2

Exploration under sparse reward is a long-standing challenge of model-free reinforcement learning. The state-of-the-art methods address this challenge by introducing intrinsic rewa…

cs.MM2020

Short Video-based Advertisements Evaluation System: Self-Organizing Learning Approach

Yunjie Zhang, Fei Tao, Xudong Liu +6

With the rising of short video apps, such as TikTok, Snapchat and Kwai, advertisement in short-term user-generated videos (UGVs) has become a trending form of advertising. Predicti…

cs.AI2020

Themes Informed Audio-visual Correspondence Learning

Runze Su, Fei Tao, Xudong Liu +6

The applications of short-term user-generated video (UGV), such as Snapchat, and Youtube short-term videos, booms recently, raising lots of multimodal machine learning tasks. Among…

cs.DC2019

: Decentralization Meets Error-Compensated Compression

Hanlin Tang, Xiangru Lian, Shuang Qiu +4

Communication is a key bottleneck in distributed training. Recently, an \emph{error-compensated} compression technology was particularly designed for the \emph{centralized} learnin…