320 citations · 736 across the 17 of their papers we have counts for
5 papers · 1 filter
U-rank: Utility-oriented Learning to Rank with Implicit Feedback
Xinyi Dai, Jiawei Hou, Qing Liu +6
Learning to rank with implicit feedback is one of the most important tasks in many real-world information systems where the objective is some specific utility, e.g., clicks and rev…
Learning to Infer User Hidden States for Online Sequential Advertising
Zhaoqing Peng, Junqi Jin, Lan Luo +11
To drive purchase in online advertising, it is of the advertiser's great interest to optimize the sequential advertising strategy whose performance and interpretability are both im…
Off-Policy Reinforcement Learning for Efficient and Effective GAN Architecture Search
Yuan Tian, Qin Wang, Zhiwu Huang +5
In this paper, we introduce a new reinforcement learning (RL) based neural architecture search (NAS) methodology for effective and efficient generative adversarial network (GAN) ar…
A Deep Recurrent Survival Model for Unbiased Ranking
Jiarui Jin, Yuchen Fang, Weinan Zhang +7
Position bias is a critical problem in information retrieval when dealing with implicit yet biased user feedback data. Unbiased ranking methods typically rely on causality models a…
Multi-Agent Interactions Modeling with Correlated Policies
Minghuan Liu, Ming Zhou, Weinan Zhang +4
In multi-agent systems, complex interacting behaviors arise due to the high correlations among agents. However, previous work on modeling multi-agent interactions from demonstratio…