320 citations · 2.7k across the 93 of their papers we have counts for
19 papers · 1 filter
Improving Unsupervised Domain Adaptation with Variational Information Bottleneck
Yuxuan Song, Lantao Yu, Zhangjie Cao +5
Domain adaptation aims to leverage the supervision signal of source domain to obtain an accurate model for target domain, where the labels are not available. To leverage and adapt…
Sequential Recommendation with Dual Side Neighbor-based Collaborative Relation Modeling
Jiarui Qin, Kan Ren, Yuchen Fang +2
Sequential recommendation task aims to predict user preference over items in the future given user historical behaviors. The order of user behaviors implies that there are resource…
Multi-Agent Reinforcement Learning for Order-dispatching via Order-Vehicle Distribution Matching
Ming Zhou, Jiarui Jin, Weinan Zhang +6
Improving the efficiency of dispatching orders to vehicles is a research hotspot in online ride-hailing systems. Most of the existing solutions for order-dispatching are centralize…
Signal Instructed Coordination in Cooperative Multi-agent Reinforcement Learning
Liheng Chen, Hongyi Guo, Yali Du +7
In many real-world problems, a team of agents need to collaborate to maximize the common reward. Although existing works formulate this problem into a centralized learning with dec…
Bi-level Actor-Critic for Multi-agent Coordination
Haifeng Zhang, Weizhe Chen, Zeren Huang +4
Coordination is one of the essential problems in multi-agent systems. Typically multi-agent reinforcement learning (MARL) methods treat agents equally and the goal is to solve the…
Learning to Advertise for Organic Traffic Maximization in E-Commerce Product Feeds
Dagui Chen, Junqi Jin, Weinan Zhang +7
Most e-commerce product feeds provide blended results of advertised products and recommended products to consumers. The underlying advertising and recommendation platforms share si…