3 papers
cs.IR2026
TokenMinds: Pretrained User Tokens and Embeddings for User Understanding in Large Recommender Systems
Qingyun Liu, Bo Yan, Yang Liu +15
User modeling in industrial recommender systems typically produces dense embeddings, which suffer from representational constraints inherent to fixed-dimensional vectors. An emergi…
cs.LG2026
Continual Policy Consolidation for Lifelong Robot Learning
Qijun He, Yuxuan Li, Mingqi Yuan +4
Building a generalist robot policy requires continuously integrating new skills while preserving previously acquired behaviors. Directly optimizing a single policy over a growing t…
cs.LG2024
ME-IGM: Individual-Global-Max in Maximum Entropy Multi-Agent Reinforcement Learning
Wen-Tse Chen, Yuxuan Li, Shiyu Huang +2
Multi-agent credit assignment is a fundamental challenge for cooperative multi-agent reinforcement learning (MARL), where a team of agents learn from shared reward signals. The Ind…