2 papers
cs.LG2026
Bilevel Optimization over Saddle Points of Zero-Sum Markov Games
Zihao Zheng, Irwin King, Songtao Lu
Reinforcement learning (RL) often has a hierarchical structure, where an upper-level (UL) learner selects model parameters and a lower-level (LL) decision-making process responds,…
cs.IR2026
Tencent Advertising Algorithm Challenge 2025: All-Modality Generative Recommendation
Junwei Pan, Wei Xue, Chao Zhou +20
Generative recommender systems are rapidly emerging as a new paradigm for recommendation, where collaborative identifiers and/or multi-modal content are mapped into discrete token…