2 papers
cs.CV2026
Modality-Balanced Collaborative Distillation for Multi-Modal Domain Generalization
Xiaohan Wang, Zhangtao Cheng, Ting Zhong +2
Weight Averaging (WA) has emerged as a powerful technique for enhancing generalization by promoting convergence to a flat loss landscape, which correlates with stronger out-of-dist…
cs.IR2025
Reward Balancing Revisited: Enhancing Offline Reinforcement Learning for Recommender Systems
Wenzheng Shu, Yanxiang Zeng, Yongxiang Tang +6
Offline reinforcement learning (RL) has emerged as a prevalent and effective methodology for real-world recommender systems, enabling learning policies from historical data and cap…