2 papers
cs.IR2025
Counterfactual Multi-player Bandits for Explainable Recommendation Diversification
Yansen Zhang, Bowei He, Xiaokun Zhang +3
Existing recommender systems tend to prioritize items closely aligned with users' historical interactions, inevitably trapping users in the dilemma of ``filter bubble''. Recent eff…
cs.CL2025
A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?
Qiyuan Zhang, Fuyuan Lyu, Zexu Sun +10
As enthusiasm for scaling computation (data and parameters) in the pretraining era gradually diminished, test-time scaling (TTS), also referred to as ``test-time computing'' has em…