2 papers
cs.LG2026
On the effectiveness of reward functions in reinforcement learning for confidence calibration of large language models
Chee Heng Tan, Zhuoyi Lin, Mehul Motani +1
In this paper, we consider the setting where large language models (LLMs) are trained using reinforcement learning (RL) to simultaneously improve reasoning accuracy and verbalize i…
cs.IR2025
Do Reviews Matter for Recommendations in the Era of Large Language Models?
Chee Heng Tan, Huiying Zheng, Jing Wang +5
With the advent of large language models (LLMs), the landscape of recommender systems is undergoing a significant transformation. Traditionally, user reviews have served as a criti…