3 papers
cs.AI2026
PAFO: Pareto Fairness Optimization for Personalized Reward Modeling
Xiaoyan Zhao, Haoting Ni, Yang Zhang +3
Large language models (LLMs) increasingly rely on reward models to align their outputs with diverse user preferences. While personalized reward models aim to capture such heterogen…
cs.AI2026
Experience Transfer for Multimodal LLM Agents in Minecraft Game
Chenghao Li, Jun Liu, Songbo Zhang +7
Multimodal LLM agents operating in complex game environments must continually reuse past experience to solve new tasks efficiently. In this work, we propose Echo, a transfer-orient…
cs.CL2025
SteerX: Disentangled Steering for LLM Personalization
Xiaoyan Zhao, Ming Yan, Yilun Qiu +5
Large language models (LLMs) have shown remarkable success in recent years, enabling a wide range of applications, including intelligent assistants that support users' daily life a…