3 papers
cs.AI2026
VISA: Value Injection via Shielded Adaptation for Personalized LLM Alignment
Jiawei Chen, Tianzhuo Yang, Guoxi Zhang +3
Aligning Large Language Models (LLMs) with nuanced human values remains a critical challenge, as existing methods like Reinforcement Learning from Human Feedback (RLHF) often handl…
cs.LG2025
Goal Discovery with Causal Capacity for Efficient Reinforcement Learning
Yan Yu, Yaodong Yang, Zhengbo Lu +3
Causal inference is crucial for humans to explore the world, which can be modeled to enable an agent to efficiently explore the environment in reinforcement learning. Existing rese…
cs.AI2025
Model Evolution Framework with Genetic Algorithm for Multi-Task Reinforcement Learning
Yan Yu, Wengang Zhou, Yaodong Yang +3
Multi-task reinforcement learning employs a single policy to complete various tasks, aiming to develop an agent with generalizability across different scenarios. Given the shared c…