Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Personalizing Embodied Multimodal Large Language Model Agents over Long-term User Interactions
Jeongeun Lee, Chanyoung Park, Dongha Lee
Multimodal large language model (MLLM)-based embodied agents have shown strong potential for solving complex tasks in physical environments. However, personalized assistance requir…
cs.AI2026
PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization
Wonjoong Kim, Yeonjun In, Sangwu Park +2
A significant hurdle for current LLMs is the execution of complex, multi-stage tasks. Group Relative Policy Optimization (GRPO) has been emerging as a leading choice, but its relia…