3 papers
cs.CV2026
Thinking with Patterns: Breaking the Perceptual Bottleneck in Visual Planning via Pattern Induction
Yichang Jian, Boyuan Xiao, Zhenyuan Huang +2
Planning from raw visual input remains a significant challenge for current Vision-Language Models (VLMs), when the complexity of input is beyond their one-step perception capabilit…
cs.LG2024
Beyond Simple Sum of Delayed Rewards: Non-Markovian Reward Modeling for Reinforcement Learning
Yuting Tang, Xin-Qiang Cai, Jing-Cheng Pang +3
Reinforcement Learning (RL) empowers agents to acquire various skills by learning from reward signals. Unfortunately, designing high-quality instance-level rewards often demands si…
cs.LG2024
Reinforcement Learning from Bagged Reward
Yuting Tang, Xin-Qiang Cai, Yao-Xiang Ding +3
In Reinforcement Learning (RL), it is commonly assumed that an immediate reward signal is generated for each action taken by the agent, helping the agent maximize cumulative reward…