3 papers
cs.LG2026
Retrospective In-Context Learning for Temporal Credit Assignment with Large Language Models
Wen-Tse Chen, Jiayu Chen, Fahim Tajwar +4
Learning from self-sampled data and sparse environmental feedback remains a fundamental challenge in training self-evolving agents. Temporal credit assignment mitigates this issue…
cs.LG2026
Accelerating Diffusion Planners in Offline RL via Reward-Aware Consistency Trajectory Distillation
Xintong Duan, Yutong He, Fahim Tajwar +3
Although diffusion models have achieved strong results in decision-making tasks, their slow inference speed remains a key limitation. While consistency models offer a potential sol…
cs.LG2025
State Combinatorial Generalization In Decision Making With Conditional Diffusion Models
Xintong Duan, Yutong He, Fahim Tajwar +3
Many real-world decision-making problems are combinatorial in nature, where states (e.g., surrounding traffic of a self-driving car) can be seen as a combination of basic elements…