3 papers
cs.RO2026
TFP: Temporally Conditioned Memory-Fusion Policies for Visuomotor Learning
Yushen Liang, Yue Peng, Baosheng Jin +6
Vision--Language--Action (VLA) policies such as and OpenVLA perform well on many manipulation tasks, but they are often reactive: the next action is predicted from the c…
cs.CV2026
Histogram-constrained Image Generation
Haoming Liu, Yuanhe Guo, Yijia Cao +2
Diffusion models have emerged as a dominant paradigm in generative modeling, enabling high-fidelity sampling from complex data distributions. Despite impressive capabilities, contr…
cs.LG2026
From Navigation to Refinement: Revealing the Two-Stage Nature of Flow-based Diffusion Models through Oracle Velocity
Haoming Liu, Jinnuo Liu, Yanhao Li +5
Flow-based diffusion models have emerged as a leading paradigm for training generative models across images and videos. However, their memorization-generalization behavior remains…