3 papers
cs.CV2026
PixelU: A U-Shaped Transformer for Efficient End-to-End Pixel Diffusion
Zipeng Guo, Lichen Ma, Yu He +4
End-to-end pixel-space diffusion models bypass the lossy compression of Latent Diffusion Models (LDMs) but struggle to jointly model low-frequency semantics and high-frequency sign…
cs.CV2025
RePainter: Empowering E-commerce Object Removal via Spatial-matting Reinforcement Learning
Zipeng Guo, Lichen Ma, Xiaolong Fu +13
In web data, product images are central to boosting user engagement and advertising efficacy on e-commerce platforms, yet the intrusive elements such as watermarks and promotional…
cs.RO2025
SAIL: Faster-than-Demonstration Execution of Imitation Learning Policies
Nadun Ranawaka Arachchige, Zhenyang Chen, Wonsuhk Jung +8
Offline Imitation Learning (IL) methods such as Behavior Cloning are effective at acquiring complex robotic manipulation skills. However, existing IL-trained policies are confined…