30 papers
WaterGen: Decoupling Scene and Medium in Underwater Image Generation
Jiayi Wu, Tianfu Wang, Tianyi Xiong +6
Underwater computer vision tasks, such as detection, restoration, and segmentation, are limited by the scarcity of large-scale and diverse training data. We introduce WaterGen, a m…
ForceBand: Learning Forceful Manipulation with sEMG
Botao He, Zhi Wang, Linna Kuang +8
Human demonstrations are a scalable data source for learning robot manipulation policies. However, common sources of human demonstration data, such as motion-capture trajectories a…
Real2SAM2Real: Generative 3D Caches as Complementary Context for Video Diffusion
Jiayi Wu, Haoming Cai, Cornelia Fermuller +2
While Video Diffusion Models (VDMs) excel at synthesizing high-fidelity videos, enabling precise camera and scene control remains challenging. Existing methods predominantly rely o…
HumanEgo: Zero-Shot Robot Learning from Minutes of Human Egocentric Videos
Zhi Wang, Botao He, Kelin Yu +4
Human egocentric video captures rich manipulation demonstrations without any robot hardware, yet transferring these skills to robots remains challenging due to the embodiment gap b…
NeuroAI and Beyond: Bridging Between Advances in Neuroscience and ArtificialIntelligence
Anthony Zador, Jean-Marc Fellous, Terrence Sejnowski +28
Neuroscience and Artificial Intelligence (AI) have made impressive progress in recent years but remain only loosely interconnected. Based on a workshop convened by the National Sci…
From Inpainting to Layer Decomposition: Repurposing Generative Inpainting Models for Image Layer Decomposition
Jingxi Chen, Yixiao Zhang, Xiaoye Qian +4
Images can be viewed as layered compositions, foreground objects over background, with potential occlusions. This layered representation enables independent editing of elements, of…