2 papers
cs.CV2026
Edit2Interp: Adapting Image Foundation Models from Spatial Editing to Video Frame Interpolation with Few-Shot Learning
Nasrin Rahimi, Mısra Yavuz, Burak Can Biner +6
Pre-trained image editing models exhibit strong spatial reasoning and object-aware transformation capabilities acquired from billions of image-text pairs, yet they possess no expli…
cs.CV2024
GAN Based Top-Down View Synthesis in Reinforcement Learning Environments
Usama Younus, Vinoj Jayasundara, Shivam Mishra +1
Human actions are based on the mental perception of the environment. Even when all the aspects of an environment are not visible, humans have an internal mental model that can gene…