2 papers
cs.CV2026
Edit2Interp: Adapting Image Foundation Models from Spatial Editing to Video Frame Interpolation with Few-Shot Learning
Nasrin Rahimi, Mısra Yavuz, Burak Can Biner +6
Pre-trained image editing models exhibit strong spatial reasoning and object-aware transformation capabilities acquired from billions of image-text pairs, yet they possess no expli…
cs.CV2024
Causal Transformer for Fusion and Pose Estimation in Deep Visual Inertial Odometry
Yunus Bilge Kurt, Ahmet Akman, A. Aydın Alatan
In recent years, transformer-based architectures become the de facto standard for sequence modeling in deep learning frameworks. Inspired by the successful examples, we propose a c…