activity
20242026
collaborators

7 papers

cs.CV2026

Edit2Interp: Adapting Image Foundation Models from Spatial Editing to Video Frame Interpolation with Few-Shot Learning

Nasrin Rahimi, Mısra Yavuz, Burak Can Biner +6

Pre-trained image editing models exhibit strong spatial reasoning and object-aware transformation capabilities acquired from billions of image-text pairs, yet they possess no expli…

cs.CV2026

Track-On2: Enhancing Online Point Tracking with Memory

Görkay Aydemir, Weidi Xie, Fatma Güney

In this paper, we consider the problem of long-term point tracking, which requires consistent identification of points across video frames under significant appearance changes, mot…

cs.CV2026

Real-World Point Tracking with Verifier-Guided Pseudo-Labeling

Görkay Aydemir, Fatma Güney, Weidi Xie

Models for long-term point tracking are typically trained on large synthetic datasets. The performance of these models degrades in real-world videos due to different characteristic…

cs.CV2025

Online Long-term Point Tracking in the Foundation Model Era

Görkay Aydemir

Point tracking aims to identify the same physical point across video frames and serves as a geometry-aware representation of motion. This representation supports a wide range of ap…

cs.CV2025

Track-On: Transformer-based Online Point Tracking with Memory

Görkay Aydemir, Xiongyi Cai, Weidi Xie +1

In this paper, we consider the problem of long-term point tracking, which requires consistent identification of points across multiple frames in a video, despite changes in appeara…

cs.CV2024

Robust Bird's Eye View Segmentation by Adapting DINOv2

Merve Rabia Barın, Görkay Aydemir, Fatma Güney

Extracting a Bird's Eye View (BEV) representation from multiple camera images offers a cost-effective, scalable alternative to LIDAR-based solutions in autonomous driving. However,…