activity
20242026
collaborators

10 papers

cs.CV2026

Learn Once, Edit Anywhere: Visual Direction Transfer for Diffusion Models

Yusuf Dalva, Hidir Yesiltepe, Pinar Yanardag

The rapid advancement of diffusion models has enabled the generation of high-fidelity images from textual prompts, yet achieving precise, disentangled control over specific attribu…

cs.CV2026

DTG-Restore: Training-Free Diffusion Refinement for Generative Video Super-Resolution

Hidir Yesiltepe, Koutilya PNVR, Gaurav Pathak +4

Recent progress in video diffusion models has enabled remarkable generative fidelity, yet leveraging these priors for restoration remains limited by the strong coupling between con…

cs.CV2026

VideoMLA: Low-Rank Latent KV Cache for Minute-Scale Autoregressive Video Diffusion

Hidir Yesiltepe, Jiazhen Hu, Tuna Han Salih Meral +4

Long-rollout causal video diffusion has converged on a fixed-size sliding-window KV cache, with recent progress innovating within this layout by changing which tokens occupy the wi…

cs.CV2026

Aligning Latent Geometry for Spherical Flow Matching in Image Generation

Tuna Han Salih Meral, Kaan Oktay, Hidir Yesiltepe +2

Latent flow matching for image generation usually transports Gaussian noise to variational autoencoder latents along linear paths. Both endpoints, however, concentrate in thin sphe…

cs.CV2026

Infinity-RoPE: Action-Controllable Infinite Video Generation Emerges From Autoregressive Self-Rollout

Hidir Yesiltepe, Tuna Han Salih Meral, Adil Kaan Akan +2

Current autoregressive video diffusion models are constrained by three core bottlenecks: (i) the finite temporal horizon imposed by the base model's 3D Rotary Positional Embedding…

cs.CV2025

Dynamic View Synthesis as an Inverse Problem

Hidir Yesiltepe, Pinar Yanardag

In this work, we address dynamic view synthesis from monocular videos as an inverse problem in a training-free setting. By redesigning the noise initialization phase of a pre-train…