activity
20242026
collaborators

5 papers

cs.CV2026

Pusa V1.0: Unlocking Temporal Control in Pretrained Video Diffusion Models via Vectorized Timestep Adaptation

Yaofang Liu, Yumeng Ren, Aitor Artola +9

The rapid advancement of video diffusion models has been hindered by fundamental limitations in temporal modeling, particularly the rigid synchronization of frame evolution imposed…

cs.CV2025

Blind Adaptive Local Denoising for CEST Imaging

Chu Chen, Aitor Artola, Yang Liu +4

Chemical Exchange Saturation Transfer (CEST) MRI enables molecular-level visualization of low-concentration metabolites by leveraging proton exchange dynamics. However, its clinica…

cs.CV2025

Improving OCR using internal document redundancy

Diego Belzarena, Seginus Mowlavi, Aitor Artola +9

Current OCR systems are based on deep learning models trained on large amounts of data. Although they have shown some ability to generalize to unseen data, especially in detection…

cs.CV2025

Improving Diffusion Generative Models via Truncated Karhunen--Loève Expansion

Yumeng Ren, Yaofang Liu, Aitor Artola +3

Pretrained diffusion models exhibit a well-known training-sampling mismatch, often attributed to exposure bias and related distribution-shift effects. We provide a quantitative int…

cs.CV2024

Redefining Temporal Modeling in Video Diffusion: The Vectorized Timestep Approach

Yaofang Liu, Yumeng Ren, Xiaodong Cun +5

Diffusion models have revolutionized image generation, and their extension to video generation has shown promise. However, current video diffusion models~(VDMs) rely on a scalar ti…