10 papers
FadeMem: Distance-Aware Memory Consolidation for Autoregressive Video Diffusion
Yu Lu, Junjie Yang, Piotr Koniusz +2
Autoregressive video generators synthesize long videos by generating successive temporal segments, but their historical KV cache grows with video length. Existing bounded-cache met…
Video Understanding by Design: How Datasets Shape Video Models
Lei Wang, Syuan-Hao Li, Piotr Koniusz +1
Research in video understanding has advanced rapidly, driven by increasingly diverse datasets and more powerful model architectures. While existing surveys typically organize progr…
Uncertainty-DTW for Sequences and Visual Tokens
Lei Wang, Syuan-Hao Li, Yongsheng Gao +1
Aligning structured data is a fundamental problem in computer vision and machine learning, underlying tasks such as time series analysis, human action recognition, and visual repre…
Subspace Kernel Learning on Tensor Sequences
Lei Wang, Xi Ding, Yongsheng Gao +1
Learning from structured multi-way data, represented as higher-order tensors, requires capturing complex interactions across tensor modes while remaining computationally efficient.…
Learning Time in Static Classifiers
Xi Ding, Lei Wang, Piotr Koniusz +1
Real-world visual data rarely presents as isolated, static instances. Instead, it often evolves gradually over time through variations in pose, lighting, object state, or scene con…
Graph Your Own Prompt
Xi Ding, Lei Wang, Piotr Koniusz +1
We propose Graph Consistency Regularization (GCR), a novel framework that injects relational graph structures, derived from model predictions, into the learning process to promote…