35 citations · 35 across the 1 of their papers we have counts for
3 papers
cs.CV2024
Image Conductor: Precision Control for Interactive Video Synthesis
Yaowei Li, Xintao Wang, Zhaoyang Zhang +5
Filmmaking and animation production often require sophisticated techniques for coordinating camera transitions and object movements, typically involving labor-intensive real-world…
cs.CV2023★ 33 cited
T2I-Adapter: Learning Adapters to Dig out More Controllable Ability for Text-to-Image Diffusion Models
Chong Mou, Xintao Wang, Liangbin Xie +5
The incredible generative ability of large-scale text-to-image (T2I) models has demonstrated strong power of learning complex structures and meaningful semantics. However, relying…
cs.CV2022★ 35 cited
Rethinking Alignment in Video Super-Resolution Transformers
Shuwei Shi, Jinjin Gu, Liangbin Xie +3
The alignment of adjacent frames is considered an essential operation in video super-resolution (VSR). Advanced VSR models, including the latest VSR Transformers, are generally equ…