2 papers
cs.CV2025
Comp-Attn: Present-and-Align Attention for Compositional Video Generation
Hongyu Zhang, Yufan Deng, Shenghai Yuan +5
In the domain of text-to-video (T2V) generation, reliably synthesizing compositional content involving multiple subjects with intricate relations is still underexplored. The main c…
cs.CV2025
UI2V-Bench: An Understanding-based Image-to-video Generation Benchmark
Ailing Zhang, Lina Lei, Dehong Kong +7
Generative diffusion models are developing rapidly and attracting increasing attention due to their wide range of applications. Image-to-Video (I2V) generation has become a major f…