5 papers
Towards Efficient SDRTV-to-HDRTV by Learning from Image Formation
Xiangyu Chen, Zheyuan Li, Zhengwen Zhang +6
Modern displays can render video content with high dynamic range (HDR) and wide color gamut (WCG). However, most resources are still in standard dynamic range (SDR). Therefore, tra…
VEnhancer: Generative Space-Time Enhancement for Video Generation
Jingwen He, Tianfan Xue, Dongyang Liu +6
We present VEnhancer, a generative space-time enhancement framework that improves the existing text-to-video results by adding more details in spatial domain and synthetic detailed…
Lumina-T2X: Transforming Text into Any Modality, Resolution, and Duration via Flow-based Large Diffusion Transformers
Peng Gao, Le Zhuo, Dongyang Liu +17
Sora unveils the potential of scaling Diffusion Transformer for generating photorealistic images and videos at arbitrary resolutions, aspect ratios, and durations, yet it still lac…
Towards Real-world Video Face Restoration: A New Benchmark
Ziyan Chen, Jingwen He, Xinqi Lin +2
Blind face restoration (BFR) on images has significantly progressed over the last several years, while real-world video face restoration (VFR), which is more challenging for more c…
DiffBIR: Towards Blind Image Restoration with Generative Diffusion Prior
Xinqi Lin, Jingwen He, Ziyan Chen +6
We present DiffBIR, a general restoration pipeline that could handle different blind image restoration tasks in a unified framework. DiffBIR decouples blind image restoration probl…