56 citations · 97 across the 12 of their papers we have counts for
12 papers
DetDiffusion: Synergizing Generative and Perceptive Models for Enhanced Data Generation and Perception
Yibo Wang, Ruiyuan Gao, Kai Chen +8
Current perceptive models heavily depend on resource-intensive datasets, prompting the need for innovative solutions. Leveraging recent advances in diffusion models, synthetic data…
A Neural-network Enhanced Video Coding Framework beyond ECM
Yanchen Zhao, Wenxuan He, Chuanmin Jia +7
In this paper, a hybrid video compression framework is proposed that serves as a demonstrative showcase of deep learning-based approaches extending beyond the confines of tradition…
LKFormer: Large Kernel Transformer for Infrared Image Super-Resolution
Feiwei Qin, Kang Yan, Changmiao Wang +3
Given the broad application of infrared technology across diverse fields, there is an increasing emphasis on investigating super-resolution techniques for infrared images within th…
DMV3D: Denoising Multi-View Diffusion using 3D Large Reconstruction Model
Yinghao Xu, Hao Tan, Fujun Luan +8
We propose \textbf{DMV3D}, a novel 3D generation approach that uses a transformer-based 3D large reconstruction model to denoise multi-view diffusion. Our reconstruction model inco…
Survey on Deep Face Restoration: From Non-blind to Blind and Beyond
Wenjie Li, Mei Wang, Kai Zhang +6
Face restoration (FR) is a specialized field within image restoration that aims to recover low-quality (LQ) face images into high-quality (HQ) face images. Recent advances in deep…
Designs and Implementations in Neural Network-based Video Coding
Yue Li, Junru Li, Chaoyi Lin +9
The past decade has witnessed the huge success of deep learning in well-known artificial intelligence applications such as face recognition, autonomous driving, and large language…