3 papers
cs.CV2025
DAE-Talker: High Fidelity Speech-Driven Talking Face Generation with Diffusion Autoencoder
Chenpeng Du, Qi Chen, Tianyu He +5
While recent research has made significant progress in speech-driven talking face generation, the quality of the generated video still lags behind that of real recordings. One reas…
cs.CV2024
IRGen: Generative Modeling for Image Retrieval
Yidan Zhang, Ting Zhang, Dong Chen +11
While generative modeling has become prevalent across numerous research fields, its integration into the realm of image retrieval remains largely unexplored and underjustified. In…
cs.CV2024
Towards Lightweight Super-Resolution with Dual Regression Learning
Yong Guo, Mingkui Tan, Zeshuai Deng +5
Deep neural networks have exhibited remarkable performance in image super-resolution (SR) tasks by learning a mapping from low-resolution (LR) images to high-resolution (HR) images…