47 citations · 88 across the 13 of their papers we have counts for
17 papers
DiffDet4SAR: Diffusion-based Aircraft Target Detection Network for SAR Images
Zhou Jie, Xiao Chao, Peng Bo +4
Aircraft target detection in SAR images is a challenging task due to the discrete scattering points and severe background clutter interference. Currently, methods with convolution-…
Enhancing Information Maximization with Distance-Aware Contrastive Learning for Source-Free Cross-Domain Few-Shot Learning
Huali Xu, Li Liu, Shuaifeng Zhi +4
Existing Cross-Domain Few-Shot Learning (CDFSL) methods require access to source domain data to train a model in the pre-training phase. However, due to increasing concerns about d…
Lightweight Pixel Difference Networks for Efficient Visual Representation Learning
Zhuo Su, Jiehua Zhang, Longguang Wang +4
Recently, there have been tremendous efforts in developing lightweight Deep Neural Networks (DNNs) with satisfactory accuracy, which can enable the ubiquitous deployment of DNNs in…
Realistic Speech-to-Face Generation with Speech-Conditioned Latent Diffusion Model with Face Prior
Jinting Wang, Li Liu, Jun Wang +1
Speech-to-face generation is an intriguing area of research that focuses on generating realistic facial images based on a speaker's audio speech. However, state-of-the-art methods…
Emotional Talking Head Generation based on Memory-Sharing and Attention-Augmented Networks
Jianrong Wang, Yaxin Zhao, Li Liu +3
Given an audio clip and a reference face image, the goal of the talking head generation is to generate a high-fidelity talking head video. Although some audio-driven methods of gen…
A Comprehensive Survey on Segment Anything Model for Vision and Beyond
Chunhui Zhang, Li Liu, Yawen Cui +4
Artificial intelligence (AI) is evolving towards artificial general intelligence, which refers to the ability of an AI system to perform a wide range of tasks and exhibit a level o…