2 citations · 2 across the 4 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
ReactDiff: Latent Diffusion for Facial Reaction Generation
Jiaming Li, Sheng Wang, Xin Wang +4
Given the audio-visual clip of the speaker, facial reaction generation aims to predict the listener's facial reactions. The challenge lies in capturing the relevance between video…
cs.CV2024
Two-stage Cytopathological Image Synthesis for Augmenting Cervical Abnormality Screening
Zhenrong Shen, Manman Fei, Xin Wang +4
Automatic thin-prep cytologic test (TCT) screening can assist pathologists in finding cervical abnormality towards accurate and efficient cervical cancer diagnosis. Current automat…
cs.CV2023
MeLo: Low-rank Adaptation is Better than Fine-tuning for Medical Image Diagnosis
Yitao Zhu, Zhenrong Shen, Zihao Zhao +5
The common practice in developing computer-aided diagnosis (CAD) models based on transformer architectures usually involves fine-tuning from ImageNet pre-trained weights. However,…