Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Reconstruction-Shift Discrimination via Mask-Guided Latent Diffusion for Medical Anomaly Detection
Yibo Wan, Jinyu Cai, Yunhe Zhang +2
Unsupervised medical anomaly detection learns normal anatomical patterns from healthy training images and identifies deviations at test time. Reconstruction-based and diffusion-bas…
cs.CV2025
Audio-Driven Talking Face Video Generation with Joint Uncertainty Learning
Yifan Xie, Fei Ma, Yi Bin +2
Talking face video generation with arbitrary speech audio is a significant challenge within the realm of digital human technology. The previous studies have emphasized the signific…
cs.CV2024
Exploring Deeper! Segment Anything Model with Depth Perception for Camouflaged Object Detection
Zhenni Yu, Xiaoqin Zhang, Li Zhao +2
This paper introduces a new Segment Anything Model with Depth Perception (DSAM) for Camouflaged Object Detection (COD). DSAM exploits the zero-shot capability of SAM to realize pre…