5 papers
ATL-Diff: Audio-Driven Talking Head Generation with Early Landmarks-Guide Noise Diffusion
Hoang-Son Vo, Quang-Vinh Nguyen, Seungwon Kim +3
Audio-driven talking head generation requires precise synchronization between facial animations and audio signals. This paper introduces ATL-Diff, a novel approach addressing synch…
Anatomical Attention Alignment representation for Radiology Report Generation
Quang Vinh Nguyen, Minh Duc Nguyen, Thanh Hoang Son Vo +2
Automated Radiology report generation (RRG) aims at producing detailed descriptions of medical images, reducing radiologists' workload and improving access to high-quality diagnost…
Polyp-SES: Automatic Polyp Segmentation with Self-Enriched Semantic Model
Quang Vinh Nguyen, Thanh Hoang Son Vo, Sae-Ryung Kang +1
Automatic polyp segmentation is crucial for effective diagnosis and treatment in colonoscopy images. Traditional methods encounter significant challenges in accurately delineating…
KAN-Based Fusion of Dual-Domain for Audio-Driven Facial Landmarks Generation
Hoang-Son Vo-Thanh, Quang-Vinh Nguyen, Soo-Hyung Kim
Audio-driven talking face generation is a widely researched topic due to its high applicability. Reconstructing a talking face using audio significantly contributes to fields such…
Adaptation of Distinct Semantics for Uncertain Areas in Polyp Segmentation
Quang Vinh Nguyen, Van Thong Huynh, Soo-Hyung Kim
Colonoscopy is a common and practical method for detecting and treating polyps. Segmenting polyps from colonoscopy image is useful for diagnosis and surgery progress. Nevertheless,…