collaborators

5 papers

cs.CV2025

ATL-Diff: Audio-Driven Talking Head Generation with Early Landmarks-Guide Noise Diffusion

Hoang-Son Vo, Quang-Vinh Nguyen, Seungwon Kim +3

Audio-driven talking head generation requires precise synchronization between facial animations and audio signals. This paper introduces ATL-Diff, a novel approach addressing synch…

cs.CV2025

Anatomical Attention Alignment representation for Radiology Report Generation

Quang Vinh Nguyen, Minh Duc Nguyen, Thanh Hoang Son Vo +2

Automated Radiology report generation (RRG) aims at producing detailed descriptions of medical images, reducing radiologists' workload and improving access to high-quality diagnost…

cs.CV2024

Polyp-SES: Automatic Polyp Segmentation with Self-Enriched Semantic Model

Quang Vinh Nguyen, Thanh Hoang Son Vo, Sae-Ryung Kang +1

Automatic polyp segmentation is crucial for effective diagnosis and treatment in colonoscopy images. Traditional methods encounter significant challenges in accurately delineating…

cs.CV2024

KAN-Based Fusion of Dual-Domain for Audio-Driven Facial Landmarks Generation

Hoang-Son Vo-Thanh, Quang-Vinh Nguyen, Soo-Hyung Kim

Audio-driven talking face generation is a widely researched topic due to its high applicability. Reconstructing a talking face using audio significantly contributes to fields such…

cs.CV2024

Adaptation of Distinct Semantics for Uncertain Areas in Polyp Segmentation

Quang Vinh Nguyen, Van Thong Huynh, Soo-Hyung Kim

Colonoscopy is a common and practical method for detecting and treating polyps. Segmenting polyps from colonoscopy image is useful for diagnosis and surgery progress. Nevertheless,…