most citedTransformer with Leveraged Masked Autoencoder for video-based Pain Assessment

3 citations · 5 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CV2025

ATL-Diff: Audio-Driven Talking Head Generation with Early Landmarks-Guide Noise Diffusion

Hoang-Son Vo, Quang-Vinh Nguyen, Seungwon Kim +3

Audio-driven talking head generation requires precise synchronization between facial animations and audio signals. This paper introduces ATL-Diff, a novel approach addressing synch…

cs.LG2025★ 1 cited

Latent Behavior Diffusion for Sequential Reaction Generation in Dyadic Setting

Minh-Duc Nguyen, Hyung-Jeong Yang, Soo-Hyung Kim +2

The dyadic reaction generation task involves synthesizing responsive facial reactions that align closely with the behaviors of a conversational partner, enhancing the naturalness a…

cs.CV2025

Anatomical Attention Alignment representation for Radiology Report Generation

Quang Vinh Nguyen, Minh Duc Nguyen, Thanh Hoang Son Vo +2

Automated Radiology report generation (RRG) aims at producing detailed descriptions of medical images, reducing radiologists' workload and improving access to high-quality diagnost…

cs.CV2024★ 3 cited

Transformer with Leveraged Masked Autoencoder for video-based Pain Assessment

Minh-Duc Nguyen, Hyung-Jeong Yang, Soo-Hyung Kim +2

Accurate pain assessment is crucial in healthcare for effective diagnosis and treatment; however, traditional methods relying on self-reporting are inadequate for populations unabl…

cs.CV2024★ 1 cited

Leveraging WaveNet for Dynamic Listening Head Modeling from Speech

Minh-Duc Nguyen, Hyung-Jeong Yang, Seung-Won Kim +2

The creation of listener facial responses aims to simulate interactive communication feedback from a listener during a face-to-face conversation. Our goal is to generate believable…