4 papers
Acoustic scattering AI for non-invasive object classifications: A case study on hair assessment
Long-Vu Hoang, Tuan Nguyen, Tran Huy Dat
This paper presents a novel non-invasive object classification approach using acoustic scattering, demonstrated through a case study on hair assessment. When an incident wave inter…
Fine-Grained Frame Modeling in Multi-head Self-Attention for Speech Deepfake Detection
Tuan Dat Phuong, Duc-Tuan Truong, Long-Vu Hoang +1
Transformer-based models have shown strong performance in speech deepfake detection, largely due to the effectiveness of the multi-head self-attention (MHSA) mechanism. MHSA provid…
Qwen vs. Gemma Integration with Whisper: A Comparative Study in Multilingual SpeechLLM Systems
Tuan Nguyen, Long-Vu Hoang, Huy-Dat Tran
This paper presents our system for the MLC-SLM Challenge 2025, focusing on multilingual speech recognition and language modeling with large language models (LLMs). Our approach com…
Pushing the Performance of Synthetic Speech Detection with Kolmogorov-Arnold Networks and Self-Supervised Learning Models
Tuan Dat Phuong, Long-Vu Hoang, Huy Dat Tran
Recent advancements in speech synthesis technologies have led to increasingly advanced spoofing attacks, posing significant challenges for automatic speaker verification systems. W…