1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.LG2025
MELCOT: A Hybrid Learning Architecture with Marginal Preservation for Matrix-Valued Regression
Khang Tran, Hieu Cao, Thinh Pham +3
Regression is essential across many domains but remains challenging in high-dimensional settings, where existing methods often lose spatial structure or demand heavy storage. In th…
eess.AS2024
Towards Unsupervised Speaker Diarization System for Multilingual Telephone Calls Using Pre-trained Whisper Model and Mixture of Sparse Autoencoders
Phat Lam, Lam Pham, Truong Nguyen +5
Existing speaker diarization systems typically rely on large amounts of manually annotated data, which is labor-intensive and difficult to obtain, especially in real-world scenario…
cs.CL2023★ 1 cited
XPhoneBERT: A Pre-trained Multilingual Model for Phoneme Representations for Text-to-Speech
Linh The Nguyen, Thinh Pham, Dat Quoc Nguyen
We present XPhoneBERT, the first multilingual model pre-trained to learn phoneme representations for the downstream text-to-speech (TTS) task. Our XPhoneBERT has the same model arc…