3 papers · 1 filter
THAI Speech Emotion Recognition (THAI-SER) corpus
Jilamika Wongpithayadisai, Chompakorn Chaksangchaichot, Soravitt Sangnark +7
We present the first sizeable corpus of Thai speech emotion recognition, THAI-SER, containing 41 hours and 36 minutes (27,854 utterances) from 100 recordings made in different reco…
Amplifying Artifacts with Speech Enhancement in Voice Anti-spoofing
Thanapat Trachu, Thanathai Lertpetchpun, Ekapol Chuangsuwanich
Spoofed utterances always contain artifacts introduced by generative models. While several countermeasures have been proposed to detect spoofed utterances, most primarily focus on…
Thunder : Unified Regression-Diffusion Speech Enhancement with a Single Reverse Step using Brownian Bridge
Thanapat Trachu, Chawan Piansaddhayanon, Ekapol Chuangsuwanich
Diffusion-based speech enhancement has shown promising results, but can suffer from a slower inference time. Initializing the diffusion process with the enhanced audio generated by…