1 citations · 1 across the 4 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2026
What the Waveform Knows: Transparent-first Speech and Audio Intelligence with Caption Studio
Cheng Siong Chin, Jianhua Zhang, Mohan Venkateshkumar
Caption Studio is a transparency-first speech and audio intelligence platform that transforms spoken audio and video into structured, searchable content through automated transcrip…
cs.SD2020★ 1 cited
Non-Negative Matrix Factorization-Convolutional Neural Network (NMF-CNN) For Sound Event Detection
Teck Kai Chan, Cheng Siong Chin, Ye Li
The main scientific question of this year DCASE challenge, Task 4 - Sound Event Detection in Domestic Environments, is to investigate the types of data (strongly labeled synthetic…