5 citations · 10 across the 2 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2019★ 5 cited
Audio-Linguistic Embeddings for Spoken Sentences
Albert Haque, Michelle Guo, Prateek Verma +1
We propose spoken sentence embeddings which capture both acoustic and linguistic content. While existing works operate at the character, phoneme, or word level, our method learns l…
cs.SD2018
Automatic Documentation of ICD Codes with Far-Field Speech Recognition
Albert Haque, Corinna Fukushima
Documentation errors increase healthcare costs and cause unnecessary patient deaths. As the standard language for diagnoses and billing, ICD codes serve as the foundation for medic…
cs.SD2018
Conditional End-to-End Audio Transforms
Albert Haque, Michelle Guo, Prateek Verma
We present an end-to-end method for transforming audio from one style to another. For the case of speech, by conditioning on speaker identities, we can train a single model to tran…