11 citations · 11 across the 1 of their papers we have counts for
3 papers
A Generic Multi-modal Dynamic Gesture Recognition System using Machine Learning
Gautham Krishna G, Karthik Subramanian Nathan, Yogesh Kumar B +3
Human computer interaction facilitates intelligent communication between humans and computers, in which gesture recognition plays a prominent role. This paper proposes a machine le…
Reducing Bias in Production Speech Models
Eric Battenberg, Rewon Child, Adam Coates +13
Replacing hand-engineered pipelines with end-to-end deep learning systems has enabled strong results in applications like speech and object recognition. However, the causality and…
Deep Speaker: an End-to-End Neural Speaker Embedding System
Chao Li, Xiaokong Ma, Bing Jiang +6
We present Deep Speaker, a neural speaker embedding system that maps utterances to a hypersphere where speaker similarity is measured by cosine similarity. The embeddings generated…