244 citations · 320 across the 17 of their papers we have counts for
4 papers · 1 filter
Coincidence, Categorization, and Consolidation: Learning to Recognize Sounds with Minimal Supervision
Aren Jansen, Daniel P. W. Ellis, Shawn Hershey +4
Humans do not acquire perceptual abilities in the way we train machines. While machine learning algorithms typically operate on large collections of randomly-chosen, explicitly-lab…
Differentiable Consistency Constraints for Improved Deep Speech Enhancement
Scott Wisdom, John R. Hershey, Kevin Wilson +4
In recent years, deep networks have led to dramatic improvements in speech enhancement by framing it as a data-driven pattern recognition problem. In many modern enhancement system…
Exploring Tradeoffs in Models for Low-latency Speech Enhancement
Kevin Wilson, Michael Chinen, Jeremy Thorpe +5
We explore a variety of neural networks configurations for one- and two-channel spectrogram-mask-based speech enhancement. Our best model improves on previous state-of-the-art perf…
Unsupervised Learning of Semantic Audio Representations
Aren Jansen, Manoj Plakal, Ratheet Pandya +5
Even in the absence of any explicit semantic annotation, vast collections of audio recordings provide valuable information for learning the categorical structure of sounds. We cons…