1 citations · 1 across the 2 of their papers we have counts for
3 papers
Hyperbolic Audio Source Separation
Darius Petermann, Gordon Wichern, Aswin Subramanian +1
We introduce a framework for audio source separation using embeddings on a hyperbolic manifold that compactly represent the hierarchical relationship between sound sources and time…
(2.5+1)D Spatio-Temporal Scene Graphs for Video Question Answering
Anoop Cherian, Chiori Hori, Tim K. Marks +1
Spatio-temporal scene-graph approaches to video-based reasoning tasks, such as video question-answering (QA), typically construct such graphs for every video frame. These approache…
Phasebook and Friends: Leveraging Discrete Representations for Source Separation
Jonathan Le Roux, Gordon Wichern, Shinji Watanabe +2
Deep learning based speech enhancement and source separation systems have recently reached unprecedented levels of quality, to the point that performance is reaching a new ceiling.…