From the 1 of 506 papers with an AI index.
29.3k citations
- Max Planck Institute for the Science of LightDE14 papers
- University of ArizonaUS14 papers
- Beijing Institute of TechnologyCN11 papers
- Complexity and Topology in Quantum MatterDE11 papers
- Ruhr University BochumDE11 papers
- TU Dortmund UniversityDE11 papers
- Leipzig UniversityDE10 papers
- University of WürzburgDE9 papers
- Centre National de la Recherche ScientifiqueFR8 papers
- Technical University of MunichDE8 papers
- Institute of PhysicsPL7 papers
- Aarhus UniversityDK6 papers
4 papers · 2 filters
Combining TF-GridNet and Mixture Encoder for Continuous Speech Separation for Meeting Transcription
Peter Vieting, Simon Berger, Thilo von Neumann +3
Many real-life applications of automatic speech recognition (ASR) require processing of overlapped speech. A common method involves first separating the speech into overlap-free st…
Post-Processing Independent Evaluation of Sound Event Detection Systems
Janek Ebbers, Reinhold Haeb-Umbach, Romain Serizel
Due to the high variation in the application requirements of sound event detection (SED) systems, it is not sufficient to evaluate systems only in a single operating mode. Therefor…
A Teacher-Student approach for extracting informative speaker embeddings from speech mixtures
Tobias Cord-Landwehr, Christoph Boeddeker, Cătălin Zorilă +2
We introduce a monaural neural speaker embeddings extractor that computes an embedding for each speaker present in a speech mixture. To allow for supervised training, a teacher-stu…
TS-SEP: Joint Diarization and Separation Conditioned on Estimated Speaker Embeddings
Christoph Boeddeker, Aswin Shanmugam Subramanian, Gordon Wichern +2
Since diarization and source separation of meeting data are closely related tasks, we here propose an approach to perform the two objectives jointly. It builds upon the target-spea…