most citedBeamformer-Guided Target Speaker Extraction

2 citations · 3 across the 9 of their papers we have counts for

collaborators

9 papers

eess.AS2023

Simulating room transfer functions between transducers mounted on audio devices using a modified image source method

Zeyu Xu, Adrian Herzog, Alexander Lodermeyer +2

The image source method (ISM) is often used to simulate room acoustics due to its ease of use and computational efficiency. The standard ISM is limited to simulations of room impul…

eess.AS2023

Data-driven 3D Room Geometry Inference with a Linear Loudspeaker Array and a Single Microphone

Cagdas Tuna, Altan Akat, H. Nazim Bicer +2

Knowing the room geometry may be very beneficial for many audio applications, including sound reproduction, acoustic scene analysis, and sound source localization. Room geometry in…

eess.AS2023

Predicting Preferred Dialogue-to-Background Loudness Difference in Dialogue-Separated Audio

Luca Resti, Martin Strauss, Matteo Torcoli +2

Dialogue Enhancement (DE) enables the rebalancing of dialogue and background sounds to fit personal preferences and needs in the context of broadcast audio. When individual audio s…

eess.AS2023

Better Together: Dialogue Separation and Voice Activity Detection for Audio Personalization in TV

Matteo Torcoli, Emanuël A. P. Habets

In TV services, dialogue level personalization is key to meeting user preferences and needs. When dialogue and background sounds are not separately available from the production st…

eess.AS20232 cited

Beamformer-Guided Target Speaker Extraction

Mohamed Elminshawi, Srikanth Raj Chetupalli, Emanuël A. P. Habets

We propose a Beamformer-guided Target Speaker Extraction (BG-TSE) method to extract a target speaker's voice from a multi-channel recording informed by the direction of arrival of…

eess.AS2023

Multi-Microphone Speaker Separation by Spatial Regions

Julian Wechsler, Srikanth Raj Chetupalli, Wolfgang Mack +1

We consider the task of region-based source separation of reverberant multi-microphone recordings. We assume pre-defined spatial regions with a single active source per region. The…