1 paper · 1 filter
Akam Rahimi, Triantafyllos Afouras, Andrew Zisserman
We present a transformer-based architecture for voice separation of a target speaker from multiple other speakers and ambient noise. We achieve this by using two separate neural ne…