3 papers
eess.AS2025
Audio-Visual Speech Enhancement for Spatial Audio - Spatial-VisualVoice and the MAVE Database
Danielle Yaffe, Ferdinand Campe, Prachi Sharma +2
Audio-visual speech enhancement (AVSE) has been found to be particularly useful at low signal-to-noise (SNR) ratios due to the immunity of the visual features to acoustic noise. Ho…
cs.CL2025
How desirable is alignment between LLMs and linguistically diverse human users?
Pia Knoeferle, Sebastian Möller, Dorothea Kolossa +2
We discuss how desirable it is that Large Language Models (LLMs) be able to adapt or align their language behavior with users who may be diverse in their language use. User diversi…
cs.CV2025
Extending Information Bottleneck Attribution to Video Sequences
Veronika Solopova, Lucas Schmidt, Dorothea Kolossa
We introduce VIBA, a novel approach for explainable video classification by adapting Information Bottlenecks for Attribution (IBA) to video sequences. While most traditional explai…