477 citations · 491 across the 13 of their papers we have counts for
8 papers · 1 filter
SyncFusion: Multimodal Onset-synchronized Video-to-Audio Foley Synthesis
Marco Comunità, Riccardo F. Gramaccioni, Emilian Postolache +3
Sound design involves creatively selecting, recording, and editing sound effects for various media like cinema, video games, and virtual/augmented reality. One of the most time-con…
Hypercomplex Multimodal Emotion Recognition from EEG and Peripheral Physiological Signals
Eleonora Lopez, Eleonora Chiarantano, Eleonora Grassucci +1
Multimodal emotion recognition from physiological signals is receiving an increasing amount of attention due to the impossibility to control them at will unlike behavioral reaction…
Dual Quaternion Rotational and Translational Equivariance in 3D Rigid Motion Modelling
Guilherme Vieira, Eleonora Grassucci, Marcos Eduardo Valle +1
Objects' rigid motions in 3D space are described by rotations and translations of a highly-correlated set of points, each with associated coordinates that real-valued netwo…
PHYDI: Initializing Parameterized Hypercomplex Neural Networks as Identity Functions
Matteo Mancanelli, Eleonora Grassucci, Aurelio Uncini +1
Neural models based on hypercomplex algebra systems are growing and prolificating for a plethora of applications, ranging from computer vision to natural language processing. Hand…
Diffusion models for audio semantic communication
Eleonora Grassucci, Christian Marinoni, Andrea Rodriguez +1
Directly sending audio signals from a transmitter to a receiver across a noisy channel may absorb consistent bandwidth and be prone to errors when trying to recover the transmitted…
Enhancing Semantic Communication with Deep Generative Models -- An ICASSP Special Session Overview
Eleonora Grassucci, Yuki Mitsufuji, Ping Zhang +1
Semantic communication is poised to play a pivotal role in shaping the landscape of future AI-driven communication systems. Its challenge of extracting semantic information from th…