2 papers
cs.CV2026
CrossFlowDG: Bridging the Modality Gap with Cross-modal Flow Matching for Domain Generalization
Antonios Kritikos, Nikolaos Spanos, Athanasios Voulodimos
Domain generalization (DG) aims to maintain performance under domain shift, which in computer vision appears primarily as stylistic variations that cause models to overfit to domai…
cs.CL2025
MEDUSA: A Multimodal Deep Fusion Multi-Stage Training Framework for Speech Emotion Recognition in Naturalistic Conditions
Georgios Chatzichristodoulou, Despoina Kosmopoulou, Antonios Kritikos +5
SER is a challenging task due to the subjective nature of human emotions and their uneven representation under naturalistic conditions. We propose MEDUSA, a multimodal framework wi…