1 paper · 1 filter
Shreeram Suresh Chandra, Lucas Goncalves, Junchen Lu +2
Current emotion-based contrastive language-audio pretraining (CLAP) methods typically learn by naïvely aligning audio samples with corresponding text prompts. Consequently, this a…