379 citations · 452 across the 33 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2024
Tango 2: Aligning Diffusion-based Text-to-Audio Generations through Direct Preference Optimization
Navonil Majumder, Chia-Yu Hung, Deepanway Ghosal +3
Generative multimodal content is increasingly prevalent in much of the content creation arena, as it has the potential to allow artists and media personnel to create pre-production…
cs.SD2019
MuSE-ing on the Impact of Utterance Ordering On Crowdsourced Emotion Annotations
Mimansa Jaiswal, Zakaria Aldeneh, Cristian-Paul Bara +4
Emotion recognition algorithms rely on data annotated with high quality labels. However, emotion expression and perception are inherently subjective. There is generally not a singl…