50 citations · 53 across the 2 of their papers we have counts for
2 papers
cs.SD2023★ 50 cited
Noise2Music: Text-conditioned Music Generation with Diffusion Models
Qingqing Huang, Daniel S. Park, Tao Wang +12
We introduce Noise2Music, where a series of diffusion models is trained to generate high-quality 30-second music clips from text prompts. Two types of diffusion models, a generator…
cs.CL2023★ 3 cited
MAQA: A Multimodal QA Benchmark for Negation
Judith Yue Li, Aren Jansen, Qingqing Huang +3
Multimodal learning can benefit from the representation power of pretrained Large Language Models (LLMs). However, state-of-the-art transformer based LLMs often ignore negations in…