429 citations · 639 across the 30 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2023
Zero-Shot Audio Captioning via Audibility Guidance
Tal Shaharabany, Ariel Shaulov, Lior Wolf
The task of audio captioning is similar in essence to tasks such as image and video captioning. However, it has received much less attention. We propose three desiderata for captio…
cs.SD2023★ 9 cited
AudioToken: Adaptation of Text-Conditioned Diffusion Models for Audio-to-Image Generation
Guy Yariv, Itai Gat, Lior Wolf +2
In recent years, image generation has shown a great leap in performance, where diffusion models play a central role. Although generating high-quality images, such models are mainly…