1 citations · 2 across the 7 of their papers we have counts for
1 paper · 2 filters
Trung X. Pham, Tri Ton, Chang D. Yoo
We introduce MDSGen, a novel framework for vision-guided open-domain sound generation optimized for model parameter size, memory consumption, and inference speed. This framework in…