1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CL2024★ 1 cited
Fast Prompt Alignment for Text-to-Image Generation
Khalil Mrini, Hanlin Lu, Linjie Yang +2
Text-to-image generation has advanced rapidly, yet aligning complex textual prompts with generated visuals remains challenging, especially with intricate object relationships and f…
cs.MM2024
VMAS: Video-to-Music Generation via Semantic Alignment in Web Music Videos
Yan-Bo Lin, Yu Tian, Linjie Yang +2
We present a framework for learning to generate background music from video inputs. Unlike existing works that rely on symbolic musical annotations, which are limited in quantity a…