2 papers
cs.SD2026
TMD-Bench: A Multi-Level Evaluation Paradigm for Music-Dance Co-Generation
Xiaoda Yang, Majun Zhang, Changhao Pan +10
Unified audio-visual generation is rapidly gaining industrial and creative relevance, enabling applications in virtual production and interactive media. However, when moving from g…
cs.CV2025
Disentangling 3D from Large Vision-Language Models for Controlled Portrait Generation
Nick Yiwen Huang, Akin Caliskan, Berkay Kicanaoglu +2
We consider the problem of disentangling 3D from large vision-language models, which we show on generative 3D portraits. This allows free-form text control of appearance attributes…