2 papers
cs.CV2024
Two-in-One: Unified Multi-Person Interactive Motion Generation by Latent Diffusion Transformer
Boyuan Li, Xihua Wang, Ruihua Song +1
Multi-person interactive motion generation, a critical yet under-explored domain in computer character animation, poses significant challenges such as intricate modeling of inter-h…
cs.CL2024
SpeechComposer: Unifying Multiple Speech Tasks with Prompt Composition
Yihan Wu, Soumi Maiti, Yifan Peng +6
Recent advancements in language models have significantly enhanced performance in multiple speech-related tasks. Existing speech language models typically utilize task-dependent pr…