1 paper
Hebeizi Li, Zihao Liang, Benyuan Sun +4
While state-of-the-art audio-video generation models like Veo3 and Sora2 demonstrate remarkable capabilities, their closed-source nature makes their architectures and training para…