3 papers
cs.CV2025
Towards Interactive Intelligence for Digital Humans
Yiyi Cai, Xuangeng Chu, Xiwei Gao +16
We introduce Interactive Intelligence, a novel paradigm of digital human that is capable of personality-aligned expression, adaptive interaction, and self-evolution. To realize thi…
cs.CV2025
FloodDiffusion: Tailored Diffusion Forcing for Streaming Motion Generation
Yiyi Cai, Yuhan Wu, Kunhang Li +3
We present FloodDiffusion, a new framework for text-driven, streaming human motion generation. Given time-varying text prompts, FloodDiffusion generates text-aligned, seamless moti…
eess.AS2025
Shallow Flow Matching for Coarse-to-Fine Text-to-Speech Synthesis
Dong Yang, Yiyi Cai, Yuki Saito +2
We propose Shallow Flow Matching (SFM), a novel mechanism that enhances flow matching (FM)-based text-to-speech (TTS) models within a coarse-to-fine generation paradigm. Unlike con…