2 citations · 2 across the 2 of their papers we have counts for
3 papers
cs.CL2026
FireRedAudio: A General-Purpose Audio Language Model with Decoupled Continuous Representations for Understanding and Generation
Feiyu Shen, Fenglong Xie, Junjie Li +13
A unified audio model must recognize and understand linguistic, paralinguistic, and environmental information while supporting speech synthesis and editing. A key challenge is repr…
cs.SD2025★ 2 cited
FireRedChat: A Pluggable, Full-Duplex Voice Interaction System with Cascaded and Semi-Cascaded Implementations
Junjie Chen, Yao Hu, Junjie Li +12
Full-duplex voice interaction allows users and agents to speak simultaneously with controllable barge-in, enabling lifelike assistants and customer service. Existing solutions are…
cs.SD2024
Melody-Guided Music Generation
Shaopeng Wei, Manzhen Wei, Haoyu Wang +2
We present the Melody-Guided Music Generation (MG2) model, a novel approach using melody to guide the text-to-music generation that, despite a simple method and limited resources,…