activity
20242026
collaborators

6 papers

cs.IR2026

Bringing Reasoning to Generative Recommendation Through the Lens of Cascaded Ranking

Xinyu Lin, Pengyuan Liu, Wenjie Wang +5

Generative Recommendation (GR) has become a promising end-to-end approach with high FLOPS utilization for resource-efficient recommendation. Despite the effectiveness, we show that…

cs.SD2025

DiTSinger: Scaling Singing Voice Synthesis with Diffusion Transformer and Implicit Alignment

Zongcai Du, Guilin Deng, Xiaofeng Guo +8

Recent progress in diffusion-based Singing Voice Synthesis (SVS) demonstrates strong expressiveness but remains limited by data scarcity and model scalability. We introduce a two-s…

cs.CL2025

Towards Data-efficient Customer Intent Recognition with Prompt-based Learning Paradigm

Hengyu Luo, Peng Liu, Stefan Esping

Recognizing customer intent accurately with language models based on customer-agent conversational data is essential in today's digital customer service marketplace, but it is ofte…

cs.CL2025

Yi-Lightning Technical Report

Alan Wake, Bei Chen, C. X. Lv +41

This technical report presents Yi-Lightning, our latest flagship large language model (LLM). It achieves exceptional performance, ranking 6th overall on Chatbot Arena, with particu…

cs.CL2025

Yi: Open Foundation Models by 01.AI

01. AI, :, Alex Young +30

We introduce the Yi model family, a series of language and multimodal models that demonstrate strong multi-dimensional capabilities. The Yi model family is based on 6B and 34B pret…

cs.SD2024

RFWave: Multi-band Rectified Flow for Audio Waveform Reconstruction

Peng Liu, Dongyang Dai, Zhiyong Wu

Recent advancements in generative modeling have significantly enhanced the reconstruction of audio waveforms from various representations. While diffusion models are adept at this…