5 papers
Evaluating and Rewarding LALMs for Expressive Role-Play TTS via Mean Continuation Log-Probability
Yong Ren, Jingbei Li, Haiyang Sun +6
Recent advances in Large Audio Language Models (LALMs) have extended Text-to-Speech (TTS) to interactive role-play scenarios, which demand high expressiveness and strict adherence…
Enhancing Conversational Recommender Systems with Tree-Structured Knowledge and Pretrained Language Models
Yongwen Ren, Chao Wang, Peng Du +3
Recent advances in pretrained language models (PLMs) have significantly improved conversational recommender systems (CRS), enabling more fluent and context-aware interactions. To f…
Step-Audio 2 Technical Report
Boyong Wu, Chao Yan, Chen Hu +106
This paper presents Step-Audio 2, an end-to-end multi-modal large language model designed for industry-strength audio understanding and speech conversation. By integrating a latent…
Evaluating Position Bias in Large Language Model Recommendations
Ethan Bito, Yongli Ren, Estrid He
Large Language Models (LLMs) are being increasingly explored as general-purpose tools for recommendation tasks, enabling zero-shot and instruction-following capabilities without th…
Evaluating Large Language Models on Financial Report Summarization: An Empirical Study
Xinqi Yang, Scott Zang, Yong Ren +2
In recent years, Large Language Models (LLMs) have demonstrated remarkable versatility across various applications, including natural language understanding, domain-specific knowle…