Publications (105)
Gradient-Adaptive Policy Optimization: Towards Multi-Objective Alignment of Large Language Models
Chengao Li, Hanyu Zhang, Yunkun Xu +3
Reinforcement Learning from Human Feedback (RLHF) has emerged as a powerful technique for aligning large language models (LLMs) with human preferences. However, effectively alignin…
Personalized Prompt for Sequential Recommendation
Yiqing Wu, Ruobing Xie, Yongchun Zhu +4
Pre-training models have shown their power in sequential recommendation. Recently, prompt has been widely explored and verified for tuning in NLP pre-training, which could help to…
Interactive Text-to-Speech System via Joint Style Analysis
Yang Gao, Weiyi Zheng, Zhaojun Yang +3
While modern TTS technologies have made significant advancements in audio quality, there is still a lack of behavior naturalness compared to conversing with people. We propose a st…
Atom Responding Machine for Dialog Generation
Ganbin Zhou, Ping Luo, Jingwu Chen +3
Recently, improving the relevance and diversity of dialogue system has attracted wide attention. For a post x, the corresponding response y is usually diverse in the real-world cor…
Deep Semantic Inference over the Air: An Efficient Task-Oriented Communication System
Chenyang Wang, Roger Olsson, Stefan Forsström +1
Empowered by deep learning, semantic communication marks a paradigm shift from transmitting raw data to conveying task-relevant meaning, enabling more efficient and intelligent wir…
Tree-Structured Neural Machine for Linguistics-Aware Sentence Generation
Ganbin Zhou, Ping Luo, Rongyu Cao +4
Different from other sequential data, sentences in natural language are structured by linguistic grammars. Previous generative conversational models with chain-structured decoder i…