13 citations · 37 across the 4 of their papers we have counts for
4 papers
Through the Lens of Core Competency: Survey on Evaluation of Large Language Models
Ziyu Zhuang, Qiguang Chen, Longxuan Ma +7
From pre-trained language model (PLM) to large language model (LLM), the field of natural language processing (NLP) has witnessed steep performance gains and wide practical uses. T…
U-NEED: A Fine-grained Dataset for User Needs-Centric E-commerce Conversational Recommendation
Yuanxing Liu, Weinan Zhang, Baohua Dong +8
Conversational recommender systems (CRSs) aim to understand the information needs and preferences expressed in a dialogue to recommend suitable items to the user. Most of the exist…
Second Thoughts are Best: Learning to Re-Align With Human Values from Text Edits
Ruibo Liu, Chenyan Jia, Ge Zhang +3
We present Second Thought, a new learning paradigm that enables language models (LMs) to re-align with human values. By modeling the chain-of-edits between value-unaligned and valu…
SelF-Eval: Self-supervised Fine-grained Dialogue Evaluation
Longxuan Ma, Ziyu Zhuang, Weinan Zhang +2
This paper introduces a novel Self-supervised Fine-grained Dialogue Evaluation framework (SelF-Eval). The core idea is to model the correlation between turn quality and the entire…