34 citations · 94 across the 14 of their papers we have counts for
7 papers · 1 filter
Reinforcement Learning Enhanced LLMs: A Survey
Shuhe Wang, Shengyu Zhang, Jie Zhang +7
Reinforcement learning (RL) enhanced large language models (LLMs), particularly exemplified by DeepSeek-R1, have exhibited outstanding performance. Despite the effectiveness in imp…
Sim-GPT: Text Similarity via GPT Annotated Data
Shuhe Wang, Beiming Cao, Shengyu Zhang +5
Due to the lack of a large collection of high-quality labeled sentence pairs with textual similarity scores, existing approaches for Semantic Textual Similarity (STS) mostly rely o…
Sentiment Analysis through LLM Negotiations
Xiaofei Sun, Xiaoya Li, Shengyu Zhang +5
A standard paradigm for sentiment analysis is to rely on a singular LLM and makes the decision in a single round under the framework of in-context learning. This framework suffers…
Interpreting Deep Learning Models in Natural Language Processing: A Review
Xiaofei Sun, Diyi Yang, Xiaoya Li +6
Neural network models have achieved state-of-the-art performances in a wide range of natural language processing (NLP) tasks. However, a long-standing criticism against neural netw…
OpenViDial 2.0: A Larger-Scale, Open-Domain Dialogue Generation Dataset with Visual Contexts
Shuhe Wang, Yuxian Meng, Xiaoya Li +3
In order to better simulate the real human conversation process, models need to generate dialogue utterances based on not only preceding textual contexts but also visual contexts.…
Folden: -Fold Ensemble for Out-Of-Distribution Detection
Xiaoya Li, Jiwei Li, Xiaofei Sun +5
Out-of-Distribution (OOD) detection is an important problem in natural language processing (NLP). In this work, we propose a simple yet effective framework Folden, which mimics…