4 papers · 1 filter
On the Stability of Prompt Ranking in Large Language Model Evaluation
Shaoshuai Du, Penghao Liang, Yixian Shen +3
Prompt-based interaction has become a dominant paradigm for using large language models (LLMs), where multiple candidate prompts are evaluated and the top-ranked one is selected fo…
Zero-Shot End-to-End Relation Extraction in Chinese: A Comparative Study of Gemini, LLaMA and ChatGPT
Shaoshuai Du, Yiyi Tao, Yixian Shen +4
This study investigates the performance of various large language models (LLMs) on zero-shot end-to-end relation extraction (RE) in Chinese, a task that integrates entity recogniti…
Comparative Analysis of Listwise Reranking with Large Language Models in Limited-Resource Language Contexts
Yanxin Shen, Lun Wang, Chuanqi Shi +4
Large Language Models (LLMs) have demonstrated significant effectiveness across various NLP tasks, including text ranking. This study assesses the performance of large language mod…
Robustness of Large Language Models Against Adversarial Attacks
Yiyi Tao, Yixian Shen, Hang Zhang +4
The increasing deployment of Large Language Models (LLMs) in various applications necessitates a rigorous evaluation of their robustness against adversarial attacks. In this paper,…