4 papers · 1 filter
AAAR-1.0: Assessing AI's Potential to Assist Research
Renze Lou, Hanzi Xu, Sijia Wang +15
Numerous studies have assessed the proficiency of AI systems, particularly large language models (LLMs), in facilitating everyday tasks such as email writing, question answering, a…
LLMs' Classification Performance is Overclaimed
Hanzi Xu, Renze Lou, Jiangshu Du +6
In many classification tasks designed for AI or human to solve, gold labels are typically included within the label space by default, often posed as "which of the following is corr…
MUFFIN: Curating Multi-Faceted Instructions for Improving Instruction-Following
Renze Lou, Kai Zhang, Jian Xie +5
In the realm of large language models (LLMs), enhancing instruction-following capability often involves curating expansive training data. This is achieved through two primary schem…
X-Shot: A Unified System to Handle Frequent, Few-shot and Zero-shot Learning Simultaneously in Classification
Hanzi Xu, Muhao Chen, Lifu Huang +2
In recent years, few-shot and zero-shot learning, which learn to predict labels with limited annotated instances, have garnered significant attention. Traditional approaches often…