8 papers
NICE: A Theory-Grounded Diagnostic Benchmark for Social Intelligence of LLMs
Yunjin Qi, Zhaojun Jiang, Xuan Wu +10
As large language models (LLMs) are increasingly applied in social contexts such as emotional companionship and customer service, measuring their social intelligence has become cri…
MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Tail Knowledge
Jie He, Nan Hu, Wanqiu Long +2
Large language models (LLMs) have demonstrated impressive capabilities in various reasoning tasks but face significant challenges with complex, knowledge-intensive multi-hop querie…
Meta-RTL: Reinforcement-Based Meta-Transfer Learning for Low-Resource Commonsense Reasoning
Yu Fu, Jie He, Yifan Yang +2
Meta learning has been widely used to exploit rich-resource source tasks to improve the performance of low-resource target tasks. Unfortunately, most existing meta learning approac…
BUCA: A Binary Classification Approach to Unsupervised Commonsense Question Answering
Jie He, Simon Chi Lok U, VÃctor Gutiérrez-Basulto +1
Unsupervised commonsense reasoning (UCR) is becoming increasingly popular as the construction of commonsense reasoning datasets is expensive, and they are inevitably limited in the…
MetaXCR: Reinforcement-Based Meta-Transfer Learning for Cross-Lingual Commonsense Reasoning
Jie He, Yu Fu
Commonsense reasoning (CR) has been studied in many pieces of domain and has achieved great progress with the aid of large datasets. Unfortunately, most existing CR datasets are bu…
Evaluating Discourse Cohesion in Pre-trained Language Models
Jie He, Wanqiu Long, Deyi Xiong
Large pre-trained neural models have achieved remarkable success in natural language process (NLP), inspiring a growing body of research analyzing their ability from different aspe…