33 citations · 117 across the 27 of their papers we have counts for
6 papers · 1 filter
AutoFAIR : Automatic Data FAIRification via Machine Reading
Tingyan Ma, Wei Liu, Bin Lu +4
The explosive growth of data fuels data-driven research, facilitating progress across diverse domains. The FAIR principles emerge as a guiding standard, aiming to enhance the finda…
Is Reference Necessary in the Evaluation of NLG Systems? When and Where?
Shuqian Sheng, Yi Xu, Luoyi Fu +4
The majority of automatic metrics for evaluating NLG systems are reference-based. However, the challenge of collecting human annotation results in a lack of reliable references in…
GeoGalactica: A Scientific Large Language Model in Geoscience
Zhouhan Lin, Cheng Deng, Le Zhou +18
Large language models (LLMs) have achieved huge success for their general knowledge and ability to solve a wide spectrum of tasks in natural language processing (NLP). Due to their…
Exploring and Verbalizing Academic Ideas by Concept Co-occurrence
Yi Xu, Shuqian Sheng, Bo Xue +3
Researchers usually come up with new ideas only after thoroughly comprehending vast quantities of literature. The difficulty of this procedure is exacerbated by the fact that the n…
PK-Chat: Pointer Network Guided Knowledge Driven Generative Dialogue Model
Cheng Deng, Bo Tong, Luoyi Fu +4
In the research of end-to-end dialogue systems, using real-world knowledge to generate natural, fluent, and human-like utterances with correct answers is crucial. However, domain-s…
Text Classification in the Wild: a Large-scale Long-tailed Name Normalization Dataset
Jiexing Qi, Shuhao Li, Zhixin Guo +5
Real-world data usually exhibits a long-tailed distribution,with a few frequent labels and a lot of few-shot labels. The study of institution name normalization is a perfect applic…