152 citations · 154 across the 3 of their papers we have counts for
3 papers
cs.CL2023★ 2 cited
CORECODE: A Common Sense Annotated Dialogue Dataset with Benchmark Tasks for Chinese Large Language Models
Dan Shi, Chaobin You, Jiantao Huang +2
As an indispensable ingredient of intelligence, commonsense reasoning is crucial for large language models (LLMs) in real-world scenarios. In this paper, we propose CORECODE, a dat…
cs.CL2023
NumHG: A Dataset for Number-Focused Headline Generation
Jian-Tao Huang, Chung-Chi Chen, Hen-Hsen Huang +1
Headline generation, a key task in abstractive summarization, strives to condense a full-length article into a succinct, single line of text. Notably, while contemporary encoder-de…
cs.LG2022★ 152 cited
Twin Contrastive Learning for Online Clustering
Yunfan Li, Mouxing Yang, Dezhong Peng +3
This paper proposes to perform online clustering by conducting twin contrastive learning (TCL) at the instance and cluster level. Specifically, we find that when the data is projec…