3 papers
cs.CL2026
Building a Multimodal Dataset of Academic Paper for Keyword Extraction
Jingyu Zhang, Xinyi Yan, Yi Xiang +2
Up to this point, keyword extraction task typically relies solely on textual data. Neglecting visual details and audio features from image and audio modalities leads to deficiencie…
cs.CL2026
Detection and Interpretability Analysis of Quotation Errors by Large Language Models
Bei Huang, Yingyi Zhang, Shenghao Huang +1
Purpose - Quotation error refers to the inconsistency between cited information and its original source. This phenomenon leads to a series of negative impacts, such as misinterpret…
cs.CL2026
SciNLP: A Domain-Specific Benchmark for Full-Text Scientific Entity and Relation Extraction in NLP
Decheng Duan, Yingyi Zhang, Jitong Peng +1
Structured information extraction from scientific literature is crucial for capturing core concepts and emerging trends in specialized fields. While existing datasets aid model dev…