8 citations · 10 across the 6 of their papers we have counts for
6 papers
Beyond Catalogue Counts: the Dataset Visibility Asymmetry in Low-Resource Multilingual NLP
Zhiyin Tan, Changxu Duan
Multilingual NLP often relies on dataset counts from centralized catalogues to characterize which languages are resource-rich or resource-poor. However, these catalogues record onl…
Diagnosing Structural Failures in LLM-Based Evidence Extraction for Meta-Analysis
Zhiyin Tan, Jennifer D'Souza
Systematic reviews and meta-analyses rely on converting narrative articles into structured, numerically grounded study records. Despite rapid advances in large language models (LLM…
Semantically Orthogonal Framework for Citation Classification: Disentangling Intent and Content
Changxu Duan, Zhiyin Tan
Understanding the role of citations is essential for research assessment and citation-aware digital libraries. However, existing citation classification frameworks often conflate c…
Multi-Disciplinary Dataset Discovery from Citation-Verified Literature Contexts
Zhiyin Tan, Changxu Duan
Identifying suitable datasets for a research question remains challenging because existing dataset search engines rely heavily on metadata quality and keyword overlap, which often…
Toward Purpose-oriented Topic Model Evaluation enabled by Large Language Models
Zhiyin Tan, Jennifer D'Souza
This study presents a framework for automated evaluation of dynamically evolving topic models using Large Language Models (LLMs). Topic modeling is essential for organizing and ret…
Bridging the Evaluation Gap: Leveraging Large Language Models for Topic Model Evaluation
Zhiyin Tan, Jennifer D'Souza
This study presents a framework for automated evaluation of dynamically evolving topic taxonomies in scientific literature using Large Language Models (LLMs). In digital library sy…