3 citations · 4 across the 4 of their papers we have counts for
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025★ 1 cited
LETToT: Label-Free Evaluation of Large Language Models On Tourism Using Expert Tree-of-Thought
Ruiyan Qi, Congding Wen, Weibo Zhou +3
Evaluating large language models (LLMs) in specific domain like tourism remains challenging due to the prohibitive cost of annotated benchmarks and persistent issues like hallucina…
cs.CL2025
D-SCoRE: Document-Centric Segmentation and CoT Reasoning with Structured Export for QA-CoT Data Generation
Weibo Zhou, Lingbo Li, Shangsong Liang
The scarcity and high cost of high-quality domain-specific question-answering (QA) datasets limit supervised fine-tuning of large language models (LLMs). We introduce $\textbf{D-SC…
cs.CL2025★ 4 cited
What Level of Automation is "Good Enough"? A Benchmark of Large Language Models for Meta-Analysis Data Extraction
Lingbo Li, Anuradha Mathrani, Teo Susnjak
Automating data extraction from full-text randomised controlled trials (RCTs) for meta-analysis remains a significant challenge. This study evaluates the practical performance of t…