62 citations · 63 across the 2 of their papers we have counts for
2 papers
cs.CL2024★ 1 cited
DARG: Dynamic Evaluation of Large Language Models via Adaptive Reasoning Graph
Zhehao Zhang, Jiaao Chen, Diyi Yang
The current paradigm of evaluating Large Language Models (LLMs) through static benchmarks comes with significant limitations, such as vulnerability to data contamination and a lack…
cs.CL2023★ 62 cited
Can Large Language Models Transform Computational Social Science?
Caleb Ziems, William Held, Omar Shaikh +3
Large Language Models (LLMs) are capable of successfully performing many language processing tasks zero-shot (without training data). If zero-shot LLMs can also reliably classify a…