6 citations · 6 across the 2 of their papers we have counts for
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2024★ 6 cited
Are Large Language Models Good Essay Graders?
Anindita Kundu, Denilson Barbosa
We evaluate the effectiveness of Large Language Models (LLMs) in assessing essay quality, focusing on their alignment with human grading. More precisely, we evaluate ChatGPT and Ll…
cs.CL2024
Semantic Graphs for Syntactic Simplification: A Revisit from the Age of LLM
Peiran Yao, Kostyantyn Guzhva, Denilson Barbosa
Symbolic sentence meaning representations, such as AMR (Abstract Meaning Representation) provide expressive and structured semantic graphs that act as intermediates that simplify d…
cs.CL2024
Accurate and Nuanced Open-QA Evaluation Through Textual Entailment
Peiran Yao, Denilson Barbosa
Open-domain question answering (Open-QA) is a common task for evaluating large language models (LLMs). However, current Open-QA evaluations are criticized for the ambiguity in ques…