Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
PIER: A Novel Metric for Evaluating What Matters in Code-Switching
Enes Yavuz Ugan, Ngoc-Quan Pham, Leonard Bärmann +1
Code-switching, the alternation of languages within a single discourse, presents a significant challenge for Automatic Speech Recognition. Despite the unique nature of the task, pe…
cs.CL2024
SciEx: Benchmarking Large Language Models on Scientific Exams with Human Expert Grading and Automatic Grading
Tu Anh Dinh, Carlos Mullov, Leonard Bärmann +15
With the rapid development of Large Language Models (LLMs), it is crucial to have benchmarks which can evaluate the ability of LLMs on different domains. One common use of LLMs is…