3 papers
cs.RO2025
Episodic Memory Verbalization using Hierarchical Representations of Life-Long Robot Experience
Leonard Bärmann, Chad DeChant, Joana Plewnia +4
Verbalization of robot experience, i.e., summarization of and question answering about a robot's past, is a crucial ability for improving human-robot interaction. Previous works ap…
cs.CL2025
PIER: A Novel Metric for Evaluating What Matters in Code-Switching
Enes Yavuz Ugan, Ngoc-Quan Pham, Leonard Bärmann +1
Code-switching, the alternation of languages within a single discourse, presents a significant challenge for Automatic Speech Recognition. Despite the unique nature of the task, pe…
cs.CL2024
SciEx: Benchmarking Large Language Models on Scientific Exams with Human Expert Grading and Automatic Grading
Tu Anh Dinh, Carlos Mullov, Leonard Bärmann +15
With the rapid development of Large Language Models (LLMs), it is crucial to have benchmarks which can evaluate the ability of LLMs on different domains. One common use of LLMs is…