6 papers
Learning to Forget -- Hierarchical Episodic Memory for Lifelong Robot Deployment
Leonard Bärmann, Joana Plewnia, Alex Waibel +1
Robots must verbalize their past experiences when users ask "Where did you put my keys?" or "Why did the task fail?" Yet maintaining life-long episodic memory (EM) from continuous…
Episodic Memory Verbalization using Hierarchical Representations of Life-Long Robot Experience
Leonard Bärmann, Chad DeChant, Joana Plewnia +4
Verbalization of robot experience, i.e., summarization of and question answering about a robot's past, is a crucial ability for improving human-robot interaction. Previous works ap…
Mask-Free Audio-driven Talking Face Generation for Enhanced Visual Quality and Identity Preservation
Dogucan Yaman, Fevziye Irem Eyiokur, Leonard Bärmann +2
Audio-Driven Talking Face Generation aims at generating realistic videos of talking faces, focusing on accurate audio-lip synchronization without deteriorating any identity-related…
PIER: A Novel Metric for Evaluating What Matters in Code-Switching
Enes Yavuz Ugan, Ngoc-Quan Pham, Leonard Bärmann +1
Code-switching, the alternation of languages within a single discourse, presents a significant challenge for Automatic Speech Recognition. Despite the unique nature of the task, pe…
Incremental Learning of Humanoid Robot Behavior from Natural Interaction and Large Language Models
Leonard Bärmann, Rainer Kartmann, Fabian Peller-Konrad +3
Natural-language dialog is key for intuitive human-robot interaction. It can be used not only to express humans' intents, but also to communicate instructions for improvement if a…
Audio-Visual Speech Representation Expert for Enhanced Talking Face Video Generation and Evaluation
Dogucan Yaman, Fevziye Irem Eyiokur, Leonard Bärmann +3
In the task of talking face generation, the objective is to generate a face video with lips synchronized to the corresponding audio while preserving visual details and identity inf…