collaborators

16 papers

cs.AI2026

Rethinking Reward Models for Multi-Domain Test-Time Scaling

Dong Bok Lee, Seanie Lee, Sangwoo Park +12

The reliability of large language models (LLMs) during test-time scaling is often assessed with \emph{external verifiers} or \emph{reward models} that distinguish correct reasoning…

cs.CL2026

Evaluation Pitfalls and Challenges in Multimedia Event Extraction

Philipp Seeberger, Steffen Freisinger, Tobias Bocklet +1

Multimedia event extraction aims to jointly identify events and their arguments across multiple modalities, such as text and images, to support more comprehensive event understandi…

eess.AS2026

Mitigating Scoring Errors and Compensating for Nonverbal Subtests in Speech-Based Dementia Assessment

Franziska Braun, Christopher Witzl, Andreas Erzigkeit +4

Early detection of cognitive impairment relies on neuropsychological tests to minimize subjectivity by assessing multiple cognitive domains. Speech-based evaluation can support dia…

cs.SD2026

Time vs. Layer: Locating Predictive Cues for Dysarthric Speech Descriptors in wav2vec 2.0

Natalie Engert, Dominik Wagner, Korbinian Riedhammer +1

Wav2vec 2.0 (W2V2) has shown strong performance in pathological speech analysis by effectively capturing the characteristics of atypical speech. Despite its success, it remains unc…

cs.CL2026

Reading Between the Waves: Robust Topic Segmentation Using Inter-Sentence Audio Features

Steffen Freisinger, Philipp Seeberger, Tobias Bocklet +1

Spoken content, such as online videos and podcasts, often spans multiple topics, which makes automatic topic segmentation essential for user navigation and downstream applications.…

cs.CL2026

Towards Multi-Level Transcript Segmentation: LoRA Fine-Tuning for Table-of-Contents Generation

Steffen Freisinger, Philipp Seeberger, Thomas Ranzenberger +2

Segmenting speech transcripts into thematic sections benefits both downstream processing and users who depend on written text for accessibility. We introduce a novel approach to hi…