1 citations · 1 across the 3 of their papers we have counts for
3 papers
Reusing Latent Speech Representations for Query-Conditioned Topic Localization in Transcripts
Steffen Freisinger, Philipp Seeberger, Thomas Ranzenberger +2
Long transcripts are costly inputs for downstream NLP systems and often contain irrelevant context. We study query-conditioned topic localization: predicting the sentence span in a…
The Spoken Wikipedia Presentation Corpus
Thomas Ranzenberger, Steffen Freisinger, Tobias Bocklet +1
We present the Spoken Wikipedia Presentation Corpus, an extension of the Spoken Wikipedia Corpora featuring LLM-generated slide decks for multimodal ASR. Slides are created from LL…
Towards Multi-Level Transcript Segmentation: LoRA Fine-Tuning for Table-of-Contents Generation
Steffen Freisinger, Philipp Seeberger, Thomas Ranzenberger +2
Segmenting speech transcripts into thematic sections benefits both downstream processing and users who depend on written text for accessibility. We introduce a novel approach to hi…