◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Florian Metze

Carnegie Mellon University

57 papers hereh-index 459.8k citations304 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author25
  • last author29

Across the 54 of 57 papers where every author was matched, so the position is known.

fields
  • cs.CL37
  • cs.CV7
  • cs.SD5
  • eess.AS4
  • cs.LG2
  • cs.CR1
affiliations
  • Carnegie Mellon University
  • FACEBOOK
Homepage
same name
  • Florian Metze — 13 papers, h 6
  • Florian Metze — 2 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20172022
most citedVisual Features for Context-Aware Speech Recognition

38 citations · 137 across the 21 of their papers we have counts for

collaborators
Showing eess.ASShow all

4 papers · 1 filter

eess.AS2020

ASR Error Correction and Domain Adaptation Using Machine Translation

Anirudh Mani, Shruti Palaskar, Nimshi Venkat Meripo +2

Off-the-shelf pre-trained Automatic Speech Recognition (ASR) systems are an increasingly viable service for companies of any size building speech-based products. While these ASR sy…

eess.AS2019

Cross-Attention End-to-End ASR for Two-Party Conversations

Suyoun Kim, Siddharth Dalmia, Florian Metze

We present an end-to-end speech recognition model that learns interaction between two speakers based on the turn-changing information. Unlike conventional speech recognition models…

eess.AS2018

Acoustic-to-Word Recognition with Sequence-to-Sequence Models

Shruti Palaskar, Florian Metze

Acoustic-to-Word recognition provides a straightforward solution to end-to-end speech recognition without needing external decoding, language model re-scoring or lexicon. While cha…

eess.AS2018

End-to-End Multimodal Speech Recognition

Shruti Palaskar, Ramon Sanabria, Florian Metze

Transcription or sub-titling of open-domain videos is still a challenging domain for Automatic Speech Recognition (ASR) due to the data's challenging acoustics, variable signal pro…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.