◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Hasim Sak

10 papers hereh-index 3310.1k citations56 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author5
  • last author4

Across the 9 of 10 papers where every author was matched, so the position is known.

fields
  • cs.CL3
  • eess.AS3
  • cs.SD2
  • cs.CV1
  • cs.LG1

identity via Semantic Scholar / OpenAlex

activity
20162024
most citedTransformer Transducer: A Streamable Speech Recognition Model with Transformer Encoders and RNN-T Loss

26 citations · 42 across the 6 of their papers we have counts for

collaborators
Showing cs.SDShow all

2 papers · 1 filter

cs.SD2024

Clustering and Mining Accented Speech for Inclusive and Fair Speech Recognition

Jaeyoung Kim, Han Lu, Soheil Khorram +3

Modern automatic speech recognition (ASR) systems are typically trained on more than tens of thousands hours of speech data, which is one of the main factors for their great succes…

cs.SD2020

Transformer Transducer: One Model Unifying Streaming and Non-streaming Speech Recognition

Anshuman Tripathi, Jaeyoung Kim, Qian Zhang +2

In this paper we present a Transformer-Transducer model architecture and a training technique to unify streaming and non-streaming speech recognition models into one model. The mod…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.