◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Kai Yu

63 papers hereh-index 5411.5k citations394 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author10
  • last author50

Across the 60 of 63 papers where every author was matched, so the position is known.

fields
  • cs.CL36
  • cs.SD8
  • eess.AS7
  • cs.LG3
  • quant-ph3
  • cs.CV2
same name
  • Kai Yu — 39 papers, h 13
  • Kai Yu — 33 papers, h 9
  • Kai Yu — 28 papers, h 9
  • Kai Yu — 22 papers, h 6
  • Kai Yu — 8 papers, h 2
  • Kai Yu — 7 papers, h 8

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20162025
most citedEnd-to-end spoofing detection with raw waveform CLDNNs

67 citations · 293 across the 40 of their papers we have counts for

collaborators
Showing 2018 · cs.CLShow all

4 papers · 2 filters

cs.CL2018

End-to-End Monaural Multi-speaker ASR System without Pretraining

Xuankai Chang, Yanmin Qian, Kai Yu +1

Recently, end-to-end models have become a popular approach as an alternative to traditional hybrid models in automatic speech recognition (ASR). The multi-speaker speech separation…

cs.CL2018

Towards Universal Dialogue State Tracking

Liliang Ren, Kaige Xie, Lu Chen +1

Dialogue state tracking is the core part of a spoken dialogue system. It estimates the beliefs of possible user's goals at every dialogue turn. However, for most current approaches…

cs.CL2018

Sequence Discriminative Training for Deep Learning based Acoustic Keyword Spotting

Zhehuai Chen, Yanmin Qian, Kai Yu

Speech recognition is a sequence prediction problem. Besides employing various deep learning approaches for framelevel classification, sequence-level discriminative training has be…

cs.CL2018

On Modular Training of Neural Acoustics-to-Word Model for LVCSR

Zhehuai Chen, Qi Liu, Hao Li +1

End-to-end (E2E) automatic speech recognition (ASR) systems directly map acoustics to words using a unified model. Previous works mostly focus on E2E training a single model which…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.