◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

H. Liao

8 papers hereh-index 202.1k citations35 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author5
  • last author3

Across the 8 of 8 papers where every author was matched, so the position is known.

fields
  • eess.AS4
  • cs.CL2
  • cs.CV1
  • stat.ML1
same name
  • H. Liao — 217 papers, h 68
  • H. Liao — 98 papers
  • H. Liao — 56 papers, h 7
  • H. Liao — 42 papers, h 42
  • H. Liao — 21 papers, h 15
  • H. Liao — 11 papers

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20172023
most citedNeural Language Modeling with Visual Features

23 citations · 58 across the 6 of their papers we have counts for

collaborators
Showing eess.ASShow all

4 papers · 1 filter

eess.AS2023

On Robustness to Missing Video for Audiovisual Speech Recognition

Oscar Chang, Otavio Braga, Hank Liao +2

It has been shown that learning audiovisual features can lead to improved speech recognition performance over audio-only features, especially for noisy speech. However, in many com…

eess.AS2022

End-to-End Multi-Person Audio/Visual Automatic Speech Recognition

Otavio Braga, Takaki Makino, Olivier Siohan +1

Traditionally, audio-visual automatic speech recognition has been studied under the assumption that the speaking face on the visual signal is the face matching the audio. However,…

eess.AS2019★ 15 cited

Recurrent Neural Network Transducer for Audio-Visual Speech Recognition

Takaki Makino, Hank Liao, Yannis Assael +4

This work presents a large-scale audio-visual speech recognition system based on a recurrent neural network transducer (RNN-T) architecture. To support the development of such a sy…

eess.AS2019★ 13 cited

A comparison of end-to-end models for long-form speech recognition

Chung-Cheng Chiu, Wei Han, Yu Zhang +11

End-to-end automatic speech recognition (ASR) models, including both attention-based models and the recurrent neural network transducer (RNN-T), have shown superior performance com…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.