activity
20232025
collaborators
Showing eess.ASShow all

5 papers · 1 filter

eess.AS2024

Speech Recognition Rescoring with Large Speech-Text Foundation Models

Prashanth Gurunath Shivakumar, Jari Kolehmainen, Aditya Gourav +4

Large language models (LLM) have demonstrated the ability to understand human language by leveraging large amount of text data. Automatic speech recognition (ASR) systems are often…

eess.AS2023

Discriminative Speech Recognition Rescoring with Pre-trained Language Models

Prashanth Gurunath Shivakumar, Jari Kolehmainen, Yile Gu +3

Second pass rescoring is a critical component of competitive automatic speech recognition (ASR) systems. Large language models have demonstrated their ability in using pre-trained…

eess.AS2023

Personalization for BERT-based Discriminative Speech Recognition Rescoring

Jari Kolehmainen, Yile Gu, Aditya Gourav +4

Recognition of personalized content remains a challenge in end-to-end speech recognition. We explore three novel approaches that use personalized content in a neural rescoring step…

eess.AS2023

Scaling Laws for Discriminative Speech Recognition Rescoring Models

Yile Gu, Prashanth Gurunath Shivakumar, Jari Kolehmainen +3

Recent studies have found that model performance has a smooth power-law relationship, or scaling laws, with training data and model size, for a wide range of problems. These scalin…

eess.AS2023

Distillation Strategies for Discriminative Speech Recognition Rescoring

Prashanth Gurunath Shivakumar, Jari Kolehmainen, Yile Gu +3

Second-pass rescoring is employed in most state-of-the-art speech recognition systems. Recently, BERT based models have gained popularity for re-ranking the n-best hypothesis by ex…