activity
20222026
collaborators
Showing cs.CLShow all

5 papers · 1 filter

cs.CL2025

SUTA-LM: Bridging Test-Time Adaptation and Language Model Rescoring for Robust ASR

Wei-Ping Huang, Guan-Ting Lin, Hung-yi Lee

Despite progress in end-to-end ASR, real-world domain mismatches still cause performance drops, which Test-Time Adaptation (TTA) aims to mitigate by adjusting models during inferen…

cs.CL20253 cited

Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization

Tzu-Quan Lin, Wei-Ping Huang, Hao Tang +1

Fine-tuning speech representation models can enhance performance on specific tasks but often compromises their cross-task generalization ability. This degradation is often caused b…

cs.CL2024

Building a Taiwanese Mandarin Spoken Language Model: A First Attempt

Chih-Kai Yang, Yu-Kuan Fu, Chen-An Li +18

This technical report presents our initial attempt to build a spoken large language model (LLM) for Taiwanese Mandarin, specifically tailored to enable real-time, speech-to-speech…

cs.CL2024

Maximizing Data Efficiency for Cross-Lingual TTS Adaptation by Self-Supervised Representation Mixing and Embedding Initialization

Wei-Ping Huang, Sung-Feng Huang, Hung-yi Lee

This paper presents an effective transfer learning framework for language adaptation in text-to-speech systems, with a focus on achieving language adaptation using minimal labeled…

cs.CL2022

On the Utility of Self-supervised Models for Prosody-related Tasks

Guan-Ting Lin, Chi-Luen Feng, Wei-Ping Huang +5

Self-Supervised Learning (SSL) from speech data has produced models that have achieved remarkable performance in many tasks, and that are known to implicitly represent many aspects…