activity
20102023
most citedSelf-Supervised Speech Representation Learning: A Review

371 citations · 718 across the 115 of their papers we have counts for

collaborators
Showing 2023Show all

17 papers · 1 filter

cs.SD2023★ 7 cited

The defender's perspective on automatic speaker verification: An overview

Haibin Wu, Jiawen Kang, Lingwei Meng +2

Automatic speaker verification (ASV) plays a critical role in security-sensitive environments. Regrettably, the reliability of ASV has been undermined by the emergence of spoofing…

cs.CL2023

Revealing the Blind Spot of Sentence Encoder Evaluation by HEROS

Cheng-Han Chiang, Yung-Sung Chuang, James Glass +1

Existing sentence textual similarity benchmark datasets only use a single number to summarize how similar the sentence encoder's decision is to humans'. However, it is unclear what…

cs.CL2023

Improving Non-autoregressive Translation Quality with Pretrained Language Model, Embedding Distillation and Upsampling Strategy for CTC

Shen-sian Syu, Juncheng Xie, Hung-yi Lee

Non-autoregressive approaches aim to improve the inference speed of translation models, particularly those that generate output in a one-pass forward manner. However, these approac…

eess.AS2023★ 10 cited

SpeechGen: Unlocking the Generative Power of Speech Language Models with Prompts

Haibin Wu, Kai-Wei Chang, Yuan-Kuei Wu +1

Large language models (LLMs) have gained considerable attention for Artificial Intelligence Generated Content (AIGC), particularly with the emergence of ChatGPT. However, the direc…

cs.CL2023★ 7 cited

How to Estimate Model Transferability of Pre-Trained Speech Models?

Zih-Ching Chen, Chao-Han Huck Yang, Bo Li +6

In this work, we introduce a "score-based assessment" framework for estimating the transferability of pre-trained speech models (PSMs) for fine-tuning target tasks. We leverage upo…

cs.CL2023★ 2 cited

Improving Cascaded Unsupervised Speech Translation with Denoising Back-translation

Yu-Kuan Fu, Liang-Hsuan Tseng, Jiatong Shi +4

Most of the speech translation models heavily rely on parallel data, which is hard to collect especially for low-resource languages. To tackle this issue, we propose to build a cas…