activity
20182026
most citedTMS: A Temporal Multi-scale Backbone Design for Speaker Embedding

5 citations · 12 across the 12 of their papers we have counts for

collaborators
Showing eess.ASShow all

11 papers · 1 filter

eess.AS2023

Neural domain alignment for spoken language recognition based on optimal transport

Xugang Lu, Peng Shen, Yu Tsao +1

Domain shift poses a significant challenge in cross-domain spoken language recognition (SLR) by reducing its effectiveness. Unsupervised domain adaptation (UDA) algorithms have bee…

eess.AS2023

Hierarchical Cross-Modality Knowledge Transfer with Sinkhorn Attention for CTC-based ASR

Xugang Lu, Peng Shen, Yu Tsao +1

Due to the modality discrepancy between textual and acoustic modeling, efficiently transferring linguistic knowledge from a pretrained language model (PLM) to acoustic encoding for…

eess.AS2023

Cross-modal Alignment with Optimal Transport for CTC-based ASR

Xugang Lu, Peng Shen, Yu Tsao +1

Temporal connectionist temporal classification (CTC)-based automatic speech recognition (ASR) is one of the most successful end to end (E2E) ASR frameworks. However, due to the tok…

eess.AS20222 cited

Partial Coupling of Optimal Transport for Spoken Language Identification

Xugang Lu, Peng Shen, Yu Tsao +1

In order to reduce domain discrepancy to improve the performance of cross-domain spoken language identification (SLID) system, as an unsupervised domain adaptation (UDA) method, we…

eess.AS2022

A Novel Temporal Attentive-Pooling based Convolutional Recurrent Architecture for Acoustic Signal Enhancement

Tassadaq Hussain, Wei-Chien Wang, Mandar Gogate +5

In acoustic signal processing, the target signals usually carry semantic information, which is encoded in a hierarchal structure of short and long-term contexts. However, the backg…

eess.AS20212 cited

Siamese Neural Network with Joint Bayesian Model Structure for Speaker Verification

Xugang Lu, Peng Shen, Yu Tsao +1

Generative probability models are widely used for speaker verification (SV). However, the generative models are lack of discriminative feature selection ability. As a hypothesis te…