activity
20242026
collaborators

5 papers

eess.AS2026

Enhancing Speaker Verification with w2v-BERT 2.0 and Knowledge Distillation guided Structured Pruning

Ze Li, Ming Cheng, Ming Li

Large-scale self-supervised Pre-Trained Models (PTMs) have shown significant improvements in the speaker verification (SV) task by providing rich feature representations. In this p…

cs.LG2025

LFM2 Technical Report

Alexander Amini, Anna Banaszak, Harold Benoit +30

We present LFM2, a family of Liquid Foundation Models designed for efficient on-device deployment and strong task capabilities. Using hardware-in-the-loop architecture search under…

eess.AS2025

The DKU System for Multi-Speaker Automatic Speech Recognition in MLC-SLM Challenge

Yuke Lin, Ming Cheng, Ze Li +1

We present the DKU system for Task 2 of the MLC-SLM Challenge, which aims to perform multi-speaker automatic speech recognition directly from raw audio without Oracle speaker label…

eess.AS2025

Diarization-Aware Multi-Speaker Automatic Speech Recognition via Large Language Models

Yuke Lin, Ming Cheng, Ze Li +2

Multi-speaker automatic speech recognition (MS-ASR) faces significant challenges in transcribing overlapped speech, a task critical for applications like meeting transcription and…

eess.AS2024

The Database and Benchmark for the Source Speaker Tracing Challenge 2024

Ze Li, Yuke Lin, Tian Yao +6

Voice conversion (VC) systems can transform audio to mimic another speaker's voice, thereby attacking speaker verification (SV) systems. However, ongoing studies on source speaker…