activity
20172022
most citedAdvanced LSTM: A Study about Better Time Dependency Modeling in Emotion Recognition

8 citations · 10 across the 2 of their papers we have counts for

collaborators

5 papers

cs.SD20222 cited

The Microsoft System for VoxCeleb Speaker Recognition Challenge 2022

Gang Liu, Tianyan Zhou, Yong Zhao +4

In this report, we describe our submitted system for track 2 of the VoxCeleb Speaker Recognition Challenge 2022 (VoxSRC-22). We fuse a variety of good-performing models ranging fro…

eess.AS2020

Microsoft Speaker Diarization System for the VoxCeleb Speaker Recognition Challenge 2020

Xiong Xiao, Naoyuki Kanda, Zhuo Chen +10

This paper describes the Microsoft speaker diarization system for monaural multi-talker recordings in the wild, evaluated at the diarization track of the VoxCeleb Speaker Recogniti…

cs.CL2018

An Online Attention-based Model for Speech Recognition

Ruchao Fan, Pan Zhou, Wei Chen +2

Attention-based end-to-end models such as Listen, Attend and Spell (LAS), simplify the whole pipeline of traditional automatic speech recognition (ASR) systems and become popular i…

eess.AS2018

An Ensemble Framework of Voice-Based Emotion Recognition System for Films and TV Programs

Fei Tao, Gang Liu, Qingen Zhao

Employing voice-based emotion recognition function in artificial intelligence (AI) product will improve the user experience. Most of researches that have been done only focus on th…

cs.LG20178 cited

Advanced LSTM: A Study about Better Time Dependency Modeling in Emotion Recognition

Fei Tao, Gang Liu

Long short-term memory (LSTM) is normally used in recurrent neural network (RNN) as basic recurrent unit. However,conventional LSTM assumes that the state at current time step depe…