7 papers
Out-of-Distribution Detection Based on Total Variation Estimation
Dabiao Ma, Zhiba Su, Jian Yang +1
This paper introduces a novel approach to securing machine learning model deployments against potential distribution shifts in practical applications, the Total Variation Out-of-Di…
TS-PEFT: Unveiling Token-Level Redundancy in Parameter-Efficient Fine-Tuning
Dabiao Ma, Ziming Dai, Zhimin Xin +3
Current Parameter-Efficient Fine-Tuning (PEFT) methods typically operate under an implicit assumption: Once a target module is selected, every token passing through it contributes…
QvTAD: Differential Relative Attribute Learning for Voice Timbre Attribute Detection
Zhiyu Wu, Jingyi Fang, Yufei Tang +3
Voice Timbre Attribute Detection (vTAD) plays a pivotal role in fine-grained timbre modeling for speech generation tasks. However, it remains challenging due to the inherently subj…
SpecWav-Attack: Leveraging Spectrogram Resizing and Wav2Vec 2.0 for Attacking Anonymized Speech
Yuqi Li, Yuanzhong Zheng, Zhongtian Guo +3
This paper presents SpecWav-Attack, an adversarial model for detecting speakers in anonymized speech. It leverages Wav2Vec2 for feature extraction and incorporates spectrogram resi…
Qieemo: Speech Is All You Need in the Emotion Recognition in Conversations
Jinming Chen, Jingyi Fang, Yuanzhong Zheng +2
Emotion recognition plays a pivotal role in intelligent human-machine interaction systems. Multimodal approaches benefit from the fusion of diverse modalities, thereby improving th…
SFE-Net: Harnessing Biological Principles of Differential Gene Expression for Improved Feature Selection in Deep Learning Networks
Yuqi Li, Yuanzhong Zheng, Yaoxuan Wang +2
In the realm of DeepFake detection, the challenge of adapting to various synthesis methodologies such as Faceswap, Deepfakes, Face2Face, and NeuralTextures significantly impacts th…