4 papers
Perceived Femininity in Singing Voice: Analysis and Prediction
Yuexuan Kong, Viet-Anh Tran, Romain Hennequin
This paper focuses on the often-overlooked aspect of perceived voice femininity in singing voices. While existing research has examined perceived voice femininity in speech, the sa…
Multi-Class-Token Transformer for Multitask Self-supervised Music Information Retrieval
Yuexuan Kong, Vincent Lostanlen, Romain Hennequin +2
Contrastive learning and equivariant learning are effective methods for self-supervised learning (SSL) for audio content analysis. Yet, their application to music information retri…
Emergent musical properties of a transformer under contrastive self-supervised learning
Yuexuan Kong, Gabriel Meseguer-Brocal, Vincent Lostanlen +2
In music information retrieval (MIR), contrastive self-supervised learning for general-purpose representation models is effective for global tasks such as automatic tagging. Howeve…
S-KEY: Self-supervised Learning of Major and Minor Keys from Audio
Yuexuan Kong, Gabriel Meseguer-Brocal, Vincent Lostanlen +2
STONE, the current method in self-supervised learning for tonality estimation in music signals, cannot distinguish relative keys, such as C major versus A minor. In this article, w…