activity
20222024
most citedAutomatic music mixing with deep learning and out-of-domain data

12 citations · 17 across the 15 of their papers we have counts for

collaborators

15 papers

cs.SD2024

DisMix: Disentangling Mixtures of Musical Instruments for Source-level Pitch and Timbre Manipulation

Yin-Jyun Luo, Kin Wai Cheuk, Woosung Choi +8

Existing work on pitch and timbre disentanglement has been mostly focused on single-instrument music audio, excluding the cases where multiple instruments are presented. To fill th…

cs.SD2024

GRAFX: An Open-Source Library for Audio Processing Graphs in PyTorch

Sungho Lee, Marco Martínez-Ramírez, Wei-Hsiang Liao +4

We present GRAFX, an open-source library designed for handling audio processing graphs in PyTorch. Along with various library functionalities, we describe technical details on the…

cs.SD2024

SpecMaskGIT: Masked Generative Modeling of Audio Spectrograms for Efficient Audio Synthesis and Beyond

Marco Comunità, Zhi Zhong, Akira Takahashi +7

Recent advances in generative models that iteratively synthesize audio clips sparked great success to text-to-audio synthesis (TTA), but with the cost of slow synthesis speed and h…

cs.SD2024

Improving Unsupervised Clean-to-Rendered Guitar Tone Transformation Using GANs and Integrated Unaligned Clean Data

Yu-Hua Chen, Woosung Choi, Wei-Hsiang Liao +5

Recent years have seen increasing interest in applying deep learning methods to the modeling of guitar amplifiers or effect pedals. Existing methods are mainly based on the supervi…

cs.CL2024

ComperDial: Commonsense Persona-grounded Dialogue Dataset and Benchmark

Hiromi Wakaki, Yuki Mitsufuji, Yoshinori Maeda +5

We propose a new benchmark, ComperDial, which facilitates the training and evaluation of evaluation metrics for open-domain dialogue systems. ComperDial consists of human-scored re…

cs.SD2024

MR-MT3: Memory Retaining Multi-Track Music Transcription to Mitigate Instrument Leakage

Hao Hao Tan, Kin Wai Cheuk, Taemin Cho +2

This paper presents enhancements to the MT3 model, a state-of-the-art (SOTA) token-based multi-instrument automatic music transcription (AMT) model. Despite SOTA performance, MT3 h…