3 citations · 6 across the 6 of their papers we have counts for
6 papers
CosyAudio: Improving Audio Generation with Confidence Scores and Synthetic Captions
Xinfa Zhu, Wenjie Tian, Xinsheng Wang +4
Text-to-Audio (TTA) generation is an emerging area within AI-generated content (AIGC), where audio is created from natural language descriptions. Despite growing interest, developi…
EDSep: An Effective Diffusion-Based Method for Speech Source Separation
Jinwei Dong, Xinsheng Wang, Qirong Mao
Generative models have attracted considerable attention for speech separation tasks, and among these, diffusion-based methods are being explored. Despite the notable success of dif…
Wavelength-shifting light traps for SWGO and other applications
M. Pihet, M. Mariotti, C. Arcaro
Wavelength-shifting (WLS) materials contain molecules that absorb light and reemit at longer wavelengths. They can be used for light detection because they provide a large effectiv…
Dark Matter searches in Dwarf Galaxies with the Southern Wide-field Gamma-ray Observatory
Micael Andrade, Aion Viana
Dark matter is thought to make up most of the matter density of the Universe, yet its true nature remains uncertain. Among dark matter theories, Weakly Interacting Massive Particle…
An update on site search activities for SWGO
M. Santander, U. Barres de Almeida, J. A. Bellido +19
The Southern Wide-field Gamma-ray Observatory (SWGO) is a project by scientists and engineers from 14 countries and 78 institutions to design and build the first wide-field, ground…
MSM-VC: High-fidelity Source Style Transfer for Non-Parallel Voice Conversion by Multi-scale Style Modeling
Zhichao Wang, Xinsheng Wang, Qicong Xie +4
In addition to conveying the linguistic content from source speech to converted speech, maintaining the speaking style of source speech also plays an important role in the voice co…