Publications (8)
UniMoE-Audio: Unified Speech and Music Generation with Dynamic-Capacity MoE
Zhenyu Liu, Yunxin Li, Xuanyu Zhang +13
Recent advances in unified multimodal models indicate a clear trend towards comprehensive content generation. However, the auditory domain remains a significant challenge, with mus…
The study and design of RF coupler for Chinese ADS HWR Superconducting Cavity
Fanbo Meng, Xu Chen, Weimin Pan +4
RF power coupler is a key component of the superconducting accelerating system in Chinese ADS proton linac injector I, which is used to transmit 15kW RF power from the power source…
Improving Adversarial Waveform Generation based Singing Voice Conversion with Harmonic Signals
Haohan Guo, Zhiping Zhou, Fanbo Meng +1
Adversarial waveform generation has been a popular approach as the backend of singing voice conversion (SVC) to generate high-quality singing audio. However, the instability of GAN…
Hierarchical Acoustic-Semantic Modeling: Modality Separation and Semantic Coherence for Full-Duplex SLMs
Zhenyu Liu, Xuanyu Zhang, Yunxin Li +10
The paper identifies gradient conflicts between acoustic and semantic modeling as the cause of modality interference in full-duplex spoken language models and proposes Lychee-FD, a…
Design of a HOM-Damped 166.6 MHz Compact Quarter-Wave beta=1 Superconducting Cavity for High Energy Photon Source
Xinying Zhang, Jin Dai, Lin Guo +7
Superconducting cavities with low RF frequencies and heavy damping of higher order modes (HOM) are desired for the main accelerator of High Energy Photon Source (HEPS), a 6 GeV syn…
ChoreoNet: Towards Music to Dance Synthesis with Choreographic Action Unit
Zijie Ye, Haozhe Wu, Jia Jia +4
Dance and music are two highly correlated artistic forms. Synthesizing dance motions has attracted much attention recently. Most previous works conduct music-to-dance synthesis via…
View-Invariant Skeleton-based Action Recognition via Global-Local Contrastive Learning
Cunling Bian, Wei Feng, Fanbo Meng +1
Skeleton-based human action recognition has been drawing more interest recently due to its low sensitivity to appearance changes and the accessibility of more skeleton data. Howeve…
Coupler Conditioning and High Power Testing of ADS Spoke Cavity
Xu Chen, Fanbo Meng, Tongming Huang +3
Power couplers, used in China-ADS proton linac injector I, are required to transfer 6kW RF power to the superconducting Spoke cavities. At present, first two couplers of coaxial de…