papers

Publications (8)

cs.SD2025

UniMoE-Audio: Unified Speech and Music Generation with Dynamic-Capacity MoE

Zhenyu Liu, Yunxin Li, Xuanyu Zhang +13

Recent advances in unified multimodal models indicate a clear trend towards comprehensive content generation. However, the auditory domain remains a significant challenge, with mus…

physics.acc-ph2013

The study and design of RF coupler for Chinese ADS HWR Superconducting Cavity

Fanbo Meng, Xu Chen, Weimin Pan +4

RF power coupler is a key component of the superconducting accelerating system in Chinese ADS proton linac injector I, which is used to transmit 15kW RF power from the power source…

cs.SD2022

Improving Adversarial Waveform Generation based Singing Voice Conversion with Harmonic Signals

Haohan Guo, Zhiping Zhou, Fanbo Meng +1

Adversarial waveform generation has been a popular approach as the backend of singing voice conversion (SVC) to generate high-quality singing audio. However, the instability of GAN…

cs.CL2026

Hierarchical Acoustic-Semantic Modeling: Modality Separation and Semantic Coherence for Full-Duplex SLMs

Zhenyu Liu, Xuanyu Zhang, Yunxin Li +10

The paper identifies gradient conflicts between acoustic and semantic modeling as the cause of modality interference in full-duplex spoken language models and proposes Lychee-FD, a…

#full-duplex spoken language models#modality interference#hierarchical parameter separation#semantic alignment
physics.acc-ph2021

Design of a HOM-Damped 166.6 MHz Compact Quarter-Wave beta=1 Superconducting Cavity for High Energy Photon Source

Xinying Zhang, Jin Dai, Lin Guo +7

Superconducting cavities with low RF frequencies and heavy damping of higher order modes (HOM) are desired for the main accelerator of High Energy Photon Source (HEPS), a 6 GeV syn…

cs.CV2020

ChoreoNet: Towards Music to Dance Synthesis with Choreographic Action Unit

Zijie Ye, Haozhe Wu, Jia Jia +4

Dance and music are two highly correlated artistic forms. Synthesizing dance motions has attracted much attention recently. Most previous works conduct music-to-dance synthesis via…

cs.CV2022

View-Invariant Skeleton-based Action Recognition via Global-Local Contrastive Learning

Cunling Bian, Wei Feng, Fanbo Meng +1

Skeleton-based human action recognition has been drawing more interest recently due to its low sensitivity to appearance changes and the accessibility of more skeleton data. Howeve…

physics.ins-det2013

Coupler Conditioning and High Power Testing of ADS Spoke Cavity

Xu Chen, Fanbo Meng, Tongming Huang +3

Power couplers, used in China-ADS proton linac injector I, are required to transfer 6kW RF power to the superconducting Spoke cavities. At present, first two couplers of coaxial de…