112 citations · 280 across the 26 of their papers we have counts for
7 papers · 1 filter
Haha-Pod: An Attempt for Laughter-based Non-Verbal Speaker Verification
Yuke Lin, Xiaoyi Qin, Ning Jiang +2
It is widely acknowledged that discriminative representation for speaker verification can be extracted from verbal speech. However, how much speaker information that non-verbal voc…
SdCT-GAN: Reconstructing CT from Biplanar X-Rays with Self-driven Generative Adversarial Networks
Shuangqin Cheng, Qingliang Chen, Qiyi Zhang +4
Computed Tomography (CT) is a medical imaging modality that can generate more informative 3D images than 2D X-rays. However, this advantage comes at the expense of more radiation e…
BiSinger: Bilingual Singing Voice Synthesis
Huali Zhou, Yueqian Lin, Yao Shi +2
Although Singing Voice Synthesis (SVS) has made great strides with Text-to-Speech (TTS) techniques, multilingual singing voice modeling remains relatively unexplored. This paper pr…
SlideSpeech: A Large-Scale Slide-Enriched Audio-Visual Corpus
Haoxu Wang, Fan Yu, Xian Shi +3
Multi-Modal automatic speech recognition (ASR) techniques aim to leverage additional modalities to improve the performance of speech recognition systems. While existing approaches…
Photonic time-delayed reservoir computing based on series coupled microring resonators with high memory capacity
Yijia Li, Ming Li, MingYi Gao +7
On-chip microring resonators (MRRs) have been proposed to construct the time-delayed reservoir computing (RC), which offers promising configurations available for computation with…
Explicit Topology Optimization of Conforming Voronoi Foams
Ming Li, Jingqiao Hu, Wei Chen +2
Topology optimization is able to maximally leverage the high DOFs and mechanical potentiality of porous foams but faces three fundamental challenges: conforming to free-form outer…