1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.HC2024
Robust Dual-Modal Speech Keyword Spotting for XR Headsets
Zhuojiang Cai, Yuhan Ma, Feng Lu
While speech interaction finds widespread utility within the Extended Reality (XR) domain, conventional vocal speech keyword spotting systems continue to grapple with formidable ch…
cs.CV2022★ 1 cited
Layout-Bridging Text-to-Image Synthesis
Jiadong Liang, Wenjie Pei, Feng Lu
The crux of text-to-image synthesis stems from the difficulty of preserving the cross-modality semantic consistency between the input text and the synthesized image. Typical method…