22 citations · 60 across the 19 of their papers we have counts for
Showing 2023Show all
2 papers · 1 filter
cs.CL2023
Improving Joint Speech-Text Representations Without Alignment
Cal Peyser, Zhong Meng, Ke Hu +5
The last year has seen astonishing progress in text-prompted image generation premised on the idea of a cross-modal representation space in which the text and image domains are rep…
cs.CL2023
A Comparison of Semi-Supervised Learning Techniques for Streaming ASR at Scale
Cal Peyser, Michael Picheny, Kyunghyun Cho +3
Unpaired text and audio injection have emerged as dominant methods for improving ASR performance in the absence of a large labeled corpus. However, little guidance exists on deploy…