5 citations · 7 across the 7 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2022
Self-Supervised Audio-and-Text Pre-training with Extremely Low-Resource Parallel Data
Yu Kang, Tianqiao Liu, Hang Li +2
Multimodal pre-training for audio-and-text has recently been proved to be effective and has significantly improved the performance of many downstream speech understanding tasks. Ho…
cs.SD2021★ 5 cited
CTAL: Pre-training Cross-modal Transformer for Audio-and-Language Representations
Hang Li, Yu Kang, Tianqiao Liu +2
Existing audio-language task-specific predictive approaches focus on building complicated late-fusion mechanisms. However, these models are facing challenges of overfitting with li…
cs.SD2021
A Multimodal Machine Learning Framework for Teacher Vocal Delivery Evaluation
Hang Li, Yu Kang, Yang Hao +3
The quality of vocal delivery is one of the key indicators for evaluating teacher enthusiasm, which has been widely accepted to be connected to the overall course qualities. Howeve…