5 citations · 6 across the 4 of their papers we have counts for
4 papers
Self-Supervised Audio-and-Text Pre-training with Extremely Low-Resource Parallel Data
Yu Kang, Tianqiao Liu, Hang Li +2
Multimodal pre-training for audio-and-text has recently been proved to be effective and has significantly improved the performance of many downstream speech understanding tasks. Ho…
CTAL: Pre-training Cross-modal Transformer for Audio-and-Language Representations
Hang Li, Yu Kang, Tianqiao Liu +2
Existing audio-language task-specific predictive approaches focus on building complicated late-fusion mechanisms. However, these models are facing challenges of overfitting with li…
A Multimodal Machine Learning Framework for Teacher Vocal Delivery Evaluation
Hang Li, Yu Kang, Yang Hao +3
The quality of vocal delivery is one of the key indicators for evaluating teacher enthusiasm, which has been widely accepted to be connected to the overall course qualities. Howeve…
Multimodal Learning For Classroom Activity Detection
Hang Li, Yu Kang, Wenbiao Ding +4
Classroom activity detection (CAD) focuses on accurately classifying whether the teacher or student is speaking and recording both the length of individual utterances during a clas…