16 citations · 16 across the 2 of their papers we have counts for
2 papers
cs.CV2020
Data-efficient Alignment of Multimodal Sequences by Aligning Gradient Updates and Internal Feature Distributions
Jianan Wang, Boyang Li, Xiangyu Fan +2
The task of video and text sequence alignment is a prerequisite step toward joint understanding of movie videos and screenplays. However, supervised methods face the obstacle of li…
eess.AS2020★ 16 cited
Semi-supervised learning using teacher-student models for vocal melody extraction
Sangeun Kum, Jing-Hua Lin, Li Su +1
The lack of labeled data is a major obstacle in many music information retrieval tasks such as melody extraction, where labeling is extremely laborious or costly. Semi-supervised l…