3 citations · 3 across the 2 of their papers we have counts for
1 paper · 1 filter
Zhepei Wang, Cem Subakan, Krishna Subramani +4
Recent advances in using language models to obtain cross-modal audio-text representations have overcome the limitations of conventional training approaches that use predefined labe…