1 paper
Yuxuan Chen, Peize He, Haoyuan Yu +1
A universal audio representation should capture fine-grained speech cues and high-level semantics for environmental sounds and music in a single encoder. Existing encoders often ex…