337 citations · 3.4k across the 96 of their papers we have counts for
1 paper · 2 filters
Rodrigo Castellon, Chris Donahue, Percy Liang
We demonstrate that language models pre-trained on codified (discretely-encoded) music audio learn representations that are useful for downstream MIR tasks. Specifically, we explor…