1 paper
Jaeyeon Kim, Jaeyoon Jung, Jinjoo Lee +1
We propose EnCLAP, a novel framework for automated audio captioning. EnCLAP employs two acoustic representation models, EnCodec and CLAP, along with a pretrained language model, BA…