1 paper
Linhao Dong, Zhecheng An, Peihao Wu +3
Speech or text representation generated by pre-trained models contains modal-specific information that could be combined for benefiting spoken language understanding (SLU) tasks. I…