2 papers
cs.CL2022
Knowledge Distillation Transfer Sets and their Impact on Downstream NLU Tasks
Charith Peris, Lizhen Tan, Thomas Gueudre +3
Teacher-student knowledge distillation is a popular technique for compressing today's prevailing large language models into manageable sizes that fit low-latency downstream applica…
cs.CL2020
Using multiple ASR hypotheses to boost i18n NLU performance
Charith Peris, Gokmen Oz, Khadige Abboud +3
Current voice assistants typically use the best hypothesis yielded by their Automatic Speech Recognition (ASR) module as input to their Natural Language Understanding (NLU) module,…