2 papers
cs.CL2024
Seal: Advancing Speech Language Models to be Few-Shot Learners
Shuyu Lei, Lingen Liu, Jiaolong Yang +4
Existing auto-regressive language models have demonstrated a remarkable capability to perform a new task with just a few examples in prompt, without requiring any additional traini…
cs.CL2023
GRASS: Unified Generation Model for Speech-to-Semantic Tasks
Aobo Xia, Shuyu Lei, Yushu Yang +2
This paper explores the instruction fine-tuning technique for speech-to-semantic tasks by introducing a unified end-to-end (E2E) framework that generates target text conditioned on…