1 citations · 7 across the 57 of their papers we have counts for
4 papers · 1 filter
Multi-Level Optimal Transport for Universal Cross-Tokenizer Knowledge Distillation on Language Models
Xiao Cui, Mo Zhu, Yulei Qin +3
Knowledge distillation (KD) has become a prevalent technique for compressing large language models (LLMs). Existing KD methods are constrained by the need for identical tokenizers…
Trustworthy Alignment of Retrieval-Augmented Large Language Models via Reinforcement Learning
Zongmeng Zhang, Yufeng Shi, Jinhua Zhu +4
Trustworthiness is an essential prerequisite for the real-world application of large language models. In this paper, we focus on the trustworthiness of language models with respect…
Semi-Supervised Spoken Language Glossification
Huijie Yao, Wengang Zhou, Hao Zhou +1
Spoken language glossification (SLG) aims to translate the spoken language text into the sign language gloss, i.e., a written record of sign language. In this work, we present a fr…
Cross-Lingual Transfer for Natural Language Inference via Multilingual Prompt Translator
Xiaoyu Qiu, Yuechen Wang, Jiaxin Shi +2
Based on multilingual pre-trained models, cross-lingual transfer with prompt learning has shown promising effectiveness, where soft prompt learned in a source language is transferr…