Showing cs.CLShow all
2 papers · 1 filter
cs.CL2024
Reverse Modeling in Large Language Models
Sicheng Yu, Yuanchen Xu, Cunxiao Du +5
Humans are accustomed to reading and writing in a forward manner, and this natural bias extends to text understanding in auto-regressive large language models (LLMs). This paper in…
cs.CL2024
GliDe with a CaPE: A Low-Hassle Method to Accelerate Speculative Decoding
Cunxiao Du, Jing Jiang, Xu Yuanchen +8
Speculative decoding is a relatively new decoding framework that leverages small and efficient draft models to reduce the latency of LLMs. In this study, we introduce GliDe and CaP…