1 paper · 1 filter
Xiaoshu Chen, Sihang Zhou, Ke Liang +2
Compressing long chains of thought (CoT) into compact latent tokens is crucial for efficient reasoning with large language models (LLMs). Recent studies employ autoencoders to achi…