1 paper
Guoxin Ma, Yibing Liu, Chengzhengxu Li +7
Context compression aims to shorten long context inputs with minimal information loss for LLM inference acceleration. While existing methods have shown promise, they typically rely…