1 paper · 1 filter
Jianbo Li, Yi Jiang, Sendong Zhao +3
Retrieval-Augmented Generation (RAG) helps LLMs stay accurate, but feeding long documents into a prompt makes the model slow and expensive. This has motivated context compression,…