Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Predictive Prefetching for Retrieval-Augmented Generation
Wuyang Zhang, Shichao Pei
Retrieval-Augmented Generation (RAG) improves factual grounding in large language models but suffers from substantial latency due to synchronous retrieval. While recent work explor…
cs.CL2024
Abstract2Appendix: Academic Reviews Enhance LLM Long-Context Capabilities
Shengzhi Li, Kittipat Kampa, Rongyu Lin +2
Large language models (LLMs) have shown remarkable performance across various tasks, yet their ability to handle long-context reading remains challenging. This study explores the e…