Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
From Prefix Cache to Fusion RAG Cache: Accelerating LLM Inference in Retrieval-Augmented Generation
Jiahao Wang, Weiyu Xie, Mingxing Zhang +10
Retrieval-Augmented Generation enhances Large Language Models by integrating external knowledge, which reduces hallucinations but increases prompt length. This increase leads to hi…
cs.CL2024
Process-Driven Autoformalization in Lean 4
Jianqiao Lu, Yingjia Wan, Zhengying Liu +10
Autoformalization, the conversion of natural language mathematics into formal languages, offers significant potential for advancing mathematical reasoning. However, existing effort…