2 papers
cs.CL2025
EfficientLLM: Efficiency in Large Language Models
Zhengqing Yuan, Weixiang Sun, Yixin Liu +13
Large Language Models (LLMs) have driven significant progress, yet their growing parameter counts and context windows incur prohibitive compute, energy, and monetary costs. We intr…
cs.CL2025
ELITE: Embedding-Less retrieval with Iterative Text Exploration
Zhangyu Wang, Siyuan Gao, Rong Zhou +2
Large Language Models (LLMs) have achieved impressive progress in natural language processing, but their limited ability to retain long-term context constrains performance on docum…