1 paper · 1 filter
Jiamu Zhang, Liang Wu, Kelly Wan +2
Large language models (LLMs) increasingly read long inputs in the agentic era, from whole documents and codebases to conversations across many turns. Their inference memory is then…