1 paper · 1 filter
Kalle Kujanpää, Ning Liu, Shahnawaz Alam +4
Production LLM agents often waste latency and reliability by regenerating code for the same procedural steps on every request. We replace this inference-time coding loop with an ag…