Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Look Before You Leap: Pre-Action Verification for LLM Agents
Asaad Althoubi
An LLM agent acts on the world by emitting actions: shell commands to run, edits to apply. A wrong action does not always fail loudly; it can fail silently, producing a plausible b…
cs.LG2026
DistillCache: KL-Guided Adaptive KV-Cache Eviction for Memory-Efficient LLM Inference
Asaad Althoubi
Transformer-based large language models (LLMs) achieve strong performance across many tasks, but their Key-Value (KV) cache grows linearly with sequence length, creating a severe m…