3 papers
cs.AI2026
Reasoning Concentrates Errors, and Self-Consistency Never Notices
Asaad Althoubi
Self-consistency assumes that independent samples disagree when a model is unsure, so agreement is evidence of correctness. Holding weights fixed and toggling only a reasoning mode…
cs.LG2026
Look Before You Leap: Pre-Action Verification for LLM Agents
Asaad Althoubi
An LLM agent acts on the world by emitting actions: shell commands to run, edits to apply. A wrong action does not always fail loudly; it can fail silently, producing a plausible b…
cs.LG2026
DistillCache: KL-Guided Adaptive KV-Cache Eviction for Memory-Efficient LLM Inference
Asaad Althoubi
Transformer-based large language models (LLMs) achieve strong performance across many tasks, but their Key-Value (KV) cache grows linearly with sequence length, creating a severe m…