2 papers
cs.AI2026
Don't Overthink It: Inter-Rollout Action Agreement as a Free Adaptive-Compute Signal for LLM Agents
Khushal Sethi
Inference-time compute scaling has emerged as a powerful technique for improving the reliability of large language model (LLM) agents, but existing methods apply compute uniformly:…
cs.NI2020
NV-Fogstore : Device-aware hybrid caching in fog computing environments
Khushal Sethi, Manan Suri
Edge caching via the placement of distributed storages throughout the network is a promising solution to reduce latency and network costs of content delivery. With the advent of th…