2 papers
cs.NI2026
NetArena: Dynamic Benchmarks for AI Agents in Network Automation
Yajie Zhou, Jiajun Ruan, Eric S. Wang +4
As AI agents expand into high-stakes domains like network system operations, evaluating their real-world reliability becomes increasingly critical. However, existing benchmarks ris…
cs.LG2025
Glinthawk: A Two-Tiered Architecture for Offline LLM Inference
Pouya Hamadanian, Sadjad Fouladi
We introduce Glinthawk, an architecture for offline Large Language Model (LLM) inference. By leveraging a two-tiered structure, Glinthawk optimizes the utilization of the high-end…