3 papers
cs.DC2025
Dynamic Rebatching for Efficient Early-Exit Inference with DREX
Xuting Liu, Daniel Alexander, Siva Kesava Reddy Kakarla +2
Early-Exit (EE) is a Large Language Model (LLM) architecture that accelerates inference by allowing easier tokens to be generated using only a subset of the model's layers. However…
cs.AI2025
Robust Heuristic Algorithm Design with LLMs
Pantea Karimi, Dany Rouhana, Pooria Namyar +3
We posit that we can generate more robust and performant heuristics if we augment approaches using LLMs for heuristic design with tools that explain why heuristics underperform and…
cs.SE2025
Extremal Testing for Network Software using LLMs
Rathin Singha, Harry Qian, Srinath Saikrishnan +4
Physicists often manually consider extreme cases when testing a theory. In this paper, we show how to automate extremal testing of network software using LLMs in two steps: first,…