2 papers
cs.SE2026
Attention to Detail: Evaluating Energy, Performance, and Accuracy Trade-offs Across vLLM Configurations
Nada Zine, Tristan Coignion, Vincenzo Stoico +4
Large Language Models are reshaping how software is developed and maintained. They are typically deployed in production using inference engines such as vLLM, which can efficiently…
cs.LG2026
Pimp My LLM: Leveraging Variability Modeling to Tune Inference Hyperparameters
Nada Zine, Clément Quinton, Romain Rouvoy
Large Language Models (LLMs) are being increasingly used across a wide range of tasks. However, their substantial computational demands raise concerns about the energy efficiency a…