4 papers
Beyond Test-Time Compute Strategies: Advocating Energy-per-Token in LLM Inference
Patrick Wilhelm, Thorsten Wittkopp, Odej Kao
Large Language Models (LLMs) demonstrate exceptional performance across diverse tasks but come with substantial energy and computational costs, particularly in request-heavy scenar…
Monitoring Emergent Reward Hacking During Generation via Internal Activations
Patrick Wilhelm, Thorsten Wittkopp, Odej Kao
Fine-tuned large language models can exhibit reward-hacking behavior arising from emergent misalignment, which is difficult to detect from final outputs alone. While prior work has…
Beyond Microservices: Testing Web-Scale RCA Methods on GPU-Driven LLM Workloads
Dominik Scheinert, Alexander Acker, Thorsten Wittkopp +6
Large language model (LLM) services have become an integral part of search, assistance, and decision-making applications. However, unlike traditional web or microservices, the hard…
A layered architecture for log analysis in complex IT systems
Thorsten Wittkopp
In the evolving IT landscape, stability and reliability of systems are essential, yet their growing complexity challenges DevOps teams in implementation and maintenance. Log analys…