40 citations · 68 across the 5 of their papers we have counts for
9 papers
Toward an End-to-End Auto-tuning Framework in HPC PowerStack
Xingfu Wu, Aniruddha Marathe, Siddhartha Jana +7
Efficiently utilizing procured power and optimizing performance of scientific applications under power and energy constraints are challenging. The HPC PowerStack defines a software…
A Machine Learning Approach to Online Fault Classification in HPC Systems
Alessio Netti, Zeynep Kiziltan, Ozalp Babaoglu +3
As High-Performance Computing (HPC) systems strive towards the exascale goal, failure rates both at the hardware and software levels will increase significantly. Thus, detecting an…
pAElla: Edge-AI based Real-Time Malware Detection in Data Centers
Antonio Libri, Andrea Bartolini, Luca Benini
The increasing use of Internet-of-Things (IoT) devices for monitoring a wide spectrum of applications, along with the challenges of "big data" streaming support they often require…
COUNTDOWN Slack: a Run-time Library to Reduce Energy Footprint in Large-scale MPI Applications
Daniele Cesarini, Andrea Bartolini, Andrea Borghesi +3
The power consumption of supercomputers is a major challenge for system owners, users, and society. It limits the capacity of system installations, it requires large cooling infras…
Online Anomaly Detection in HPC Systems
Andrea Borghesi, Antonio Libri, Luca Benini +1
Reliability is a cumbersome problem in High Performance Computing Systems and Data Centers evolution. During operation, several types of fault conditions or anomalies can arise, ra…
A LoRaWAN Wireless Sensor Network for Data Center Temperature Monitoring
Tommaso Polonelli, Davide Brunelli, Andrea Bartolini +1
High-performance computing installations, which are at the basis of web and cloud servers as well as supercomputers, are constrained by two main conflicting requirements: IT power…