19 citations · 26 across the 5 of their papers we have counts for
14 papers · 1 filter
Workload-Aware DRAM Error Prediction using Machine Learning
Lev Mukhanov, Konstantinos Tovletoglou, Hans Vandierendonck +2
The aggressive scaling of technology may have helped to meet the growing demand for higher memory capacity and density, but has also made DRAM cells more prone to errors. Such a re…
Implementing Efficient Message Logging Protocols as MPI Application Extensions
Kiril Dichev, Dimitrios S. Nikolopoulos
Message logging protocols are enablers of local rollback, a more efficient alternative to global rollback, for fault tolerant MPI applications. Until now, message logging MPI imple…
RADS: Real-time Anomaly Detection System for Cloud Data Centres
Sakil Barbhuiya, Zafeirios Papazachos, Peter Kilpatrick +1
Cybersecurity attacks in Cloud data centres are increasing alongside the growth of the Cloud services market. Existing research proposes a number of anomaly detection systems for d…
VEBO: A Vertex- and Edge-Balanced Ordering Heuristic to Load Balance Parallel Graph Processing
Jiawen Sun, Hans Vandierendonck, Dimitrios S. Nikolopoulos
Graph partitioning drives graph processing in distributed, disk-based and NUMA-aware systems. A commonly used partitioning goal is to balance the number of edges per partition in c…
Energy-efficient localised rollback after failures via data flow analysis
Kiril Dichev, Kirk Cameron, Dimitrios Nikolopoulos
Exascale systems will suffer failures hourly. HPC programmers rely mostly on application-level checkpoint and a global rollback to recover. In recent years, techniques reducing the…
Intra-node Memory Safe GPU Co-Scheduling
Carlos Reano, Federico Silla, Dimitrios S. Nikolopoulos +1
GPUs in High-Performance Computing systems remain under-utilised due to the unavailability of schedulers that can safely schedule multiple applications to share the same GPU. The r…