collaborators

12 papers

cs.AR2026

MailoHLS: Multi-Adapter Structure-Aware Learning for Pareto-Driven HLS Pragma Optimization

Elena Vouvali, Dimosthenis Masouros, Aggelos Ferikoglou +2

High-Level Synthesis (HLS) enables rapid development of FPGA accelerators, yet achieving high-quality results (QoR) remains challenging due to the large and irregular design space…

cs.DC2026

Multi-Partner Project: Multi-GPU Performance Portability Analysis for CFD Simulations at Scale

Panagiotis-Eleftherios Eleftherakis, George Anagnostopoulos, Anastassis Kapetanakis +10

As heterogeneous supercomputing architectures leveraging GPUs become increasingly central to high-performance computing (HPC), it is crucial for computational fluid dynamics (CFD)…

cs.LG2025

Neural expressiveness for beyond importance model compression

Angelos-Christos Maroudis, Sotirios Xydis

Neural Network Pruning has been established as driving force in the exploration of memory and energy efficient solutions with high throughput both during training and at test time.…

cs.DC2025

SLO-aware GPU Frequency Scaling for Energy Efficient LLM Inference Serving

Andreas Kosmas Kakolyris, Dimosthenis Masouros, Petros Vavaroutsos +2

As Large Language Models (LLMs) gain traction, their reliance on power-hungry GPUs places ever-increasing energy demands, raising environmental and monetary concerns. Inference dom…

cs.LG2025

MaRVIn: A Cross-Layer Mixed-Precision RISC-V Framework for DNN Inference, from ISA Extension to Hardware Acceleration

Giorgos Armeniakos, Alexis Maras, Sotirios Xydis +1

The evolution of quantization and mixed-precision techniques has unlocked new possibilities for enhancing the speed and energy efficiency of NNs. Several recent studies indicate th…

cs.DC2025

SynergAI: Edge-to-Cloud Synergy for Architecture-Driven High-Performance Orchestration for AI Inference

Foteini Stathopoulou, Aggelos Ferikoglou, Manolis Katsaragakis +3

The rapid evolution of Artificial Intelligence (AI) and Machine Learning (ML) has significantly heightened computational demands, particularly for inference-serving workloads. Whil…