4 papers · 1 filter
Multi-Resource List Scheduling of Moldable Parallel Jobs under Precedence Constraints
Lucas Perotin, Hongyang Sun, Padma Raghavan
The scheduling literature has traditionally focused on a single type of resource (e.g., computing nodes). However, scientific applications in modern High-Performance Computing (HPC…
Deep-Edge: An Efficient Framework for Deep Learning Model Update on Heterogeneous Edge
Anirban Bhattacharjee, Ajay Dev Chhokra, Hongyang Sun +4
Deep Learning (DL) model-based AI services are increasingly offered in a variety of predictive analytics services such as computer vision, natural language processing, speech recog…
FECBench: A Holistic Interference-aware Approach for Application Performance Modeling
Yogesh D. Barve, Shashank Shekhar, Ajay Dev Chhokra +5
Services hosted in multi-tenant cloud platforms often encounter performance interference due to contention for non-partitionable resources, which in turn causes unpredictable behav…
BARISTA: Efficient and Scalable Serverless Serving System for Deep Learning Prediction Services
Anirban Bhattacharjee, Ajay Dev Chhokra, Zhuangwei Kang +3
Pre-trained deep learning models are increasingly being used to offer a variety of compute-intensive predictive analytics services such as fitness tracking, speech and image recogn…