11 papers
ADASCALE: An Adaptive Scaling and Placement Framework for Microservices Under Dynamics
Ming Chen, Muhammed Tawfiqul Islam, Maria Rodriguez Read +1
Microservice applications are increasingly deployed across cloud--edge environments, where heterogeneous nodes and time-varying inter-node delays amplify the impact of placement de…
iDynamics: A Configurable Emulation Framework for Evaluating Microservice Scheduling Policies under Controllable Cloud-Edge Dynamics
Ming Chen, Muhammed Tawfiqul Islam, Maria Rodriguez Read +1
This paper presents iDynamics, a configurable emulation framework that exposes these dynamics as controllable experimental factors while running real microservice code on a Kuberne…
GraphFlash: Enabling Fast and Elastic Graph Processing on Serverless Infrastructure
Chen Zhao, Parsa Poorsistani, Mohammad Goudarzi +2
Graph processing systems are essential for analyzing large-scale data with complex relationships, yet most existing frameworks rely on statically provisioned clusters, resulting in…
Adaptive Management of Microservices in Dynamic Computing Environments: A Taxonomy and Future Directions
Ming Chen, Muhammed Tawfiqul Islam, Maria Rodriguez Read +1
Microservice-based cloud applications face changing workloads, evolving request paths, variable network conditions, interference, and failures. These dynamics couple autoscaling, p…
PipeLive: Efficient Live In-place Pipeline Parallelism Reconfiguration for Dynamic LLM Serving
Xu Bai, Muhammed Tawfiqul Islam, Chen Wang +1
Pipeline parallelism (PP) is widely used to partition layers of large language models (LLMs) across GPUs, enabling scalable inference for large models. However, existing systems re…
LLM-Driven Intent-Based Privacy-Aware Orchestration Across the Cloud-Edge Continuum
Zijie Su, Muhammed Tawfiqul Islam, Mohammad Goudarzi +1
With the rapid advancement of large language models (LLMs), efficiently serving LLM inference under limited GPU resources has become a critical challenge. Recently, an increasing n…