5 papers
iDDS: Intelligent Distributed Dispatch and Scheduling for Workflow Orchestration
Wen Guan, Tadashi Maeno, Aleksandr Alekseev +10
The intelligent Distributed Dispatch and Scheduling (iDDS) service is a versatile workflow orchestration system designed for large-scale, distributed scientific computing. iDDS ext…
Machine Learning-Driven Predictive Resource Management in Complex Science Workflows
Tasnuva Chowdhury, Tadashi Maeno, Fatih Furkan Akman +23
The collaborative efforts of large communities in science experiments, often comprising thousands of global members, reflect a monumental commitment to exploration and discovery. R…
Towards an Introspective Dynamic Model of Globally Distributed Computing Infrastructures
Ozgur O. Kilic, David K. Park, Yihui Ren +18
Large-scale scientific collaborations like ATLAS, Belle II, CMS, DUNE, and others involve hundreds of research institutes and thousands of researchers spread across the globe. Thes…
Alternative Mixed Integer Linear Programming Optimization for Joint Job Scheduling and Data Allocation in Grid Computing
Shengyu Feng, Jaehyung Kim, Yiming Yang +18
This paper presents a novel approach to the joint optimization of job scheduling and data allocation in grid computing environments. We formulate this joint optimization problem as…
AI Surrogate Model for Distributed Computing Workloads
David K. Park, Yihui Ren, Ozgur O. Kilic +18
Large-scale international scientific collaborations, such as ATLAS, Belle II, CMS, and DUNE, generate vast volumes of data. These experiments necessitate substantial computational…