6 papers
Scalable AI-assisted Workflow Management for Detector Design Optimization Using Distributed Computing
Derek Anderson, Amit Bashyal, Markus Diefenthaler +13
The Production and Distributed Analysis (PanDA) system, originally developed for the ATLAS experiment at the CERN Large Hadron Collider (LHC), has evolved into a robust platform fo…
iDDS: Intelligent Distributed Dispatch and Scheduling for Workflow Orchestration
Wen Guan, Tadashi Maeno, Aleksandr Alekseev +10
The intelligent Distributed Dispatch and Scheduling (iDDS) service is a versatile workflow orchestration system designed for large-scale, distributed scientific computing. iDDS ext…
Data Management System Analysis for Distributed Computing Workloads
Kuan-Chieh Hsu, Sairam Sri Vatsavai, Ozgur O. Kilic +20
Large-scale international collaborations such as ATLAS rely on globally distributed workflows and data management to process, move, and store vast volumes of data. ATLAS's Producti…
CGSim: A Simulation Framework for Large Scale Distributed Computing Environment
Sairam Sri Vatsavai, Raees Khan, Kuan-Chieh Hsu +20
Large-scale distributed computing infrastructures such as the Worldwide LHC Computing Grid (WLCG) require comprehensive simulation tools for evaluating performance, testing new alg…
Towards an Introspective Dynamic Model of Globally Distributed Computing Infrastructures
Ozgur O. Kilic, David K. Park, Yihui Ren +18
Large-scale scientific collaborations like ATLAS, Belle II, CMS, DUNE, and others involve hundreds of research institutes and thousands of researchers spread across the globe. Thes…
Alternative Mixed Integer Linear Programming Optimization for Joint Job Scheduling and Data Allocation in Grid Computing
Shengyu Feng, Jaehyung Kim, Yiming Yang +18
This paper presents a novel approach to the joint optimization of job scheduling and data allocation in grid computing environments. We formulate this joint optimization problem as…