activity
20212026
most citedA Programming Model for GPU Load Balancing

11 citations · 22 across the 6 of their papers we have counts for

collaborators

8 papers

cs.DS2026

The Einsum-Enabled Design Space for Graph Algorithms: A BFS Case Study

Toluwanimi O. Odemuyiwa, Serban D. Porumbescu, Muhammad Osama +2

We propose a principled approach to reasoning about various graph algorithm implementations. We leverage the extended general Einsum notation (EDGE) which allows us to factor compl…

cs.GR2026

Fast Sparse Matrix Permutation for Mesh-Based Direct Solvers

Behrooz Zarebavami, Ahmed H. Mahmoud, Ana Dodik +5

We present a fast sparse matrix permutation algorithm tailored to linear systems arising from triangle meshes. Our approach produces nested-dissection-style permutations while sign…

cs.DC2023★ 1 cited

BOBA: A Parallel Lightweight Graph Reordering Algorithm with Heavyweight Implications

Matthew Drescher, Muhammad A. Awad, Serban D. Porumbescu +1

We describe a simple parallel-friendly lightweight graph reordering algorithm for COO graphs (edge lists). Our ``Batched Order By Attachment'' (BOBA) algorithm is linear in the num…

cs.DC2023★ 11 cited

A Programming Model for GPU Load Balancing

Muhammad Osama, Serban D. Porumbescu, John D. Owens

We propose a GPU fine-grained load-balancing abstraction that decouples load balancing from work processing and aims to support both static and dynamic schedules with a programmabl…

cs.DC2022★ 9 cited

Essentials of Parallel Graph Analytics

Muhammad Osama, Serban D. Porumbescu, John D. Owens

We identify the graph data structure, frontiers, operators, an iterative loop structure, and convergence conditions as essential components of graph analytics systems based on the…

cs.DC2021★ 1 cited

Atos: A Task-Parallel GPU Dynamic Scheduling Framework for Dynamic Irregular Computations

Yuxin Chen, Benjamin Brock, Serban Porumbescu +3

We present Atos, a task-parallel GPU dynamic scheduling framework that is especially suited to dynamic irregular applications. Compared to the dominant Bulk Synchronous Parallel (B…