Elasticity in Parallel Sparse Triangular Solve
Distributed, Parallel, and Cluster Computing
2026-07-02 v1
Abstract
We introduce stale synchronous parallel as a mode of execution in parallel sparse triangular linear system solve and present a general directed-acyclic-graph scheduler capable of producing such schedules. Stale-synchronous-parallel schedules allow the overlap of synchronisation and compute which results in a geometric-mean speed-up of - of our scheduler, ElasticDivide, over state-of-the-art synchronous scheduler GrowLocal on an ARM machine using 48 cores. On an x86 machine using 48 cores, we report geometric-mean speed-ups of - over SpMP.
Cite
@article{arxiv.2607.02324,
title = {Elasticity in Parallel Sparse Triangular Solve},
author = {Raphael S. Steiner and Christos K. Matzoros and Pál András Papp and Toni Böhnlein and A. N. Yzelman},
journal= {arXiv preprint arXiv:2607.02324},
year = {2026}
}
Comments
23 pages