Distributed, Parallel, and Cluster Computing · Computer Science
Simulating Stellar Merger using HPX/Kokkos on A64FX on Supercomputer Fugaku
Patrick Diehl, Gregor Daiß, Kevin Huck, Dominic Marcello +3
2023-09-18
Distributed, Parallel, and Cluster Computing · Computer Science
Asynchronous-Many-Task Systems: Challenges and Opportunities -- Scaling an AMR Astrophysics Code on Exascale machines using Kokkos and HPX
Gregor Daiß, Patrick Diehl, Jiakun Yan, John K. Holmen +8
2025-09-29
Distributed, Parallel, and Cluster Computing · Computer Science
Stellar Mergers with HPX-Kokkos and SYCL: Methods of using an Asynchronous Many-Task Runtime System with SYCL
Gregor Daiß, Patrick Diehl, Hartmut Kaiser, Dirk Pflüger
2023-05-10
Distributed, Parallel, and Cluster Computing · Computer Science
From Task-Based GPU Work Aggregation to Stellar Mergers: Turning Fine-Grained CPU Tasks into Portable GPU Kernels
Gregor Daiß, Patrick Diehl, Dominic Marcello, Alireza Kheirkhahan +2
2023-03-07
Instrumentation and Methods for Astrophysics · Physics
Octo-Tiger: A New, 3D Hydrodynamic Code for Stellar Mergers that uses HPX Parallelisation
Dominic C. Marcello, Sagiv Shiber, Orsola De Marco, Juhan Frank +4
2021-08-12
Distributed, Parallel, and Cluster Computing · Computer Science
From Piz Daint to the Stars: Simulation of Stellar Mergers using High-Level Abstractions
Gregor Daiß, Parsa Amini, John Biddiscombe, Patrick Diehl +6
2019-11-28
Distributed, Parallel, and Cluster Computing · Computer Science
Evaluating HPX and Kokkos on RISC-V using an Astrophysics Application Octo-Tiger
Parick Diehl, Gregor Daiss, Steven R. Brandt, Alireza Kheirkhahan +3
2023-09-20
Distributed, Parallel, and Cluster Computing · Computer Science
Octo-Tiger's New Hydro Module and Performance Using HPX+CUDA on ORNL's Summit
Patrick Diehl, Gregor Daiß, Dominic Marcello, Kevin Huck +4
2021-10-22
Distributed, Parallel, and Cluster Computing · Computer Science
Distributed, combined CPU and GPU profiling within HPX using APEX
Patrick Diehl, Gregor Daiss, Kevin Huck, Dominic Marcello +5
2022-10-13
Distributed, Parallel, and Cluster Computing · Computer Science
Preparing for HPC on RISC-V: Examining Vectorization and Distributed Performance of an Astrophyiscs Application with HPX and Kokkos
Patrick Diehl, Panagiotis Syskakis, Gregor Daiß, Steven R. Brandt +6
2025-01-15
Performance · Computer Science
Performance Measurements within Asynchronous Task-based Runtime Systems: A Double White Dwarf Merger as an Application
Patrick Diehl, Dominic Marcello, Parsa Amini, Hartmut Kaiser +8
2021-06-10
Distributed, Parallel, and Cluster Computing · Computer Science
Comparing the Performance of Different x86 SIMD Instruction Sets for a Medical Imaging Application on Modern Multi- and Manycore Chips
Johannes Hofmann, Jan Treibig, Georg Hager, Gerhard Wellein
2014-01-30
Distributed, Parallel, and Cluster Computing · Computer Science
Accelerating X-Ray Tracing for Exascale Systems using Kokkos
Felix Wittwer, Nicholas K. Sauter, Derek Mendez, Billy K. Poon +6
2022-05-18
Solar and Stellar Astrophysics · Physics
Hydrodynamic simulations of WD-WD mergers and the origin of RCB stars
Sagiv Shiber, Orsola De Marco, Patrick M. Motl, Bradley Munson +9
2024-11-08
Distributed, Parallel, and Cluster Computing · Computer Science
Leveraging SIMD for Accelerating Large-number Arithmetic
Subhrajit Das, Abhishek Bichhawat, Yuvraj Patel
2026-04-27
Instrumentation and Methods for Astrophysics · Physics
Astrophysical Particle Simulations on Heterogeneous CPU-GPU Systems
Naohito Nakasato, Go Ogiya, Yohei Miki, Masao Mori +1
2012-06-07
Distributed, Parallel, and Cluster Computing · Computer Science
Exploiting long vectors with a CFD code: a co-design show case
Marc Blancafort, Roger Ferrer, Guillaume Houzeaux, Marta Garcia-Gasulla +1
2024-11-05
Materials Science · Physics
A Kokkos-Accelerated Moment Tensor Potential Implementation for LAMMPS
Zijian Meng, Karim Zongo, Edmanuel Torres, Christopher Maxwell +2
2025-10-02
Performance · Computer Science
A high-performance and portable implementation of the SISSO method for CPUs and GPUs
Sebastian Eibl, Yi Yao, Matthias Scheffler, Markus Rampp +2
2025-02-28
Distributed, Parallel, and Cluster Computing · Computer Science
Performance Engineering for a Medical Imaging Application on the Intel Xeon Phi Accelerator
Johannes Hofmann, Jan Treibig, Georg Hager, Gerhard Wellein
2014-01-16
Distributed, Parallel, and Cluster Computing · Computer Science
MMStencil: Optimizing High-order Stencils on Multicore CPU using Matrix Unit
Yinuo Wang, Tianqi Mao, Lin Gan, Wubing Wan +7
2025-07-16
Distributed, Parallel, and Cluster Computing · Computer Science
Rapid Exploration of Optimization Strategies on Advanced Architectures using TestSNAP and LAMMPS
Rahulkumar Gayatri, Stan Moore, Evan Weinberg, Nicholas Lubbers +4
2020-11-26
Computational Physics · Physics
Porting CMS Heterogeneous Pixel Reconstruction to Kokkos
Taylor Childers, Matti J. Kortelainen, Martin Kwok, Alexei Strelchenko +1
2021-04-15
Hardware Architecture · Computer Science
Exploring the Efficiency of 3D-Stacked AI Chip Architecture for LLM Inference with Voxel
Yiqi Liu, Noelle Crawford, Michael Wang, Jilong Xue +1
2026-04-30