Mathematical Software · Computer Science
Comparing OpenMP Implementations With Applications Across A64FX Platforms
Benjamin Michalowicz, Eric Raut, Yan Kang, Tony Curtis +2
2021-07-26
Performance · Computer Science
Investigating Applications on the A64FX
Adrian Jackson, Michèle Weiland, Nick Brown, Andrew Turner +1
2020-09-25
Instrumentation and Methods for Astrophysics · Physics
Ookami: An A64FX Computing Resource
A. C. Calder, E. Siegmann, C. Feldman, S. Chheda +7
2023-11-09
Distributed, Parallel, and Cluster Computing · Computer Science
Productivity meets Performance: Julia on A64FX
Mosè Giordano, Milan Klöwer, Valentin Churavy
2022-10-20
Distributed, Parallel, and Cluster Computing · Computer Science
Ookami: Deployment and Initial Experiences
Andrew Burford, Alan C. Calder, David Carlson, Barbara Chapman +14
2021-06-17
Hardware Architecture · Computer Science
A "New Ara" for Vector Computing: An Open Source Highly Efficient RISC-V V 1.0 Vector Processor Design
Matteo Perotti, Matheus Cavalcante, Nils Wistoff, Renzo Andri +2
2025-01-10
Performance · Computer Science
ECM modeling and performance tuning of SpMV and Lattice QCD on A64FX
Christie Alappat, Nils Meyer, Jan Laukemann, Thomas Gruber +3
2021-08-05
Distributed, Parallel, and Cluster Computing · Computer Science
mpiQulacs: A Distributed Quantum Computer Simulator for A64FX-based Cluster Systems
Satoshi Imamura, Masafumi Yamazaki, Takumi Honda, Akihiko Kasagi +4
2022-03-31
Distributed, Parallel, and Cluster Computing · Computer Science
Simulating Stellar Merger using HPX/Kokkos on A64FX on Supercomputer Fugaku
Patrick Diehl, Gregor Daiß, Kevin Huck, Dominic Marcello +3
2023-09-18
Distributed, Parallel, and Cluster Computing · Computer Science
Making use of supercomputers in financial machine learning
Philippe Cotte, Pierre Lagier, Vincent Margot, Christophe Geissler
2022-03-02
Distributed, Parallel, and Cluster Computing · Computer Science
Enabling OpenMP Task Parallelism on Multi-FPGAs
R. Nepomuceno, R. Sterle, G. Valarini, M. Pereira +2
2021-03-23
Distributed, Parallel, and Cluster Computing · Computer Science
Preparing for the Future -- Rethinking Proxy Apps
Satoshi Matsuoka, Jens Domke, Mohamed Wahib, Aleksandr Drozd +4
2022-04-18
Distributed, Parallel, and Cluster Computing · Computer Science
Evaluation of Programming Models and Performance for Stencil Computation on Current GPU Architectures
Baodi Shan, Mauricio Araya-Polo
2024-08-13
Distributed, Parallel, and Cluster Computing · Computer Science
Performance Evaluation of Parallel Sortings on the Supercomputer Fugaku
Tomoyuki Tokuue, Tomoaki Ishiyama
2023-09-08
Distributed, Parallel, and Cluster Computing · Computer Science
Benchmarking OpenCL, OpenACC, OpenMP, and CUDA: programming productivity, performance, and energy consumption
Suejb Memeti, Lu Li, Sabri Pllana, Joanna Kolodziej +1
2017-04-19
Distributed, Parallel, and Cluster Computing · Computer Science
Parallel FFTW on RISC-V: A Comparative Study including OpenMP, MPI, and HPX
Alexander Strack, Christopher Taylor, Dirk Pflüger
2025-06-11
Distributed, Parallel, and Cluster Computing · Computer Science
Portability and Scalability of OpenMP Offloading on State-of-the-art Accelerators
Yehonatan Fridman, Guy Tamir, Gal Oren
2023-05-16
Distributed, Parallel, and Cluster Computing · Computer Science
Comparison of OpenMP & OpenCL Parallel Processing Technologies
Krishnahari Thouti, S. R. Sathe
2012-11-12
Distributed, Parallel, and Cluster Computing · Computer Science
Benchmarking with Supernovae: A Performance Study of the FLASH Code
Joshua Martin, Catherine Feldman, Eva Siegmann, Tony Curtis +6
2024-08-30
Distributed, Parallel, and Cluster Computing · Computer Science
An efficient MPI/OpenMP parallelization of the Hartree-Fock method for the second generation of Intel Xeon Phi processor
Vladimir Mironov, Yuri Alexeev, Kristopher Keipert, Michael D'mello +2
2017-08-15
Distributed, Parallel, and Cluster Computing · Computer Science
A Further Study of Linux Kernel Hugepages on A64FX with FLASH, an Astrophysical Simulation Code
Catherine Feldman, Smeet Chheda, Alan C. Calder, Eva Siegmann +3
2023-09-19
Performance · Computer Science
Performance Modeling of Streaming Kernels and Sparse Matrix-Vector Multiplication on A64FX
Christie L. Alappat, Jan Laukemann, Thomas Gruber, Georg Hager +3
2021-08-05
Distributed, Parallel, and Cluster Computing · Computer Science
Optimizing the hybrid parallelization of BHAC
Salvatore Cielo, Oliver Porth, Luigi Iapichino, Anupam Karmakar +2
2021-08-30