English

A GPU accelerated Barnes-Hut Tree Code for FLASH4

Instrumentation and Methods for Astrophysics 2015-11-30 v2 Astrophysics of Galaxies Computational Physics

Abstract

We present a GPU accelerated CUDA-C implementation of the Barnes Hut (BH) tree code for calculating the gravitational potential on octree adaptive meshes. The tree code algorithm is implemented within the FLASH4 adaptive mesh refinement (AMR) code framework and therefore fully MPI parallel. We describe the algorithm and present test results that demonstrate its accuracy and performance in comparison to the algorithms available in the current FLASH4 version. We use a MacLaurin spheroid to test the accuracy of our new implementation and use spherical, collapsing cloud cores with effective AMR to carry out performance tests also in comparison with previous gravity solvers. Depending on the setup and the GPU/CPU ratio, we find a speedup for the gravity unit of at least a factor of 3 and up to 60 in comparison to the gravity solvers implemented in the FLASH4 code. We find an overall speedup factor for full simulations of at least factor 1.6 up to a factor of 10

Keywords

Cite

@article{arxiv.1509.07370,
  title  = {A GPU accelerated Barnes-Hut Tree Code for FLASH4},
  author = {Gunther Lukat and Robi Banerjee},
  journal= {arXiv preprint arXiv:1509.07370},
  year   = {2015}
}

Comments

For further information see: http://www.hs.uni-hamburg.de/gpubh

R2 v1 2026-06-22T11:04:35.555Z