English
Related papers

Related papers: N-Body Simulations on GPUs

200 papers

Matrix multiplication is a foundational operation in scientific computing and machine learning, yet its computational complexity makes it a significant bottleneck for large-scale applications. The shift to parallel architectures, primarily…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-07-30 Mufakir Qamar Ansari , Mudabir Qamar Ansari

General Purpose Graphics Processing Unit (GPGPU) computing plays a transformative role in deep learning and machine learning by leveraging the computational advantages of parallel processing. Through the power of Compute Unified Device…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-11-20 Ming Li , Ziqian Bi , Tianyang Wang , Yizhu Wen , Qian Niu , Xinyuan Song , Zekun Jiang , Junyu Liu , Benji Peng , Sen Zhang , Xuanhe Pan , Jiawei Xu , Jinlang Wang , Keyu Chen , Caitlyn Heqi Yin , Pohsun Feng , Ming Liu

We present a GPU-accelerated cosmological simulation code, PhotoNs-GPU, based on algorithm of Particle Mesh Fast Multipole Method (PM-FMM), and focus on the GPU utilization and optimization. A proper interpolated method for truncated…

Instrumentation and Methods for Astrophysics · Physics 2021-12-28 Qiao Wang , Chen Meng

A modern graphics processing unit (GPU) is able to perform massively parallel scientific computations at low cost. We extend our implementation of the checkerboard algorithm for the two dimensional Ising model [T. Preis et al., J. Comp.…

Computational Physics · Physics 2010-07-22 Benjamin Block , Peter Virnau , Tobias Preis

We investigate the performance of Opticks, a NVIDIA OptiX API 7.5 GPU-accelerated photon propagation tool compared with a single-threaded Geant4 simulation. We compare the simulations using an improved model of the NEXT-CRAB-0 gaseous time…

Instrumentation and Detectors · Physics 2025-11-20 NEXT Collaboration , I. Parmaksiz , K. Mistry , E. Church , C. Adams , J. Asaadi , J. Baeza-Rubio , K. Bailey , N. Byrnes , B. J. P. Jones , I. A. Moya , K. E. Navarro , D. R. Nygren , P. Oyedele , L. Rogers , F. Samaniego , K. Stogsdill , H. Almazán , V. Álvarez , B. Aparicio , A. I. Aranburu , L. Arazi , I. J. Arnquist , F. Auria-Luna , S. Ayet , C. D. R. Azevedo , F. Ballester , M. del Barrio-Torregrosa , A. Bayo , J. M. Benlloch-Rodríguez , F. I. G. M. Borges , A. Brodolin , S. Cárcel , A. Castillo , L. Cid , C. A. N. Conde , T. Contreras , F. P. Cossío , R. Coupe , E. Dey , G. Díaz , C. Echevarria , M. Elorza , J. Escada , R. Esteve , R. Felkai , L. M. P. Fernandes , P. Ferrario , A. L. Ferreira , F. W. Foss , Z. Freixa , J. García-Barrena , J. J. Gómez-Cadenas , J. W. R. Grocott , R. Guenette , J. Hauptman , C. A. O. Henriques , J. A. Hernando Morata , P. Herrero-Gómez , V. Herrero , C. Hervés Carrete , Y. Ifergan , F. Kellerer , L. Larizgoitia , A. Larumbe , P. Lebrun , F. Lopez , N. López-March , R. Madigan , R. D. P. Mano , A. P. Marques , J. Martín-Albo , G. Martínez-Lema , M. Martínez-Vara , R. L. Miller , J. Molina-Canteras , F. Monrabal , C. M. B. Monteiro , F. J. Mora , P. Novella , A. Nuñez , E. Oblak , J. Palacio , B. Palmeiro , A. Para , A. Pazos , J. Pelegrin , M. Pérez Maneiro , M. Querol , J. Renner , I. Rivilla , C. Rogero , B. Romeo , C. Romo-Luque , V. San Nacienciano , F. P. Santos , J. M. F. dos Santos , M. Seemann , I. Shomroni , P. A. O. C. Silva , A. Simón , S. R. Soleti , M. Sorel , J. Soto-Oton , J. M. R. Teixeira , S. Teruel-Pardo , J. F. Toledo , C. Tonnelé , S. Torelli , J. Torrent , A. Trettin , A. Usón , P. R. G. Valle , J. F. C. A. Veloso , J. Waiton , A. Yubero-Navarro

Energy-efficiency is a key concern for neural network applications. To alleviate this issue, hardware acceleration using FPGAs or GPUs can provide better energy-efficiency than general-purpose processors. However, further improvement of the…

Distributed, Parallel, and Cluster Computing · Computer Science 2021-06-29 Seyed Morteza Nabavinejad , Behzad Salami

The availability of low cost sensors has led to an unprecedented growth in the volume of spatial data. However, the time required to evaluate even simple spatial queries over large data sets greatly hampers our ability to interactively…

Databases · Computer Science 2020-04-09 Harish Doraiswamy , Juliana Freire

One of the current challenges in physically-based simulations, and, more specifically, fluid simulations, is to produce visually appealing results at interactive rates, capable of being used in multiple forms of media. In recent times, a…

Graphics · Computer Science 2024-04-17 Pedro Centeno , João Madeiras Pereira

We use the graphics processing unit (GPU) for fast calculations of helicity amplitudes of physics processes. As our first attempt, we compute $u\bar{u}\to n\gamma$ ($n=2$ to 8) processes in $pp$ collisions at $\sqrt{s} = 14$TeV by…

Computational Physics · Physics 2010-10-12 K. Hagiwara , J. Kanzaki , N. Okamura , D. Rainwater , T. Stelzer

We present an approach to molecular-dynamics simulations of ferrofluids on graphics processing units (GPUs). Our numerical scheme is based on a GPU-oriented modification of the Barnes-Hut (BH) algorithm designed to increase the parallelism…

Computational Physics · Physics 2013-04-30 A. Yu. Polyakov , T. V. Lyutyy , S. Denisov , V. V. Reva , P. Hanggi

This paper represents the first investigation of the suitability and performance of Graphcore Intelligence Processing Units (IPUs) for deep learning applications in cosmology. It presents the benchmark between a Nvidia V100 GPU and a…

Computational Physics · Physics 2021-06-07 Bastien Arcelin

GPU computing is popular due to the calculation potential of a single card. The N-body integrator GENGA is built to for this, but it suffers a performance penalty on consumer-grade GPUs due to their truncated double precision (FP64)…

Earth and Planetary Astrophysics · Physics 2023-09-18 R. Brasser , S. L. Grimm , P. Hatalova , J. G. Stadel

The modern trend in High-Performance Computing (HPC) involves the use of accelerators such as Graphics Processing Units (GPUs) alongside Central Processing Units (CPUs) to speed up numerical operations in various applications. Leading…

Mathematical Software · Computer Science 2025-07-25 Giulio Malenza , Giovanni Stabile , Filippo Spiga , Robert Birke , Marco Aldinucci

Parallel computing can offer an enormous advantage regarding the performance for very large applications in almost any field: scientific computing, computer vision, databases, data mining, and economics. GPUs are high performance many-core…

Distributed, Parallel, and Cluster Computing · Computer Science 2015-11-24 Bogdan Oancea , Tudorel Andrei , Raluca Mariana Dragoescu

This work studies the porting and optimization of the tensor network simulator QTensor on GPUs, with the ultimate goal of simulating quantum circuits efficiently at scale on large GPU supercomputers. We implement NumPy, PyTorch, and CuPy…

Quantum Physics · Physics 2022-04-14 Danylo Lykov , Angela Chen , Huaxuan Chen , Kristopher Keipert , Zheng Zhang , Tom Gibbs , Yuri Alexeev

Particle tracking simulations with space charge effects are very important for high-intensity proton rings. Since they include not only Hamilton mechanics of a single particle but constructing charge densities and solving Poisson equations…

Accelerator Physics · Physics 2021-09-01 Yoshinori Kurimoto

Study of general purpose computation by GPU (Graphics Processing Unit) can improve the image processing capability of micro-computer system. This paper studies the parallelism of the different stages of decimation in time radix 2 FFT…

Mathematical Software · Computer Science 2015-06-01 Feifei Shen , Zhenjian Song , Congrui Wu , Jiaqi Geng , Qingyun Wang

Scientific computing in the exascale era demands increased computational power to solve complex problems across various domains. With the rise of heterogeneous computing architectures the need for vendor-agnostic, performance portability…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-11-05 Johansell Villalobos , Josef Ruzicka , Silvio Rizzi

In this work, we survey the role of GPUs in real-time systems. Originally designed for parallel graphics workloads, GPUs are now widely used in time-critical applications such as machine learning, autonomous vehicles, and robotics due to…

It is shown micromagnetic and atomistic spin dynamics simulations can use multiple GPUs in order to reduce computation time, but also to allow for a larger simulation size than is possible on a single GPU. Whilst interactions which depend…

Mesoscale and Nanoscale Physics · Physics 2023-10-12 Serban Lepadatu