中文
相关论文

相关论文: N-Body Simulations on GPUs

200 篇论文

Matrix multiplication is a foundational operation in scientific computing and machine learning, yet its computational complexity makes it a significant bottleneck for large-scale applications. The shift to parallel architectures, primarily…

分布式、并行与集群计算 · 计算机科学 2025-07-30 Mufakir Qamar Ansari , Mudabir Qamar Ansari

General Purpose Graphics Processing Unit (GPGPU) computing plays a transformative role in deep learning and machine learning by leveraging the computational advantages of parallel processing. Through the power of Compute Unified Device…

We present a GPU-accelerated cosmological simulation code, PhotoNs-GPU, based on algorithm of Particle Mesh Fast Multipole Method (PM-FMM), and focus on the GPU utilization and optimization. A proper interpolated method for truncated…

天体物理仪器与方法 · 物理学 2021-12-28 Qiao Wang , Chen Meng

A modern graphics processing unit (GPU) is able to perform massively parallel scientific computations at low cost. We extend our implementation of the checkerboard algorithm for the two dimensional Ising model [T. Preis et al., J. Comp.…

计算物理 · 物理学 2010-07-22 Benjamin Block , Peter Virnau , Tobias Preis

We investigate the performance of Opticks, a NVIDIA OptiX API 7.5 GPU-accelerated photon propagation tool compared with a single-threaded Geant4 simulation. We compare the simulations using an improved model of the NEXT-CRAB-0 gaseous time…

仪器与探测器 · 物理学 2025-11-20 NEXT Collaboration , I. Parmaksiz , K. Mistry , E. Church , C. Adams , J. Asaadi , J. Baeza-Rubio , K. Bailey , N. Byrnes , B. J. P. Jones , I. A. Moya , K. E. Navarro , D. R. Nygren , P. Oyedele , L. Rogers , F. Samaniego , K. Stogsdill , H. Almazán , V. Álvarez , B. Aparicio , A. I. Aranburu , L. Arazi , I. J. Arnquist , F. Auria-Luna , S. Ayet , C. D. R. Azevedo , F. Ballester , M. del Barrio-Torregrosa , A. Bayo , J. M. Benlloch-Rodríguez , F. I. G. M. Borges , A. Brodolin , S. Cárcel , A. Castillo , L. Cid , C. A. N. Conde , T. Contreras , F. P. Cossío , R. Coupe , E. Dey , G. Díaz , C. Echevarria , M. Elorza , J. Escada , R. Esteve , R. Felkai , L. M. P. Fernandes , P. Ferrario , A. L. Ferreira , F. W. Foss , Z. Freixa , J. García-Barrena , J. J. Gómez-Cadenas , J. W. R. Grocott , R. Guenette , J. Hauptman , C. A. O. Henriques , J. A. Hernando Morata , P. Herrero-Gómez , V. Herrero , C. Hervés Carrete , Y. Ifergan , F. Kellerer , L. Larizgoitia , A. Larumbe , P. Lebrun , F. Lopez , N. López-March , R. Madigan , R. D. P. Mano , A. P. Marques , J. Martín-Albo , G. Martínez-Lema , M. Martínez-Vara , R. L. Miller , J. Molina-Canteras , F. Monrabal , C. M. B. Monteiro , F. J. Mora , P. Novella , A. Nuñez , E. Oblak , J. Palacio , B. Palmeiro , A. Para , A. Pazos , J. Pelegrin , M. Pérez Maneiro , M. Querol , J. Renner , I. Rivilla , C. Rogero , B. Romeo , C. Romo-Luque , V. San Nacienciano , F. P. Santos , J. M. F. dos Santos , M. Seemann , I. Shomroni , P. A. O. C. Silva , A. Simón , S. R. Soleti , M. Sorel , J. Soto-Oton , J. M. R. Teixeira , S. Teruel-Pardo , J. F. Toledo , C. Tonnelé , S. Torelli , J. Torrent , A. Trettin , A. Usón , P. R. G. Valle , J. F. C. A. Veloso , J. Waiton , A. Yubero-Navarro

Energy-efficiency is a key concern for neural network applications. To alleviate this issue, hardware acceleration using FPGAs or GPUs can provide better energy-efficiency than general-purpose processors. However, further improvement of the…

分布式、并行与集群计算 · 计算机科学 2021-06-29 Seyed Morteza Nabavinejad , Behzad Salami

The availability of low cost sensors has led to an unprecedented growth in the volume of spatial data. However, the time required to evaluate even simple spatial queries over large data sets greatly hampers our ability to interactively…

数据库 · 计算机科学 2020-04-09 Harish Doraiswamy , Juliana Freire

One of the current challenges in physically-based simulations, and, more specifically, fluid simulations, is to produce visually appealing results at interactive rates, capable of being used in multiple forms of media. In recent times, a…

图形学 · 计算机科学 2024-04-17 Pedro Centeno , João Madeiras Pereira

We use the graphics processing unit (GPU) for fast calculations of helicity amplitudes of physics processes. As our first attempt, we compute $u\bar{u}\to n\gamma$ ($n=2$ to 8) processes in $pp$ collisions at $\sqrt{s} = 14$TeV by…

计算物理 · 物理学 2010-10-12 K. Hagiwara , J. Kanzaki , N. Okamura , D. Rainwater , T. Stelzer

We present an approach to molecular-dynamics simulations of ferrofluids on graphics processing units (GPUs). Our numerical scheme is based on a GPU-oriented modification of the Barnes-Hut (BH) algorithm designed to increase the parallelism…

计算物理 · 物理学 2013-04-30 A. Yu. Polyakov , T. V. Lyutyy , S. Denisov , V. V. Reva , P. Hanggi

This paper represents the first investigation of the suitability and performance of Graphcore Intelligence Processing Units (IPUs) for deep learning applications in cosmology. It presents the benchmark between a Nvidia V100 GPU and a…

计算物理 · 物理学 2021-06-07 Bastien Arcelin

GPU computing is popular due to the calculation potential of a single card. The N-body integrator GENGA is built to for this, but it suffers a performance penalty on consumer-grade GPUs due to their truncated double precision (FP64)…

地球与行星天体物理 · 物理学 2023-09-18 R. Brasser , S. L. Grimm , P. Hatalova , J. G. Stadel

The modern trend in High-Performance Computing (HPC) involves the use of accelerators such as Graphics Processing Units (GPUs) alongside Central Processing Units (CPUs) to speed up numerical operations in various applications. Leading…

数学软件 · 计算机科学 2025-07-25 Giulio Malenza , Giovanni Stabile , Filippo Spiga , Robert Birke , Marco Aldinucci

Parallel computing can offer an enormous advantage regarding the performance for very large applications in almost any field: scientific computing, computer vision, databases, data mining, and economics. GPUs are high performance many-core…

分布式、并行与集群计算 · 计算机科学 2015-11-24 Bogdan Oancea , Tudorel Andrei , Raluca Mariana Dragoescu

This work studies the porting and optimization of the tensor network simulator QTensor on GPUs, with the ultimate goal of simulating quantum circuits efficiently at scale on large GPU supercomputers. We implement NumPy, PyTorch, and CuPy…

量子物理 · 物理学 2022-04-14 Danylo Lykov , Angela Chen , Huaxuan Chen , Kristopher Keipert , Zheng Zhang , Tom Gibbs , Yuri Alexeev

Particle tracking simulations with space charge effects are very important for high-intensity proton rings. Since they include not only Hamilton mechanics of a single particle but constructing charge densities and solving Poisson equations…

加速器物理 · 物理学 2021-09-01 Yoshinori Kurimoto

Study of general purpose computation by GPU (Graphics Processing Unit) can improve the image processing capability of micro-computer system. This paper studies the parallelism of the different stages of decimation in time radix 2 FFT…

数学软件 · 计算机科学 2015-06-01 Feifei Shen , Zhenjian Song , Congrui Wu , Jiaqi Geng , Qingyun Wang

Scientific computing in the exascale era demands increased computational power to solve complex problems across various domains. With the rise of heterogeneous computing architectures the need for vendor-agnostic, performance portability…

分布式、并行与集群计算 · 计算机科学 2025-11-05 Johansell Villalobos , Josef Ruzicka , Silvio Rizzi

In this work, we survey the role of GPUs in real-time systems. Originally designed for parallel graphics workloads, GPUs are now widely used in time-critical applications such as machine learning, autonomous vehicles, and robotics due to…

It is shown micromagnetic and atomistic spin dynamics simulations can use multiple GPUs in order to reduce computation time, but also to allow for a larger simulation size than is possible on a single GPU. Whilst interactions which depend…

介观与纳米尺度物理 · 物理学 2023-10-12 Serban Lepadatu