English
Related papers

Related papers: A multi-GPU benchmark for 2D Marchenko Imaging

200 papers

We present a case-study on the utility of graphics cards to perform massively parallel simulation of advanced Monte Carlo methods. Graphics cards, containing multiple Graphics Processing Units (GPUs), are self-contained parallel…

Computation · Statistics 2015-05-05 Anthony Lee , Christopher Yau , Michael B. Giles , Arnaud Doucet , Christopher C. Holmes

The advanced magnetic resonance (MR) image reconstructions such as the compressed sensing and subspace-based imaging are considered as large-scale, iterative, optimization problems. Given the large number of reconstructions required by the…

Computational Engineering, Finance, and Science · Computer Science 2020-06-26 Tianjian Lu , Thibault Marin , Yue Zhuo , Yi-Fan Chen , Chao Ma

Modern parallel computing devices, such as the graphics processing unit (GPU), have gained significant traction in scientific and statistical computing. They are particularly well-suited to data-parallel algorithms such as the particle…

Computation · Statistics 2015-06-12 Lawrence M. Murray , Anthony Lee , Pierre E. Jacob

A spectral fitter based on the graphics processor unit (GPU) has been developed for Borexino solar neutrino analysis. It is able to shorten the fitting time to a superior level compared to the CPU fitting procedure. In Borexino solar…

Data Analysis, Statistics and Probability · Physics 2020-01-22 X. F. Ding , M. Agostini , K. Altenmuller , S. Appel , V. Atroshchenko , Z. Bagdasarian , D. Basilico , G. Bellini , J. Benziger , D. Bick , G. Bonfini , D. Bravo , B. Caccianiga , F. Calaprice , A. Caminata , S. Caprioli , M. Carlini , P. Cavalcante , A. Chepurnov , K. Choi , L. Collica , D. D'Angelo , S. Davini , A. Derbin , A. Di Ludovico , L. Di Noto , I. Drachnev , K. Fomenko , A. Formozov , D. Franco , F. Froborg , F. Gabriele , C. Galbiati , C. Ghiano , M. Giammarchi , A. Goretti , M. Gromov , D. Guffanti , C. Hagner , T. Houdy , E. Hungerford , Aldo Ianni , Andrea Ianni , A. Jany , D. Jeschke , V. Kobychev , D. Korablev , G. Korga , D. Kryn , M. Laubenstein , E. Litvinovich , F. Lombardi , P. Lombardi , L. Ludhova , G. Lukyanchenko , L. Lukyanchenko , I. Machulin , G. Manuzio , S. Marcocci , J. Martyn , E. Meroni , M. Meyer , L. Miramonti , M. Misiaszek , V. Muratova , B. Neumair , L. Oberauer , B. Opitz , V. Orekhov , F. Ortica , M. Pallavicini , L. Papp , O. Penek , N. Pilipenko , A. Pocar , A. Porcelli , G. Ranucci , A. Razeto , A. Re , M. Redchuk , A. Romani , R. Roncin , N. Rossi , S. Schonert , D. Semenov , M. Skorokhvatov , O. Smirnov , A. Sotnikov , L. F. F. Stokes , Y. Suvorov , R. Tartaglia , G. Testera , J. Thurn , M. Toropova , E. Unzhakov , A. Vishneva , R. B. Vogelaar , F. von Feilitzsch , H. Wang , S. Weinz , M. Wojcik , M. Wurm , Z. Yokley , O. Zaimidoroga , S. Zavatarelli , K. Zuber , G. Zuzel

This article presents a systematic quantitative performance analysis for large finite element computations on extreme scale computing systems. Three parallel iterative solvers for the Stokes system, discretized by low order tetrahedral…

Computational Engineering, Finance, and Science · Computer Science 2015-11-09 Björn Gmeiner , Markus Huber , Lorenz John , Ulrich Rüde , Barbara Wohlmuth

A parallel algorithm for the implementation of the recursive Green's function technique, which is extensively applied in the coherent scattering formalism, is developed. The algorithm performs a domain decomposition of the scattering region…

Mesoscale and Nanoscale Physics · Physics 2009-11-11 P. S. Drouvelis , P. Schmelcher , P. Bastian

We apply a recently proposed method for the acceleration of model predictive control (MPC) to 36 MPC implementations, which result from combining six sample receding horizon control problems with six quadratic programming solvers. We…

Optimization and Control · Mathematics 2015-05-27 Michael Jost , Gabriele Pannocchia , Martin Mönnigmann

We present the methodology of a photon-conserving, spatially-adaptive, ray-tracing radiative transfer algorithm, designed to run on multiple parallel Graphic Processing Units (GPUs). Each GPU has thousands computing cores, making them…

Cosmology and Nongalactic Astrophysics · Physics 2018-10-17 Blake Hartley , Massimo Ricotti

Computational inference of causal relationships underlying complex networks, such as gene-regulatory pathways, is NP-complete due to its combinatorial nature when permuting all possible interactions. Markov chain Monte Carlo (MCMC) has been…

Distributed, Parallel, and Cluster Computing · Computer Science 2012-10-19 Yu Wang , Weikang Qian , Shuchang Zhang , Bo Yuan

This work proposes a GPU tensor core approach that encodes the arithmetic reduction of $n$ numbers as a set of chained $m \times m$ matrix multiply accumulate (MMA) operations executed in parallel by GPU tensor cores. The asymptotic running…

Distributed, Parallel, and Cluster Computing · Computer Science 2020-01-17 Cristóbal A. Navarro , Roberto Carrasco , Ricardo J. Barrientos , Javier A. Riquelme , Raimundo Vega

We present a new adaptive parallel algorithm for the challenging problem of multi-dimensional numerical integration on massively parallel architectures. Adaptive algorithms have demonstrated the best performance, but efficient many-core…

Distributed, Parallel, and Cluster Computing · Computer Science 2021-06-24 Ioannis Sakiotis , Kamesh Arumugam , Marc Paterno , Desh Ranjan , Balša Terzić , Mohammad Zubair

Inverse rendering methods have achieved remarkable performance in reconstructing high-fidelity 3D objects with disentangled geometries, materials, and environmental light. However, they still face huge challenges in reflective surface…

Computer Vision and Pattern Recognition · Computer Science 2024-11-22 Tengjie Zhu , Zhuo Chen , Jingnan Gao , Yichao Yan , Xiaokang Yang

This paper presents a Graphics Processing Units (GPUs) acceleration method of an iterative scheme for gas-kinetic model equations. Unlike the previous GPU parallelization of explicit kinetic schemes, this work features a fast converging…

Computational Physics · Physics 2020-01-08 Lianhua Zhu , Peng Wang , Songze Chen , Zhaoli Guo , Yonghao Zhang

GPUs are the most popular platform for accelerating HPC workloads, such as artificial intelligence and science simulations. However, most microarchitectural research in academia relies on GPU core pipeline designs based on architectures…

Hardware Architecture · Computer Science 2025-10-30 Rodrigo Huerta , Mojtaba Abaie Shoushtary , José-Lorenzo Cruz , Antonio González

An efficient error reconciliation scheme is important for post-processing of quantum key distribution (QKD). Recently, a multi-matrix low-density parity-check codes based reconciliation algorithm which can provide remarkable perspectives…

Quantum Physics · Physics 2020-01-23 Yu Guo , Chaohui Gao , Dong Jiang , Lijun Chen

Ptychography is an emerging imaging technique that is able to provide wavelength-limited spatial resolution from specimen with extended lateral dimensions. As a scanning microscopy method, a typical two-dimensional image requires a number…

We present a multi-GPU extension of the 3D Gaussian Splatting (3D-GS) pipeline for scientific visualization. Building on previous work that demonstrated high-fidelity isosurface reconstruction using Gaussian primitives, we incorporate a…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-09-08 Mengjiao Han , Andres Sewell , Joseph Insley , Janet Knowles , Victor A. Mateevitsi , Michael E. Papka , Steve Petruzza , Silvio Rizzi

We introduce a new Markov-Chain Monte Carlo (MCMC) approach designed for efficient sampling of highly correlated and multimodal posteriors. Parallel tempering, though effective, is a costly technique for sampling such posteriors. Our…

Instrumentation and Methods for Astrophysics · Physics 2014-10-01 Benjamin Farr , Vicky Kalogera , Erik Luijten

Gaussian Processes have become an indispensable part of the spatial statistician's toolbox but are unsuitable for analyzing large dataset because of the significant time and memory needed to fit the associated model exactly. Vecchia…

Computation · Statistics 2025-07-18 Zachary James , Joseph Guinness

Matrix multiplication is a fundamental operation in both training of neural networks and inference. To accelerate matrix multiplication, Graphical Processing Units (GPUs) provide it implemented in hardware. Due to the increased throughput…

Mathematical Software · Computer Science 2026-04-07 Faizan A. Khattak , Mantas Mikaitis
‹ Prev 1 4 5 6 7 8 10 Next ›