中文
相关论文

相关论文: Parallel Pairwise Correlation Computation On Intel…

200 篇论文

Limits on power dissipation have pushed CPUs to grow in parallel processing capabilities rather than clock rate, leading to the rise of "manycore" or GPU-like processors. In order to achieve the best performance, applications must be able…

The Massive Parallel Computation (MPC) model is a theoretical framework for popular parallel and distributed platforms such as MapReduce, Hadoop, or Spark. We consider the task of computing a large matching or small vertex cover in this…

数据结构与算法 · 计算机科学 2018-07-24 Krzysztof Onak

GPU-based HPC clusters are attracting more scientific application developers due to their extensive parallelism and energy efficiency. In order to achieve portability among a variety of multi/many core architectures, a popular choice for an…

分布式、并行与集群计算 · 计算机科学 2023-04-10 Ali TehraniJamsaz , Alok Mishra , Akash Dutta , Abid M. Malik , Barbara Chapman , Ali Jannesari

Gaussian process (GP) models are widely used to analyze spatially referenced data and to predict values at locations without observations. In contrast to many algorithmic procedures, GP models are based on a statistical framework, which…

统计计算 · 统计学 2020-01-01 Florian Gerber , Douglas W. Nychka

Prior work on Automatically Scalable Computation (ASC) suggests that it is possible to parallelize sequential computation by building a model of whole-program execution, using that model to predict future computations, and then…

分布式、并行与集群计算 · 计算机科学 2018-09-21 Peter Kraft , Amos Waterland , Daniel Y Fu , Anitha Gollamudi , Shai Szulanski , Margo Seltzer

Principal component analysis (PCA) is a key statistical technique for multivariate data analysis. For large data sets the common approach to PCA computation is based on the standard NIPALS-PCA algorithm, which unfortunately suffers from…

定量方法 · 定量生物学 2008-11-10 M. Andrecut

We study parallel particle-in-cell (PIC) methods for low-temperature plasmas (LTPs), which discretize kinetic formulations that capture the time evolution of the probability density function of particles as a function of position and…

计算工程、金融与科学 · 计算机科学 2025-08-12 James Almgren-Bell , Nader Al Awar , Dilip S Geethakrishnan , Milos Gligoric , George Biros

We discuss the scalable parallel solution of the Poisson equation within a Particle-In-Cell (PIC) code for the simulation of electron beams in particle accelerators of irregular shape. The problem is discretized by Finite Differences.…

计算物理 · 物理学 2010-04-21 A. Adelmann , P. Arbenz , Y. Ineichen

ALICE is the dedicated heavy ion experiment at the LHC at CERN and records lead-lead collisions at a rate of up to 50 kHz. The detector with the highest data rate of up to 3.4 TB/s is the TPC. ALICE performs the full online TPC processing…

仪器与探测器 · 物理学 2025-11-24 David Rohr

Particle-in-Cell (PIC) Monte Carlo (MC) simulations are central to plasma physics but face increasing challenges on heterogeneous HPC systems due to excessive data movement, synchronization overheads, and inefficient utilization of multiple…

Spectral clustering is one of the most popular graph clustering algorithms, which achieves the best performance for many scientific and engineering applications. However, existing implementations in commonly used software platforms such as…

分布式、并行与集群计算 · 计算机科学 2018-02-14 Yu Jin , Joseph F. JaJa

We introduce generalized spatially coupled parallel concatenated codes (GSC-PCCs), a class of spatially coupled turbo-like codes obtained by coupling parallel concatenated codes (PCCs) with a fraction of information bits repeated before the…

信息论 · 计算机科学 2021-05-04 Min Qiu , Xiaowei Wu , Jinhong Yuan , Alexandre Graell i Amat

The Gyrokinetic Toroidal Code at Princeton (GTC-P) is a highly scalable and portable particle-in-cell (PIC) code. It solves the 5D Vlasov-Poisson equation featuring efficient utilization of modern parallel computer architectures at the…

分布式、并行与集群计算 · 计算机科学 2015-10-20 Bei Wang , Stephane Ethier , William Tang , Khaled Ibrahim , Kamesh Madduri , Samuel Williams , Leonid Oliker

Leading HPC systems achieve their status through use of highly parallel devices such as NVIDIA GPUs or Intel Xeon Phi many-core CPUs. The concept of performance portability across such architectures, as well as traditional CPUs, is vital…

分布式、并行与集群计算 · 计算机科学 2016-11-10 Alan Gray , Kevin Stratford

Large-scale HPC simulations of plasma dynamics in fusion devices require efficient parallel I/O to avoid slowing down the simulation and to enable the post-processing of critical information. Such complex simulations lacking parallel I/O…

Heterogeneous computing is emerging as a mandatory requirement for power-efficient system design. With this aim, modern heterogeneous platforms like Zynq All-Programmable SoC, that integrates ARM-based SMP and programmable logic, have been…

分布式、并行与集群计算 · 计算机科学 2015-08-28 Daniel Jiménez-González , Carlos Álvarez , Antonio Filgueras , Xavier Martorell , Jan Langer , Juanjo Noguera , Kees Vissers

LDPC (Low Density Parity Check) codes are among the most powerful and widely adopted modern error correcting codes. The iterative decoding algorithms required for these codes involve high computational complexity and high processing…

硬件体系结构 · 计算机科学 2011-05-16 Carlo Condo , Guido Masera

New computing paradigms are required to solve the most challenging computational problems where no exact polynomial time solution exists.Probabilistic Ising Accelerators has gained promise on these problems with the ability to model complex…

分布式、并行与集群计算 · 计算机科学 2024-09-17 Saavan Patel , Philip Canoza , Adhiraj Datar , Steven Lu , Chirag Garg , Sayeef Salahuddin

We evaluate the second-generation Intel Xeon Phi coprocessor based on the Intel Many Integrated Core (MIC) architecture, aka the Knights Landing or KNL, for simulating neutrino oscillations in (core-collapse) supernovae. For this purpose we…

计算物理 · 物理学 2019-12-24 Vahid Noormofidi , Susan R. Atlas , Huaiyu Duan

For LHC Run 3, the ALICE Time Projection Chamber was upgraded to operate in continuous readout mode. Interaction rates of up to 50 kHz in Pb-Pb collisions require real-time processing of more than 3 TB/s of raw detector data. This…

仪器与探测器 · 物理学 2026-03-18 J. Alme , T. Alt , C. Andrei , V. Anguelov , H. Appelshäuser , M. Arslandok , R. Averbeck , M. Ball , G. G. Barnaföldi , P. Becht , R. Bellwied , A. Berdnikova , B. Blidaru , L. Boldizsár , L. Bratrud , P. Braun-Munzinger , M. Bregant , C. L. Britton , H. Büsching , H. Caines , P. Chatzidaki , P. Christiansen , T. M. Cormier , L. Döpper , R. Ehlers , L. Fabbietti , F. Flor , J. J. Gaardhøje , M. G. Munhoz , C. Garabatos , P. Gasik , Á. Gera , P. Glässel , N. Grünwald , T. Gündem , T. Gunji , H. Hamagaki , J. W. Harris , P. Hauer , E. Hellbär , H. Helstrup , A. Herghelegiu , H. D. Hernandez Herrera , Y. Hou , C. Hughes , M. Ivanov , J. Jäger , Y. Ji , J. Jung , M. Jung , B. Ketzer , S. Kirsch , M. Kleiner , A. G. Knospe , M. Korwieser , M. Kowalski , L. Lautner , M. Lesch , C. Lippmann , G. Mantzaridis , R. D. Majka , A. Marin , C. Markert , S. Masciocchi , A. Matyja , M. Meres , D. L. Mihaylov , D. Miśkowiec , R. H. Munzer , H. Murakami , K. Münning , A. Nassirpour , C. Nattrass , B. S. Nielsen , W. A. V. Noije , A. C. Oliveira Da Silva , A. Oskarsson , K. Oyama , L. Österman , Y. Pachmayer , G. Paić , M. Petris , M. Petrovici , M. Planinic , J. Rasson , K. F. Read , A. Rehman , R. Renfordt , A. Riedel , K. Røed , D. Röhrich , E. Rubio , A. Rusu , S. Sadhu , B. C. S. Sanches , J. Schambach , A. Schmah , C. Schmidt , A. Schmier , K. Schweda , D. Sekihata , D. Silvermyr , B. Sitar , N. Smirnov , H. K. Soltveit , C. Sonnabend , S. P. Sorensen , J. Stachel , L. Šerkšnytė , G. Tambave , K. Ullaland , B. Ulukutlu , D. Varga , O. Vazquez Rueda , B. Voss , J. Wiechula , B. Windelband , J. Wilkinson , J. Witte , A. Yadav , F. Zanone , S. Zhu