English
Related papers

Related papers: Parallel implementation of the Density Matrix Reno…

200 papers

We propose a density matrix renormalization group approach to tackle a two-state system coupled to a bosonic bath with continuous spectrum. In this approach, the optimized phonon scheme is applied to several hundred phonon modes which are…

Strongly Correlated Electrons · Physics 2008-05-31 Hang Wong , Zhi-De Chen

We have studied transition metal clusters from a quantum information theory perspective using the density-matrix renormalization group (DMRG) method. We demonstrate the competition between entanglement and interaction localization. We also…

Quantum Physics · Physics 2015-05-19 G. Barcza , Ö. Legeza , K. H. Marti , M. Reiher

We present an approach for the calculation of spin density distributions for molecules that require very large active spaces for a qualitatively correct description of their electronic structure. Our approach is based on the density-matrix…

Chemical Physics · Physics 2012-06-29 Katharina Boguslawski , Konrad H. Marti , Örs Legeza , Markus Reiher

We propose an initialization procedure for the density-matrix renormalization group (DMRG): {\it the recursive sweep method}. In a conventional DMRG calculation, the infinite-algorithm, where two new sites are added to the system at each…

Strongly Correlated Electrons · Physics 2007-05-23 Masaki Tezuka

Real-world node embedding applications often contain hundreds of billions of edges with high-dimension node features. Scaling node embedding systems to efficiently support these applications remains a challenging problem. In this paper we…

Distributed, Parallel, and Cluster Computing · Computer Science 2021-08-19 Wanjing Wei , Yangzihao Wang , Pin Gao , Shijie Sun , Donghai Yu

Performance optimization can be a daunting task especially as the hardware architecture becomes more and more complex. This paper takes a kernel from the Materials Science code BerkeleyGW, and demonstrates a few performance analysis and…

Distributed, Parallel, and Cluster Computing · Computer Science 2020-09-24 Charlene Yang

Tensor Core is a mixed-precision matrix-matrix multiplication unit on NVIDIA GPUs with a theoretical peak performance of more than 300 TFlop/s on Ampere architectures. Tensor Cores were developed in response to the high demand of dense…

Distributed, Parallel, and Cluster Computing · Computer Science 2023-10-19 Hiroyuki Ootomo , Rio Yokota

This study systematically tests a computational power reuse scheme proposed by the open source community disabling specific instruction sets (Fused Multiply Add instructions) through CUDA source code modifications on the NVIDIA CMP 170HX…

Hardware Architecture · Computer Science 2025-05-09 Xing Kangwei

We present a Gauss-Newton-Krylov solver for large deformation diffeomorphic image registration. We extend the publicly available CLAIRE library to multi-node multi-graphics processing unit (GPUs) systems and introduce novel algorithmic…

Distributed, Parallel, and Cluster Computing · Computer Science 2020-12-25 Malte Brunn , Naveen Himthani , George Biros , Miriam Mehl , Andreas Mang

The solution of eigenproblems is often a key computational bottleneck that limits the tractable system size of numerical algorithms, among them electronic structure theory in chemistry and in condensed matter physics. Large eigenproblems…

General Matrix Multiplication (GEMM) has a wide range of applications in scientific simulation and artificial intelligence. Although traditional libraries can achieve high performance on large regular-shaped GEMMs, they often behave not…

Distributed, Parallel, and Cluster Computing · Computer Science 2022-08-12 Shangfei Yin , Qinglin Wang , Ruochen Hao , Tianyang Zhou , Songzhu Mei , Jie Liu

Neural Radiance Fields (NeRF) enables 3D scene reconstruction from several 2D images but incurs high rendering latency via its point-sampling design. 3D Gaussian Splatting (3DGS) improves on NeRF with explicit scene representation and an…

Hardware Architecture · Computer Science 2026-04-07 Haomin Li , Bowen Zhu , Fangxin Liu , Zongwu Wang , Xinran Liang , Li Jiang , Haibing Guan

The increasing scale and complexity of integrated circuit design have led to increased challenges in Electronic Design Automation (EDA). Graph Neural Networks (GNNs) have emerged as a promising approach to assist EDA design as circuits can…

Machine Learning · Computer Science 2025-08-26 Yuebo Luo , Shiyang Li , Junran Tao , Kiran Thorat , Xi Xie , Hongwu Peng , Nuo Xu , Caiwen Ding , Shaoyi Huang

We present novel algorithmic solutions together with implementation details utilizing non-Abelian symmetries in order to boost the current limits of tensor network state algorithms on high performance computing infrastructure. In our…

Computational Physics · Physics 2023-10-02 Andor Menczer , Örs Legeza

The density matrix renormalization group (DMRG) method has already proved itself as a very efficient and accurate computational method, which can treat large active spaces and capture the major part of strong correlation. Its application on…

Chemical Physics · Physics 2022-10-31 Pavel Beran , Katarzyna Pernal , Fabijan Pavosevic , Libor Veis

During the past 15 years, the density matrix renormalization group (DMRG) has become increasingly important for ab initio quantum chemistry. The underlying matrix product state (MPS) ansatz is a low-rank decomposition of the full…

Strongly Correlated Electrons · Physics 2014-05-22 Sebastian Wouters

We propose a high-performance GPU solver for inverse homogenization problems to design high-resolution 3D microstructures. Central to our solver is a favorable combination of data structures and algorithms, making full use of the parallel…

Optimization and Control · Mathematics 2023-05-26 Di Zhang , Xiaoya Zhai , Ligang Liu , Xiao-Ming Fu

In this work, we simulate the electron dynamics in molecular systems with the Time-Dependent Density Matrix Renormalization Group (TD-DMRG) algorithm. We leverage the generality of the so-called tangent-space TD-DMRG formulation and design…

Chemical Physics · Physics 2021-06-08 Alberto Baiardi

Finite element schemes based on discontinuous Galerkin methods possess features amenable to massively parallel computing accelerated with general purpose graphics processing units (GPUs). However, the computational performance of such…

Computational Physics · Physics 2016-04-20 Axel Modave , Amik St-Cyr , Tim Warburton

High-order gas-kinetic scheme (HGKS) has become a workable tool for the direct numerical simulation (DNS) of turbulence. In this paper, to accelerate the computation, HGKS is implemented with the graphical processing unit (GPU) using the…

Numerical Analysis · Mathematics 2023-01-25 Yuhang Wang , Guiyu Cao , Liang Pan