English
Related papers

Related papers: Thermal Analysis for NVIDIA GTX480 Fermi GPU Archi…

200 papers

We show that efficient simulations of the Kardar-Parisi-Zhang interface growth in 2 + 1 dimensions and of the 3-dimensional Kinetic Monte Carlo of thermally activated diffusion can be realized both on GPUs and modern CPUs. In this article…

Distributed, Parallel, and Cluster Computing · Computer Science 2014-01-21 Jeffrey Kelling , Géza Ódor , Máté Ferenc Nagy , Henrik Schulz , Karl-Heinz Heinig

Collaborative filtering (CF) has been proven to be one of the most effective techniques for recommendation. Among all CF approaches, SimpleX is the state-of-the-art method that adopts a novel loss function and a proper number of negative…

Distributed, Parallel, and Cluster Computing · Computer Science 2023-05-04 Chengming Zhang , Shaden Smith , Baixi Sun , Jiannan Tian , Jonathan Soifer , Xiaodong Yu , Shuaiwen Leon Song , Yuxiong He , Dingwen Tao

We present a physics-based neural network framework for the discovery of constitutive models in fully coupled thermomechanics. In contrast to classical formulations based on the Helmholtz energy, we adopt the internal energy and a…

Computational Engineering, Finance, and Science · Computer Science 2026-05-25 Hagen Holthusen , Paul Steinmann , Ellen Kuhl

Deploying multiple models within shared GPU clusters is a key strategy to improve resource efficiency in large language model (LLM) serving. Existing multi-LLM serving systems improve GPU utilization at the cost of degraded inference…

Distributed, Parallel, and Cluster Computing · Computer Science 2026-05-22 Chiheng Lou , Sheng Qi , Rui Kang , Yong Zhang , Chen Sun , Pengcheng Wang , Xuanzhe Liu , Xin Jin

Thermal aware routing and placement algorithms are important in industry. Currently, there are reasonably fast Green's function based algorithms that calculate the temperature distribution in a chip made from a stack of different materials.…

General Physics · Physics 2008-01-08 Virginia Martín Hériz , J. -H. Park , T. Kemper , S. -M. Kang , A. Shakouri

Sparse Matrix-Matrix Multiplication (SpMM) is a fundamental kernel across scientific computing and machine learning. While prior work accelerates SpMM using Tensor Cores, no existing sparse kernel exploits the asynchronous features of…

Distributed, Parallel, and Cluster Computing · Computer Science 2026-04-21 Jie Liu , Huanzhi Pu , Zhiru Zhang

The edge computing paradigm has emerged to handle cloud computing issues such as scalability, security and low response time among others. This new computing trend heavily relies on ubiquitous embedded systems on the edge. Performance and…

Distributed, Parallel, and Cluster Computing · Computer Science 2019-01-28 Mohammad Hosseinabady , Mohd Amiruddin Bin Zainol , Jose Nunez-Yanez

We study the use of the Evans Nonequilibrium Molecular Dynamics (NEMD) heat flow algorithm for the computation of the heat conductivity in one-dimensional lattices. For the well-known Fermi-Pasta-Ulam (FPU) model, it is shown that when the…

chao-dyn · Physics 2009-10-31 Fei Zhang , Dennis J. Isbister , Denis J. Evans

Graph Pattern Mining (GPM) is an important, rapidly evolving, and computation demanding area. GPM computation relies on subgraph enumeration, which consists in extracting subgraphs that match a given property from an input graph. Graphics…

Distributed, Parallel, and Cluster Computing · Computer Science 2022-12-12 Samuel Ferraz , Vinicius Dias , Carlos H. C. Teixeira , George Teodoro , Wagner Meira

In this work, we develop a combined convolutional neural networks (CNNs) and finite element method (FEM) to examine the effective thermal properties of composite phase change materials (CPCMs) consisting of paraffin and copper foam. In this…

Computational Physics · Physics 2021-03-25 Felix Kolodziejczyk , Bohayra Mortazavi , Timon Rabczuk , Xiaoying Zhuang

As the landscape of deep neural networks evolves, heterogeneous dataflow accelerators, in the form of multi-core architectures or chiplet-based designs, promise more flexibility and higher inference performance through scalability. So far,…

Hardware Architecture · Computer Science 2025-10-08 Arne Symons , Linyan Mei , Steven Colleman , Pouya Houshmand , Sebastian Karl , Marian Verhelst

Graphics Processing Units (GPUs) are now powerful and flexible systems adapted and used for other purposes than graphics calculations (General Purpose computation on GPU -- GPGPU). We present here a prototype to be integrated into…

Distributed, Parallel, and Cluster Computing · Computer Science 2007-06-13 Sylvain Collange , Marc Daumas , David Defour

According to the increasing complexity of network application and internet traffic, network processor as a subset of embedded processors have to process more computation intensive tasks. By scaling down the feature size and emersion of chip…

Hardware Architecture · Computer Science 2012-04-13 Mehdi Alipour , Hojjat Taghdisi

Rapid growth in artificial intelligence (AI) workloads is driving up data center power densities, increasing the need for advanced thermal management. Direct-to-chip liquid cooling can remove heat efficiently at the source, but many cold…

Systems and Control · Electrical Eng. & Systems 2026-04-14 Zheng Liu

This work builds on the previous introduction [1] of a coupled experimental-computational system devised to fully characterize the thermal behavior of complex 3D submicron electronic devices. The new system replaces the laser-based surface…

Materials Science · Physics 2007-09-13 Peter E. Raad , Pavel L. Komarov , M. Burzo

Adjacent GEMM problems that differ by a single 128-element step in N can show 30% different throughput on the same GPU. This pervasive performance ruggedness - invisible to roofline analysis and peak-FLOPs intuition, yet dominant for every…

Performance · Computer Science 2026-05-29 Aditya Chatterjee

With electric power systems becoming more compact and increasingly powerful, the relevance of thermal stress especially during overload operation is expected to increase ceaselessly. Whenever critical temperatures cannot be measured…

Machine Learning · Computer Science 2022-11-03 Wilhelm Kirchgässner , Oliver Wallscheid , Joachim Böcker

Thermodynamic trade-off relations dictate fundamental limits on the performance of thermodynamic tasks through costs such as heat dissipation. Here, we propose a framework called thermodynamic recycling to circumvent these limits in quantum…

Quantum Physics · Physics 2026-04-28 Nobumasa Ishida , Yoshihiko Hasegawa

Proactive maintenance strategies, such as Predictive Maintenance (PdM), play an important role in the operation of Nuclear Power Plants (NPPs), particularly due to their capacity to reduce offline time by preventing unexpected shutdowns…

We present a new technique of VLSI chip-level thermal analysis. We extend a newly developed method of solving two dimensional Laplace equations to thermal analysis of four adjacent materials on a mother board. We implement our technique in…

General Physics · Physics 2008-01-08 K. Nakabayashi , T. Nakabayashi , K. Nakajima