中文
相关论文

相关论文: Low-ordered Orthogonal Voxel Finite Element with I…

200 篇论文

When modelling discontinuities (interfaces) using the finite element method, the standard approach is to use a conforming finite-element mesh in which the mesh matches the interfaces. However, this approach can prove cumbersome if the…

计算工程、金融与科学 · 计算机科学 2024-06-06 Jedrzej Dobrzanski , Kajetan Wojtacki , Stanislaw Stupkiewicz

We present a design and implementation of the Thomas algorithm optimized for hardware acceleration on an FPGA, the Thomas Core. The hardware-based algorithm combined with the custom data flow and low level parallelism available in an FPGA…

计算金融 · 定量金融 2015-10-16 Samuel Palmer

We review a scalable two- and three-dimensional computer code for low-temperature plasma simulations in multi-material complex geometries. Our approach is based on embedded boundary (EB) finite volume discretizations of the minimal…

计算物理 · 物理学 2019-05-01 Robert Marskar

It is well known that the solution of topology optimization problems may be affected both by the geometric properties of the computational mesh, which can steer the minimization process towards local (and non-physical) minima, and by the…

数值分析 · 数学 2016-12-28 Paola F. Antonietti , Matteo Bruggi , Simone Scacchi , Marco Verani

Recent developments in vortex particle methods for simulating three-dimensional incompressible flows are presented. A lightweight, dynamic Large-Eddy Simulation model is tested, featuring a dynamic procedure that relies solely on Lagrangian…

流体动力学 · 物理学 2026-01-13 Flavio A. C. Martins , Alexander van Zuijlen , Carlos J. Simao Ferreira

We propose an efficient finite-element analysis of the vector wave equation in a class of relatively general curved polygons. The proposed method is suitable for an accurate and efficient calculation of the propagation constants of…

计算物理 · 物理学 2016-03-15 Ehsan Khodapanah

Large-scale deep learning benefits from an emerging class of AI accelerators. Some of these accelerators' designs are general enough for compute-intensive applications beyond AI and Cloud TPU is one such example. In this paper, we…

分布式、并行与集群计算 · 计算机科学 2019-11-19 Kun Yang , Yi-Fan Chen , Georgios Roumpos , Chris Colby , John Anderson

In recent years, high performance scientific computing on graphics processing units (GPUs) have gained widespread acceptance. These devices are designed to offer massively parallel threads for running code with general purpose. There are…

数学软件 · 计算机科学 2018-02-13 Tao Cui , Xiaohu Guo , Hui Liu

An effective strategy for accelerating the calculation of convex hulls for point sets is to filter the input points by discarding interior points. In this paper, we present such a straightforward and efficient preprocessing approach by…

计算几何 · 计算机科学 2014-05-30 Gang Mei

Neural Radiance Fields (NeRF) enables 3D scene reconstruction from several 2D images but incurs high rendering latency via its point-sampling design. 3D Gaussian Splatting (3DGS) improves on NeRF with explicit scene representation and an…

硬件体系结构 · 计算机科学 2026-04-07 Haomin Li , Bowen Zhu , Fangxin Liu , Zongwu Wang , Xinran Liang , Li Jiang , Haibing Guan

Tensor Cores have been an important unit to accelerate Fused Matrix Multiplication Accumulation (MMA) in all NVIDIA GPUs since Volta Architecture. To program Tensor Cores, users have to use either legacy wmma APIs or current mma APIs.…

硬件体系结构 · 计算机科学 2022-11-29 Wei Sun , Ang Li , Tong Geng , Sander Stuijk , Henk Corporaal

Fast Fourier Transform (FFT) is an essential tool in scientific and engineering computation. The increasing demand for mixed-precision FFT has made it possible to utilize half-precision floating-point (FP16) arithmetic for faster speed and…

分布式、并行与集群计算 · 计算机科学 2021-04-26 Binrui Li , Shenggan Cheng , James Lin

We present a GPU-accelerated cosmological simulation code, PhotoNs-GPU, based on algorithm of Particle Mesh Fast Multipole Method (PM-FMM), and focus on the GPU utilization and optimization. A proper interpolated method for truncated…

天体物理仪器与方法 · 物理学 2021-12-28 Qiao Wang , Chen Meng

Dense 3D convolutions provide high accuracy for perception but are too computationally expensive for real-time robotic systems. Existing tri-plane methods rely on 2D image features with interpolation, point-wise queries, and implicit MLPs,…

机器人学 · 计算机科学 2025-09-19 Sibaek Lee , Jiung Yeon , Hyeonwoo Yu

Many research works have been performed on implementation of Vitrerbi decoding algorithm on GPU instead of FPGA because this platform provides considerable flexibility in addition to great performance. Recently, the recently-introduced…

分布式、并行与集群计算 · 计算机科学 2020-11-30 Alireza Mohammadidoost , Matin Hashemi

This paper introduces an efficient and generic framework for finite-element simulations under an implicit time integration scheme. Being compatible with generic constitutive models, a fast matrix assembly method exploits the fact that…

分布式、并行与集群计算 · 计算机科学 2023-06-12 Ziqiu Zeng , Hadrien Courtecuisse

A simple method for improving cache efficiency of serial and parallel explicit finite procedure with application to casting solidification simulation over three-dimensional complex geometries is presented. The method is based on division of…

分布式、并行与集群计算 · 计算机科学 2010-05-19 Ruhollah Tavakoli

A novel approach is presented for fast generation of synthetic seismograms due to microseismic events, using heterogeneous marine velocity models. The partial differential equations (PDEs) for the 3D elastic wave equation have been…

地球物理 · 物理学 2017-05-16 Saptarshi Das , Xi Chen , Michael P. Hobson

We present a finite element method (FEM) solver for computation of optical resonance modes in VCSELs. We perform a convergence study and demonstrate that high accuracies for 3D setups can be attained on standard computers. We also…

光学 · 物理学 2012-03-02 M. Rozova , J. Pomplun , L. Zschiedrich , F. Schmidt , S. Burger

This paper presents an accurate density computation approach for large dark matter simulations, based on a recently introduced phase-space tessellation technique and designed for massively parallel, heterogeneous cluster architectures. We…

计算物理 · 物理学 2017-08-28 Ralf Kaehler