中文
相关论文

相关论文: Indicating Asynchronous Array Multipliers

200 篇论文

Edge computing must be capable of executing computationally intensive algorithms, such as Deep Neural Networks (DNNs) while operating within a constrained computational resource budget. Such computations involve Matrix Vector…

硬件体系结构 · 计算机科学 2023-10-24 Arani Roy , Kaushik Roy

The design of isophoric phased arrays composed of two-sized square-shaped tiles that fully cover rectangular apertures is dealt with. The number and the positions of the tiles within the array aperture are optimized to fit desired…

信号处理 · 电气工程与系统科学 2021-02-05 P. Rocca , N. Anselmi , A. Polo , A. Massa

Despite the prominence of neural network approaches in the field of recommender systems, simple methods such as matrix factorization with quadratic loss are still used in industry for several reasons. These models can be trained with…

信息检索 · 计算机科学 2022-05-24 Dmitrii Beloborodov , Andrei Zimovnov , Petr Molodyk , Dmitrii Kirillov

Deploying mixed-precision neural networks on edge devices is friendly to hardware resources and power consumption. To support fully mixed-precision neural network inference, it is necessary to design flexible hardware accelerators for…

硬件体系结构 · 计算机科学 2025-02-04 Liang Zhao , Kunming Shao , Fengshi Tian , Tim Kwang-Ting Cheng , Chi-Ying Tsui , Yi Zou

Semaphores are a widely used and foundational synchronization and coordination construct used for shared memory multithreaded programming. They are a keystone concept, in the sense that most other synchronization constructs can be…

分布式、并行与集群计算 · 计算机科学 2025-04-23 Dave Dice , Alex Kogan

The Alternating Direction Method of Multipliers (ADMM) has gained a lot of attention for solving large-scale and objective-separable constrained optimization. However, the two-block variable structure of the ADMM still limits the practical…

最优化与控制 · 数学 2020-03-24 Kresimir Mihic , Mingxi Zhu , Yinyu Ye

One of the most common, but at the same time expensive operations in linear algebra, is multiplying two matrices $A$ and $B$. With the rapid development of machine learning and increases in data volume, performing fast matrix intensive…

信息论 · 计算机科学 2020-11-20 Neophytos Charalambides , Mert Pilanci , Alfred Hero

A novel compressive-sensing based signal multiplexing scheme is proposed in this paper to further improve the multiplexing gain for multiple input multiple output (MIMO) system. At the transmitter side, a Gaussian random measurement matrix…

信息论 · 计算机科学 2016-04-05 Chanzi Liu , Qingchun Chen , Xiaohu Tang

Multidimensional Retiming is one of the most important optimization techniques to improve timing parameters of nested loops. It consists in exploring the iterative and recursive structures of loops to redistribute computation nodes on cycle…

编程语言 · 计算机科学 2012-05-22 Yaroub Elloumi , Mohamed Akil , Mohamed Hedi Bedoui

We present a superconducting micro-resonator array fabrication method that is scalable, reconfigurable, and has been optimized for high multiplexing factors. The method uses uniformly sized tiles patterned on stepper photolithography…

In many applications one wants to identify identical subtrees of a program syntax tree. This identification should ideally be robust to alpha-renaming of the program, but no existing technique has been shown to achieve this with good…

编程语言 · 计算机科学 2021-05-07 Krzysztof Maziarz , Tom Ellis , Alan Lawrence , Andrew Fitzgibbon , Simon Peyton Jones

A novel parallel algorithm for matrix multiplication is presented. The hyper-systolic algorithm makes use of a one-dimensional processor abstraction. The procedure can be implemented on all types of parallel systems. It can handle…

数学软件 · 计算机科学 2007-05-23 Thomas Lippert , Nikolay Petkov , Paolo Palazzari , Klaus Schilling

Factorization and multiplication of dense matrices and tensors are critical, yet extremely expensive pieces of the scientific toolbox. Careful use of low rank approximation can drastically reduce the computation and memory requirements of…

性能 · 计算机科学 2023-11-15 Sameer Deshmukh , Rio Yokota , George Bosilca

The escalating demands of compute-intensive applications urgently necessitate the adoption of optical interconnect technologies to overcome bottlenecks in scaling computing systems. This requires fully exploiting the inherent parallelism of…

Computation-in-Memory (CiM) is attracting attention as a technology that can perform MAC calculations required for AI accelerators, at high speed with low power consumption. However, there is a problem regarding power consumption and…

硬件体系结构 · 计算机科学 2025-07-21 Fuyuki Kihara , Seiji Uenohara , Satoshi Awamura , Naoko Misawa , Chihiro Matsui , Ken Takeuchi

The simulation of systems that act on multiple time scales is challenging. A stable integration of the fast dynamics requires a highly accurate approximation whereas for the simulation of the slow part, a coarser approximation is accurate…

数值分析 · 数学 2024-06-21 Sina Ober-Blöbaum , Theresa Wenger , Tobias Gail , Sigrid Leyendecker

Multiplexers based on the modulation of superconducting quantum interference devices are now regularly used in multi-kilopixel arrays of superconducting detectors for astrophysics, cosmology, and materials analysis. Over the next decade,…

天体物理仪器与方法 · 物理学 2012-05-02 K. D. Irwin , H. M. Cho , W. B. Doriese , J. W. Fowler , G. C. Hilton , M. D. Niemack , C. D. Reintsema , D. R. Schmidt , J. N. Ullom , L. R. Vale

In what ways could cellular massive MIMO be improved? This technology has already been shown to bring huge performance gains. However, coverage holes and difficulties to transmit multiple streams to multi-antenna users because of…

信号处理 · 电气工程与系统科学 2025-01-08 Sara Willhammar , Hiroki Iimori , Joao Vieira , Lars Sundström , Fredrik Tufvesson , Erik G. Larsson

In this work, we optimally solve the problem of multiplierless design of second-order Infinite Impulse Response filters with minimum number of adders. Given a frequency specification, we design a stable direct form filter with…

硬件体系结构 · 计算机科学 2022-05-11 Rémi Garcia , Anastasia Volkova , Martin Kumm , Alexandre Goldsztejn , Jonas Kühle

Recognizing the explosive increase in the use of AI-based applications, several industrial companies developed custom ASICs (e.g., Google TPU, IBM RaPiD, Intel NNP-I/NNP-T) and constructed a hyperscale cloud infrastructure with them. These…

硬件体系结构 · 计算机科学 2025-03-03 Seock-Hwan Noh , Seungpyo Lee , Banseok Shin , Sehun Park , Yongjoo Jang , Jaeha Kung
‹ 上一页 1 8 9 10 下一页 ›