English
Related papers

Related papers: MCFlash: Bulk Bitwise Processing in 3D NAND with D…

200 papers

With the versatile manipulation capability, programmable metasurfaces are rapidly advancing in their intelligence, integration, and commercialization levels. However, as the programmable metasurfaces scale up, their control configuration…

Micro/nano electro-mechanical systems (MEMS/NEMS) are constantly attracting an increasing attention for their relevant technological applications in fields ranging from biology, medicine, ecology, energy to industry. Most of the…

Materials Science · Physics 2021-03-10 Omar Tricinci , Marco Carlotti , Andrea Desii , Fabian Meder , Virgilio Mattoli

As the amount of data produced in society continues to grow at an exponential rate, modern applications are incurring significant performance and energy penalties due to high data movement between the CPU and memory/storage. While…

Hardware Architecture · Computer Science 2024-03-12 Ryan Wong , Nikita Kim , Kevin Higgs , Sapan Agarwal , Engin Ipek , Saugata Ghose , Ben Feinberg

Deep neural networks (DNN) use a wide range of network topologies to achieve high accuracy within diverse applications. This model diversity makes it impossible to identify a single "dataflow" (execution schedule) to perform optimally…

Hardware Architecture · Computer Science 2024-06-24 Man Shi , Steven Colleman , Charlotte VanDeMieroop , Antony Joseph , Maurice Meijer , Wim Dehaene , Marian Verhelst

Network-on-Chip (NoC) paradigm has been proposed as an auspicious solution to handle the strict communication requirements between the increasingly large number of cores on a single multi and many-core chips. However, NoC systems are…

Hardware Architecture · Computer Science 2020-03-24 Khanh N. Dang , Yuichi Okuyama , Abderazek Ben Abdallah

We report the realization of a completely controllable high-speed nanomechanical memory element fabricated from single-crystal silicon wafers. This element consists of a doubly-clamped suspended nanomechanical beam structure, which can be…

Other Condensed Matter · Physics 2007-05-23 Robert L. Badzey , Guiti Zolfagharkhani , Alexei Gaidarzhy , Pritiraj Mohanty

In-memory computing is a promising alternative to traditional computer designs, as it helps overcome performance limits caused by the separation of memory and processing units. However, many current approaches struggle with unreliable…

Modern multicore processors are employing large last-level caches, for example Intel's E7-8800 processor uses 24MB L3 cache. Further, with each CMOS technology generation, leakage energy has been dramatically increasing and hence, leakage…

Hardware Architecture · Computer Science 2013-10-17 Sparsh Mittal

GPUs are the heart of the latest generations of supercomputers. We efficiently accelerate a compressible multiphase flow solver via OpenACC on NVIDIA and AMD Instinct GPUs. Optimization is accomplished by specifying the directive clauses…

Processing-using-DRAM has been proposed for a limited set of basic operations (i.e., logic operations, addition). However, in order to enable full adoption of processing-using-DRAM, it is necessary to provide support for more complex…

Flash-algorithm track-reconstruction routines with speed factors 3000-4000 in excess those of traditional iterative routines are presented. The methods were successfully tested in the alignment of the Test Beam setup for the ATLAS Pixel…

High Energy Physics - Experiment · Physics 2016-08-31 M. Dima

In this paper we report a semianalytical model of the mass transfer impedance in microfluidic electrochemical chips (MEC). It is based on the molar advection diffusion equation in a microfluidic channel with a Poiseuille flow and an…

Chemical Physics · Physics 2022-10-18 Stéphane Chevalier , Marine Garcia , Alain Sommier , Jean-Christophe Batsale

Large-scale GPU clusters are widely-used to speed up both latency-critical (online) and best-effort (offline) deep learning (DL) workloads. However, most DL clusters either dedicate each GPU to one workload or share workloads in time,…

Distributed, Parallel, and Cluster Computing · Computer Science 2023-03-27 Yihao Zhao , Xin Liu , Shufan Liu , Xiang Li , Yibo Zhu , Gang Huang , Xuanzhe Liu , Xin Jin

Y-Flash memristors utilize the mature technology of single polysilicon floating gate non-volatile memories (NVM). It can be operated in a two-terminal configuration similar to the other emerging memristive devices, i.e., resistive…

Today's microprocessors have grown significantly in complexity and functionality. Most of today's processors provide at least three levels of memory hierarchy, are heavily pipelined, and support some sort of cache coherency protocol. These…

Hardware Architecture · Computer Science 2020-09-02 Mitul S Nagar , Haresh A Suthar , Chintan Panchal

We theoretically demonstrate that atomically-precise ``nanoscale processing" can be reproducibly performed by adaptive laser pulses. We present the new approach on the controlled welding of crossed carbon nanotubes, giving various…

Materials Science · Physics 2009-11-07 Petr Král

Microfluidic mixing is a fundamental functionality in most lab on a chip (LOC) systems,whereas realization of efficient mixing is challenging in microfluidic channels due to the small Reynolds numbers. Here, we design and fabricate a…

Applied Physics · Physics 2020-01-10 Wenbo Li , Wei Chu , Peng Wang , Jia Qi , Zhe Wang , Jintian Lin , Min Wang , Ya Cheng

Magnetic random-access memory (MRAM) is a promising memory technology due to its high density, non-volatility, and high endurance. However, achieving high memory fidelity incurs significant write-energy costs, which should be reduced for…

Emerging Technologies · Computer Science 2021-12-07 Yongjune Kim , Yoocharn Jeon , Hyeokjin Choi , Cyril Guyot , Yuval Cassuto

Crossbar arrays using emerging non-volatile memory technologies such as Resistive RAM (ReRAM) offer high density, fast access speed and low-power. However the bandwidth of the crossbar is limited to single-bit read/write per access to avoid…

Emerging Technologies · Computer Science 2016-06-03 Mohammad Nasim Imtiaz Khan , Swaroop Ghosh , Radha Krishna Aluru , Rashmi Jha

Estimating instruction-level throughput is critical for many applications: multimedia, low-latency networking, medical, automotive, avionic, and industrial control systems all rely on tightly calculable and accurate timing bounds of their…

Programming Languages · Computer Science 2023-05-18 Min-Yih Hsu , Felicitas Hetzelt , David Gens , Michael Maitland , Michael Franz