English
Related papers

Related papers: Preliminary Design of Scalable Hardware Integrated…

200 papers

The present paper describes the design of a cost effective, better resolution data acquisition system (DAS) which is compatible to most of the PC and laptops. A low cost DAS has been designed using PIC12F675 having 4-channel analog input…

Hardware Architecture · Computer Science 2012-07-10 N. Monoranjan Singh , K. C. Sarma

Federated fine-tuning provides a practical route to adapt large language models (LLMs) on edge devices without centralizing private data, yet in mobile deployments the training wall-clock is often bottlenecked by straggler-limited uplink…

Machine Learning · Computer Science 2026-04-29 Changyu Li , Shuanghong Huang , Jiashen Liu , Ming Lei , Jidu Xing , Kaishun Wu , Lu Wang , Fei Luo

Fluid antenna systems (FAS) exploit antenna position reconfigurability to unlock massive spatial diversity within compact form factors, making them a promising enabler for 6G user terminals (UTs). However, practical port switching incurs…

Information Theory · Computer Science 2026-05-08 Xusheng Zhu , Kai-Kit Wong , Hao Xu , Chenguang Rao , Hyundong Shin

In the big data era of observational oceanography, passive acoustics datasets are becoming too high volume to be processed on local computers due to their processor and memory limitations. As a result there is a current need for our…

Distributed, Parallel, and Cluster Computing · Computer Science 2019-06-10 Paul Nguyen Hong Duc , Dorian Cazau

Approximate matrix inversion based methods is widely used for linear massive multiple-input multiple-output (MIMO) received symbol vector detection. Such detectors typically utilize the diagonally dominant channel matrix of a massive MIMO…

Information Theory · Computer Science 2020-11-02 Shahriar Shahabuddin , Mahmoud A. Albreem , Mohammad Shahanewaz Shahabuddin , Zaheer Khan , Markku Juntti

Fast and energy-efficient low-bitwidth floating-point (FP) arithmetic is essential for Artificial Intelligence (AI) systems. Microscaling (MX) standardized formats have recently emerged as a promising alternative to baseline low-bitwidth FP…

Hardware Architecture · Computer Science 2025-05-20 Gamze İslamoğlu , Luca Bertaccini , Arpan Suravi Prasad , Francesco Conti , Angelo Garofalo , Luca Benini

Many small and medium-sized physics experiments are being conducted worldwide. These experiments have similar requirements for readout electronics, especially the back-end electronics. Some experiments need a trigger logic unit(TLU) to…

Instrumentation and Detectors · Physics 2024-07-10 Jianguo Liu , Yu Wang , Changqing Feng , Shubin Liu , Qian Chen

Integrated sensing and communication (ISAC) is a key enabler for future radio networks. This paper presents a sub-band full-duplex (SBFD) ISAC system that assigns non-overlapping OFDM subbands to sensing and communication, enabling…

Signal Processing · Electrical Eng. & Systems 2026-02-03 Bixing Yan , Kwadwo Mensah Obeng Afrane , Achiel Colpaert , Andre Kokkeler , Sofie Pollin , Yang Miao

Scaling up hardware systems has become an important tactic for improving performance as Moore's law fades. Unfortunately, simulations of large hardware systems are often a design bottleneck due to slow throughput and long build times. In…

Distributed, Parallel, and Cluster Computing · Computer Science 2024-07-31 Steven Herbst , Noah Moroze , Edgar Iglesias , Andreas Olofsson

The adoption of low-voltage direct current sections within grid architectures is emerging as a promising design option in the naval sector. This paper presents a preliminary comparative assessment of three different grid topologies, using…

Systems and Control · Electrical Eng. & Systems 2025-09-29 D. Roncagliolo , M. Gallo , D. Kaza , F. D'Agostino , A. Chiarelli , F. Silvestro

In this paper, we propose DEEPSERVE, a scalable and serverless AI platform designed to efficiently serve large language models (LLMs) at scale in cloud environments. DEEPSERVE addresses key challenges such as resource allocation, serving…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-06-10 Junhao Hu , Jiang Xu , Zhixia Liu , Yulong He , Yuetao Chen , Hao Xu , Jiang Liu , Jie Meng , Baoquan Zhang , Shining Wan , Gengyuan Dan , Zhiyu Dong , Zhihao Ren , Changhong Liu , Tao Xie , Dayun Lin , Qin Zhang , Yue Yu , Hao Feng , Xusheng Chen , Yizhou Shan

The A64FX CPU is arguably the most powerful Arm-based processor design to date. Although it is a traditional cache-based multicore processor, its peak performance and memory bandwidth rival accelerator devices. A good understanding of its…

Performance · Computer Science 2021-08-05 Christie Alappat , Nils Meyer , Jan Laukemann , Thomas Gruber , Georg Hager , Gerhard Wellein , Tilo Wettig

Mass characterisation of emerging memory devices is an essential step in modelling their behaviour for integration within a standard design flow for existing integrated circuit designers. This work develops a novel characterisation platform…

Emerging virtualized radio access networks (vRANs) demand flexible and efficient baseband processing across heterogeneous compute substrates. In this paper, we present DecodeX, a unified benchmarking framework for evaluating low-density…

Networking and Internet Architecture · Computer Science 2025-11-06 Zhenzhou Qi , Yuncheng Yao , Yiming Li , Chung-Hsuan Tung , Junyao Zheng , Danyang Zhuo , Tingjun Chen

NVFP4 has grown increasingly popular as a 4-bit format for quantizing large language models due to its hardware support and its ability to retain useful information with relatively few bits per parameter. However, the format is not without…

Computation and Language · Computer Science 2026-03-31 Jack Cook , Hyemin S. Lee , Kathryn Le , Junxian Guo , Giovanni Traverso , Anantha P. Chandrakasan , Song Han

In today's era of "scale-out", this paper makes the case that a specialized hardware architecture based on "scale-in"--placing as many specialized processors as possible along with their memory systems and interconnect links within one or…

Hardware Architecture · Computer Science 2020-09-14 Suresh Krishna , Ravi Krishna

CuboDAQ is a custom data acquisition system to read out SiPM-based detectors. It features electronic boards to digitize the SiPMs signal, an FPGA-based system-on-module board, the connectivity to transmit the data to a central server, and…

Instrumentation and Detectors · Physics 2025-01-10 Ettore Zaffaroni , Guido Haefeli

We have developed a prototype system for the ILC vertex detector based on DEPFET pixels. The system operates a 128x64 matrix (with ~35x25 square micron large pixels) and uses two dedicated microchips, the SWITCHER II chip for matrix…

Ultra-reliable and low-latency communication (URLLC) is one of the three major service classes supported by the fifth generation (5G) New Radio (NR) technical specifications. In this paper, we introduce a physical layer architecture that…

Signal Processing · Electrical Eng. & Systems 2020-02-17 Zhiwen Wu , Tianpeng Yang , Haitao Liu , Geng Shi , Kan Zheng

A new TDC-chip is under development for the COMPASS experiment at CERN. The ASIC, which exploits the 0.6 micrometer CMOS sea-of-gate technology, will allow high resolution time measurements with digitization of 75 ps, and an unprecedented…

High Energy Physics - Experiment · Physics 2007-05-23 G. Braun , H. Fischer , J. Franz , A. Grunemaier , F. H. Heinsius , K. Konigsmann , M. Schierloh , T. Schmidt , H. Schmitt , H. J. Urban