English
Related papers

Related papers: LB Scalability: Achieving the Right Balance Betwee…

200 papers

As compute power increases with time, more involved and larger simulations become possible. However, it gets increasingly difficult to efficiently use the provided computational resources. Especially in particle-based simulations with a…

Distributed, Parallel, and Cluster Computing · Computer Science 2019-08-05 Sebastian Eibl , Ulrich Rüde

Distributed multi-controller deployment is a promising method to achieve a scalable and reliable control plane of Software-Defined Networking (SDN). However, it brings a new challenge for balancing loads on the distributed controllers as…

Networking and Internet Architecture · Computer Science 2018-01-29 Tao Hu , Julong Lan , Jianhui Zhang , Wei Zhao

Low Latency, Low Loss Scalable throughput (L4S) is being proposed as the new default Internet service. L4S can be considered as an `incrementally deployable clean-slate' for new Internet flow-rate control mechanisms. Because, for a brief…

Networking and Internet Architecture · Computer Science 2019-04-17 Bob Briscoe , Koen De Schepper

Large language models (LLMs) iteratively generate text token by token, with memory usage increasing with the length of generated token sequences. Since the request generation length is generally unpredictable, it is difficult to estimate…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-03-11 Ke Cheng , Wen Hu , Zhi Wang , Hongen Peng , Jianguo Li , Sheng Zhang

The Vision of Autonomic Computing (ACV), proposed over two decades ago, envisions computing systems that self-manage akin to biological organisms, adapting seamlessly to changing environments. Despite decades of research, achieving ACV…

Artificial Intelligence · Computer Science 2024-07-22 Zhiyang Zhang , Fangkai Yang , Xiaoting Qin , Jue Zhang , Qingwei Lin , Gong Cheng , Dongmei Zhang , Saravan Rajmohan , Qi Zhang

The explosive growth of Large Language Models (LLMs), such as GPT-4 with 1.8 trillion parameters, demands a fundamental rethinking of data center architecture to ensure scalability, efficiency, and cost-effectiveness. Our work provides a…

Hardware Architecture · Computer Science 2025-09-09 Jesmin Jahan Tithi , Hanjiang Wu , Avishaii Abuhatzera , Fabrizio Petrini

Although modern, AI-centric datacenters heavily rely on SmartNICs, existing devices impose a hard trade-off. Commercial SmartNICs provide high bandwidth and easy software integration, but offer limited support for customization and data…

Hardware Architecture · Computer Science 2026-04-17 Benjamin Ramhorst , Maximilian Jakob Heer , Luhao Liu , Heejae Kim , Jonas Dann , Jin-Soo Kim , Gustavo Alonso

Virtualisation first and cloud computing later has led to a consolidation of workload in data centres that also comprises latency-sensitive application domains such as High Performance Computing and telecommunication. These types of…

Networking and Internet Architecture · Computer Science 2020-04-20 Kamil Tokmakov , Mitalee Sarker , Jörg Domaschka , Stefan Wesner

Modern distributed databases face challenges in achieving transactional consistency across distributed partitions. Traditional two-phase commit (2PC) protocols incur high coordination overhead and latency, and require complex recovery for…

Databases · Computer Science 2026-03-03 Quanqing Xu , Chen Qian , Chuanhui Yang , Fanyu Kong , Guixiang Liu , Fusheng Han , Zixiang Zhai

Network function virtualization is a promising technology to simultaneously support multiple services with diverse characteristics and requirements in the 5G and beyond networks. In particular, each service consists of a predetermined…

Networking and Internet Architecture · Computer Science 2021-06-08 Wei-Kun Chen , Ya-Feng Liu , Antonio De Domenico , Zhi-Quan Luo , Yu-Hong Dai

Modern computing workloads are often composed of parallelizable jobs. A parallelizable job can be completed more quickly when run on additional servers. However, each job can only use a limited number of servers, known as its…

Performance · Computer Science 2025-12-30 Benjamin Berg , Benjamin Moseley , Weina Wang , Mor Harchol-Balter

Atomic multicast is a communication primitive used in dependable systems to ensure consistent ordering of messages delivered to a set of replica groups. This primitive enables critical services to integrate replication and sharding (i.e.,…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-09-10 Lorenzo Martignetti , Eliã Batista , Gianpaolo Cugola , Fernando Pedone

End users face a choice between privacy and efficiency in current Large Language Model (LLM) service paradigms. In cloud-based paradigms, users are forced to compromise data locality for generation quality and processing speed. Conversely,…

Artificial Intelligence · Computer Science 2023-11-27 Yiming Wang , Yu Lin , Xiaodong Zeng , Guannan Zhang

This paper studies a 2-class, 2-server parallel server system under the recently introduced extended heavy traffic condition, which states that the underlying 'static allocation' linear program (LP) is critical, but does not require that it…

Optimization and Control · Mathematics 2022-07-19 Rami Atar , Eyal Castiel , Marty Reiman

The Low Latency, Low Loss, Scalable Throughput (L4S) architecture has the potential to reduce queuing delay when it is deployed at endpoints and routers throughout the Internet. However, it is not clear how TCP Prague, a prototype scalable…

Networking and Internet Architecture · Computer Science 2024-07-02 Fatih Berkay Sarpkaya , Ashutosh Srivastava , Fraida Fund , Shivendra Panwar

In this method, service of one load balancer can be borrowed or shared among other load balancers when any correction is needed in the estimation of the load.

Distributed, Parallel, and Cluster Computing · Computer Science 2018-11-06 Sreelekshmi S , K R Remesh Babu

Ubiquitous cell-free massive MIMO (multiple-input multiple-output) combines massive MIMO technology and user-centric transmission in a distributed architecture. All the access points (APs) in the network cooperate to jointly and coherently…

Information Theory · Computer Science 2019-09-09 Giovanni Interdonato , Pål Frenger , Erik G. Larsson

In this paper, we propose a novel cloud-native architecture for collaborative agentic network slicing. Our approach addresses the challenge of managing shared infrastructure, particularly CPU resources, across multiple network slices with…

Networking and Internet Architecture · Computer Science 2025-02-18 Juan Sebastián Camargo , Farhad Rezazadeh , Hatim Chergui , Shuaib Siddiqui , Lingjia Liu

A PAM4 based direct detection system has been standardized for short-distance data center interconnects because of its simple architecture. Performance of the PAM4 systems is limited for high dispersion values or demands complicated signal…

Signal Processing · Electrical Eng. & Systems 2020-08-26 Rashmi Kamran , Shalabh Gupta

Load balancing is among the basic primitives in distributed computing. In this paper, we consider this problem when executed locally on a network with nodes prone to failures. We show that there exist lightweight network topologies that are…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-08-05 Dariusz R. Kowalski , Jan Olkowski
‹ Prev 1 8 9 10 Next ›