English
Related papers

Related papers: Explicit Steady-State Approximations for Parallel …

200 papers

We study a two-type server queueing system where flexible Type-I servers, upon their initial interaction with jobs, decide in real time whether to process them independently or in collaboration with dedicated Type-II servers. Independent…

Optimization and Control · Mathematics 2026-01-21 Shuwen Lu , Mark E. Lewis , Jamol Pender

This paper presents a systematic review of mapping and scheduling strategies within the High-Performance Computing (HPC) compute continuum, with a particular emphasis on heterogeneous systems. It introduces a prototype workflow to establish…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-05-19 Aasish Kumar Sharma , Julian Kunkel

Randomized load balancing networks arise in a variety of applications, and allow for efficient sharing of resources, while being relatively easy to implement. We consider a network of parallel queues in which incoming jobs with independent…

Probability · Mathematics 2017-10-13 Reza Aghajani , Kavita Ramanan

We consider a multi-server queue in the Halfin-Whitt regime: as the number of servers $n$ grows without a bound, the utilization approaches 1 from below at the rate $\Theta(1/\sqrt{n})$. Assuming that the service time distribution is…

Probability · Mathematics 2008-03-19 David Gamarnik , Petar Momcilovic

In this paper, a many-sources large deviations principle (LDP) for the transient workload of a multi-queue single-server system is established where the service rates are chosen from a compact, convex and coordinate-convex rate region and…

Probability · Mathematics 2009-02-27 Vijay G. Subramanian , Tara Javidi , Somsak Kittipiyakul

The problem of minimizing mean response time of generic jobs submitted to a heterogenous distributed computer systems is considered in this paper. A static load balancing strategy, in which decision of redistribution of loads does not…

Distributed, Parallel, and Cluster Computing · Computer Science 2011-11-09 S. A. Mondal

In modern computer systems, jobs are divided into short tasks and executed in parallel. Empirical observations in practical systems suggest that the task service times are highly random and the job service time is bottlenecked by the…

Performance · Computer Science 2017-02-08 Yin Sun , C. Emre Koksal , Ness B. Shroff

Coordinating time-sensitive deliveries in environments like hospitals poses a complex challenge, particularly when managing multiple online pickup and delivery requests within strict time windows using a team of heterogeneous robots.…

Robotics · Computer Science 2025-05-14 Ashish Verma , Avinash Gautam , Tanishq Duhan , V. S. Shekhawat , Sudeept Mohan

Cloud computing is emerging as an important platform for business, personal and mobile computing applications. In this paper, we study a stochastic model of cloud computing, where jobs arrive according to a stochastic process and request…

Performance · Computer Science 2012-06-07 Siva Theja Maguluri , R Srikant , Lei Ying

To make better use of file diversity provided by random caching and improve the successful transmission probability (STP) of a file, we consider retransmissions with random discontinuous transmission (DTX) in a large-scale cache-enabled…

Information Theory · Computer Science 2018-02-13 Wanli Wen , Ying Cui , Fu-Chun Zheng , Shi Jin , Yanxiang Jiang

Motivated by distributed schedulers that combine the power-of-d-choices with late binding and systems that use replication with cancellation-on-start, we study the performance of the LL(d) policy which assigns a job to a server that…

Performance · Computer Science 2018-02-16 Tim Hellemans , Benny Van Houdt

We consider the problem of load balancing in parallel queues by transferring customers between them at discrete points in time. Holding costs accrue as customers wait in the queue, while transfer decisions incur both fixed (setup) costs and…

Optimization and Control · Mathematics 2025-08-13 Timothy C. Y. Chan , Jangwon Park , Vahid Sarhangian

This paper studies the heavy-traffic (HT) behaviour of queueing networks with a single roving server. External customers arrive at the queues according to independent renewal processes and after completing service, a customer either leaves…

Probability · Mathematics 2016-11-09 Marko Boon , Rob van der Mei , Erik Winands

We study deterministic fluid approximations of parallel service systems operating under first come first served policy (FCFS). The condition for complete resource pooling is identified in terms of the system structure and the customer…

Probability · Mathematics 2016-04-18 Yuval Nov , Gideon Weiss , Hanqin Zhang

To deliver high performance in power limited systems, architects have turned to using heterogeneous systems, either CPU+GPU or mixed CPU-hardware systems. However, in systems with different processor types and task affinities, scheduling…

Performance · Computer Science 2017-12-12 Zhuo Chen , Diana Marculescu

This paper presents a new simulation-based approach to address the stochastic Dynamic Traffic Assignment (DTA) problem, focusing on large congested networks and dynamic settings. The proposed methodology incorporates a random walk model…

Multiagent Systems · Computer Science 2023-11-22 Kaveh Khoshkhah , Mozhgan Pourmoradnasseri , Sadok Ben Yahia , Amnir Hadachi

We study how to design edge server placement and server scheduling policies under workload uncertainty for 5G networks. We introduce a new metric called resource pooling factor to handle unexpected workload bursts. Maximizing this metric…

Networking and Internet Architecture · Computer Science 2021-04-30 Shizhen Zhao , Xiao Zhang , Peirui Cao , Xinbing Wang

To support operations and passenger-facing services, transit agencies need reliable passenger load trajectories. Currently, load estimates are typically inferred from imperfect sensing systems rather than fully observed, and the accuracy of…

Machine Learning · Computer Science 2026-05-20 Yiyao Xu , Hao Zhou , Yuhang Wang , Jingran Sun

We present a framework for Multi-Robot Task Allocation (MRTA) in heterogeneous teams performing long-endurance missions in dynamic scenarios. Given the limited battery of robots, especially for aerial vehicles, we allow for robot recharges…

Robotics · Computer Science 2025-11-27 Alvaro Calvo , Jesus Capitan

We analyze the performance of redundancy in a multi-type job and multi-type server system. We assume the job dispatcher is unaware of the servers' capacities, and we set out to study under which circumstances redundancy improves the…

Networking and Internet Architecture · Computer Science 2020-12-16 Elene Anton , Urtzi Ayesta , Matthieu Jonckheere , Ina Verloop