中文
相关论文

相关论文: Reactive Failure Mitigation through Seamless Migra…

200 篇论文

The transmission grid is often comprised of several control areas that are connected by multiple tie lines in a mesh structure for reliability. It is also well-known that line failures can propagate non-locally and redundancy can exacerbate…

系统与控制 · 电气工程与系统科学 2020-04-23 Chen Liang , Linqi Guo , Alessandro Zocca , Shuyue Yu , Steven H. Low , Adam Wierman

Penetration testing is a well-established practical concept for the identification of potentially exploitable security weaknesses and an important component of a security audit. Providing a holistic security assessment for networks…

密码学与安全 · 计算机科学 2019-01-07 Patrick Speicher , Marcel Steinmetz , Jörg Hoffmann , Michael Backes , Robert Künnemann

Failures in Task-based Parallel Programming (TBPP) can severely degrade performance and result in incomplete or incorrect outcomes. Existing failure-handling approaches, including reactive, proactive, and resilient methods such as retry and…

分布式、并行与集群计算 · 计算机科学 2025-03-31 Sicheng Zhou , Zhuozhao Li , Valérie Hayot-Sasson , Haochen Pan , Maxime Gonthier , J. Gregory Pauloski , Ryan Chard , Kyle Chard , Ian Foster

In current scenario several commercial and social organizations are using computer networks for their business and management purposes. In order to meet the business requirements networks are also grow. The growth of network also promotes…

分布式、并行与集群计算 · 计算机科学 2014-02-03 Bhagvan Krishna Gupta , Ankit Mundra , Nitin Rakesh

In networks, availability is of paramount importance. As link failures are disruptive, modern networks in turn provide Fast ReRoute (FRR) mechanisms to rapidly restore connectivity. However, existing FRR approaches heavily impact…

网络与互联网体系结构 · 计算机科学 2021-11-30 Apoorv Shukla , Klaus-Tycho Foerster

Redundancy is widely used to sustain service continuity in programmable and virtualized networks; however, replicated functions often share platforms, software stacks, and control dependencies, making them vulnerable to correlated failures.…

网络与互联网体系结构 · 计算机科学 2026-05-06 Mohamed Khalafalla Hassan , Indrakshi Dey

Scaling to larger systems, with current levels of reliability, requires cost-effective methods to mitigate hardware failures. One of the main causes of hardware failure is an uncorrected error in memory, which terminates the current job and…

分布式、并行与集群计算 · 计算机科学 2024-09-06 Isaac Boixaderas , Sergi Moré , Javier Bartolome , David Vicente , Petar Radojković , Paul M. Carpenter , Eduard Ayguadé

Network Functions Virtualization (NFV) allows flexibility, scalability, agility, and easy manageability of networks by leveraging the features of virtualization and cloud computing technologies. However, softwarization of network functions…

网络与互联网体系结构 · 计算机科学 2020-12-15 Prabhu Kaliyammal Thiruvasagam , Vijeth J. Kotagi , C. Siva Ram Murthy

Network structures and models have been widely adopted, e.g., for Internet of Things, wireless sensor networks, smart grids, transportation networks, communication networks, social networks, and computer grid systems. Network reliability is…

数据结构与算法 · 计算机科学 2020-04-20 Wei-Chang Yeh

Computing at the exascale level is expected to be affected by a significantly higher rate of faults, due to increased component counts as well as power considerations. Therefore, current day numerical algorithms need to be reexamined as to…

数值分析 · 数学 2019-05-27 Mark Ainsworth , Christian Glusa

The disaggregation of base stations into discrete RAN functions introduces new threats to mobile networks, as failures in one RAN function can trigger cascading failures and interrupt entire function chains, with potential to degrade…

网络与互联网体系结构 · 计算机科学 2026-05-18 Gabriel Almeida , Jacek Kibiłda , Joao F. Santos , Kleber Vieira Cardoso

Modern ML applications increasingly rely on complex deep learning models and large datasets. There has been an exponential growth in the amount of computation needed to train the largest models. Therefore, to scale computation and data,…

机器学习 · 计算机科学 2023-09-26 Hamidreza Almasi , Harsh Mishra , Balajee Vamanan , Sathya N. Ravi

Satellite constellation systems are becoming more attractive to provide communication services worldwide, especially in areas without network connectivity. While optimizing satellite gateway placement is crucial for operators to minimize…

系统与控制 · 电气工程与系统科学 2026-02-04 Yuma Abe , Flor Ortiz , Eva Lagunas , Victor Monzon Baeza , Symeon Chatzinotas , Hiroyuki Tsuji

Massive machine-type communications (mMTC) are fundamental to the Internet of Things (IoT) framework in future wireless networks, involving the connection of a vast number of devices with sporadic transmission patterns. Traditional device…

信号处理 · 电气工程与系统科学 2025-10-22 Xinjue Wang , Esa Ollila , Sergiy A. Vorobyov

The maximum possible throughput (or the rate of job completion) of a multi-server system is typically the sum of the service rates of individual servers. Recent work shows that launching multiple replicas of a job and canceling them as soon…

分布式、并行与集群计算 · 计算机科学 2020-12-29 Gauri Joshi , Dhruva Kaushal

Ensuring reliable operation of large power systems subjected to multiple outages is a challenging task because of the combinatorial nature of the problem. Traditional approaches for security assessment are often limited by their scope…

系统与控制 · 电气工程与系统科学 2021-05-03 Reetam Sen Biswas , Anamitra Pal , Trevor Werho , Vijay Vittal

The binary-state network, a basic network, and its components are either working or failed. It is fundamental to all types of current networks, such as utility networks (gas, water, electricity, and 4G/5G), the Internet of Things (IoT),…

系统与控制 · 电气工程与系统科学 2022-12-26 Wei-Chang Yeh

Software-defined networking offers numerous benefits against the legacy networking systems through simplifying the process of network management and reducing the cost of network configuration. Currently, the management of failures in the…

网络与互联网体系结构 · 计算机科学 2019-04-02 Ali Malik , Benjamin Aziz , Mo Adda , Chih-Heng Ke

Intermittent faults are transient errors that sporadically appear and disappear. Although intermittent faults pose substantial challenges to reliability and coordination, existing studies of fault tolerance in robot swarms focus instead on…

机器人学 · 计算机科学 2025-09-24 Sinan Oğuz , Emanuele Garone , Marco Dorigo , Mary Katherine Heinrich

A natural way for cooperative tasking in multi-agent systems is through a top-down design by decomposing a global task into sub-tasks for each individual agent such that the accomplishments of these sub-tasks will guarantee the achievement…

系统与控制 · 计算机科学 2015-03-17 Mohammad Karimadini , Hai Lin