English
Related papers

Related papers: Cooperative Repair of Multiple Node Failures in Di…

200 papers

Given the scale of today's distributed storage systems, the failure of an individual node is a common phenomenon. Various metrics have been proposed to measure the efficacy of the repair of a failed node, such as the amount of data download…

Information Theory · Computer Science 2015-01-21 Gaurav Kumar Agarwal , Birenjith Sasidharan , P. Vijay Kumar

This paper presents a novel construction of $(n,k,d=n-1)$ access-optimal regenerating codes for an arbitrary sub-packetization level $\alpha$ for exact repair of any systematic node. We refer to these codes as general sub-packetized because…

Information Theory · Computer Science 2020-02-14 Katina Kralevska , Danilo Gligoroski , Harald Øverby

Minimum storage regenerating (MSR) codes are a class of maximum distance separable (MDS) array codes capable of repairing any single failed node by downloading the minimum amount of information from each of the helper nodes. However, MSR…

Information Theory · Computer Science 2024-08-30 Vinayak Ramkumar , Netanel Raviv , Itzhak Tamo

The optimal tradeoff between node storage and repair bandwidth is an important issue for distributed storage systems (DSSs). As for realistic DSSs with clusters, when repairing a failed node, it is more efficient to download more data from…

Distributed, Parallel, and Cluster Computing · Computer Science 2019-01-08 Jingzhao Wang , Tinghan Wang , Yuan Luo

The explosion in the volumes of data being stored online has resulted in distributed storage systems transitioning to erasure coding based schemes. Yet, the codes being deployed in practice are fairly short. In this work, we address what we…

Information Theory · Computer Science 2016-09-22 Parikshit Gopalan , Guangda Hu , Swastik Kopparty , Shubhangi Saraf , Carol Wang , Sergey Yekhanin

Locally repairable codes enables fast repair of node failure in a distributed storage system. The code symbols in a codeword are stored in different storage nodes, such that a disk failure can be recovered by accessing a small fraction of…

Information Theory · Computer Science 2021-12-13 Kenneth W. Shum , Jie Hao

This paper considers capacity-achieving coding for the clustered form of distributed storage that reflects practical storage networks. To reflect the clustered structure with limited cross-cluster communication bandwidths, nodes in the same…

Information Theory · Computer Science 2019-08-05 Jy-yong Sohn , Beongjun Choi , Jaekyun Moon

We propose repair pipelining, a technique that speeds up the repair performance in general erasure-coded storage. By carefully scheduling the repair of failed data in small-size units across storage nodes in a pipelined manner, repair…

Distributed, Parallel, and Cluster Computing · Computer Science 2020-11-23 Xiaolu Li , Zuoru Yang , Jinhong Li , Runhui Li , Patrick P. C. Lee , Qun Huang , Yuchong Hu

We study a generalization of the setting of regenerating codes, motivated by applications to storage systems consisting of clusters of storage nodes. There are $n$ clusters in total, with $m$ nodes per cluster. A data file is coded and…

Information Theory · Computer Science 2018-08-07 N. Prakash , Vitaly Abdrashitov , Muriel Medard

In distributed storage systems built using commodity hardware, it is necessary to have data redundancy in order to ensure system reliability. In such systems, it is also often desirable to be able to quickly repair storage nodes that fail.…

Information Theory · Computer Science 2012-01-24 Joseph C. Koo , John Gill

A network coding-based scheme is proposed to improve the energy efficiency of distributed storage systems in WSNs (wireless sensor networks), which mainly focuses on two problems: firstly, consideration is given to effective distributed…

Networking and Internet Architecture · Computer Science 2013-04-08 Wang Lei , Yang Yuwang , Zhao Wei , Lu Wei

In large scale distributed storage systems (DSS) deployed in cloud computing, correlated failures resulting in simultaneous failure (or, unavailability) of blocks of nodes are common. In such scenarios, the stored data or a content of a…

Information Theory · Computer Science 2014-06-30 Gokhan Calis , O. Ozan Koyluoglu

Distributed data storage systems are essential to deal with the need to store massive volumes of data. In order to make such a system fault-tolerant, some form of redundancy becomes crucial, incurring various overheads - most prominently in…

Distributed, Parallel, and Cluster Computing · Computer Science 2011-12-25 Frederique Oggier , Anwitaman Datta

In large scale distributed storage systems (DSS) deployed in cloud computing, correlated failures resulting in simultaneous failure (or, unavailability) of blocks of nodes are common. In such scenarios, the stored data or a content of a…

Information Theory · Computer Science 2017-02-22 Gokhan Calis , O. Ozan Koyluoglu

This paper presents a flexible irregular model for heterogeneous cloud storage systems and investigates how the cost of repairing failed nodes can be minimized. The fractional repetition code, originally designed for minimizing repair…

Information Theory · Computer Science 2016-11-18 Quan Yu , Chi Wan Sung , Terence H. Chan

In this paper we extend the notion of {\em locally repairable} codes to {\em secret sharing} schemes. The main problem that we consider is to find optimal ways to distribute shares of a secret among a set of storage-nodes (participants)…

Information Theory · Computer Science 2016-08-15 Abhishek Agarwal , Arya Mazumdar

We study exact-regenerating codes for entanglement-assisted distributed storage systems. Consider an $(n,k,d,\alpha,\beta_{\mathsf{q}},B)$ distributed system that stores a file of $B$ classical symbols across $n$ nodes with each node…

Information Theory · Computer Science 2026-05-13 Lei Hu , Mohamed Nomeir , Alptug Aytekin , Sennur Ulukus

We study the trade-off between storage overhead and inter-cluster repair bandwidth in clustered storage systems, while recovering from multiple node failures within a cluster. A cluster is a collection of $m$ nodes, and there are $n$…

Information Theory · Computer Science 2017-08-21 Vitaly Abdrashitov , N. Prakash , Muriel Médard

Maximum-distance separable (MDS) array codes with high rate and an optimal repair property were introduced recently. These codes could be applied in distributed storage systems, where they minimize the communication and disk access required…

Information Theory · Computer Science 2013-02-19 Eyal En Gad , Robert Mateescu , Filip Blagojevic , Cyril Guyot , Zvonimir Bandic

Maximum-distance-separable (MDS) codes are a class of erasure codes that are widely adopted to enhance the reliability of distributed storage systems (DSS). In (n, k) MDS coded DSS, the original data are stored into n distributed nodes in…

Information Theory · Computer Science 2017-06-16 Sheng Guan , Haibin Kan , Xin Wang