English

Industrial Computing Systems: A Case Study of Fault Tolerance Analysis

Systems and Control 2015-03-31 v1 Distributed, Parallel, and Cluster Computing

Abstract

Fault tolerance is a key factor of industrial computing systems design. But in practical terms, these systems, like every commercial product, are under great financial constraints and they have to remain in operational state as long as possible due to their commercial attractiveness. This work provides an analysis of the instantaneous failure rate of these systems at the end of their life-time period. On the basis of this analysis, we determine the effect of a critical increase in the system failure rate and the basic condition of its existence. The next step determines the maintenance scheduling which can help to avoid this effect and to extend the system life-time in fault-tolerant mode.

Keywords

Cite

@article{arxiv.1503.08715,
  title  = {Industrial Computing Systems: A Case Study of Fault Tolerance Analysis},
  author = {Andrey A. Shchurov},
  journal= {arXiv preprint arXiv:1503.08715},
  year   = {2015}
}

Comments

6 figures

R2 v1 2026-06-22T09:05:46.328Z