Fault tolerance

From Citizendium, the Citizens' Compendium

Jump to: navigation, search


This article is a stub and thus not approved.
Main Article
Talk
Related Articles  [?]
Bibliography  [?]
External Links  [?]
 
This is a draft article, under development and not meant to be cited but you can help to improve it. These unapproved articles are subject to a disclaimer.

In engineering, fault tolerance is a characteristic of a system that can have one or more subcomponents fail without the entire system failing. This does not mean that the system has no single point of failure, but that at least some parts can fail and have continued operation; a system might be robust to a component failure, but have no backup for complete physical destruction.

With the loss of some components, there might be a loss of performance or functionality, or the system could be engineered with reserve capacity such that some failures would not be noticed by users.

Simple duplication of components does not make a system fault-tolerant; there needs to be an intelligent mechanism for spreading the work around the failure.

Views
Personal tools