arrow
Return

The Failure Detector Abstraction

delete2011-02-04
delete26
delete
OA
AI
F
Felix Freiling *
R
Rachid Guerraoui
P
Petr Kuznetsov
DOI:10.1145/1883612.1883616delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
A failure detector is a fundamental abstraction in distributed computing. This article surveys this abstraction through two dimensions. First we study failure detectors as building blocks to simplify the design of reliable distributed algorithms. In particular, we illustrate how failure detectors can factor out timing assumptions to detect failures in distributed agreement algorithms. Second, we study failure detectors as computability benchmarks. That is, we survey the weakest failure detector question and illustrate how failure detectors can be used to classify problems. We also highlight some limitations of the failure detector abstraction along each of the dimensions.
Keywords:
Algorithms
Design
Reliability
Theory
Distributed system
agreement problem
consensus
atomic commit
fault tolerance
liveness
message passing
safety
synchrony
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

ACM Computing Surveys cover
ACM Computing Surveys
IF:
28
Papers:
2.4K
Citations:
3.5W

Organization

D
deutsche telekom ag
Scholars:
172
Papers: 129
Citations: 1
U
University of Mannheim
Scholars:
1.9K
Papers: 2.2K
Citations: 3.2K