An adaptive programming model for fault-tolerant distributed computing

被引:14
|
作者
Gorender, Sergio
Macedo, Raimundo Jose de Araujo
Raynal, Michel
机构
[1] Univ Fed Bahia, Dept Comp Sci, Distributed Syst Lab, BR-40170110 Salvador, BA, Brazil
[2] Univ Rennes 1, IRISA, F-35042 Rennes, France
关键词
adaptability; asynchronous/synchronous distributed system; consensus; distributed computing model; fault tolerance; quality of service;
D O I
10.1109/TDSC.2007.3
中图分类号
TP3 [计算技术、计算机技术];
学科分类号
0812 ;
摘要
The capability of dynamically adapting to distinct runtime conditions is an important issue when designing distributed systems where negotiated quality of service (QoS) cannot always be delivered between processes. Providing fault tolerance for such dynamic environments is a challenging task. Considering such a context, this paper proposes an adaptive programming model for fault-tolerant distributed computing, which provides upper-layer applications with process state information according to the current system synchrony ( or QoS). The underlying system model is hybrid, composed by a synchronous part ( where there are time bounds on processing speed and message delay) and an asynchronous part ( where there is no time bound). However, such a composition can vary over time, and, in particular, the system may become totally asynchronous ( e. g., when the underlying system QoS degrade) or totally synchronous. Moreover, processes are not required to share the same view of the system synchrony at a given time. To illustrate what can be done in this programming model and how to use it, the consensus problem is taken as a benchmark problem. This paper also presents an implementation of the model that relies on a negotiated quality of service ( QoS) for communication channels.
引用
收藏
页码:18 / 31
页数:14
相关论文
共 50 条
  • [1] A hybrid and adaptive model for fault-tolerant distributed computing
    Gorender, S
    Macêdo, R
    Raynal, M
    2005 INTERNATIONAL CONFERENCE ON DEPENDABLE SYSTEMS AND NETWORKS, PROCEEDINGS, 2005, : 412 - 421
  • [2] A Fault-Tolerant Programming Model for Distributed Interactive Applications
    Mogk, Ragnar
    Drechsler, Joscha
    Salvaneschi, Guido
    Mezini, Mira
    PROCEEDINGS OF THE ACM ON PROGRAMMING LANGUAGES-PACMPL, 2019, 3 (OOPSLA):
  • [3] Adaptive distributed and fault-tolerant systems
    Hiltunen, MA
    Schlichting, RD
    COMPUTER SYSTEMS SCIENCE AND ENGINEERING, 1996, 11 (05): : 275 - 285
  • [4] Fundamentals of fault-tolerant distributed computing in asynchronous environments
    Gärtner, FC
    ACM COMPUTING SURVEYS, 1999, 31 (01) : 1 - 26
  • [5] Reconciling fault-tolerant distributed algorithms and real-time computing
    Moser, Heinrich
    Schmid, Ulrich
    DISTRIBUTED COMPUTING, 2014, 27 (03) : 203 - 230
  • [6] Reconciling fault-tolerant distributed computing and systems-on-chip
    Fuegger, Matthias
    Schmid, Ulrich
    DISTRIBUTED COMPUTING, 2012, 24 (06) : 323 - 355
  • [7] Deterministic Fault-Tolerant Distributed Computing in Linear Time and Communication
    Chlebus, Bogdan S.
    Kowalski, Dariusz R.
    Olkowski, Jan
    PROCEEDINGS OF THE 2023 ACM SYMPOSIUM ON PRINCIPLES OF DISTRIBUTED COMPUTING, PODC 2023, 2023, : 344 - 354
  • [8] Adaptive fault-tolerant scheduling strategies for mobile cloud computing
    Lee, JongHyuk
    Gil, JoonMin
    JOURNAL OF SUPERCOMPUTING, 2019, 75 (08) : 4472 - 4488
  • [9] Adaptive fault-tolerant scheduling strategies for mobile cloud computing
    JongHyuk Lee
    JoonMin Gil
    The Journal of Supercomputing, 2019, 75 : 4472 - 4488
  • [10] Hierarchical Distributed Model-Free Adaptive Fault-Tolerant Vehicular Platooning Control
    Zhang, Peng
    Che, Wei-Wei
    INTERNATIONAL JOURNAL OF ROBUST AND NONLINEAR CONTROL, 2025, 35 (03) : 1281 - 1293