CPL - Chalmers Publication Library
| Utbildning | Forskning | Styrkeområden | Om Chalmers | In English In English Ej inloggad.

Monitoring local progress with watchdog timers deduced from global properties

Raul Barbosa (Institutionen för data- och informationsteknik, Nätverk och system (Chalmers) )
29th IEEE Symposium on Reliable Distributed Systems, SRDS 2010; New Delhi; India; 31 October 2010 through 3 November 2010 (1060-9857). p. 131-140. (2010)
[Konferensbidrag, övrigt]

Distributed systems are used in numerous applications where failures can be costly. Due to concerns that some of the nodes may become faulty, critical services are usually replicated across several nodes, which execute distributed algorithms to ensure correct service in spite of failures. To prevent replica-exhaustion, it is fundamental to detect errors and trigger appropriate recovery actions. In particular, it is important to detect situations in which nodes cease to execute the intended algorithm, e.g., when a replica is compromised by an attacker or when a hardware fault causes the node to behave erratically. This paper proposes a method for monitoring the local execution of nodes using watchdog timers. The approach consists in deducing, from the global system properties, local states that must be visited periodically by nodes that execute the intended algorithm correctly. When a node fails to trigger a watchdog before the time limit, an appropriate response can be initiated. The approach is applied to a well-known Byzantine consensus algorithm. The algorithm is modeled in the PROMELA language and the SPIN model checker is used to identify local states that must be visited periodically by correct nodes. Such states are suitable for online monitoring using watchdog timers.

Nyckelord: Distributed systems, Fault tolerance, Intrusion tolerance, Model checking, Online monitoring, Watchdogs



Denna post skapades 2011-01-11. Senast ändrad 2016-08-16.
CPL Pubid: 133026

 

Läs direkt!


Länk till annan sajt (kan kräva inloggning)


Institutioner (Chalmers)

Institutionen för data- och informationsteknik, Nätverk och system (Chalmers)

Ämnesområden

Information Technology

Chalmers infrastruktur