Sciweavers

CCGRID
2007
IEEE

Reliability Analysis of Self-Healing Network using Discrete-Event Simulation

13 years 11 months ago
Reliability Analysis of Self-Healing Network using Discrete-Event Simulation
The number of processors embedded on high performance computing platforms is continuously increasing to accommodate user desire to solve larger and more complex problems. However, as the number of components increases, so does the probability of failure. Thus, both scalable and fault-tolerance of software are important issues in this field. To ensure reliability of the software especially under the failure circumstance, the reliability analysis is needed. The discrete-event simulation technique offers an attractive alternative to traditional Markovian-based analytical models, which often have an intractably large state space. In this paper, we analyze reliability of a self-healing network developed for parallel runtime environments using discreteevent simulation. The network is designed to support transmission of messages across multiple nodes and at the same time, to protect against node and process failures. Results demonstrate the flexibility of a discrete-event simulation approa...
Thara Angskun, George Bosilca, Graham E. Fagg, Jel
Added 02 Jun 2010
Updated 02 Jun 2010
Type Conference
Year 2007
Where CCGRID
Authors Thara Angskun, George Bosilca, Graham E. Fagg, Jelena Pjesivac-Grbovic, Jack Dongarra
Comments (0)