Even though the cloud platform promises to be reliable, several availability incidents prove that it is not. How can we be sure that a parallel application finishes it´s execution even if a site is affected by a failure? This paper presents H-RADIC, an approach based on RADIC architecture, that executes parallel applications protected by RADIC in at least 3 different virtual clusters or sites. The execution state of each site is saved periodically in another site and it is recovered in case of failure. The paper details the configuration of the architecture and the experiment´s results using 3 clusters running NAS parallel applications protected with DMTCP, a very well-known distributed multi-threaded checkpoint tool. Our experiments show t...
La tolerancia a fallos es una línea de investigación que ha adquirido una importancia relevante con ...
Tese de doutoramento, Informática (Engenharia Informática), Universidade de Lisboa, Faculdade de Ciê...
En Computación de Altas Prestaciones (HPC), con el objetivo de aumentar las prestaciones se ha ido i...
Even though the cloud platform promises to be reliable, several availability incidents prove that it...
Even though the cloud platform promises to be reliable, several availability incidents prove that th...
AbstractThe increasing failure rate in High Performance Computing encourages the investigation of fa...
The increasing failure rate in High Performance Computing encourages the investigation of fault tole...
The demand for computational power has been leading the improvement of the High Performance Computin...
The demand for computational power has been leading the improvement of the High Performance Computin...
Fault tolerance has become an important issue for parallel applications in the last few years. The p...
Cloud Computing offers the possibility of computing resources, allowing remote access to software, s...
Consultable des del TDXTítol obtingut de la portada digitalitzadaLa tolerancia a fallos se ha conver...
La tolerancia a fallos se ha convertido en un requerimiento importante para los ingenieros informáti...
Los sistemas de computación de alto rendimiento (HPC) continúan creciendo exponencialmente en términ...
La tolerancia a fallos es una línea de investigación que ha adquirido una importancia relevante con ...
La tolerancia a fallos es una línea de investigación que ha adquirido una importancia relevante con ...
Tese de doutoramento, Informática (Engenharia Informática), Universidade de Lisboa, Faculdade de Ciê...
En Computación de Altas Prestaciones (HPC), con el objetivo de aumentar las prestaciones se ha ido i...
Even though the cloud platform promises to be reliable, several availability incidents prove that it...
Even though the cloud platform promises to be reliable, several availability incidents prove that th...
AbstractThe increasing failure rate in High Performance Computing encourages the investigation of fa...
The increasing failure rate in High Performance Computing encourages the investigation of fault tole...
The demand for computational power has been leading the improvement of the High Performance Computin...
The demand for computational power has been leading the improvement of the High Performance Computin...
Fault tolerance has become an important issue for parallel applications in the last few years. The p...
Cloud Computing offers the possibility of computing resources, allowing remote access to software, s...
Consultable des del TDXTítol obtingut de la portada digitalitzadaLa tolerancia a fallos se ha conver...
La tolerancia a fallos se ha convertido en un requerimiento importante para los ingenieros informáti...
Los sistemas de computación de alto rendimiento (HPC) continúan creciendo exponencialmente en términ...
La tolerancia a fallos es una línea de investigación que ha adquirido una importancia relevante con ...
La tolerancia a fallos es una línea de investigación que ha adquirido una importancia relevante con ...
Tese de doutoramento, Informática (Engenharia Informática), Universidade de Lisboa, Faculdade de Ciê...
En Computación de Altas Prestaciones (HPC), con el objetivo de aumentar las prestaciones se ha ido i...