Matches in SemOpenAlex for { <https://semopenalex.org/work/W2123500455> ?p ?o ?g. }
- W2123500455 abstract "Dense matrix factorizations, such as LU, Cholesky and QR, are widely used for scientific applications that require solving systems of linear equations, eigenvalues and linear least squares problems. Such computations are normally carried out on supercomputers, whose ever-growing scale induces a fast decline of the Mean Time To Failure (MTTF). This paper proposes a new hybrid approach, based on Algorithm-Based Fault Tolerance (ABFT), to help matrix factorizations algorithms survive fail-stop failures. We consider extreme conditions, such as the absence of any reliable component and the possibility of loosing both data and checksum from a single failure. We will present a generic solution for protecting the right factor, where the updates are applied, of all above mentioned factorizations. For the left factor, where the panel has been applied, we propose a scalable checkpointing algorithm. This algorithm features high degree of checkpointing parallelism and cooperatively utilizes the checksum storage leftover from the right factor protection. The fault-tolerant algorithms derived from this hybrid solution is applicable to a wide range of dense matrix factorizations, with minor modifications. Theoretical analysis shows that the fault tolerance overhead sharply decreases with the scaling in the number of computing units and the problem size. Experimental results of LU and QR factorization on the Kraken (Cray XT5) supercomputer validate the theoretical evaluation and confirm negligible overhead, with- and without-errors." @default.
- W2123500455 created "2016-06-24" @default.
- W2123500455 creator A5008117654 @default.
- W2123500455 creator A5010055736 @default.
- W2123500455 creator A5010408125 @default.
- W2123500455 creator A5054873210 @default.
- W2123500455 creator A5075517045 @default.
- W2123500455 date "2012-02-25" @default.
- W2123500455 modified "2023-10-18" @default.
- W2123500455 title "Algorithm-based fault tolerance for dense matrix factorizations" @default.
- W2123500455 cites W2001495258 @default.
- W2123500455 cites W2072072075 @default.
- W2123500455 cites W2083606889 @default.
- W2123500455 cites W2083613288 @default.
- W2123500455 cites W2096504919 @default.
- W2123500455 cites W2151984682 @default.
- W2123500455 cites W2158344138 @default.
- W2123500455 cites W2165009364 @default.
- W2123500455 cites W2296772319 @default.
- W2123500455 cites W4231150350 @default.
- W2123500455 cites W4239025233 @default.
- W2123500455 doi "https://doi.org/10.1145/2145816.2145845" @default.
- W2123500455 hasPublicationYear "2012" @default.
- W2123500455 type Work @default.
- W2123500455 sameAs 2123500455 @default.
- W2123500455 citedByCount "77" @default.
- W2123500455 countsByYear W21235004552012 @default.
- W2123500455 countsByYear W21235004552013 @default.
- W2123500455 countsByYear W21235004552014 @default.
- W2123500455 countsByYear W21235004552015 @default.
- W2123500455 countsByYear W21235004552016 @default.
- W2123500455 countsByYear W21235004552017 @default.
- W2123500455 countsByYear W21235004552018 @default.
- W2123500455 countsByYear W21235004552019 @default.
- W2123500455 countsByYear W21235004552020 @default.
- W2123500455 countsByYear W21235004552022 @default.
- W2123500455 crossrefType "proceedings-article" @default.
- W2123500455 hasAuthorship W2123500455A5008117654 @default.
- W2123500455 hasAuthorship W2123500455A5010055736 @default.
- W2123500455 hasAuthorship W2123500455A5010408125 @default.
- W2123500455 hasAuthorship W2123500455A5054873210 @default.
- W2123500455 hasAuthorship W2123500455A5075517045 @default.
- W2123500455 hasConcept C106487976 @default.
- W2123500455 hasConcept C111919701 @default.
- W2123500455 hasConcept C11413529 @default.
- W2123500455 hasConcept C120314980 @default.
- W2123500455 hasConcept C121332964 @default.
- W2123500455 hasConcept C123213974 @default.
- W2123500455 hasConcept C158693339 @default.
- W2123500455 hasConcept C159985019 @default.
- W2123500455 hasConcept C162372511 @default.
- W2123500455 hasConcept C173608175 @default.
- W2123500455 hasConcept C188060507 @default.
- W2123500455 hasConcept C192562407 @default.
- W2123500455 hasConcept C2779960059 @default.
- W2123500455 hasConcept C34727166 @default.
- W2123500455 hasConcept C41008148 @default.
- W2123500455 hasConcept C42355184 @default.
- W2123500455 hasConcept C44363057 @default.
- W2123500455 hasConcept C46085209 @default.
- W2123500455 hasConcept C48044578 @default.
- W2123500455 hasConcept C62520636 @default.
- W2123500455 hasConcept C63540848 @default.
- W2123500455 hasConcept C77088390 @default.
- W2123500455 hasConcept C83283714 @default.
- W2123500455 hasConceptScore W2123500455C106487976 @default.
- W2123500455 hasConceptScore W2123500455C111919701 @default.
- W2123500455 hasConceptScore W2123500455C11413529 @default.
- W2123500455 hasConceptScore W2123500455C120314980 @default.
- W2123500455 hasConceptScore W2123500455C121332964 @default.
- W2123500455 hasConceptScore W2123500455C123213974 @default.
- W2123500455 hasConceptScore W2123500455C158693339 @default.
- W2123500455 hasConceptScore W2123500455C159985019 @default.
- W2123500455 hasConceptScore W2123500455C162372511 @default.
- W2123500455 hasConceptScore W2123500455C173608175 @default.
- W2123500455 hasConceptScore W2123500455C188060507 @default.
- W2123500455 hasConceptScore W2123500455C192562407 @default.
- W2123500455 hasConceptScore W2123500455C2779960059 @default.
- W2123500455 hasConceptScore W2123500455C34727166 @default.
- W2123500455 hasConceptScore W2123500455C41008148 @default.
- W2123500455 hasConceptScore W2123500455C42355184 @default.
- W2123500455 hasConceptScore W2123500455C44363057 @default.
- W2123500455 hasConceptScore W2123500455C46085209 @default.
- W2123500455 hasConceptScore W2123500455C48044578 @default.
- W2123500455 hasConceptScore W2123500455C62520636 @default.
- W2123500455 hasConceptScore W2123500455C63540848 @default.
- W2123500455 hasConceptScore W2123500455C77088390 @default.
- W2123500455 hasConceptScore W2123500455C83283714 @default.
- W2123500455 hasLocation W21235004551 @default.
- W2123500455 hasOpenAccess W2123500455 @default.
- W2123500455 hasPrimaryLocation W21235004551 @default.
- W2123500455 hasRelatedWork W2052455844 @default.
- W2123500455 hasRelatedWork W2059724662 @default.
- W2123500455 hasRelatedWork W2112426221 @default.
- W2123500455 hasRelatedWork W2113085198 @default.
- W2123500455 hasRelatedWork W2123500455 @default.
- W2123500455 hasRelatedWork W2127054029 @default.
- W2123500455 hasRelatedWork W2135765421 @default.
- W2123500455 hasRelatedWork W2162739943 @default.
- W2123500455 hasRelatedWork W2341410909 @default.