Publication:
Performance-reliability tradeoff analysis for multithreaded applications

dc.contributor.authorsOz I., Topcuoglu H.R., Kandemir M., Tosun O.
dc.date.accessioned2022-03-15T02:09:40Z
dc.date.accessioned2026-01-11T17:56:53Z
dc.date.available2022-03-15T02:09:40Z
dc.date.issued2012
dc.description.abstractModern architectures become more susceptible to transient errors with the scale down of circuits. This makes reliability an increasingly critical concern in computer systems. In general, there is a tradeoff between system reliability and performance of multithreaded applications running on multicore architectures. In this paper, we conduct a performance-reliability analysis for different parallel versions of three data-intensive applications including FFT, Jacobi Kernel, andWater Simulation. We measure the performance of these programs by counting execution clock cycles, while the system reliability is measured by Thread Vulnerability Factor (TVF) which is a recently-proposed metric. TVF measures the vulnerability of a thread to hardware faults at a high level. We carry out experiments by executing parallel implementations on multicore architectures and collect data about the performance and vulnerability. Our experimental evaluation indicates that the choice is clear for FFT application and Jacobi Kernel. Transpose algorithm for FFT application results in less than 5% performance loss while the vulnerability increases by 20% compared to binary-exchange algorithm. Unrolled Jacobi code reduces execution time up to 50% with no significant change on vulnerability values. However, the tradeoff is more interesting for Water Simulation where nsquared version reduces the vulnerability values significantly by worsening the performance with similar rates compared to faster but more vulnerable spatial version. © 2012 EDAA.
dc.identifier.doi10.1109/date.2012.6176624
dc.identifier.isbn9783981080186
dc.identifier.issn15301591
dc.identifier.urihttps://hdl.handle.net/11424/247249
dc.language.isoeng
dc.publisherInstitute of Electrical and Electronics Engineers Inc.
dc.relation.ispartofProceedings -Design, Automation and Test in Europe, DATE
dc.rightsinfo:eu-repo/semantics/closedAccess
dc.subjectDistributed Algorithms
dc.subjectMulti-Core Architectures and Support
dc.subjectReliable Parallel
dc.titlePerformance-reliability tradeoff analysis for multithreaded applications
dc.typeconferenceObject
dspace.entity.typePublication
oaire.citation.endPage898
oaire.citation.startPage893
oaire.citation.titleProceedings -Design, Automation and Test in Europe, DATE

Files