REVIEW 1 cited by
MPGemmFI: A Fault Injection Technique for Mixed Precision GEMM in ML Applications
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Emerging deep learning workloads urgently need fast general matrix multiplication (GEMM). To meet such demand, one of the critical features of machine-learning-specific accelerators such as NVIDIA Tensor Cores, AMD Matrix Cores, and Google TPUs is the support of mixed-precision enabled GEMM. For DNN models, lower-precision FP data formats and computation offer acceptable correctness but significant performance, area, and memory footprint improvement. While promising, the mixed-precision computation on error resilience remains unexplored. To this end, we develop a fault injection framework that systematically injects fault into the mixed-precision computation results. We investigate how the faults affect the accuracy of machine learning applications. Based on the error resilience characteristics, we offer lightweight error detection and correction solutions that significantly improve the overall model accuracy if the models experience hardware faults. The solutions can be efficiently integrated into the accelerator's pipelines.
Forward citations
Cited by 1 Pith paper
-
Aharanov-Bohm Type Arbitrage and Homological Obstructions in Financial Markets
Non-trivial holonomy of the multiplicative distortion induced by the conditional expectation functor on a market filtration corresponds to Aharonov-Bohm arbitrage realizable as a self-financing trading strategy under ...
Discussion (0). Continue with ORCID to comment.