8 citations · 19 across the 10 of their papers we have counts for
Showing 2011Show all
2 papers · 1 filter
cs.DC2011★ 2 cited
A New and Efficient Algorithm-Based Fault Tolerance Scheme for A Million Way Parallelism
Erlin Yao, Mingyu Chen, Rui Wang +2
Fault tolerance overhead of high performance computing (HPC) applications is becoming critical to the efficient utilization of HPC systems at large scale. HPC applications typicall…
cs.AR2011★ 1 cited
HMTT: A Hybrid Hardware/Software Tracing System for Bridging Memory Trace's Semantic Gap
Yungang Bao, Jinyong Zhang, Yan Zhu +4
Memory trace analysis is an important technology for architecture research, system software (i.e., OS, compiler) optimization, and application performance improvements. Hardware-sn…