Keeping checkpoint/restart viable for exascale systems.
Description:
Next-generation exascale systems, those capable of performing a quintillion (10{sup 18}) operations per second, are expected to be delivered in the next 8-10 years. These systems, which will be 1,000 times faster than current systems, will be of unprecedented scale. As these systems continue to grow in size, faults will become increasingly common, even over the course of small calculations. Therefore, issues such as fault tolerance and reliability will limit application scalability. Current tec…
more
Date:
September 1, 2011
Creator:
Riesen, Rolf E.; Bridges, Patrick G. (IBM Research, Ireland, Mulhuddart, Dublin); Stearley, Jon R.; Laros, James H., III; Oldfield, Ron A.; Arnold, Dorian (University of New Mexico, Albuquerque, NM) et al.
Item Type:
Refine your search to only
Report
Partner:
UNT Libraries Government Documents Department