Search NASASearch

NASA NTRS · 20020051487

Efficacy of Code Optimization on Cache-Based Processors

Abstract

In this paper a number of techniques for improving the cache performance of a representative piece of numerical software is presented. Target machines are popular processors from several vendors: MIPS R5000 (SGI Indy), MIPS R8000 (SGI PowerChallenge), MIPS R10000 (SGI Origin), DEC Alpha EV4 + EV5 (Cray T3D & T3E), IBM RS6000 (SP Wide-node), Intel PentiumPro (Ames' Whitney), Sun UltraSparc (NERSC's NOW). The optimizations all attempt to increase the locality of memory accesses. But they meet with rather varied and often counterintuitive success on the different computing platforms. We conclude that it may be genuinely impossible to obtain portable performance on the current generation of cache-based machines. At the least, it appears that the performance of modern commodity processors cannot be described with parameters defining the cache alone.

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

VanderWijngaart, Rob F., Saphir, William C., Chancellor, Marisa K.. 1997-01-01. Efficacy of Code Optimization on Cache-Based Processors. https://ntrs.nasa.gov/citations/20020051487

Cite the original work for its findings. Save a collection to share your selection of sources.