Search NASASearch

DOE OSTI · 2587170

A Synthesis Methodology for Intelligent Memory Interfaces in Accelerator Systems

Abstract

Domain-specific systems improve the performance of a specific set of applications compared to general-purpose processing systems by deploying custom hardware accelerators. These hardware accelerators are generated using high-level synthesis (HLS) tools. The HLS tools enable a comprehensive design space exploration to optimize the compute performance of the generated accelerators. However, they often ignore the challenges of implementing the accelerators in a system-on-chip, particularly how the accelerators access memory. Our work introduces a buffering system design that improves accelerators' memory accesses by intelligently employing burst transactions to prefetch useful data from external memory to on-chip local buffers. Our design is dynamic, parametric, and transparent to the accelerators generated by HLS tools. We derive the buffering system parameters using appropriate compiler-based analysis passes and memory channel latency constraints. The proposed buffering system design results in, on average, 8.8x performance improvements while lowering memory channel utilization on average by 53.2% for a set of PolyBench kernels.

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Limaye, Ankur M. (ORCID:0000000194062584), Bohm Agostini, Nicolas, Barone, Claudio, Castellana, Vito G. (ORCID:0000000335167903), Fiorito, Michele, Ferrandi, Fabrizio, Marquez, Andres (ORCID:0000000243131882), Tumeo, Antonino. 2025-03-04. A Synthesis Methodology for Intelligent Memory Interfaces in Accelerator Systems. https://www.osti.gov/biblio/2587170

Cite the original work for its findings. Save a collection to share your selection of sources.