DOE OSTI · 1813800
Offloading Calculations to Computational Storage Devices: Spark and HDFS [Slides]
Abstract
The objective is to evaluate the capabilities of multiple CSDs (provided by NDG Systems) using Hadoop Filesystem and Apache Spark. The independent variables are: number of CSDs, 0, 1, 2, 4, or 6; size of dataset, 1 GB, 5 GB, 10 GB; type of dataset, one large file with all of the data, 10 files, 100 files. The dependent variables are: job time; execution time. the constants are operations on the dataset.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Cunningham, IV, Clyburn, Goldstein, Justin James, Hammock, Charles Warren, Janz, Jacob Benjamin, Liu, Ralph, Rimerman, Mitchell. 2021-08-12. Offloading Calculations to Computational Storage Devices: Spark and HDFS [Slides]. https://doi.org/10.2172/1813800
Cite the original work for its findings. Save a collection to share your selection of sources.