Search NASA⌕ Search

SEARCH · Search NASA

Results for “data storage”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 289 records · Page 16

CO 2 rock physics modeling for reliable monitoring of geologic carbon storage

Monitoring, verification, and accounting (MVA) are crucial to ensure safe and long-term geologic carbon storage. Seismic monitoring is a key MVA technique that utilizes seismic data to infer elastic properties of CO 2 -saturated rocks. Reliable accounting of CO 2 in subsurface storage reservoirs and potential leakage zones requires an accurate rock physics model. However, the widely used CO 2 rock physics model based on the conventional Biot-Gassmann equation can substantially underestimate the influence of CO 2 saturation on seismic waves, leading to inaccurate accounting. We develop an accurate CO 2 rock physics model by accounting for both effects of the stress dependence of seismic velocities in porous rocks and CO 2 weakening on the rock framework. We validate our CO 2 rock physics model using the Kimberlina-1.2 model (a previously proposed geologic carbon storage site in California) and create time-lapse elastic property models with our new rock physics method. We compare the results with those obtained using the conventional Biot-Gassmann equation. Our innovative approach produces larger changes in elastic properties than the Biot-Gassmann results. Using our CO 2 rock physics model can replicate shear-wave speed reductions observed in the laboratory. Our rock physics model enhances the accuracy of time-lapse elastic-wave modeling and enables reliable CO 2 accounting using seismic monitoring.

58 GEOSCIENCES↗

Leveraging Pre-Built Catalogs and Object-Level Scheduling to Eliminate I/O Bottlenecks in HPC Environments

Modern High-Performance Computing (HPC) environments face mounting challenges due to the shift from large to small file datasets, along with an increasing number of users and parallelized applications. As HPC systems rely on Parallel File Systems (PFS), such as Lustre for data processing, performance bottlenecks stemming from Object Storage Target (OST) contention have become a significant concern. Existing solutions, such as LADS with its object-level scheduling approach, fall short in large-scale HPC environments due to their inability to effectively address metadata I/O bottlenecks and the growing number of I/O processes. This study highlights the pressing need for a comprehensive solution that tackles both OST contention and metadata I/O challenges in diverse HPC workloads. To address these challenges, we propose SwiftLoad, an object-level I/O scheduling framework that leverages a metadata catalog to enhance the performance and efficiency of parallel HPC utilities. The adoption of the metadata catalog mitigates the metadata I/O bottlenecks that commonly occur in HPC utilities, a challenge that is particularly pronounced in object-level I/O scheduling. SwiftLoad addresses OST contention and the uneven distribution of I/O processes across different OSTs through mathematical modeling and incorporates a Loader Configuration Module to regulate the number of I/O processes. Evaluated with two representative utilities—data deduplication profiling and data augmentation—SwiftLoad achieved performance improvements of up to 5.63x and 11.0x, respectively, on a production supercomputer.

HPC↗

Grid Resiliency with a 100% Renewable Microgrid

San Diego Gas & Electric Company (SDG&E) installed America’s first and largest utility-scale microgrid in Borrego Springs in 2013. The first generation Borrego Springs Microgrid utilized diesel generators to form and stabilize the microgrid island, with support from grid-scale batteries and local solar photovoltaic (PV) generation. In this project, SDG&E in partnership with National Renewable Energy Laboratory (NREL) demonstrated through modeling, simulation and utility field testing that blackstart and islanding of the microgrid can be led with 100% renewable, inverter based resources (IBRs), to help reduce community reliance on conventional generation resources. Through equipment upgrades, grid-forming island leader capability was transitioned to a battery IBR instead of the Borrego Springs Microgrid diesel generators. A new microgrid controller was integrated to the microgrid and programmed to control and manage multiple energy storage systems. Synchrophasor and other power quality data verified autonomous, high-speed response of the IBRs through blackstart, islanding, and load step testing. Results of project field evaluations provide distribution systems operators (DSO) with increased confidence that renewable, IBR can replace traditional generators to blackstart and island microgrids and rapidly establish stable island frequency with rapid changes in peak power demand. Importantly, the project validated the integration feasibility of a distributed energy resource management system (DERMS) controller that manages multiple grid-forming and grid-following IBRs, establishing a standard design interface to reduce the complexity of integrating new DERs in the future and supporting replication by the industry. As a result of learnings in this project, SDG&E has implemented the microgrid controller strategy at multiple other microgrid sites, thereby validating the replicability of the solution. Hardware-in-the-loop (HIL) simulations including power and controller HIL hardware — along with electromagnetic transient (EMT) simulations of Borrego Springs Microgrid —informed adjustments to inverter parameters and were important to characterize the performance of the IBRs in relevant operating conditions before deployment. The EMT and HIL simulations of islanding the entire community are important contributions in providing confidence in IBR performance prior to future islanding of the community in the field. High-fidelity EMT and/or HIL simulation of IBRs can de-risk field operations, and its relevance and importance as a tool is increasing as distribution grids and microgrids become more complex and dynamic with an increasing proportion of renewable generation, distributed energy storage, and two-way power and energy flows.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Online and Offline Identification of False Data Injection Attacks in Battery Sensors Using a Single Particle Model

The cells in battery energy storage systems are monitored, protected, and controlled by battery management systems whose sensors are susceptible to cyberattacks. False data injection attacks (FDIAs) targeting batteries’ voltage sensors affect cell protection functions and the estimation of critical battery states like the state of charge (SoC). Inaccurate SoC estimation could result in battery overcharging and over discharging, which can have disastrous consequences on grid operations. This paper proposes a three-pronged online and offline method to detect, identify, and classify FDIAs corrupting the voltage sensors of a battery stack. To accurately model the dynamics of the series-connected cells a single particle model is used and to estimate the SoC, the unscented Kalman filter is employed. FDIA detection, identification, and classification was accomplished using a tuned cumulative sum (CUSUM) algorithm, which was compared with a baseline method, the chi-squared error detector. Online simulations and offline batch simulations were performed to determine the effectiveness of the proposed approach. Throughout the batch simulations, the CUSUM algorithm detected attacks, with no false positives, in 99.83% of cases, identified the corrupted sensor in 97% of cases, and determined if the attack was positively or negatively biased in 97% of cases.

25 ENERGY STORAGE↗

Streaming Data in HPC Workflows Using ADIOS

The “IO Wall” problem, in which the gap between computation rate and data access rate grows continuously, poses significant problems to scientific workflows which have traditionally relied upon using the filesystem for intermediate storage between workflow stages. One way to avoid this problem in scientific workflows is to stream data directly from producers to consumers and avoiding storage entirely. However, the manner in which this is accomplished is key to both performance and usability. This paper presents the Sustainable Staging Transport, an approach which allows direct streaming between traditional file writers and readers with few application changes. SST is an ADIOS “engine”, accessible via standard ADIOS APIs, and because ADIOS allows engines to be chosen at run-time, many existing file-oriented ADIOS workflows can utilize SST for direct application-to-application communication without any source code changes. This paper describes the design of SST and presents performance results from various applications that use SST, for feeding model training with simulation data with substantially higher bandwidth than the theoretical limits of Frontier’s file system, for strong coupling of separately developed applications for multiphysics multiscale simulation, or for in situ analysis and visualization of data to complete all data processing shortly after the simulation finishes.

Podhorszki, Norbert [ORNL] (ORCID:000000019647542X↗

Exploring the Future Energy Value of Long-Duration Energy Storage

Long-duration energy storage is commonly viewed as a key technology for providing flexibility to the grid and broader energy systems over a multidecadal time frame. However, prior work has typically used present-day grid infrastructures to characterize the relationship between the duration and arbitrage value of storage in electricity markets. This study leverages established National Renewable Energy Laboratory grid planning and operations tools, analysis, and data to execute a price-taker model of an energy storage system for several 8760 h price series representative of current and future contiguous United States grid infrastructures with varying shares of variable renewable energy (VRE). We find that the total value of energy storage typically increases with VRE shares, but any increase in the relative value of longer storage durations over time depends on the region and grid mix. Some regions see incremental value increasing notably, up to 20–40 h in 2050, while others do not. The negative effect of lower roundtrip efficiency on value is also found to be scenario-dependent, with the energy value in higher VRE scenarios being less sensitive to roundtrip efficiency and more supportive of longer storage durations. Long-duration storage value and deployment potential are a function of evolving electricity sector infrastructure, markets, and policy, making it critical to consistently revisit potential long-duration storage contributions to the grid.

14 SOLAR ENERGY↗

LANL Meteorological Program: 2023 Data Completeness/Quality Report

Los Alamos National Laboratory (LANL) operates seven mesa-top instrumented meteorology towers: Technical Area (TA) 6, TA-49, TA-53, TA-54, TA-63, TA-54B, and TA-16B. An additional instrumented tower is located in Mortandad Canyon (TA-5 MDCN), and there is a rain gauge at North Community (NCOM), located within the town of Los Alamos. The 10 meter (m) towers at TA-63, TA-54B, and TA-16B have been in testing since they were installed in 2021, and will be included in a future data completeness report. A description of the meteorology monitoring network, prior to the installation of the TA-63, TA-54B, and TA-16B is found in Dewart and Boggs (2014). Four of the mesa-top towers (e.g., TA-6, TA-49, TA-53, and TA-54) are instrumented at the 1.2 m, 11.5 m, 23 m, and 46 m levels. In addition, the TA-6 tower is instrumented at the 92 m level. The TA-5 MDCN tower is 10 m in height and is instrumented at 1.2 m and 10 m. Data are collected and averaged every 15 minutes. Range checking is performed on each measurement every 15 minutes; data that are beyond normal ranges are eliminated from the data set and replaced by a code for missing data. In addition, data are reviewed weekly by qualified meteorologists to identify bad data not identified by the range checking technique. The data steward eliminates these data from the data set and replaces them with a code for missing data. The instrument technicians also review that data and schedule instrument replacement, as required. All instruments are calibrated at a frequency that meets the criteria identified in ANSI/ANS-3.11-2015. Data completeness is determined by the number of total 15-minute records available versus the number of possible measurements for the entire year. As a rule, the meteorologists do not attempt to estimate data that are eliminated as bad data. Original datalogger records, including bad data, can be recalled from program archival storage.

54 ENVIRONMENTAL SCIENCES↗

MDLoader: A Hybrid Model-Driven Data Loader for Distributed Graph Neural Network Training

Scalable data management is essential for processing large scientific dataset on HPC platforms for distributed deep learning. In-memory distributed storage is preferred for its speed, enabling rapid, random, and frequent data access required by stochastic optimizers. Processes use one-sided or collective communication to fetch remote data, with optimal performance depending on (i) dataset characteristics, (ii) training scale, and (iii) interconnection network. Empirical analysis shows collective communication excels with larger mini-batch sizes and/or fewer processes, whereas one-sided communication outperforms at larger scales. We propose MDLoader, a hybrid in-memory data loader for distributed graph neural network training. MDLoader features a model-driven performance estimator that dynamically selects between one-sided and collective communication at the beginning of training using Tree of Parzen Estimators (TPE). Evaluations on NERSC Perlmutter and OLCF Summit show MDLoader outperforms single-backend loaders by up to 2.83 × and predicts the suitable communication method with 96.3% (Perlmutter) and 94.3% (Summit) success rate.

Bae, Jonghyun↗

A Novel and Scalable Method for Microencapsulating Salt Hydrate Phase Change Materials in Core–Shell Fibers

Phase change materials (PCMs) are in high demand for applications such as thermal energy storage in buildings, electronics cooling, and thermal management of electric vehicle batteries and data centers. Among these materials, salt hydrate PCMs are particularly attractive due to their high thermal energy storage capacity and low cost. However, they suffer from two major issues: leakage in the melted phase and phase segregation during phase transitions. Microencapsulation is the primary process capable of addressing both of these challenges. However, there is no reliable or scalable method available for microencapsulating salt hydrate PCMs. As a result, the full potential of salt hydrates for building and data center applications has yet to be realized. In this work, we present an innovative method for the microencapsulation of salt hydrate PCMs using a co‐axial pushing technique. This process creates core–shell fibers, with the salt hydrate as the core and a polymer as the shell. Our approach demonstrates strong potential for scalable microencapsulation of salt hydrate PCMs. In conclusion, achieving scalability could enable their widespread use in applications such as data center cooling, battery thermal management, and building climate control.

Sharma, Jaswinder [Oak Ridge National Laboratory (↗

Prospecting for Critical Minerals and Rare Earth Elements from Marcellus Shale in the Western Portion of the Appalachian Basin with Non-Destructive Core Characterization

Identification of sources for domestic critical minerals and rare earth elements (CM/REE) has been deemed essential for the energy transition by the United States Department of Energy (DOE). The U.S. DOE’s National Energy Technology Laboratory’s (NETL) Geomaterials Characterization Laboratory has performed non-destructive core characterizations on energy-relevant rock cores for the past decade. During this time, NETL has published over 36 technical reports and made the associated data publicly available. Much of this work focuses on unconventional shale gas, subsurface carbon storage systems, and carbon-ore. These efforts provide cm-scale petrophysical and elemental data, photographic documentation, detailed core descriptions, and computed tomography (CT) data for each well. This provides a first phase prospecting resource for CM/REE resources and can provide a map for pin-pointing intervals and lithologies for further development. Using historical core characterization data from 12 Marcellus wells from the western portion of the Appalachian Basin, this study builds an improved understanding of the chemostratigraphy of the basin. X-ray fluorescence (XRF) and CT images were used to determine lithologic intervals and potential ore bodies for further analysis, including benchtop digestion and inductively coupled plasma mass spectrometry (ICP-MS) to better understand the CM/REE enrichments.

Paronish, Thomas J.↗

What to Support When You’re Compressing

Over the last nearly 20 years, lossy compression has become an essential aspect of HPC applications’ data pipelines, allowing them to overcome limitations in storage capacity and bandwidth and, in some cases, increase computational throughput and capacity. However, with the adoption of lossy compression comes the requirement to assess and control the impact lossy compression has on scientific outcomes. In this work, we take a major step forward in describing the state of practice and by characterizing workloads. We examine applications’ needs and compressors’ capabilities across 9 different supercomputing application domains. We present 24 takeaways that provide best practices for applications, operational impacts for facilities achieving compressed data, and gaps in application needs not addressed by production compressors that point towards opportunities for future compression research.

Error-Bounded Lossy Compression↗

A Convolution Neural Network for Voltage Event Classification at a Photovoltaic Inverter

This paper presents a convolutional neural network (CNN) developed to identify voltage events in photovoltaic (PV) inverters. The CNN is trained on synthetic data generated using the IEEE 13-bus distribution feeder model and evaluated on field measured data collected from Energy Northwest’s Horn Rapids Solar, Storage, and Training (HRSST) facility. The study focuses on two common voltage events: faults and voltage sags. The CNN is configured to analyze voltage and current waveforms from three-phase PV systems, demonstrating excellent accuracy during training. Field data from the HRSST facility is employed to assess its real-world performance, where the CNN achieves perfect identification of faults and voltage sags in a sample of nine events. This work highlights the potential of the proposed method to enhance PV protection schemes, providing a robust foundation for improved voltage event detection and grid reliability.

Cornachione, Matthew A.↗

Immobilization of Urease for continuous flow conversion of waste urea

An efficient and robust system for the urease catalyzed conversion of urea to ammonia has been developed using urease immobilized on modified agarose beads. Two different immobilization strategies, adsorption and covalent binding were studies using six different types of modified agarose beads. The immobilization of urease on each of the beads was studied at different concentrations and times using immobilization efficiency as a measure of success. The data from these experiments was used to identify potential candidates for immobilization scale up and implementation into the continuous flow system. The enzyme was then immobilized on the candidates in a packed bed reactor and the optimal flow rate and storage stability was determined. Future work will utilize the data obtained from these experiments to expand to other resins and enzymes.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

The Artificial Scientist: in-Transit Machine Learning of Plasma Simulations

Large-scale simulations or scientific experiments produce petabytes of data per run. This poses massive challenges for I/O and storage when scientific analysis workflows are run manually offline. Unsupervised deep learning-based techniques to extract patterns and non-linear relations from these large amounts of data provide a way to build scientific understanding from raw data, reducing the need for manual pre-selection of analysis steps, but require exascale compute and memory to process the full dataset available. In this paper, we demonstrate a heterogeneous streaming workflow in which plasma simulation data is streamed directly to a Machine Learning (ML) application training a model on the simulation data in-transit, completely circumventing the capacity-constrained filesystem bottleneck. This workflow employs openPMD to provide a high level interface to describe scientific data and also uses ADIOS2, to transfer volumes of data that exceed the capabilities of the filesystem. We employ experience replay to avoid catastrophic forgetting in learning from this non-steady state process in a continual manner and adapt it to improve model convergence while learning in-transit. As a proof-of-concept, we approach the ill-posed inverse problem of predicting particle dynamics from radiation in a particle-incell (PIConGPU) simulation of the Kelvin-Helmholtz instability (KHI). We detail hardware-software co-design challenges as we scale PIConGPU to full Frontier, the Top-1 system as of June 2024 Top500 list.

Kelling, Jeffrey [Helmholtz-Zentrum Dresden Rossen↗

Predicting U 3 O 8 powder processing conditions: An AI/ML approach analyzing deep learning embeddings of SEM micrographs

High-resolution SEM images of uranium-oxide powders encode micro- and nanoscale clues to their synthesis route and calcination temperature. We trained a ResNet-50 model on 11 commercial-scale U₃O₈ classes, ammonium diuranate (ADU) or uranyl peroxide (H₂O₂) precursors calcined at temperatures ranging from 400 to 750 °C and added a 256-D projection head before the classifier to analyze the learned representation. The best of eight seeds reached 92.4 % accuracy on reserved testing data, but our focus is the structure of the embedding space rather than the accuracy and labels. We quantify class relatedness in the original 256-D space using centroid similarity and distributional distances, and we use Uniform Manifold Approximation Projection (UMAP) for visualization. ‘Unknown’ images from different preparation methods, SEM operators, and from the literature localized near the expected classes under a nearest-centroid analysis without retraining, as well as clustered in similar UMAP space. In conclusion, this embedding-centered workflow complements black-box classification by providing quantitative, similarity-based comparisons of U₃O₈ morphologies and reduces storage space by up to 98 % for image data used in millisecond vector search comparisons.

36 MATERIALS SCIENCE↗

Basin-Scale Structural Features Database: Spatial Datasets to Support Carbon Storage Resource Assessments

Presentation slides on "Basin-Scale Structural Features Database: Spatial Datasets to Support Carbon Storage Resource Assessments" for CCUS 2025 Annual Meeting. The Basin-Scale Structural Features database contains a series of basin-scale spatial datasets representing structural features, including faults, fractures, folds, and earthquakes. Designed to support carbon storage feasibility and resources assessments for Carbon Capture and Storage (CCS) projects, the database leverages publicly available data resources from authoritative sources (e.g. US Geological Survey, State Geologic Surveys), and aims to help users better understand basin-scale structural features, as well as potential data gaps in areas with sparse information.

basin scale↗

Site Characterization of the Highest-Priority Geologic Formations for CO2 Storage in Wyoming

The project Site Characterization of the Highest-Priority Geologic Formations for CO2 Storage in Wyoming is one of 9 site characterization projects that were implemented as part of ARRA (American Recovery and Reinvestment Act). Data from this project was used to improve resolution of data in NATCARB in the area of study. Data related to this study has already been incorporated in NATCARB Atlas. The Wyoming Carbon Underground Storage Project (WY-CUSP) consisted of CO2 storage site characterization and evaluation, focusing on Wyoming’s most promising CO2 storage reservoirs (the Pennsylvanian Weber/Tensleep Sandstone and Mississippian Madison Limestone) and premier CO2 storage site (Rock Springs Uplift). Results from the WY-CUSP project suggest the two reservoirs could store up to 17,000 million tons of CO2. The WY-CUSP team drilled a stratigraphic test well and acquired a 3-D seismic survey covering 25 square miles of the Rock Springs Uplift site. The team retrieved 916 feet of core from the 12,810-foot-deep well, along with a complete log suite, borehole images, fluid samples, and other data. Project partners (1) provided continuous visual documentation of the core, including grain size, mineralogy, facies distribution, and porosity; (2) performed continuous permeability and velocity scans of selected reservoir intervals; and (3) chemically analyzed the fluid samples. WY-CUSP scientists integrated seismic attributes with observations from log suites, a VSP survey, core, fluid samples, and laboratory analyses, including continuous permeability scans. From these integrations, researchers constructed 3-D spatial distribution volumes of reservoir and seal properties that represent geological heterogeneity at the targeted CO2 storage site. The WY-CUSP team used this data to perform new CO2 plume migration simulations. Baker Hughes, Inc., completed a series of small-scale, in-situ water injectivity measurements. A database was formed when observations, analyses, and experiments from the stratigraphic test well were integrated. Correlation of these data allowed petrophysical parameters to be extrapolated from the test well out into the storage domain (5x5 mile 3-D seismic survey volume). This resulted in an improved, realistic understanding of performance assessments for potential CO2 storage scenarios. The WY-CUSP team worked on (1) improving CO2 storage resource estimates, (2) establishing long-term integrity and permanence of confining layers, (3) designing a profitable strategy for pressure management, and (4) evaluating the utilization of stored CO2 at the Rock Spring Uplift. Finally, Baker Hughes developed a microseismic baseline for the test site using in-bore geophones to complete field operations.

3-D seismic↗

DaYu: Optimizing Distributed Scientific Workflows by Decoding Dataflow Semantics and Dynamics

The combination of ever-growing scientific datasets and distributed workflow complexity creates I/O performance bottlenecks due to data volume, velocity, and variety. Although the increasing use of descriptive data formats (e.g., HDF5, netCDF) helps organize these datasets, it also creates obscure bottlenecks due to the need to translate high level operations into file addresses and then into low-level I/O operations. To address this challenge, we introduce DaYu, a method and toolset for analyzing (a) semantic relationships between logical datasets and file addresses, (b) how dataset operations translate into I/O, and (c) the combination across entire workflows. DaYu's analysis and visualization enables identification of critical bottlenecks and reasoning about remediation. We describe our methodology and propose optimization guidelines. Evaluation on scientific workflows demonstrates up to 3.7x performance improvements in I/O time for obscure bottlenecks. The time and storage overhead for DaYu's time-ordered data is typically under 0.2% of runtime and 0.25% of data volume, respectively.

Tang, Meng↗