Search NASA⌕ Search

SEARCH · Search NASA

Results for “DATA STORAGE”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 307 records · Page 17

Machine Learning Applications in Analyzing the Role of Shale Barriers and Baffles for CO2 Storage

This study uses machine learning to analyze microseismic data from the Illinois Basin Decatur Project (IBDP) and quantify CO₂ plume extents. By leveraging well logs, microseismic records, and CO₂ injection metrics, the research predicts subsurface CO₂ plume dynamics. Findings show vertical clustering of microseismic events near the injection well, with CO₂ periodically breaching barriers due to buoyancy. K-Means clustering performed best, achieving the highest Silhouette Score and lowest Davies-Bouldin Index. This capability is crucial for real-time monitoring and management of CO₂ sequestration sites, validated against physical models and IBDP data, reinforcing CO₂ geological sequestration's viability and enhancing management tools.

Carr, Timothy↗

Myna

The additive manufacturing (AM) community has been developing digital factory tools over the past decade to better leverage the multi-modal process data coming out of the advanced manufacturing process. As a result, numerous databases of additive manufacturing process data exist in the literature and in the archival storage of disparate research groups. While some efforts have been made to create a standard ontology for storing and sharing AM data, in practice a variety of data structures are used to store AM build data, even within a single institution. This causes many problems for maintainability and extensibility when attempting to integrate computational modeling tools with experimental data to either validate models or to provide further insight into results and trends. Myna is a Python-based framework that aims to decrease the effort needed to connect individual computational models to the variety of AM process data that exist in different research groups and institutions. This type of software is sometimes referred to as "middleware" or “glueware,” in that it connects disparate databases and applications into a single computational ecosystem. Instead of maintaining unique interfaces between each application and each database, developers can create a single interface from each application to Myna and thereby gain access to the implemented database connections. Similarly, developing a database connection in Myna provides access to the developed simulation applications. This framework greatly simplifies the maintainability of model applications that rely on experimental data. Using external simulation tools, users will also be able to run pre-configured workflows using the built-in workflow manager. Several examples of input files are provided with Myna for different workflows, including melt pool geometry predictions and detailed melt pool and solidification microstructure predictions.

Knapp, GerryL. [Oak Ridge National Laboratory (ORN↗

One Earth Energy Seismic Interpretation

The objectives of the Illinois Storage Corridor (ISC) project are to accelerate commercial deployment of carbon capture utilization and storage at two individual sites and receive approvals for Underground Injection Control (UIC) Class VI permits for construction at each site (ISC Project Narrative, 2020). As part of this project, and as part of the subsurface geologic characterization, 2D seismic data was acquired at both sites. This report summarizes the findings from the 2D and 3D seismic interpretation at the One Earth Energy site near Gibson City, Illinois. The seismic data confirms the stratigraphic continuity of the Mt. Simon Arkose Zone storage interval and the Eau Claire confining unit across the project area. The seismic data also indicates that there are faults that transect the Mt. Simon Arkose Zone Sandstone storage reservoir within the modeled CO 2 plume (for more detailed information, see Faults and Fractures section of One Earth Energy Class VI Permit applications). However, the seismic data also shows that there are no faults within the modeled CO 2 plume that transect the confining unit Eau Claire Formation. The faults that transect the Mt. Simon Arkose Zone Sandstone storage reservoir all tip out in the Lower Mt. Simon Formation and do not reach the overlying Eau Claire confining unit. A small 3D survey acquired around the One Earth Energy #1 characterization well confirms these findings.

20 FOSSIL-FUELED POWER PLANTS↗

Eureka: Enabling Fine-Grained Access and Range Queries on Compressed Scientific Data via Data-Index Co-Compression

Handling large-scale scientific data in high-performance computing (HPC) environments poses significant challenges, including excessive I/O, high storage costs, and slow query performance. Traditional approaches often require full data decompression and scans, making them impractical for real-time or interactive analysis. To address these limitations, we introduce Eureka, a unified data-index co-compression framework that enables fine-grained access and efficient range queries on compressed scientific datasets. Eureka integrates spatial domain decomposition with block-wise error-bounded lossy compression to support selective decompression. It constructs a hierarchical AVL-tree index during compression to capture block-level value ranges, enabling fast pruning during query execution. To reduce metadata overhead, the index itself is also compressed while ensuring recall-preserving results. Experiments on six diverse HPC simulation datasets show that Eureka achieves up to 25x data compression and over 300x index compression, surpassing state-of-the-art compressors such as SZ3 and ZFP in rate-distortion performance. Additionally, Eureka delivers over 30x speedup for low-selectivity range queries, making it a scalable and efficient solution for modern scientific data analysis.

Yan, Ning↗

Field Test Report Neutron Scintillator Array Dry Storage Cask Scanner FY2024

During two weeks of Field Testing at the Idaho National Laboratory INTEC Cask Farm in July and August 2024, the LLNL Dry Storage Cask Scanner Array was lifted on top of an MC-10 dry storage fuel cask and operated to acquire neutron and gamma-ray data from the 24 fuel bundle positions. Neutron and gamma-ray data acquisition scans across the top of the cask of varying dwell times were performed July 15-18, 2024 and August 19-22, 2024 to evaluate the ability of the scanner data to reveal asymmetries in the fuel positions that reflect asymmetries in the MC-10 cask fuel bundle loading. The MC-10 cask 24 position fuel bundle loading at the INTEC Cask Farm is well documented, including the locations of six empty fuel bundle positions. This loading presents an opportunity to test the ability of the scanner system to detect diversion of spent fuel bundles as well as to validate the MC-10 cask MCNP modeling. The cask scanner array consists of six Stilbene crystal scintillator detectors and a linear actuator frame that moves the six detectors across the MC-10 dry storage cask to obtain data above each of the 24 fuel bundle positions. The detectors are connected to a pulse-shape discrimination data acquisition system capable of generating separate neutron and gamma-ray spectra for each detector and for each scan position. From the prior single detector Field Test in 2021 and iteration with MCNP modeling, the neutron and gamma-ray data were analyzed in multiple energy regions to identify an analysis method that would provide the strongest and most consistent signature of the asymmetric MC-10 cask fuel loading1 . From both the 2021 Field Test and the current Field Test results, the neutron capture gamma-ray count rate around 2.2 MeV provides the strongest signature of the asymmetric MC-10 cask fuel loading and has qualitative agreement with MCNP calculations. Counting all gamma-rays produces a similar signature. Neutrons emerging from the cask top are moderated and captured by the hydrogen in the polyethylene moderator and scintillator detector, producing a 2.2 MeV gamma ray which is seen in the scintillator gamma-ray spectrum. The count rate in the 2.2 MeV gamma-ray region is ~50 c/s, which is ~1000x higher than the ~0.05 n/s rate in the > 4MeV neutron region, and ~50x greater than the ~1 n/s rate in the neutrons > 500 keV region. Analysis of the 2.2 MeV neutron-capture Compton-scattered gamma-rays produces a statistically significant signature of the INTEC Cask Farm MC-10 asymmetric fuel loading. MCNP simulations indicate that the average neutron energy spectrum offers the potential to detect a large asymmetry from several missing bundles as well as individual missing fuel bundles. Testing this feature will require measurements on a cask with single missing elements.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

SWARM: Reimagining scientific workflow management systems in a distributed world

Modern scientific workflows process massive amounts of data from diverse instruments and sensors, leveraging geographically distributed, heterogeneous compute and storage resources—from leadership-class systems to edge devices—connected by high-performance networks. The diversity of resources introduces challenges in harnessing their full potential, with resilience issues arising across applications, system software, networks, storage, and hardware. Today, workflow management systems (WMS) coordinate the execution of computation and data management tasks across target resources. However, WMS’s centralized nature makes them vulnerable to faults and scalability issues that may result in failures of entire computational campaigns. In conclusion, this paper introduces a novel agentic framework for workflow management, fully distributing and decentralizing the WMS functions and modeling them as swarm intelligence agents infused with advanced artificial intelligence solutions and traditional distributed computing algorithms that can make coordinated decisions in the presence of failures of the underlying cyberinfrastructure.

Swarm intelligence↗

Managing negative values is reservoir inflow computation: A case study

Reservoir inflow is conventionally estimated using the water balance method, which involves the reservoir release and the change in storage during the period considered. As a result, the estimated inflow may sometimes be negative as the errors involved in each input variable build-up to the output. In our study, the fleet data was provided by the Tennessee Valley Authority (TVA) for their Norris Hydropower facility. Unlike the flow release data, which was readily accessible, the change in storage had to be calculated using the reservoir elevation and volume relationship. The original inflow estimates produced a wide range of negative values with large outliers, making it difficult to visualize the current trends. This paper describes a methodology to remove the negative values encountered during the inflow computation, and the results were analyzed by correlating with the nearby streamflow gaging stations.

Shibu, Asha↗

Reservoir Storage Capacity Change (ResCap)

Overview Storage capacity is an essential reservoir metric that is directly linked to various water management and energy objectives. Accurate reporting and tracking of change in storage over time is crucial for the safe and reliable operation of the associated dam. While storage information is available for many reservoirs through the National Inventory of Dams, additional details, e.g. water elevation levels as well as changes over time are not included. This dataset contains reservoir storage capacities based on conducted surveys in CONUS. To represent changes in a reservoir’s storage over time, the storage capacity as determined by the first and last conducted survey is listed. The level of detail of surveys can vary greatly and improved with technological advancements. Therefore, the type of survey and year when it was conducted is noted. To ensure a fair comparison of storage capacities, the water elevation level along with the corresponding operation of the dam is reported. Structural changes, e.g. heightening of a dam will have an influence on the storage capacity and are therefore also mentioned. A total of 739 different reservoir storage capacity comparisons are listed, with some reservoirs represented more than once (storage capacity comparison at different water elevation levels). Methodology Data were acquired from USBR reservoir survey reports, TWDB lake survey reports and elevation-area-capacity tables, the RSI Web Portal and the NID (USACE, 2024). Initial storage capacity along with year and type of survey record is compared to the most recent reported storage capacity, survey type and year. Comparison elevation in feet as well as comparison elevation type were either extracted from survey reports (USBR, TWDB) or the Web Portal (RSI) and in some cases cross-referenced with data from other sources (Water Management Data, USACE, Water Data for Texas, TWDB).

Chu, Antonia [ORNL] (ORCID:0009000510540427)↗

I/O in Machine Learning Applications on HPC Systems: A 360-degree Survey

Growing interest in Artificial Intelligence (AI) has resulted in a surge in demand for faster methods of Machine Learning (ML) model training and inference. This demand for speed has prompted the use of high performance computing (HPC) systems that excel in managing distributed workloads. Because data is the main fuel for AI applications, the performance of the storage and I/O subsystem of HPC systems is critical. In the past, HPC applications accessed large portions of data written by simulations or experiments or ingested data for visualizations or analysis tasks. ML workloads perform small reads spread across a large number of random files. This shift of I/O access patterns poses several challenges to modern parallel storage systems. In this paper, we survey I/O in ML applications on HPC systems, and target literature within a 6-year time window from 2019 to 2024. We define the scope of the survey, provide an overview of the common phases of ML, review available profilers and benchmarks, examine the I/O patterns encountered during offline data preparation, training, and inference, and explore I/O optimizations utilized in modern ML frameworks and proposed in recent literature. Lastly, we seek to expose research gaps that could spawn further R&D.

97 MATHEMATICS AND COMPUTING↗

The Role of Snowmelt and Subsurface Heterogeneity in Headwater Hydrology of a Mountainous Catchment in Colorado: A Model‐Data Integration Approach

Mountainous headwater streams are sustained by both snowmelt‐driven streamflow and groundwater discharge in the Upper Colorado River Basin. However, predicting headwater stream discharge magnitude and peak flow timing is challenging in mountainous terrains, where snowmelt rates vary with vegetation type and elevation, and heterogeneous subsurface physical properties influence groundwater storage and its release. We used a model‐data integration approach to investigate the roles of snowmelt and subsurface structure in stream discharge and groundwater level. We ran an ensemble of 100 integrated surface‐subsurface hydrologic models for a mountainous headwater catchment near Crested Butte, Colorado, USA. We also evaluated and calibrated these models against observed data sets, including snow depth measurements using distributed temperature probes, stream discharge, and groundwater levels. Calibration with multiple data sources using neural density estimators has further constrained uncertainty in subsurface properties and snowmelt rates. Results indicated that observed slower snowmelt rates in evergreen forests delayed the peak flow and baseflow onset. In upstream areas with lower subsurface permeability, water was stored within the subsurface but was not released as interflow or shallow groundwater flow, and thereby not contributing to downstream streamflow during recession limb periods. Double peaks in groundwater occurred in areas with spatial subsurface heterogeneity, in our case due to the contrast between granodiorite and Mancos shale. These process‐based insights into groundwater and snowmelt dynamics in mountainous headwaters will help improve predictions of headwater hydrology.

Wang, Lijing [University of Connecticut, Storrs, C↗

Learning from Arctic Microgrids: Cost and Resiliency Projections for Renewable Energy Expansion with Hydrogen and Battery Storage

Electricity in rural Alaska is provided by more than 200 standalone microgrid systems powered predominantly by diesel generators. Incorporating renewable energy generation and storage to these systems can reduce their reliance on costly imported fuel and improve sustainability; however, uncertainty remains about optimal grid architectures to minimize cost, including how and when to incorporate long-duration energy storage. This study implements a novel, multi-pronged approach to assess the techno-economic feasibility of future energy pathways in the community of Kotzebue, which has already successfully deployed solar photovoltaics, wind turbines, and battery storage systems. Using real community load, resource, and generation data, we develop a series of comparison models using the HOMER Pro software tool to evaluate microgrid architectures to meet over 90% of the annual community electricity demand with renewable generation, considering both battery and hydrogen energy storage. We find that near-term planned capacity expansions in the community could enable over 50% renewable generation and reduce the total cost of energy. Additional build-outs to reach 75% renewable generation are shown to be competitive with current costs, but further capacity expansion is not currently economical. We additionally include a cost sensitivity analysis and a storage capacity sizing assessment that suggest hydrogen storage may be economically viable if battery costs increase, but large-scale seasonal storage via hydrogen is currently unlikely to be cost-effective nor practical for the region considered. While these findings are based on data and community priorities in Kotzebue, we expect this approach to be relevant to many communities in the Arctic and Sub-Arctic regions working to improve energy reliability, sustainability, and security.

25 ENERGY STORAGE↗

CO 2 rock physics modeling for reliable monitoring of geologic carbon storage

Monitoring, verification, and accounting (MVA) are crucial to ensure safe and long-term geologic carbon storage. Seismic monitoring is a key MVA technique that utilizes seismic data to infer elastic properties of CO 2 -saturated rocks. Reliable accounting of CO 2 in subsurface storage reservoirs and potential leakage zones requires an accurate rock physics model. However, the widely used CO 2 rock physics model based on the conventional Biot-Gassmann equation can substantially underestimate the influence of CO 2 saturation on seismic waves, leading to inaccurate accounting. We develop an accurate CO 2 rock physics model by accounting for both effects of the stress dependence of seismic velocities in porous rocks and CO 2 weakening on the rock framework. We validate our CO 2 rock physics model using the Kimberlina-1.2 model (a previously proposed geologic carbon storage site in California) and create time-lapse elastic property models with our new rock physics method. We compare the results with those obtained using the conventional Biot-Gassmann equation. Our innovative approach produces larger changes in elastic properties than the Biot-Gassmann results. Using our CO 2 rock physics model can replicate shear-wave speed reductions observed in the laboratory. Our rock physics model enhances the accuracy of time-lapse elastic-wave modeling and enables reliable CO 2 accounting using seismic monitoring.

58 GEOSCIENCES↗

Leveraging Pre-Built Catalogs and Object-Level Scheduling to Eliminate I/O Bottlenecks in HPC Environments

Modern High-Performance Computing (HPC) environments face mounting challenges due to the shift from large to small file datasets, along with an increasing number of users and parallelized applications. As HPC systems rely on Parallel File Systems (PFS), such as Lustre for data processing, performance bottlenecks stemming from Object Storage Target (OST) contention have become a significant concern. Existing solutions, such as LADS with its object-level scheduling approach, fall short in large-scale HPC environments due to their inability to effectively address metadata I/O bottlenecks and the growing number of I/O processes. This study highlights the pressing need for a comprehensive solution that tackles both OST contention and metadata I/O challenges in diverse HPC workloads. To address these challenges, we propose SwiftLoad, an object-level I/O scheduling framework that leverages a metadata catalog to enhance the performance and efficiency of parallel HPC utilities. The adoption of the metadata catalog mitigates the metadata I/O bottlenecks that commonly occur in HPC utilities, a challenge that is particularly pronounced in object-level I/O scheduling. SwiftLoad addresses OST contention and the uneven distribution of I/O processes across different OSTs through mathematical modeling and incorporates a Loader Configuration Module to regulate the number of I/O processes. Evaluated with two representative utilities—data deduplication profiling and data augmentation—SwiftLoad achieved performance improvements of up to 5.63x and 11.0x, respectively, on a production supercomputer.

HPC↗

Grid Resiliency with a 100% Renewable Microgrid

San Diego Gas & Electric Company (SDG&E) installed America’s first and largest utility-scale microgrid in Borrego Springs in 2013. The first generation Borrego Springs Microgrid utilized diesel generators to form and stabilize the microgrid island, with support from grid-scale batteries and local solar photovoltaic (PV) generation. In this project, SDG&E in partnership with National Renewable Energy Laboratory (NREL) demonstrated through modeling, simulation and utility field testing that blackstart and islanding of the microgrid can be led with 100% renewable, inverter based resources (IBRs), to help reduce community reliance on conventional generation resources. Through equipment upgrades, grid-forming island leader capability was transitioned to a battery IBR instead of the Borrego Springs Microgrid diesel generators. A new microgrid controller was integrated to the microgrid and programmed to control and manage multiple energy storage systems. Synchrophasor and other power quality data verified autonomous, high-speed response of the IBRs through blackstart, islanding, and load step testing. Results of project field evaluations provide distribution systems operators (DSO) with increased confidence that renewable, IBR can replace traditional generators to blackstart and island microgrids and rapidly establish stable island frequency with rapid changes in peak power demand. Importantly, the project validated the integration feasibility of a distributed energy resource management system (DERMS) controller that manages multiple grid-forming and grid-following IBRs, establishing a standard design interface to reduce the complexity of integrating new DERs in the future and supporting replication by the industry. As a result of learnings in this project, SDG&E has implemented the microgrid controller strategy at multiple other microgrid sites, thereby validating the replicability of the solution. Hardware-in-the-loop (HIL) simulations including power and controller HIL hardware — along with electromagnetic transient (EMT) simulations of Borrego Springs Microgrid —informed adjustments to inverter parameters and were important to characterize the performance of the IBRs in relevant operating conditions before deployment. The EMT and HIL simulations of islanding the entire community are important contributions in providing confidence in IBR performance prior to future islanding of the community in the field. High-fidelity EMT and/or HIL simulation of IBRs can de-risk field operations, and its relevance and importance as a tool is increasing as distribution grids and microgrids become more complex and dynamic with an increasing proportion of renewable generation, distributed energy storage, and two-way power and energy flows.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Online and Offline Identification of False Data Injection Attacks in Battery Sensors Using a Single Particle Model

The cells in battery energy storage systems are monitored, protected, and controlled by battery management systems whose sensors are susceptible to cyberattacks. False data injection attacks (FDIAs) targeting batteries’ voltage sensors affect cell protection functions and the estimation of critical battery states like the state of charge (SoC). Inaccurate SoC estimation could result in battery overcharging and over discharging, which can have disastrous consequences on grid operations. This paper proposes a three-pronged online and offline method to detect, identify, and classify FDIAs corrupting the voltage sensors of a battery stack. To accurately model the dynamics of the series-connected cells a single particle model is used and to estimate the SoC, the unscented Kalman filter is employed. FDIA detection, identification, and classification was accomplished using a tuned cumulative sum (CUSUM) algorithm, which was compared with a baseline method, the chi-squared error detector. Online simulations and offline batch simulations were performed to determine the effectiveness of the proposed approach. Throughout the batch simulations, the CUSUM algorithm detected attacks, with no false positives, in 99.83% of cases, identified the corrupted sensor in 97% of cases, and determined if the attack was positively or negatively biased in 97% of cases.

25 ENERGY STORAGE↗

Streaming Data in HPC Workflows Using ADIOS

The “IO Wall” problem, in which the gap between computation rate and data access rate grows continuously, poses significant problems to scientific workflows which have traditionally relied upon using the filesystem for intermediate storage between workflow stages. One way to avoid this problem in scientific workflows is to stream data directly from producers to consumers and avoiding storage entirely. However, the manner in which this is accomplished is key to both performance and usability. This paper presents the Sustainable Staging Transport, an approach which allows direct streaming between traditional file writers and readers with few application changes. SST is an ADIOS “engine”, accessible via standard ADIOS APIs, and because ADIOS allows engines to be chosen at run-time, many existing file-oriented ADIOS workflows can utilize SST for direct application-to-application communication without any source code changes. This paper describes the design of SST and presents performance results from various applications that use SST, for feeding model training with simulation data with substantially higher bandwidth than the theoretical limits of Frontier’s file system, for strong coupling of separately developed applications for multiphysics multiscale simulation, or for in situ analysis and visualization of data to complete all data processing shortly after the simulation finishes.

Podhorszki, Norbert [ORNL] (ORCID:000000019647542X↗

Exploring the Future Energy Value of Long-Duration Energy Storage

Long-duration energy storage is commonly viewed as a key technology for providing flexibility to the grid and broader energy systems over a multidecadal time frame. However, prior work has typically used present-day grid infrastructures to characterize the relationship between the duration and arbitrage value of storage in electricity markets. This study leverages established National Renewable Energy Laboratory grid planning and operations tools, analysis, and data to execute a price-taker model of an energy storage system for several 8760 h price series representative of current and future contiguous United States grid infrastructures with varying shares of variable renewable energy (VRE). We find that the total value of energy storage typically increases with VRE shares, but any increase in the relative value of longer storage durations over time depends on the region and grid mix. Some regions see incremental value increasing notably, up to 20–40 h in 2050, while others do not. The negative effect of lower roundtrip efficiency on value is also found to be scenario-dependent, with the energy value in higher VRE scenarios being less sensitive to roundtrip efficiency and more supportive of longer storage durations. Long-duration storage value and deployment potential are a function of evolving electricity sector infrastructure, markets, and policy, making it critical to consistently revisit potential long-duration storage contributions to the grid.

14 SOLAR ENERGY↗