Search NASASearch

SEARCH · Search NASA

Results for “DATA STORAGE SYSTEMS”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Energy Systems Integration Facility (ESIF): World-Class Systems Integration Capabilities and Research

The Energy Systems Integration Facility (ESIF), located at the National Renewable Energy Laboratory (NREL) South Table Mountain campus, is a world-renowned user facility for research and development of modern, advanced, and clean energy technologies. ESIF is distinguished by its continuously evolving, highly integrated systems that span throughout the building, connecting research capabilities across multiple laboratories and test areas. The primary ESIF research systems include: [1] data, cyber, and control networks, [2] research electrical distribution buses (REDB), [3] thermal integration infrastructure, and [4] hydrogen systems. The data, cyber, and control networks provide monitoring, control, communication, automation, visualization, and time series data storage and tagging capabilities for research projects and ESIF systems, including facility safety functions. The REDB system consists of four dedicated AC and DC electrical power networks that can connect devices located across the facility through versatile, automatic circuit configuration to support complex power electronics experiments up to the megawatt-scale. The thermal integration infrastructure consists of three temperature-conditioned water loops that provide heating and cooling interfaces and capabilities for thermal energy research. The hydrogen systems provide megawatt-scale hydrogen production, drying, compression, high-pressure storage, and delivery to laboratory end uses, including hydrogen fuel cell vehicle fueling. The ESIF research systems interconnect and extend throughout the various lab areas of the facility to create elaborate networks composed of diverse technologies for cutting-edge research. The ESIF capabilities are operated and stewarded by the ESIF Research Operations group, who also actively upgrade and advance the systems to ensure they remain ahead of anticipated research - enabling the success of many pioneering energy integration projects. The poster, created by members of the ESIF Research Operations team, highlights and summarizes the four core integrated systems at ESIF. The poster was first presented at the internal NREL Energize Forum on May 13th, 2024, and received the "Best Poster" award.

capabilities

Grid Resiliency with a 100% Renewable Microgrid

San Diego Gas & Electric Company (SDG&E) installed America’s first and largest utility-scale microgrid in Borrego Springs in 2013. The first generation Borrego Springs Microgrid utilized diesel generators to form and stabilize the microgrid island, with support from grid-scale batteries and local solar photovoltaic (PV) generation. In this project, SDG&E in partnership with National Renewable Energy Laboratory (NREL) demonstrated through modeling, simulation and utility field testing that blackstart and islanding of the microgrid can be led with 100% renewable, inverter based resources (IBRs), to help reduce community reliance on conventional generation resources. Through equipment upgrades, grid-forming island leader capability was transitioned to a battery IBR instead of the Borrego Springs Microgrid diesel generators. A new microgrid controller was integrated to the microgrid and programmed to control and manage multiple energy storage systems. Synchrophasor and other power quality data verified autonomous, high-speed response of the IBRs through blackstart, islanding, and load step testing. Results of project field evaluations provide distribution systems operators (DSO) with increased confidence that renewable, IBR can replace traditional generators to blackstart and island microgrids and rapidly establish stable island frequency with rapid changes in peak power demand. Importantly, the project validated the integration feasibility of a distributed energy resource management system (DERMS) controller that manages multiple grid-forming and grid-following IBRs, establishing a standard design interface to reduce the complexity of integrating new DERs in the future and supporting replication by the industry. As a result of learnings in this project, SDG&E has implemented the microgrid controller strategy at multiple other microgrid sites, thereby validating the replicability of the solution. Hardware-in-the-loop (HIL) simulations including power and controller HIL hardware — along with electromagnetic transient (EMT) simulations of Borrego Springs Microgrid —informed adjustments to inverter parameters and were important to characterize the performance of the IBRs in relevant operating conditions before deployment. The EMT and HIL simulations of islanding the entire community are important contributions in providing confidence in IBR performance prior to future islanding of the community in the field. High-fidelity EMT and/or HIL simulation of IBRs can de-risk field operations, and its relevance and importance as a tool is increasing as distribution grids and microgrids become more complex and dynamic with an increasing proportion of renewable generation, distributed energy storage, and two-way power and energy flows.

24 POWER TRANSMISSION AND DISTRIBUTION

Towards an Introspective Dynamic Model of Globally Distributed Computing Infrastructures

Large-scale scientific collaborations like ATLAS, Belle II, CMS, DUNE, and others involve hundreds of research institutes and thousands of researchers spread across the globe. These experiments generate petabytes of data, with volumes soon expected to reach exabytes. Consequently, there is a growing need for computation, including structured data processing from raw data to consumer-ready derived data, extensive Monte Carlo simulation campaigns, and a wide range of end-user analysis. To manage these computational and storage demands, centralized workflow and data management systems are implemented. However, decisions regarding data placement and payload allocation are often made disjointly and via heuristic means. A significant obstacle in adopting more effective heuristic or AI-driven solutions is the absence of a quick and reliable introspective dynamic model to evaluate and refine alternative approaches. In this study, we aim to develop such an interactive system using real-world data. By examining job execution records from the PanDA workflow management system, we have pinpointed key performance indicators such as queuing time, error rate, and the extent of remote data access. The dataset includes five months of activity. Additionally, we are creating a generative AI model to simulate time series of payloads, which incorporate visible features like category, event count, and submitting group, as well as hidden features like the total computational load—derived from existing PanDA records and computing site capabilities. These hidden features, which are not visible to job allocators, whether heuristic or AI-driven, influence factors such as queuing times and data movement.

kilic, Ozgur Ozan [Brookhaven National Laboratory

Zero-Power Analog Optical Processing

The motivation behind this research is the growing challenge of handling the massive amounts of data generated by modern imaging systems. Conventional digital image processing techniques are struggling to keep pace with the demands of high-resolution and high-speed imaging systems for remote sensing due to their high-power consumption and data storage requirements. We present a novel approach based on analog photonics to address this challenge. The proposed system utilizes a silicon-photonics-based image encoder positioned after image formation and initial optical-to-electrical conversion. The photonic encoder compresses image data using a passive disordered photonic structure to perform kernel-type random projections of the raw data. The compressed data is then processed by a back-end neural network, which reconstructs the original image with high fidelity (structural similarity exceeding 90%). Our proposed approach has the potential to compress images with ~ 1000X lower power consumption compared to digital approaches with data rates exceeding 1 terapixel/second.

97 MATHEMATICS AND COMPUTING

Towards a self-driving trigger at the LHC: adaptive response in real time

Real-time data filtering and selection—or trigger—systems at high-throughput scientific facilities such as the experiments at the Large Hadron Collider must process extremely high-rate data streams under stringent bandwidth, latency, and storage constraints. Yet these systems are typically designed as static, hand-tuned menus of selection criteria grounded in prior knowledge and simulation. In this work, we further explore the concept of a self-driving trigger, an autonomous data-filtering framework that reallocates resources and adjusts thresholds dynamically in real-time to optimize signal efficiency, rate stability, and computational cost as instrumentation and environmental conditions evolve. We introduce a benchmark ecosystem to emulate realistic collider scenarios and demonstrate real-time optimization of a menu including canonical energy sum triggers as well as modern anomaly-detection algorithms that target non-standard event topologies using machine learning. Using simulated data streams and publicly available collision data from the Compact Muon Solenoid experiment, we demonstrate the capability to dynamically and automatically optimize trigger performance under specific cost objectives without manual retuning. Our adaptive strategy shifts trigger design from static menus with heuristic tuning to intelligent, automated, data-driven control, unlocking greater flexibility and discovery potential in future high-energy physics analyses.

Emami, Shaghayegh [Michigan U.] (ORCID:00090007589

THERMAL MODELING OF HANFORD CESIUM AND STRONTIUM CANISTERS DURING SIMULATED LOADING

A computational fluid dynamics (CFD) model was built to simulate planned testing of heater assemblies within a canister and overpack for the Hanford Lead Canister (HLC) project. The HLC is a canister storage system that will contain heaters to simulate the decay heat of nuclear material and provide the canister storage system with environmental conditions equivalent to the operating conditions on a dry storage pad. The HLC will be equipped with long-term data collection and monitoring systems to provide an early warning of corrosion, pitting, cracking, or other signs of canister degradation that might threaten the integrity of the containment boundary over the potentially long term of dry storage. An important part of the HLC development is to make pretest numerical predictions for the behavior of the heated canister during the simulated radiolytic decay heat testing, which simulates the dry storage system during loading operations. The simulated radiolytic decay heat test is planned for mid-2024 in a configuration that includes the heater assembly, overpack, and canister, but with the lids removed to allow loading cesium and strontium capsules into the canister. One of the goals of the test is to evaluate the thermal behavior of the canister and overpack assembly in the ambient air of the test facility, which will provide data critical to validating the thermal models and understanding how the HLC will perform as a system once deployed. To best approximate real-world conditions, the CFD model includes the full air volume of the mock-up truck bay the heated canister test will be performed in, enabling detailed investigation of how the heated canister affects airflow around it. Rigorous pre-deployment testing of the complete HLC cask and canister system is intended to be completed before the HLC is deployed in the 2028 timeframe. This study presents the pre-test temperature predictions of the simulated radiolytic decay heat test. A description of the heater assembly, canister, and overpack system is presented. The model was developed with the commercial CFD software STAR-CCM+. An uncertainty analysis was run with the CFD model to determine the uncertainty in the temperature predictions and provide a range over which the predicted temperatures are expected to vary. The uncertainty analysis was preformed by coupling STAR-CCM+ with the software Dakota, which provides advanced parametric analyses, including quantification of margins and uncertainty with computational models. This work is expected to provide insight into SNF canister behavior.

Carpenter-Graffy, Dina E.

Accelerating data acquisition with FPGA-based edge machine learning: a case study with LCLS-II

New scientific experiments and instruments generate vast amounts of data that need to be transferred for storage or further processing, often overwhelming traditional systems. Edge machine learning (EdgeML) addresses this challenge by integrating machine learning (ML) algorithms with edge computing, enabling real-time data processing directly at the point of data generation. EdgeML is particularly beneficial for environments where immediate decisions are required, or where bandwidth and storage are limited. In this paper, we demonstrate a high-speed configurable ML model in a fully customizable EdgeML system using a field programmable gate array (FPGA). Our demonstration focuses on an angular array of electron spectrometers, referred to as the ‘CookieBox,’ developed for the Linac Coherent Light Source II project. The EdgeML system captures 51.2 Gbps from a 6.4 GS s −1 analog to digital converter and is designed to integrate data pre-processing and ML inside an FPGA. Our implementation achieves an inference latency of 0.2 µs for the ML model, and a total latency of 0.4 µs for the complete EdgeML system, which includes pre-processing, data transmission, digitization, and ML inference. The modular design of the system allows it to be adapted for other instrumentation applications requiring low-latency data processing.

97 MATHEMATICS AND COMPUTING

Learning from Arctic Microgrids: Cost and Resiliency Projections for Renewable Energy Expansion with Hydrogen and Battery Storage

Electricity in rural Alaska is provided by more than 200 standalone microgrid systems powered predominantly by diesel generators. Incorporating renewable energy generation and storage to these systems can reduce their reliance on costly imported fuel and improve sustainability; however, uncertainty remains about optimal grid architectures to minimize cost, including how and when to incorporate long-duration energy storage. This study implements a novel, multi-pronged approach to assess the techno-economic feasibility of future energy pathways in the community of Kotzebue, which has already successfully deployed solar photovoltaics, wind turbines, and battery storage systems. Using real community load, resource, and generation data, we develop a series of comparison models using the HOMER Pro software tool to evaluate microgrid architectures to meet over 90% of the annual community electricity demand with renewable generation, considering both battery and hydrogen energy storage. We find that near-term planned capacity expansions in the community could enable over 50% renewable generation and reduce the total cost of energy. Additional build-outs to reach 75% renewable generation are shown to be competitive with current costs, but further capacity expansion is not currently economical. We additionally include a cost sensitivity analysis and a storage capacity sizing assessment that suggest hydrogen storage may be economically viable if battery costs increase, but large-scale seasonal storage via hydrogen is currently unlikely to be cost-effective nor practical for the region considered. While these findings are based on data and community priorities in Kotzebue, we expect this approach to be relevant to many communities in the Arctic and Sub-Arctic regions working to improve energy reliability, sustainability, and security.

25 ENERGY STORAGE

Online and Offline Identification of False Data Injection Attacks in Battery Sensors Using a Single Particle Model

The cells in battery energy storage systems are monitored, protected, and controlled by battery management systems whose sensors are susceptible to cyberattacks. False data injection attacks (FDIAs) targeting batteries’ voltage sensors affect cell protection functions and the estimation of critical battery states like the state of charge (SoC). Inaccurate SoC estimation could result in battery overcharging and over discharging, which can have disastrous consequences on grid operations. This paper proposes a three-pronged online and offline method to detect, identify, and classify FDIAs corrupting the voltage sensors of a battery stack. To accurately model the dynamics of the series-connected cells a single particle model is used and to estimate the SoC, the unscented Kalman filter is employed. FDIA detection, identification, and classification was accomplished using a tuned cumulative sum (CUSUM) algorithm, which was compared with a baseline method, the chi-squared error detector. Online simulations and offline batch simulations were performed to determine the effectiveness of the proposed approach. Throughout the batch simulations, the CUSUM algorithm detected attacks, with no false positives, in 99.83% of cases, identified the corrupted sensor in 97% of cases, and determined if the attack was positively or negatively biased in 97% of cases.

25 ENERGY STORAGE

Prospecting for Critical Minerals and Rare Earth Elements from Marcellus Shale in the Western Portion of the Appalachian Basin with Non-Destructive Core Characterization

Identification of sources for domestic critical minerals and rare earth elements (CM/REE) has been deemed essential for the energy transition by the United States Department of Energy (DOE). The U.S. DOE’s National Energy Technology Laboratory’s (NETL) Geomaterials Characterization Laboratory has performed non-destructive core characterizations on energy-relevant rock cores for the past decade. During this time, NETL has published over 36 technical reports and made the associated data publicly available. Much of this work focuses on unconventional shale gas, subsurface carbon storage systems, and carbon-ore. These efforts provide cm-scale petrophysical and elemental data, photographic documentation, detailed core descriptions, and computed tomography (CT) data for each well. This provides a first phase prospecting resource for CM/REE resources and can provide a map for pin-pointing intervals and lithologies for further development. Using historical core characterization data from 12 Marcellus wells from the western portion of the Appalachian Basin, this study builds an improved understanding of the chemostratigraphy of the basin. X-ray fluorescence (XRF) and CT images were used to determine lithologic intervals and potential ore bodies for further analysis, including benchtop digestion and inductively coupled plasma mass spectrometry (ICP-MS) to better understand the CM/REE enrichments.

Paronish, Thomas J.

Leveraging public AI tools to explore systems biology resources in mathematical modeling

Predictive mathematical modeling is an essential part of systems biology and is interconnected with information management. Systems biology information is often stored in specialized formats to facilitate data storage and analysis. These formats are not designed for easy human readability and thus require specialized software to visualize and interpret results. Therefore, comprehending modeling and underlying networks and pathways is contingent on mastering systems biology tools, which is particularly challenging for users with no or little background in data science or system biology. To address this challenge, we investigated the usage of public Artificial Intelligence (AI) tools in exploring systems biology resources in mathematical modeling. We tested public AI’s understanding of mathematics in models, related systems biology data, and the complexity of model structures. Our approach can enhance the accessibility of systems biology for non-system biologists and help them understand systems biology without a deep learning curve.

59 BASIC BIOLOGICAL SCIENCES

Label-based Virtual Directories In dCache

Traditional filesystems organize data in directories. These directories are typically a collection of files whose grouping is based on a single criterion, e.g., the starting date of an experiment, experiment name, beamline ID, measurement device, or instrument. However, each file in a directory can belong to several logical groups, such as a special event type, experiment condition, or a part of a selected dataset. dCache is a storage system developed to store large amounts of scientific data, used by many HEP and Photon Science experiments. With recent developments in dCache, we have introduced a concept of file tagging, which dynamically groups files with the same label into virtual directories. The file labels can be added, removed, renamed, and deleted through the admin interface or via REST API. The files in virtual directories are exposed through all protocols supported by dCache. This contribution will describe the details of the implementation for file tagging in dCache and present our future development plans on automatic metadata extractions, a feature that will significantly simplify data management. Additionally, we are exploring the future use of virtual directories as a way to translate scientific data catalogs into filesystem views for direct data analysis.

Sahakyan, Marina [DESY]

Toward a persistent event-streaming system for high-performance computing applications

High-performance computing (HPC) applications have traditionally relied on parallel file systems and file transfer services to manage data movement and storage. Alternative approaches have been proposed that use direct communications between application components, trading persistence and fault tolerance for speed. Event-driven architectures, as popularized in enterprise contexts, present a compelling middle ground, avoiding the performance cost and API constraints of parallel file systems while retaining persistence and offering impedance matching between application components. However, adapting streaming frameworks to HPC workloads requires addressing challenges unique to HPC systems. This paper investigates the potential for a streaming framework designed for HPC infrastructures and use cases. We introduce Mofka, a persistent event-streaming framework designed specifically for HPC environments. Mofka combines the capabilities of a traditional streaming service with optimizations tailored to the HPC context, such as support for massively multicore nodes, efficient scaling for large producer-consumer workflows, RDMA-enabled high-performance network communications, specialized network fabrics with multiple links per node, and efficient handling of large scientific data payloads. Built using the Mochi suite of HPC data service components, Mofka provides a lightweight, modular, and high-performance solution for persistent streaming in HPC systems. We present the architecture of Mofka and evaluate its performance against Kafka and Redpanda using benchmarks on diverse platforms, including Argonne's Polaris and Oak Ridge's Frontier supercomputers, showing up to 8× improvement in throughput in some scenarios. We then demonstrate its utility in several real-world applications: a tomographic reconstruction pipeline, a workflow for the discovery of metal-organic frameworks for carbon capture, and the instrumentation of Dask workflows for provenance tracking and performance analysis.

HPC

Optimizing Alabama’s CO 2 Storage in Shelby County (Project OASIS) - Deliverable Task 7: Risk Assessment

The “Optimizing Alabama’s CO 2 Storage in Shelby County: Project OASIS” CarbonSAFE Phase II Project seeks to build on regional data sets that demonstrate that the subsurface within Shelby County, Alabama has the potential to store commercial volumes of CO 2 safely, permanently, and economically. The primary target CO 2 Storage Complex is the deep Ketona Dolomite located within a 140 square mile area of the Valley and Ridge Region of Alabama. This deep saline reservoir is beneath a confining system encompassing at least 6,500 ft of shales and other low permeability sediments. Project OASIS drilled a deep stratigraphic test well to confirm the geological properties of the confining system and storage reservoir(s) within the Storage Complex. The geological data was incorporated into numerical models to establish the areal extent of the CO 2 plume and help design the storage site and its monitoring system. The project is managed by the Southern States Energy Board (SSEB), an interstate compact organization consisting of governors and state legislative leaders from sixteen southern states, Puerto Rico, and the U.S. Virgin Islands, as well as an appointee by the President of the United States. The organizational compact provides it with access to state government organizations and legislatures. The Board also maintains an Associate Members program comprised of energy resource companies, utilities, trade groups, academic R&D science and technology experts and energy consultants. Further, the SSEB staff is experienced in managing and coordinating complex energy and environmental programs, from research programs to full-scale design and demonstrations of new and innovative technologies. Project OASIS is a public-private partnership of six entities with multiple principal investigators (PIs). SSEB’s Lead PI and Co-PI are responsible for all aspects of project performance in accordance with the DOE-NETL Cooperative Agreement. SSEB has issued subgrants to Advanced Resources International, Inc., Alabama A&M University, Auburn University, Crescent Resource Innovation, and Oklahoma State University. Advanced Resources International, Inc., issued subgrants to Baker Hughes and Loudon Technical Services for field services

20 FOSSIL-FUELED POWER PLANTS

Assessing hydrogen supply chains: An integrated review of leakage and energy efficiency studies

This paper examines hydrogen leakage and efficiency across the supply chain for liquid, gaseous, and mixed hydrogen systems. These factors are crucial for assessing hydrogen's role in mitigating emissions and facilitating a clean energy transition. Drawing on a comprehensive review of existing literature and model-based analysis, the study compiles leakage rates and efficiency metrics at each stage of the supply chain: production, storage, transmission, distribution, and end-use. These data inform system scenarios that estimate the impact of leakage on overall performance and climate benefits. The analysis also identifies persistent data gaps, particularly for liquid and mixed system configurations, and outlines priorities for future research. A comparison of hydrogen system types shows that gaseous pathways generally achieve the highest efficiencies (28 %–39 %) and the lowest leakage rates (∼4.5 %) across the supply chain. Liquid hydrogen systems, while favorable for long-distance and high-volume transport due to their higher energy density, exhibit lower efficiency (∼28 %) and a greater leakage potential (∼12 %). Mixed systems, which combine gaseous and liquid elements (e.g., pipeline transmission followed by liquefaction and truck distribution), show compounded energy losses and moderate-to-high leakage rates (6.8 %–9.4 %), highlighting trade-offs associated with added system complexity. The study highlights opportunities for technological advancements, including optimizing liquefaction, enhancing insulation for storage and transportation, and refining refueling equipment. These improvements are crucial for maximizing the climate benefits of hydrogen. The results offer actionable insights for researchers, industry, and policymakers working to develop low-leakage, high-efficiency hydrogen infrastructure.

08 HYDROGEN

Performance Analysis of Data Processing in Distributed File Systems with Near Data Processing

In the era of big data, the escalating volume and velocity of data generation pose significant challenges in data processing. Traditional systems like Spark and Hadoop manage the increasing amount and velocity of data by improving data placement and processing speeds. However, they face inherent limitations due to the essential data movement required for processing. In this paper, we explore the Skyhook framework, a novel extension of the Ceph distributed system, which significantly reduces the need for data movement. We present an extensive case study using the Skyhook framework, applying it with the TPC-H and K-means clustering algorithms. More specifically, we leverage the TPC-H benchmark to distinguish between CPU-intensive and I/O-intensive tasks. We explore the integration of K-means clustering into SQL, coupled with a near-data processing system to offload the computational burden of the K-means clustering algorithm to storage nodes. We conduct a comprehensive performance evaluation of distributed data processing applications across three processing approaches: traditional layout (baseline), optimized layout, and near-data processing. Additionally, we introduce the use of the FIO tool to simulate real-world system workloads, enabling the measurement of performance metrics such as average latency and CPU utilization. Our research is a significant advance in understanding how to optimize data processing systems to meet the demands of the modern data landscape.

Hou, Shiyue

Integrated Energy-Water Data for Cross-Sector Resilience

This white paper focuses on the “energy-for-water” domain, addressing the urgent need for integrated, empirical data to support regional management, benchmarking, and research on improving efficiency and developing technologies for water and wastewater management systems. The costs and energy required for the supply, treatment, and distribution of water and wastewater lack a standard data collection mechanism and centralized database or storage infrastructure, limiting data-driven decision-making across interdependent infrastructure systems.

42 ENGINEERING