Search NASASearch

SEARCH · Search NASA

Results for “Data driven”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

Data-driven analysis of dipole strength functions using artificial neural networks

Here, we present a data-driven analysis of dipole strength functions across the nuclear chart, employing an artificial neural network to model nuclear dipole responses. We train the network on a dataset of experimentally measured dipole strength functions for 216 different nuclei. To assess its predictive capability, we test the trained model on an additional set of 10 new nuclei, where experimental data exist. We demonstrate that the artificial neural network not only accurately reproduces known data but also identifies potential inconsistencies in experimental datasets, indicating which results may warrant further review or possible rejection. For nuclei where experimental data are sparse or unavailable, the network confirms theoretical calculations, reinforcing its utility as a predictive tool in nuclear physics. Finally, utilizing the predicted electric dipole polarizability, we extract the value of the symmetry energy at saturation density and find it consistent with results from the literature.

artificial neural networks

Data-driven reduced-order models for port-Hamiltonian systems with operator inference

Hamiltonian operator inference has been developed in Sharma et al. (2022) to learn structure-preserving reduced-order models (ROMs) for Hamiltonian systems. The method constructs a low-dimensional model using only data and knowledge of the functional form of the Hamiltonian. The resulting ROMs preserve the intrinsic structure of the system, ensuring that the mechanical and physical properties of the system are maintained. In this work, we extend this approach to port-Hamiltonian systems, which generalize Hamiltonian systems by including energy dissipation, external input, and output. Based on snapshots of the system’s state and output, together with the information about the functional form of the Hamiltonian, reduced operators are inferred through optimization and are then used to construct data-driven ROMs. To further alleviate the complexity of evaluating nonlinear terms in the ROMs, a hyper-reduction method via discrete empirical interpolation is applied. Accordingly, we derive error estimates for the ROM approximations of the state and output. Lastly, we demonstrate the structure preservation, as well as the accuracy of the proposed port-Hamiltonian operator inference framework, through numerical experiments on a linear mass–spring-damper problem and a nonlinear Toda lattice problem.

97 MATHEMATICS AND COMPUTING

A Data-Driven Method for Modeling Creep-Fatigue Stress- Strain Behavior Using Neural ODEs

In this paper, we introduce a data-driven machine learning approach for modeling one-dimensional stress–strain behavior under cyclic loading, utilizing experimental data from the nickel-based Alloy 617. The study employs uniaxial creep–fatigue test data acquired under various loading histories and compares two distinct neural network-based ODE models. The first model, known as the black-box model, comprehensively describes the strain–stress relationship using a Neural ODE equation. To interpret this black-box model, we apply the Sparse Identification of Nonlinear Dynamical Systems (SINDy) technique, transforming the black-box model into an equation-based model using symbolic regression. The second model, the Neural flow rule model, incorporates Hooke’s Law for the linear elastic component, with the nonlinear part characterized by a Neural ODE. Both models are trained with experimental data to accurately reflect the observed stress–strain behavior. We conduct a detailed comparison with the standard Chaboche model, which includes three back stresses. Our results demonstrate that the neural network-based ODE models precisely capture the experimental creep–fatigue mechanical behavior, exceeding the standard Chaboche model’s accuracy. Furthermore, an interpretable model derived from the black-box neural ODE model through symbolic regression achieves accuracy comparable to the Chaboche model, enhancing its interpretability. The results highlight the potential of neural network-based ODE models to depict complex creep–fatigue behavior, eliminating the necessity for experts to define a specific, material-focused model form.

creep-fatigue

Data-Driven Energy Resilience Assessment and Enhancement in Urban Communities: A Case Study in Detroit

This paper presents a data-driven framework for assessing and enhancing energy resilience in urban communities. The resilience assessment is based on two datasets: 1) annual aggregated power outage data and 2) 15-minute interval outage data. High-impact, low-probability (HILP) events are identified within these datasets to evaluate community resilience under extreme conditions. To enhance resilience, an optimization framework utilizing mixed integer linear programming is developed to determine the optimal sizing and placement of solar photovoltaic (PV) systems and battery energy storage systems (BESS). This method offers a cost-effective and practical solution for improving energy resilience in vulnerable communities. Furthermore, a case study of the City of Detroit in Michigan demonstrates the effectiveness of the framework through simulation and validation.

Energy resilience assessment

Experimentation in Exploring Photovoltaic Inverter Dynamics Under Different Irradiance Levels Through a Data-Driven Approach

As conventional direct connections of synchronous generators are being phased out, inverter-based resources (IBRs) with grid support functions are increasingly being integrated into power systems. This transition requires the development of accurate dynamic models for IBRs to predict how power systems will adapt to varying levels of IBRs penetration, establish grid code requirements, and ensure compliance. Here, this study introduces an active probing signal-based data-driven modeling technique to accurately derive the dynamics model of a smart photovoltaic inverter operating in Volt-Watt and Freq-Watt modes, in compliance with the IEEE 1547–2018 standard. The paper focuses on investigating how the dynamics of the PV inverter model respond to fluctuations in solar irradiance, utilizing real-time digital simulator experimentation. The experimental analysis demonstrates that the amplitude of dynamics fluctuates with changes in irradiance across both operational modes and confirms the active power’s dependence on irradiance levels. Furthermore, the nature of inverter dynamics varies distinctly between the different modes of activation. Critically, our findings indicate that dynamic models require DC-gain adjustments to accommodate contrasting irradiance levels, highlighting a negative gradient linear relationship between the DC-gain of each model and the irradiance.

14 SOLAR ENERGY

A new data-driven map predicts substantial undocumented peatland areas in Amazonia

Tropical peatlands are among the most carbon-dense terrestrial ecosystems yet recorded. Collectively, they comprise a large but highly uncertain reservoir of the global carbon cycle, with wide-ranging estimates of their global area (441 025–1700 000 km 2 ) and below-ground carbon storage (105–288 Pg C). Substantial gaps remain in our understanding of peatland distribution in some key regions, including most of tropical South America. Here we compile 2413 ground reference points in and around Amazonian peatlands and use them alongside a stack of remote sensing products in a random forest model to generate the first field-data-driven model of peatland distribution across the Amazon basin. Our model predicts a total Amazonian peatland extent of 251 015 km 2 (95th percentile confidence interval: 128 671–373 359), greater than that of the Congo basin, but around 30% smaller than a recent model-derived estimate of peatland area across Amazonia. The model performs relatively well against point observations but spatial gaps in the ground reference dataset mean that model uncertainty remains high, particularly in parts of Brazil and Bolivia. For example, we predict significant peatland areas in northern Peru with relatively high confidence, while peatland areas in the Rio Negro basin and adjacent south-western Orinoco basin which have previously been predicted to hold Campinarana or white sand forests, are predicted with greater uncertainty. Similarly, we predict large areas of peatlands in Bolivia, surprisingly given the strong climatic seasonality found over most of the country. Very little field data exists with which to quantitatively assess the accuracy of our map in these regions. Data gaps such as these should be a high priority for new field sampling. This new map can facilitate future research into the vulnerability of peatlands to climate change and anthropogenic impacts, which is likely to vary spatially across the Amazon basin.

54 ENVIRONMENTAL SCIENCES

Next-Generation Energy Technologies for Connected and Automated On-Road Vehicles (NEXTCAR) - Predictive Data-Driven Vehicle Dynamics and Powertrain Control: from ECU to the Cloud (Final Scientific/Technical Report)

This project developed and demonstrated a predictive, data-driven vehicle control system designed to improve energy efficiency and driving performance. The team created intelligent self-driving car technology that optimizes fuel and electricity use by proactively planning vehicle actions. By combining Level 4 autonomous driving capabilities with vehicle-to-everything (V2X) connectivity, the system enables vehicles to adjust speed and change lanes in response to traffic signals, surrounding vehicles, and road conditions, reducing unnecessary stops and delays. In testing, the system improved vehicle fuel economy by more than 30% and reduced travel time by approximately 10%, compared to a conventional adaptive cruise control baseline. These results demonstrate the technical effectiveness of using predictive, V2X-enabled strategies, such as traffic light timing and surrounding traffic awareness, to inform real-time vehicle powertrain control and driving behavior. Additionally, a supporting cloud platform was developed to provide dispatch and route recommendations as well as to log vehicle data, demonstrating the economic feasibility of this approach at the fleet level. By optimizing dispatching and routing operations, this technology enables electric fleet operators to use their vehicles more efficiently and reduce reliance on diesel backups, lowering both operating costs and energy consumption. Overall, this project’s technology advances the future of clean, energy-efficient transportation, enabling vehicles and fleets to reduce energy waste, cut costs, and lower emissions through intelligent automation and connectivity.

33 ADVANCED PROPULSION SYSTEMS

Data-Driven Modeling and Control of Systems with Plasma-Surface Interactions (Final Technical Report)

This final technical report summarizes the activities and accomplishments in the period from February 2023 thru January 2026. The objective of the proposed research is to investigate the physical mechanisms and processes underlying the formation of structures and patterns in systems with plasma-surface interactions. In the past decades, there have been extensive studies on the interaction of glow discharges, dielectric barrier discharges, and arc discharges with confining or intervening surfaces. The advancement of the understanding of these phenomena is not only of fundamental scientific interest and relevance to the knowledge of the plasma state, but also with profound implications in various technological applications. The research will integrate theoretical, computational, and experimental work within an innovative framework of data assimilation, i.e., optimally combining model predictions with measurements. The scientific merit of this research has three aspects. Firstly, it extends the studies of plasma-surface interactions to systems with insulator surfaces and multi-layer systems, while existing studies are predominantly on electrode surfaces. Secondly, it expects to develop a novel data-driven modeling approach based on data assimilation to enhance the predictive and control capabilities, which could make transformative contributions to basic plasma research. Thirdly, it will shed new light on outstanding problems related to formation of patterns interfacing plasmas. This project also aims to launch an education and outreach initiative at Texas A&M University-Kingsville, a non-R1, minority-serving institution in South Texas. The initiative is structured as a four-tier pyramid. Tier one will be a webinar series for culture and capacity building to inform broader audience in the region about the research fields of plasma science and engineering. Tier two will be the creation and offering of an upper-level undergraduate course on introductory plasma physics, which will help with the recruitment for the upper tiers. On tier three, we will engage and mentor senior design students to conduct work toward the research goal of this project. There will also be a certificate program on general plasma science for undergrad and graduate students, part of which will be lab training at Princeton University. Tier four will be the supervision and mentoring of Ph.D. students. Therefore, this project will systematically expand the talent pipeline, broaden participation from communities historically and geographically underrepresented in DOE SC research portfolio, significantly improve the research and education capacity at the PI’s institution, and contribute to developing a diverse workforce in plasma science and engineering.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Vertical instability forecasting and controllability assessment of multi-device tokamak plasmas in DECAF with data-driven optimization

Abstract Reliable vertical position control will be an essential element of any future tokamak-based fusion power plant in order to reduce disruptions and maximize performance. We investigate methods to improve vertical controllability boundary determination in plasma operational space and demonstrate a data-driven approach based on direct pseudoinversion of operational space data that is rigorously quantitative, applicable in real-time plasma control systems, and physically intuitive to interpret. Applied to historical shot data from entire run campaigns on the MAST-U, KSTAR, and NSTX tokamaks, this approach, implemented in DECAF, improves vertical displacement event identification accuracy to 98.9%–100%. Further, we explore the application of a physics-based vertical stability metric as an early warning forecaster for vertical displacement events. The development of a linear surrogate model for the plasma current density profile, with a coefficient of determination of 0.992 on the training dataset, enables potential employment of this forecaster in real-time. The application of this approach on historical data from the MAST-U MU02 campaign yields a forecaster with 62.6% accuracy, indicating promise for this method when further refined and potentially coupled with other stability metrics.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Accelerated data-driven materials science with the Materials Project

The Materials Project was launched formally in 2011 to drive materials discovery forwards through high-throughput computation and open data. More than a decade later, the Materials Project has become an indispensable tool used by more than 600,000 materials researchers around the world. This Perspective describes how the Materials Project, as a data platform and a software ecosystem, has helped to shape research in data-driven materials science. We cover how sustainable software and computational methods have accelerated materials design while becoming more open source and collaborative in nature. Next, we present cases where the Materials Project was used to understand and discover functional materials. We then describe our efforts to meet the needs of an expanding user base, through technical infrastructure updates ranging from data architecture and cloud resources to interactive web applications. Finally, we discuss opportunities to better aid the research community, with the vision that more accessible and easy-to-understand materials data will result in democratized materials knowledge and an increasingly collaborative community.

Horton, Matthew K

End-to-end microgrid protection using distributed data-driven methods

This paper introduces an end-to-end microgrid protection framework that offers real-time system monitoring, fault-related decision making, and circuit breaker control. This is achieved through the design of distributed data-driven techniques based on the support vector machine method, where each relay is responsible for distributed data collection, fault detection, fault localization, and fault isolation. Local communication is established among neighboring relays, fostering cooperative fault localization and isolation. This decentralized design not only reduces the computational and communication requirements but also enables the adaptability of each relay under varying operational dynamics. The proposed end-to-end protection framework was validated using MATLAB/Simulink simulations on a 100% renewable microgrid, achieving an accuracy of 93.1% with response time of 0.0523 s, in protecting against a range of fault scenarios that are characterized by various types, locations, impedances, load conditions, photovoltaic power levels, and microgrid operating modes.

24 POWER TRANSMISSION AND DISTRIBUTION

Data-Driven Method for Groundwater-Level Mapping and Monitoring-Well Network Optimization at Hanford

This report summarizes the initial results and outcomes of a physics-informed, data-driven groundwater level (GWL) mapping capability for the Hanford Site. GWL mapping at Hanford is typically conducted annually and requires a significant amount of computational and expert resources, and it does not allow assessment of the informational value of specific monitoring wells. The proposed method produces spatially and temporally resolved fields consistent with sparse, irregularly sampled, and nonuniformly distributed well measurements. Implemented successfully, this capability will allow rapid mapping of groundwater levels and provide an opportunity to optimize monitoring activities (both location and sampling frequency) based on data information value evaluation. The approach integrates a diffusion-based generative model – trained on MODFLOW simulation data from the Plateau-to-River (P2R) model – with score-based data assimilation (SDA), allowing observation-conditioned mapping without retraining for each monitoring-network layout.

54 ENVIRONMENTAL SCIENCES

Data-Driven Performance Optimization of Gamma Spectrometers With Many Channels

In gamma spectrometers with variable spectroscopic performance across many channels (e.g., many pixels or voxels), a tradeoff exists between including data from successively worse-performing readout channels and increasing efficiency. Brute-force calculation of the optimal set of included channels is exponentially infeasible as the number of channels grows, and approximate methods are required. In this work, we present a data-driven framework for attempting to find near-optimal sets of included detector channels. The framework leverages non-negative matrix factorization (NMF) to learn the behavior of gamma spectra across the detector and clusters similarly-performing detector channels together. Performance comparisons are then made between spectra with channel clusters removed, which is more feasible than brute force. The framework is general and can be applied to arbitrary, user-defined performance metrics depending on the application. We apply this framework to optimizing gamma spectra measured by H3D M400 CdZnTe (CZT) spectrometers, which exhibit variable performance across their crystal volumes. In particular, we show several examples optimizing various performance metrics for uranium and plutonium gamma spectra in non-destructive assay (NDA) for nuclear safeguards, and explore trends in performance versus parameters such as clustering algorithm type. We also compare the NMF + clustering pipeline to several non-machine-learning (ML) algorithms, including several greedy algorithms. Although, we find that the NMF + clustering pipeline tends to find the best-performing set of detector voxels, significantly improving over the unoptimized spectra, but that a greedy accumulation of spectra segmented by detector depth can, in some cases, give similar performance improvements in much less computation time.

Energy resolution

Data Quality Assessment Process for Real-Time Data-Driven Traffic Microsimulation of Smart Corridor

Smart corridor digital twins are often created for the development and evaluation of emerging intelligent transportation systems and Connected and Autonomous Vehicle (CAV) technologies. However, limited guidance exists for data quality assessment for digital twin development. To address this, this paper discusses the data quality assessment utilized to develop data-driven real-time microscopic simulation models, i.e., digital twins, for two separate smart corridors: the North Avenue Smart Corridor in Atlanta, GA, and the Martin Luther King Smart Corridor in Chattanooga, Tennessee. This paper provides a summary of the author’s investigations of data requirements and data characteristics for the given smart corridor digital twin development efforts. With a focus on data, this summary includes a description of the data investigation process, key data issues observed, and strategies to address observed issues. Discussion is provided to help expand the lessons from these studies to other digital twin development efforts.

Saroj, Abhilasha [ORNL] (ORCID:0000000191178063)

Evaluation of data driven low-rank matrix factorization for accelerated solutions of the Vlasov equation

Low-rank methods have shown success in accelerating simulations of a collisionless plasma described by the Vlasov equation, but still rely on computationally costly linear algebra every time step. We propose a data-driven factorization method using artificial neural networks, specifically with convolutional layer architecture, that trains on existing simulation data. At inference time, the model outputs a low-rank decomposition of the distribution field of the charged particles, and we demonstrate that this step is faster than the standard linear algebra technique. Numerical experiments show that the method achieves comparable reconstruction accuracy for interpolation tasks, generalizing to unseen test data in a manner beyond just memorizing training data; patterns in factorization also inherently followed the same numerical trend as those within algebraic methods (e.g., truncated singular-value decomposition). However, when training on the first 70% of a time-series data and testing on the remaining 30%, the method fails to meaningfully extrapolate. Despite this limiting result, the technique may have benefits for simulations in a statistical steady-state or otherwise showing temporal stability. These results suggest that while the model offers a computationally efficient alternative for datasets with temporal stability, its current formulation is best suited for interpolation rather than for predicting future states in time-evolving systems. This study thus lays the groundwork for further refinement of neural network-based approaches to low-rank matrix factorization in high-dimensional plasma simulations.

97 MATHEMATICS AND COMPUTING

Forecasting Solar Photovoltaic Power Production: A Comprehensive Review and Innovative Data-Driven Modeling Framework

The intermittent and stochastic nature of Renewable Energy Sources (RESs) necessitates accurate power production prediction for effective scheduling and grid management. This paper presents a comprehensive review conducted with reference to a pioneering, comprehensive, and data-driven framework proposed for solar Photovoltaic (PV) power generation prediction. The systematic and integrating framework comprises three main phases carried out by seven main comprehensive modules for addressing numerous practical difficulties of the prediction task: phase I handles the aspects related to data acquisition (module 1) and manipulation (module 2) in preparation for the development of the prediction scheme; phase II tackles the aspects associated with the development of the prediction model (module 3) and the assessment of its accuracy (module 4), including the quantification of the uncertainty (module 5); and phase III evolves towards enhancing the prediction accuracy by incorporating aspects of context change detection (module 6) and incremental learning when new data become available (module 7). This framework adeptly addresses all facets of solar PV power production prediction, bridging existing gaps and offering a comprehensive solution to inherent challenges. By seamlessly integrating these elements, our approach stands as a robust and versatile tool for enhancing the precision of solar PV power prediction in real-world applications.

14 SOLAR ENERGY