Search NASA⌕ Search

SEARCH · Search NASA

Results for “Data driven”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Data driven discovery and quantification of hyperspectral leaf reflectance phenotypes across a maize diversity panel

Abstract Estimates of plant traits derived from hyperspectral reflectance data have the potential to efficiently substitute for traits, which are time or labor intensive to manually score. Typical workflows for estimating plant traits from hyperspectral reflectance data employ supervised classification models that can require substantial ground truth datasets for training. We explore the potential of an unsupervised approach, autoencoders, to extract meaningful traits from plant hyperspectral reflectance data using measurements of the reflectance of 2151 individual wavelengths of light from the leaves of maize ( Zea mays ) plants harvested from 1658 field plots in a replicated field trial. A subset of autoencoder‐derived variables exhibited significant repeatability, indicating that a substantial proportion of the total variance in these variables was explained by difference between maize genotypes, while other autoencoder variables appear to capture variation resulting from changes in leaf reflectance between different batches of data collection. Several of the repeatable latent variables were significantly correlated with other traits scored from the same maize field experiment, including one autoencoder‐derived latent variable (LV8) that predicted plant chlorophyll content modestly better than a supervised model trained on the same data. In at least one case, genome‐wide association study hits for variation in autoencoder‐derived variables were proximal to genes with known or plausible links to leaf phenotypes expected to alter hyperspectral reflectance. In aggregate, these results suggest that an unsupervised, autoencoder‐based approach can identify meaningful and genetically controlled variation in high‐dimensional, high‐throughput phenotyping data and link identified variables back to known plant traits of interest.

Tross, Michael C.↗

Data-Driven Atomic Physics: Harnessing Machine Learning and High-Repetition-Rate Experiments for Laser-driven HED

High-energy-density plasma experiments are central to progress in atomic physics, fusion energy, and national security science, but they have traditionally been constrained by slow data collection and manual, time-intensive analysis. This project targeted that bottleneck by enabling high-repetition-rate experiments to produce and interpret much larger volumes of data quickly enough to guide experiments while they run.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Data-Driven Protection Software to classify fault locations by protective zone in distribution systems with high PV penetration

The software contains (a) the source codes to generate Point-on-Wave (PoW) transient data for any feeder model in Alternative Transient Program (ATP) format. Codes provide options to change different steady state settings, including the loading condition and PV capacity and transient state setting like faults type, location and initiation time (b) data post-processing source code to converted data from native format to COMTRADE, csv, HDF5 (c) Docker container to train CNN to classify fault locations by protective zone. The container takes dataset and other training parameters (sampling rate, training epochs, batch size etc) as input to train CNN. The container writes back the trained CNN model, training and testing metrics and plots to the local workstation

Ramesh, Meghana↗

Data-driven background model for the CUORE experiment

Here, we present the model we developed to reconstruct the CUORE radioactive background based on the analysis of an experimental exposure of 1038.4 kg yr. The data reconstruction relies on a simultaneous Bayesian fit applied to energy spectra over a broad energy range. The high granularity of the CUORE detector, together with the large exposure and extended stable operations, allow for an in-depth exploration of both spatial and time dependence of backgrounds. We achieve high sensitivity to both bulk and surface activities of the materials of the setup, detecting levels as low as 10 nBq kg −1 and 0.1 nBq cm −2 , respectively. We compare the contamination levels we extract from the background model with prior radio-assay data, which informs future background risk mitigation strategies. The results of this background model play a crucial role in constructing the background budget for the CUPID experiment as it will exploit the same CUORE infrastructure.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Resilience Through Data-Driven, Intelligent Designed Control: A Formal Methods Approach

The PNNL and GTRI team developed a strategy to integrate temporal logic rule specification for detection of cyber-intrusion in the source code and control algorithms of CPS using advanced cyber-data. The GTRI team utilized its capabilities in rule synthesis and temporal logic specifications for software assurance and verification to detect and predict impact of cyber-intrusions and malware in the computational and control algorithms of cyber-physical systems. The team also developed a testing and verification approach that could be used to validate the suggested approach against a realistic use-case CPS showcasing improvements in system impact prediction performance. Temporal logic offers a compact expression of events in absolute and relative time and has a formalized translation to state machines. As such, temporal logic rules can feasibly be synthesized to any system as a rule engine, with the process being formally verified to be correct. The goal here is to utilize temporal logic rules to detect cyber-attacks and manipulations in the computational algorithms and provide real-time software assurance and verification guarantees.

97 MATHEMATICS AND COMPUTING↗

Data-driven modeling of dynamic occupant thermostat override behavior for demand response applications

Buildings consume nearly 40% of global energy and produce similar emissions. Whiletechnological advances address efficiency, occupant behavior causes energy use variations up to 300% between identical buildings. This gap between predicted and actual building performance impacts building design, operations, and grid demand management programs. Through analyses of smart thermostat data from 1,400 single-occupant homes, the researchdemonstrates that occupants respond to 8°F thermostat setpoint changes within a median of 15 minutes, while 2°F changes trigger responses within a median of 30 minutes. This highlights an understudied temporal relationship between thermostat setbacks and response time of occupant behaviors. Models of such behavior dynamics are required to incorporate occupant impacts into building performance simulation. A key contribution of this dissertation is the Thermal Frustration Theory (TFT), which positsthat thermal discomfort driven behaviors are caused by the time-accumulation of discomfort, not simply a temperature deviation threshold or a delay from an initiating event. Using a dataset of 634 thermostats, each with 25+ manual setpoint changes, a comparative analysis of TFT and comfort zone and a delayed response theories demonstrated that personalized TFT models better predict when manual setpoint change occur. This was measured by the area under the curve statistical measure (AUC); all three models perform similarly by a Matthews Correlation Coefficient measure. Higher AUC performance is especially important for modeling occupant behavior in demand response programs where false negatives of rare occupant interactions could adversely affect grid stability. EnergyPlus based simulations were conducted with TFT-derived occupant models, demonstrating the ability to identify parameters of known TFT models from only data observable with smart thermostats, even under the presence of noise from routine overrides. Overall, the dissertation highlights that thermostat interactions are neither static,instantaneous, nor driven solely by the environment. Instead, temporal accumulation of discomfort and routine-based behavior play important roles. The methodology and results offer a pathway towards more accurate modeling of human-building interactions for policy assessment, building design, and demand response programs.

Sharma, Kunind [Northeastern University] (ORCID:00↗

New Data-Driven Constraints on the Sign of Gluon Polarization in the Proton

Recently, the possible existence of negative gluon helicity Δ g has been observed to be compatible with existing empirical constraints, including from jet production in polarized proton-proton collisions at the Relativistic Heavy Ion Collider, and lattice QCD data on polarized gluon Ioffe time distributions. We perform a new global analysis of polarized parton distributions in the proton with new constraints from the high- x region of deep-inelastic scattering (DIS). A dramatic reduction in the quality of the fit for the negative Δ g replicas compared to those with positive Δ g suggests that the negative Δ g solution cannot simultaneously account for high- x polarized DIS data along with lattice and polarized jet data. Published by the American Physical Society 2024

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Global Data-Driven Determination of Baryon Transition Form Factors

Hadronic resonances emerge from strong interactions encoding the dynamics of quarks and gluons. The structure of these resonances can be probed by virtual photons parametrized in transition form factors. Here, in this study, twelve N* and Δ transition form factors at the pole are extracted from data with the center-of-mass energy from πN threshold to 1.8 GeV, and the photon virtuality 0 ≤ Q 2 /GeV 2 ≤ 8. For the first time, these results are determined from a simultaneous analysis of more than one state, i.e., ~10 5 π⁢N, η⁢N, and K⁢Λ electroproduction data. In addition, about 5×10 4 data in the hadronic sector as well as photoproduction serve as boundary conditions. For the Δ⁡(1232) and N⁡(1440) states our results are in qualitative agreement with previous studies, while the transition form factors at the poles of some higher excited states are estimated for the first time. Realistic uncertainties are determined by further exploring the parameter space.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Improving Detector Systematic Uncertainties Through Data-Driven Machine Learning

Detector simulation in liquid argon time projection chambers (LArTPCs) is a constant challenge. In particular the modeling of electrons response on wires is highly nontrivial. However, new machine learning techniques exist which can be leveraged to ameliorate these concerns. We present a novel methodology to attempt to learn from cosmic muon data in the ICARUS detector how reconstructed wire signals are influenced by features of the hits such as location, particle direction, angle relative to the wire plane, etc. A model can then be generated which can apply the learned mapping to Monte Carlo events to create a more data-like simulation sample. By creating such a sample we expect to reduce the systematic uncertainties at ICARUS due to our detector modeling.

Hausner, Harry [Fermilab] (ORCID:0000000188932280)↗

Improving Detector Systematic Uncertainties Through Data-Driven Machine Learning

Detector simulation in liquid argon time projection chambers (LArTPCs) is a constant challenge. In particular the modeling of electrons response on wires is highly nontrivial. However, new machine learning techniques exist which can be leveraged to ameliorate these concerns. We present a novel methodology to attempt to learn from cosmic muon data in the ICARUS detector how reconstructed wire signals are influenced by features of the hits such as location, particle direction, angle relative to the wire plane, etc. A model can then be generated which can apply the learned mapping to Monte Carlo events to create a more data-like simulation sample. By creating such a sample we expect to reduce the systematic uncertainties at ICARUS due to our detector modeling.

Hausner, Harry [Fermilab] (ORCID:0000000188932280)↗

Data driven investigation to understand the influence of total solids on biological biogas upgrading

In situ biogas upgrading achieves CO 2 conversion to CH 4 via hydrogenotrophic methanogenesis; however, gas-liquid mass transfer constraints limit the upgrading performance. Recognizing that optimization studies often underrepresent the effects of total solids (TS) and organic loading rate (OLR), this study undertook a holistic, statistics driven assessment of operating conditions for in situ H 2 assisted biogas upgrading, centering the analysis on TS and OLR. A dataset of 31 studies was compiled and comprised 99 observations. A rigorous analytical framework was employed, combining data standardization, fixed- and random-effects (REML) weighted regressions with cluster-robust errors, stratified analyses, and machine learning. Mixed-effects meta regression indicated that TS was the main factor explaining differences of methane fraction (CH 4 %) when considering the between studies heterogeneity. Focusing on a near-stoichiometric subset (H 2 /CO 2 ≈ 4:1), TS remained significant. Stratified results showed a stronger negative relationship between TS and CH 4 % in UASB reactors than in CSTRs, with a negative effect under mesophilic conditions and no significant effect under thermophilic conditions. A Random Forest model corroborated the statistical findings, consistently ranking H 2 /CO 2 ratio, OLR, TS, and hydrogen injection rate (HIR) as the most influential predictors. These findings delineate trends across increasing TS levels, particularly between 1% and 10%, and provide preliminary insights for TS above 15% in in situ biogas upgrading. They further provide insights for the influence of TS by reactor type and temperature, thereby advancing the evidence base for implementing biological CO 2 conversion to CH 4 in practice.

In situ biogas upgrading↗

Evaluating system responses to electric vehicle charging infrastructure expansion through data-driven simulation

Understanding the system responses to electric vehicle (EV) charging infrastructure expansion, including vehicle charging needs, station utilization, and energy consumption, is critical for effective planning to meet growing charging demand without unnecessary resource investment. This study evaluates the system responses to EV charging infrastructure expansion, focusing on charging needs, station utilization, and energy consumption. Using trip data from the National Household Travel Survey and origin–destination patterns, we simulated trip chains in downtown Atlanta with 10 % EV penetration. We assessed 32 scenarios involving different charging port power levels and siting strategies. Furthermore, we found that higher-power ports were more sensitive to placement, with concentrated expansion boosting station utilization more than uniform expansion. Adding high-power ports did not always increase peak energy consumption; in some cases, a few 400 kW ports reduced overall consumption compared to 150 kW ports by enabling faster charging and higher vehicle turnover.

Electric vehicle↗

A Physical Model Enhanced Data Driven Method for High-Resolution Residential Load Profile Generation

Residential buildings account for significant energy consumption, creating opportunities to offer grid services. As electric utilities seek to implement effective system operation strategies, understanding residential energy consumption patterns becomes essential; However, the time intervals of load profiles measured by utilities' smart meters are typically from 15 minutes to 60 minutes. The low-resolution data make it hard to extract appliance-level load information, which is critical for providing grid services. This paper presents a load profile generator designed to produce synthetic load profiles for residential buildings that emphasizes the importance of accurate representations of realistic energy consumption patterns. The generator takes realistic low-resolution residential load measurements and weather data as inputs, producing 1-minute interval profiles that match the characteristics of the original profiles. Further, this generator can be used to populate load profiles in areas where actual measurements are limited to improve the ability of utilities to analyze their distribution systems. By providing more high-resolution residential building load profiles, this tool supports electric utilities to enhance their residential building load control strategies and improve overall grid stability.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

A Data-driven Phytotechnology Framework for Identification and Remediation of Leached-Metals-Contaminated Soil Near Coal Ash Impoundments

This project developed and evaluated advanced remote sensing and machine learning approaches to identify, monitor, and address environmental contamination associated with coal combustion residual (CCR) impoundments and landfills at coal-fired power plants. In Phase I, historical and multi-temporal Sentinel-2 satellite imagery, groundwater monitoring data, and environmental variables were integrated to detect vegetation stress potentially caused by toxic metal leaching from coal ash disposal sites. Multiple vegetation and biophysical indices were analyzed to determine their effectiveness in identifying abnormal vegetation growth patterns linked to contamination. Results from case studies conducted at coal ash–impacted power plant sites in North Carolina and Virginia demonstrated that satellite-based vegetation monitoring can serve as an effective early indicator of environmental stress associated with metals such as arsenic, cadmium, cobalt, lead, lithium, radium, and thallium.

01 COAL, LIGNITE, AND PEAT↗

Dynamic data-driven multiscale modeling for predicting the degradation of a 316L stainless steel nuclear cladding material

Here, we have developed a long short-term memory stacked ensemble (LSTM-SE) surrogate modeling approach that can provide rapid predictions of microstructural evolution and the resultant mechanical properties of American Iron and Steel Institute (AISI) 316L series stainless steel (316LSS) fuel cladding under conditions of varying temperature and radiation dose rate. To acquire training data, we developed and implemented a kinetic Monte Carlo (KMC) model to simulate precipitation kinetics of M 23 C 6 , γ', and G phases within SS316L cladding. Experimentally reported precipitation kinetics of SS316L in literature were linked to the kinetic parameters of the simulated precipitation in our KMC model. The model was then used to simulate microstructure evolution under synthetically generated treatments of varying temperature and radiation dose rate, for periods of up to 3000 hours. Changes in volume fraction, number density, and particle size of precipitates were recorded, and particle area fractions were correlated using statistical methods to develop the surrogate model. Simultaneously, the mechanical properties of the simulated microstructures were evaluated using microstructure-based finite element method (FEM) analysis to determine the elastic modulus, yield stress, ultimate tensile strength, and elongation to failure of the aged microstructures. Using this approach, our surrogate model can predict precipitation behavior within 0.25% volume fraction and mechanical properties within 6% relative error from the values predicted by the KMC and FEM models using 50 training simulations as input. The trained recurrent neural network-based model can return estimations of precipitation kinetics and mechanical properties ~1000 times faster than the physics-based codes. This work demonstrates, as a proof of concept, that reactor material service lifetimes under variable service conditions can be predicted for a statistics-based model from a practicably obtainable dataset.

36 MATERIALS SCIENCE↗

Leading Energy Analysis: Data-Driven Insights That Power Our Energy Future

Energy analysts at the National Renewable Energy Laboratory (NREL) use a broad set of expertise and cutting-edge tools to capture the complexities of our interconnected energy system. Decision-makers rely on these insights to drive cost savings, enhance grid reliability, support long-term innovation, and bolster America's energy workforce and global competitiveness. This fact sheet highlights some of NREL's high-impact analyses, data, and tools, largely focusing on power grid analysis.

data↗

A Data-Driven Approach to Real-World Degradation of Backsheets

The objectives of this project are as follows: • The population behavior of fielded modules in various conditions of use • Predictions of materials in specific climatic zones • Understanding of a module’s local environment in the field on its degradation It aims to understand how and what backsheet materials of photovoltaic modules degrade in the real world field. • Field Survey Protocol This project is started from the protocol, because all the data, information, domain knowledge are from experience of the real world field surveys, which is based on the protocol. The Protocol is explored from the experience of the field survey observations. With the increasing of the field surveys, it is refined for three versions, which are Task 1.0 (Section 3.1), Task 6.0 (Section 3.6), Task 10.0 (Section 3.10) respectively. It includes a document and a training video, which is able to direct other teams to follow the same procedure with the sites surveyed during this project. The documents include the detailed information but is not limited to the terminology definition, instruments SOP, preparation items for the surveys, form for the data collections, the order of the information collection. The final version of the protocol can be found at Appendix A, see Section 5. Additional, It can also be found at Open Science Framework (OSF), see Section 3.18 for detail. • Written Waiver Request Before staring the field surveys, the request of the waiver for international surveys is completed, because the limitation of climate zone in the United States, see Appendix B in Section 6 for the request documents. Unfortunately, only 1 international site from Taiwan, China can be finished, due to the COVID-19. • Field Survey According to the protocol we built in Section 3.1, 3.6 and 3.10, 41 sites have been surveyed across seven different climate zones (Cfa, Csa, Csb, BSk, Dfa, Dfb, Am). A variety of materials, including Polyethylene Naphthalate (PEN), Polyethylene Terephthalate (PET), Polyvinyl Fluoride (PVF), Polyvinylidene Fluoride (PVDF), Acrylic PVDF, Fluoroethylene Vinyl Ether (FEVE), and Glass, were identified. These sites are located in various states including California, South Carolina, New Mexico, Maryland, Ohio, Tennessee, Florida, Massachusetts, Illinois, Minnesota, Oregon, Colorado, and Taiwan, Republic of China. The ages of the sites ranged from 2 - 38 years in service and the field size varied from 1 MW - 25 MW. All requirements for the modeling have been satisfied. Some observations like ’Edge Effect’ for the rows and Junction box heating will also be a useful knowledge to build the model. Section 3.7 provides detailed information on the sites visited during this reporting period.

14 SOLAR ENERGY↗

Towards a data-driven model of hadronization using normalizing flows

We introduce a model of hadronization based on invertible neural networks that faithfully reproduces a simplified version of the Lund string model for meson hadronization. Additionally, we introduce a new training method for normalizing flows, termed MAGIC, that improves the agreement between simulated and experimental distributions of high-level (macroscopic) observables by adjusting single-emission (microscopic) dynamics. Our results constitute an important step toward realizing a machine-learning based model of hadronization that utilizes experimental data during training. Finally, we demonstrate how a Bayesian extension to this normalizing-flow architecture can be used to provide analysis of statistical and modeling uncertainties on the generated observable distributions.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗