Search NASA⌕ Search

SEARCH · Search NASA

Results for “generative machine learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 415 records · Page 23

Unsupervised atomic data mining via multi-kernel graph autoencoders for machine learning force fields

Constructing a chemically diverse dataset while avoiding sampling bias is critical to training efficient and generalizable force fields. However, in computational chemistry and materials science, many common dataset generation techniques are prone to oversampling regions of the potential energy surface. Furthermore, these regions can be difficult to identify and isolate from each other or may not align well with human intuition, making it challenging to systematically remove bias in the dataset. While traditional clustering and pruning (down-sampling) approaches can be useful for this, they can often lead to information loss or a failure to properly identify distinct regions of the potential energy surface due to difficulties associated with the high dimensionality of atomic descriptors. In this work, we introduce the Multi-kernel Edge Attention-based Graph Autoencoder (MEAGraph) model, an unsupervised approach for analyzing atomic datasets. MEAGraph combines multiple linear kernel transformations with attention-based message passing to capture geometric sensitivity and enable effective dataset pruning without relying on labels or extensive training. Demonstrated applications on niobium, tantalum, and iron datasets show that MEAGraph efficiently groups similar atomic environments, allowing for the use of basic pruning techniques for removing sampling bias. This approach provides an effective method for representation learning and clustering that can be used for data analysis, outlier detection, and dataset optimization.

Materials science↗

Protection System Validation with Machine Learning Anomaly Classification

A poster for the Early Career Poster Session. Power system protection devices have transitioned over the past few decades from mechanical to analog devices, then to solid state and finally digital. Relays and their associated critical network of equipment have significantly increased in complexity. Even internally, relays have gained significant intricacy, with relatively simple overcurrent or differential functions now being assisted by a myriad of other functions. This is necessary as the grid becomes more complex, but it brings increased difficulty in monitoring and upkeep. Misoperation caused by improper relay settings or malicious actions is a constant challenge faced by all utilities. These improper settings can be difficult to identify and may require exhaustive post-mortem analysis, typically after a major outage event has already occurred. A mechanism is needed for monitoring the behavior of protection systems to validate that they act and perform as expected. This work presents a concept for a machine learning (ML) system capable of validating the performance of protection systems by classifying anomalous events and characterizing protection system responses based solely on available current and voltage measurements. As a first step in its development, an experimental dataset is generated, and a random forest model is implemented with high accuracy in distinguishing four power system scenarios.

24 - POWER TRANSMISSION AND DISTRIBUTION↗

High-throughput validation of phase formability and simulation accuracy of Cantor alloys

High-throughput methods enable accelerated discovery of novel materials in complex systems such as high-entropy alloys, which exhibit intricate phase stability across vast compositional spaces. Computational approaches, including Density Functional Theory (DFT) and calculation of phase diagrams (CALPHAD), facilitate screening of phase formability as a function of composition and temperature. However, the integration of computational predictions with experimental validation remains challenging in high-throughput studies. In this work, we introduce a quantitative confidence metric to assess the agreement between predictions and experimental observations, providing a quantitative measure of the confidence of machine learning models trained on either DFT or CALPHAD input in accounting for experimental evidence. The experimental dataset was generated via high-throughput in-situ synchrotron X-ray diffraction on compositionally varied FeNiMnCr alloy libraries, heated from room temperature to ~1000 °C. Agreement between the observed and predicted phases was evaluated using either temperature-independent phase classification or a model that incorporates a temperature-dependent probability of phase formation. This integrated approach demonstrates where strong overall agreement between computation and experiment exists, while also identifying key discrepancies, particularly in FCC/BCC predictions at Mn-rich regions to inform future model refinement.

36 - MATERIALS SCIENCE↗

Anion-derived contact ion pairing as a unifying principle for electrolyte design

Enabling new electrochemical technologies requires systems that can operate under ever-more demanding conditions, and progress in energy storage applications reveals tantalizing opportunities to reimagine electrolyte design for performance at extreme potentials. Here, a common thread among these innovations is the formation of significant populations of contact ion pairs (CIPs) in the electrolyte, regardless of the specific cation chemistry or solvent system. The examples summarized in this review suggest that a set of general electrolyte design rules likely exists, where the purposeful selection of anion chemistry can yield CIP structures with tunable control over reaction thermodynamics, kinetics, and interphase chemistry. Identifying the relevant descriptors for high-performance, anion-derived CIP structures can be achieved utilizing a combined experimental and computational approach, aided by machine learning and artificial intelligence, to more rapidly survey the vast combinatorial space available and to enable a new generation of electrolytes for decarbonized electrochemical processes at scale.

electrochemistry↗

Validation of new and existing methods for time-domain simulations of turbulence and loads

We seek to obtain a second-by-second match between the simulated and measured structural loads of a utility-scale wind turbine. To obtain the one-to-one load simulations, we start with the furthest upstream component of the modeling chain: the turbulent inflow. We consider new and existing methods to generate constrained-turbulence flow fields. The new method is based on large-eddy simulations (LES) and machine learning (ML). The existing methods include Kaimal-based TurbSim and the superstatistical wind field model. The inflow measurements used to constrain these simulations are obtained with a nacelle-mounted scanning lidar. We compare the flow fields for the different inflow simulation approaches and validate their associated load predictions against measurements collected in the Rotor Aero-dynamics, Aeroelastics, and Wake (RAAW) field campaign. We find that the rotor-position control developed for this study is key in enabling the time match between measurements and simulations. When this control approach is used, the load simulation performance tracks with the inflow simulation fidelity, with LES+ML yielding errors ≤ 4% for the damage-equivalent loads of flapwise bending moment, and tower fore-aft bending moments.

17 WIND ENERGY↗

Physics-Informed Neural Network (PINN) Prediction of Mixed Mass-Heat-Crystallization Limited Methane Hydrate Formation and Dissociation in Micro-Confinement

The creation and use of Physics-Informed Neural Networks (PINNs) for simulating the dynamics of methane hydrate formation and dissociation will be presented. The PINN framework's main benefit is its capacity to impose physical consistency with only a partial comprehension of the governing equations. This makes the algorithm especially useful for systems with little experimental evidence or a lack of theoretical knowledge. A strong basis for forecasting methane hydrate behavior over the verified operating ranges of 30.0-80.9 bar pressure and 1.0-4.0 K sub-cooling conditions is provided by the combination of conductive heat transfer equations and mixed mass-transfer–crystallization kinetics. PINNs were more accurate at predicting the mixed mass-heat-crystallization limited kinetics than conventional Artificial Neural Networks (ANNs), demonstrating remarkable predictive accuracy for methane hydrate production over the ANN model. The efficiency of incorporating physical limitations from first principles into machine learning frameworks for methane hydrate crystallizations is reinforced by these findings. For hydrate-related applications in energy generation, carbon sequestration, and climate modelling, our study establishes PINNs as a computational tool that is both scalable and efficient. The proven capacity to close the gap between conventional physics-based simulations and solely data-driven models creates new opportunities for expedited hydrate research and practical applications.

Hartman, Ryan L [NYU Tandon School of Engineering]↗

HGQ: High Granularity Quantization for Real-time Neural Networks on FPGAs

Neural networks with sub-microsecond inference latency are required by many critical applications. Targeting such applications deployed on FPGAs, we present High Granularity Quantization (HGQ), a quantization-aware training framework that optimizes parameter bit-widths through gradient descent. Unlike conventional methods, HGQ determines the optimal bit-width for each parameter independently, making it suitable for hardware platforms supporting heterogeneous arbitrary precision arithmetic. In our experiments, HGQ shows superior performance compared to existing network compression methods, achieving orders of magnitude reduction in resource consumption and latency while maintaining the accuracy on several benchmark tasks. These improvements enable the deployment of complex models previously infeasible due to resource or latency constraints. HGQ is open-source and is used for developing next-generation trigger systems at the CERN ATLAS and CMS experiments for particle physics, enabling the use of advanced machine learning models for real-time data selection with sub-microsecond latency.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Autonomy Verification & Validation Roadmap and Vision 2045

Advanced capabilities planned for the next generation of autonomous and increasingly autonomous air vehicles will include non-traditional components based on artificial intelligence, machine learning, and complex optimization and planning algorithms. These complex components will be used to provide enhanced safety and high-level decision-making functions. However, there are serious barriers to the deployment of autonomous aircraft in the National Airspace System (NAS). Current civil aviation certification processes are based on the concept that the correct behavior of a system or a component must be completely specified and verified prior to operation. This report from the Autonomy Verification and Validation (V&V) Roadmap and Vision 2045 project presents the most recent effort to build a comprehensive list of verification challenges and needs for autonomous aircraft, a roadmap to meet those autonomy V&V needs, the services they can enable, and point to the certification gaps they fill. To accomplish these goals, we assembled a team of world-class researchers from the aerospace industry (Boeing, Collins Aerospace, and GeneralElectric) and academia (University of Michigan, University of Texas, and Massachusetts Institute of Technology) with deep expertise in autonomy, aerospace systems, and assurance of Artificial Intelligence/machine learning systems.

Software Assurance↗

Utilizing Convolutional Neural Networks for Global Seagrass Habitat Mapping

Convolutional neural networks (CNNs) are becoming an increasingly prevalent machine learning algorithm due to their high accuracy and lack of reliance on heuristic processes. One of the major drawbacks of convolutional neural networks is their reliance on large amounts of training data in order to generate sensible results. This talk will cover how our team has utilized the strengths and overcome the weaknesses of convolutional neural networks as they apply to seagrass habitat mapping. We will share our technical CNN results over time, detail the requirements and challenges that our team overcame and explore how other teams can better incorporate a stronger seagrass component into their machine learning projects.

Convolutional↗

Decoding the proton’s gluonic density with lattice QCD-informed machine learning

We present a first machine learning-based decoding of the gluonic structure of the proton from lattice QCD using a variational autoencoder inverse mapper (VAIM). Harnessing the power of generative AI, we predict the parton distribution function (PDF) of the gluon given information on the reduced pseudo-Ioffe-time distributions (RpITDs) as calculated from an ensemble with lattice spacing a ≈ 0.09 fm and a pion mass of M π ≈ 310 MeV. The resulting gluon PDF is consistent with phenomenological global fits within uncertainties, particularly in the intermediate-to-high-x region where lattice data are most constraining. A subsequent correlation analysis confirms that the VAIM learns a meaningful latent representation, highlighting the potential of generative AI to bridge lattice QCD and phenomenological extractions within a unified analysis framework.

Gluon parton distribution function↗

A methodology for decay heat characterization in molten salt reactors

Accurate decay heat prediction in molten salt reactors (MSRs) faces dual challenges: complex operational uncertainties and the need for interpretable models compatible with engineering workflows. This work presents a hybrid machine learning and segmented polynomial methodology that addresses both requirements through three key innovations. First, a modular data architecture encodes MSR-specific operational parameters (power density: 1-100 W cm -3 , humidity: 0-0.1 wt %, air ingress: 0-0.1 mol %) with uncertainty-aware temporal discretization spanning 15 orders of magnitude. Second, region-optimized machine learning models achieve 92.3 % root mean square error (RMSE) reduction over conventional polynomials while maintaining physical interpretability through automated piecewise equation generation. Third, dual front-end interfaces accelerate safety analyses — a Jupyter environment enables researchers to explore 10,000+ parameter combinations via interactive widgets, while a Streamlit web application reduces design iteration cycles through production-grade visualization tools. Operational deployment demonstrates prediction times of only a couple hundred milliseconds for 10 4 years decay profiles, enabling real-time optimization of spent fuel container designs.

42 - ENGINEERING↗

Leverage modern artificial intelligence (AI) enabled systems for waste reduction

Manufacturing industries continue to face challenges in reducing waste, as upstream strategies such as source reduction and product redesign require a deeper understanding of processes compared to conventional recycling methods. Recent advancements in artificial intelligence (AI) and machine learning (ML) have opened new opportunities to integrate modern computational techniques with traditional waste minimization strategies. This paper explores AI-enabled approaches for product redesign, source reduction, and recycling that can significantly reduce waste generation while improving efficiency and sustainability. AI-driven material substitution and lightweighting in product design enable discovery of novel materials with optimized properties, reducing waste without compromising performance. Reinforcement learning models optimize process parameters, raw material specifications, and machine sequencing to minimize production losses, while Industrial Internet of Things (IIoT) systems paired with AI analytics enhance real-time waste tracking, predictive maintenance, and quality inspection. Furthermore, AI-based demand forecasting and production planning reduce overproduction and excess inventory, as demonstrated in industrial applications. In recycling, ML-powered pattern recognition and robotic sorting technologies achieve higher accuracy in waste segregation, directly improving recycling efficiency. Complementary solutions such as smart bins and AI-enabled waste pickup scheduling optimize collection logistics, reducing both costs and emissions. Although implementation requires upfront investment in infrastructure and training, the long-term benefits include higher material efficiency, reduced waste, improved product quality, and stronger sustainability outcomes across the supply chain. By leveraging AI-enabled systems, manufacturers can align waste minimization efforts with circular economy principles, creating scalable solutions for both industry and society.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Assurance of Reasoning Enabled Systems (ARES)

ARES was in part motivated by the determination of President’s Council of Advisors on Science and Technology (PCAST) on May 13th, 2023 that published a set of inquiries: In an era in which convincing images, audio, and text can be generated with ease on a massive scale, how can we ensure reliable access to verifiable, trustworthy information? How can we be certain that a particular piece of media is genuinely from the claimed source? What technologies, policies, and infrastructure can be developed to detect and counter AI-generated disinformation? In an effort to automatically analyze and patch/optimize code the work in this report describes various neural Machine Learning (ML) analysis engine implementations to assist in situations where source code is deficient or completely lacking to decompile (lift) binary code to ’C’. The goal is to gradually reduce human intervention. To this end, two Large Language Model (LLM) variants (Code LLama 2, LLama 3.1 and Starcoder1, Starcoder 2) where finetuned with ’before/after’ code pairs on the OpenBLAS library. LLama trained on the lowering process, Starcoder trained on the lifting process with National Security Agency’s (NSA) open-source Ghidra decompiler assist. The inferencing test results indicate correctness for only very short sequences for Starcoder 2. Moving forward, the experiments conclude with a set of recommendations of required resources and technologies

97 MATHEMATICS AND COMPUTING↗

A Fast Framework for Generating Radioactive Mixture Spectra and Its Application to Remote High-Performance Mixture Identification

Remote detection of radioactive materials in mixtures using handheld or portal detectors remains a challenge because of factors such as low concentration, environmental interference, sensor noise, and other complications. This work introduces a fast framework for generating realistic mixture spectra. Moreover, we present mixture isotope identification using data generated by the fast framework. Researchers have examined a range of conventional and recent algorithms within the fields of machine learning and deep learning. An application to uranium enrichment-level prediction has been included. Extensive simulation experiments validated the efficacy of the proposed framework.

GADRAS↗

Enhancing Air Quality Applications in the Hindu Kush-Himalayan Region Using Satellite, Model, and Machine Learning Techniques

Air pollution in the Hindu Kush Himalayan (HKH) region of South Asia is a severe issue, as increases in emissions over the past two decades have degraded air quality (AQ) across the region, which poses major threats to human health, the ecosystem, climate, and agriculture. A diversity of anthropogenic and natural emission sources including transportation, power plants, industries, open biomass burning of crop residue, forest fires, cooking and heating fires, and dust storms contribute to unhealthy AQ and transboundary pollution issues in the region. Further complicating matters is the importance of meteorology and terrain on AQ, especially in the Kathmandu Valley where extreme haze episodes frequently develop from the atmospherically stable weather conditions during the winter monsoon. This study uses state-of-the-art satellite observations and modeling capabilities in conjunction with machine learning techniques to develop a comprehensive toolkit for enhancing AQ monitoring and forecasting in HKH. The toolkit incorporates new generation satellite observations from the TROPOspheric Monitoring Instrument (TROPOMI), Geostationary Environment Monitoring Spectrometer (GEMS), and Advanced Meteorological Imager (AMI), which provide unprecedented resolution on aerosols and trace gases, including nitrogen dioxide (NO 2 ), formaldehyde (CH2O), sulfur dioxide (SO 2 ), carbon monoxide (CO), and ozone (O 3 ), and aerosol optical depth (AOD). Value-added products [e.g., Particulate matter with diameters less than 2.5 micrometers (PM2.5)] are developed from the suite of satellite observations to further improve AQ monitoring capabilities in the region. The satellite products are also used to assimilate a high-resolution chemical transport model tailored for the HKH region, which is providing daily, 54-hour AQ forecasts with horizontal grid spacings of 12- and 4-km. This presentation will provide an overview of the suite of satellite- and model-based products in the AQ toolkit and application and performance of the toolkit for AQ monitoring and forecasting in HKH.

Air Quality↗

Snow Depth from AMSR-2 Using Multispectral Satellite Data in an Artificial Neural Network

By using diffusion theory and Monte Carlo lidar radiative transfer simulations, Hu et al. (2022b) has derived snow depth from the first-, second- and third-order moments of the lidar backscattering pathlength distribution. Lu et al. (2022) calculated the snow depth by applying the methods to the satellite ICESat-2 lidar measurements over the Arctic sea ice, as well as land surfaces of Northern Hemisphere. In this paper, an artificial neural network (ANN) algorithm, employing several channels from Advanced Microwave Scanning Radiometer 2 (AMSR-2) and the humidity vertical profiles from Global Modeling and Assimilation Office (GMAO) Goddard Earth Observing System for Instrument Teams (GEOS-IT) product, is trained to determine snow depth identified by time and geolocation matched 2019 ICESat-2 snow-depth data during winter months over the Arctic sea ice. The trained ANN snow-depth was applied to 2018 AMSR-2 clear pixel data, although the algorithms perform reasonably well in thinner clouds. The validation data (different from the training set) of ANN snow depth from AMSR-2 showed a good agreement with time matched and co-located snow-depth values from ICESat-2. The bias was near zero, with mean absolute error (MAE) 0.05 cm and a root-mean-square-error (RMSE) 0.08 cm. Prior applying the trained ANN snow depth to AMSR-2 data, a cloud screening algorithm was developed with a similar approach. A separate ANN cloud mask was trained to determine an AMSR-2 pixel is clear or cloudy with time and geolocation matched 2015 CALIOP Vertical Feature Mask (VFM) over Arctic sea ice. The ANN cloud mask from AMSR-2 under-estimated cloud fraction by 3-6% compared to CALIOP . The additional research is needed to conclusively evaluate the ANN cloud mask accuracy. Finally, this paper will lay the foundation for a sustained long-term snowfall and snow-storm monitoring system. The future Cloud Aerosol LIdar for Global scale Observations of the ocean-Land Atmosphere system (CALIGOLA) mission will provide a means to calculate snow depth from the lidar backscattering pathlength distribution, benefiting from the UV, visible and infrared pulses. With the calculated snow depth as the truth one could develop a machine learning algorithm, as it was done in this paper, using a passive microwave instrument available at that time to generate a wide range of snow depth data, covering extensive spatial areas in the cross-orbit direction.

Neural Network↗

Enhancing Air Quality Applications in the Hindu Kush-Himalayan Region Using Satellite, Model, and Machine Learning Techniques

Air pollution in the Hindu Kush Himalayan (HKH) region of South Asia is a severe issue, as increases in emissions over the past two decades have degraded air quality (AQ) across the region, which poses major threats to human health, the ecosystem, climate, and agriculture. A diversity of anthropogenic and natural emission sources including transportation, power plants, industries, open biomass burning of crop residue, forest fires, cooking and heating fires, and dust storms contribute to unhealthy AQ and transboundary pollution issues in the region. Further complicating matters is the importance of meteorology and terrain on AQ, especially in the Kathmandu Valley where extreme haze episodes frequently develop from the atmospherically stable weather conditions during the winter monsoon. This study uses state-of-the-art satellite observations and modeling capabilities in conjunction with machine learning techniques to develop a comprehensive toolkit for enhancing AQ monitoring and forecasting in HKH. The toolkit incorporates new generation satellite observations from the TROPOspheric Monitoring Instrument (TROPOMI), Geostationary Environment Monitoring Spectrometer (GEMS), and Advanced Meteorological Imager (AMI), which provide unprecedented resolution on aerosols and trace gases, including nitrogen dioxide, formaldehyde, sulfur dioxide, carbon monoxide, and ozone, and aerosol optical depth. Value-added products, such as level 4 PM2.5 products, are developed from the suite of satellite observations to further improve AQ monitoring capabilities in the region. The satellite products are also used to assimilate a high-resolution chemical transport model tailored for the HKH region, which is providing daily, 54-hour AQ forecasts with horizontal grid spacings of 12- and 4-km. This presentation will provide an overview of the suite of satellite- and model-based products in the AQ toolkit and application and performance of the toolkit for AQ monitoring and forecasting in HKH.

Forecasting↗

Surrogate-driven Variance-based Sensitivity Analysis of Thermal Storage Tanks in Integrated Energy Systems

Sensitivity analysis and uncertainty quantification are essential steps for enhancing the accuracy of computational models by identifying and mitigating uncertainties. This study focuses on these steps for the Thermal Energy Delivery System at Idaho National Laboratory, specifically targeting the thermocline tank. Using a Modelica/Dymola simulation model, the study perturbed various design parameters and boundary conditions, including shape factor, porosity, outlet temperature, inlet mass flow rate, and system pressure, to predict and quantify uncertainty in the tank’s ax- ial temperature. A dataset of over 1,000 simulations was generated, and surrogate models were developed using the pyMAISE (Michigan Artificial Intelligence Standard Environment) library, which is an Automatic Machine Learning library for nuclear engineering applications. The optimal model, a feedforward neural network with two hidden layers, achieved an R2 score above 0.99 and a mean absolute error below 1 Kelvin. Sensitivity analyses using Sobol indices and Fourier amplitude sensitivity testing methods on this surrogate model revealed that the inlet mass flow rate at initial timestamps and porosity significantly impacts predicted temperatures across all sensors and time steps.

22 - GENERAL STUDIES OF NUCLEAR REACTORS↗