Search NASA⌕ Search

SEARCH · Search NASA

Results for “Machine learning models”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 307 records · Page 17

Hyperplane decision trees as piecewise linear surrogate models for chemical process design

Recent trends in chemical engineering research point towards an increasing reliance on data-driven modeling approaches. Neural networks, for instance, have proven to be accurate when data is plentiful and high-dimensional, but in many cases, they require computationally-intensive training procedures. Here, in this work, we describe hyperplane decision trees (HT) as a highly expressive and low-compute machine learning model architecture. These models are locally linear and have linear decision boundaries, resulting in a piecewise linear model of the data. This property allows them to be converted into mixed-integer linear constraints which can be globally optimized. Our open-source PyTorch implementation of this method is a fast, flexible, and accessible way to build accurate piecewise linear models of data.

Decision trees↗

Modeling of Vertical Motor-driven Pump for Simulation of a Fault Signature \\ for Condition Monitoring

As part of the ongoing effort to transition from preventive maintenance strategies to condition-based maintenance strategies in nuclear power plants, there is significant reliance on using machine learning techniques. To develop a robust machine learning model that can diagnose all the fault modes of a vertical motor-driven pump, data capturing the unique signature of each fault mode is required. In practice, it is difficult to collect or capture data that captures all the fault modes from a single plant site. So to address this situation, a computational model of a vertical motor-driven pump is developed using the multipurpose finite element software COMSOL Multiphysics. The developed model is used to generate simulated data under normal operation and is compared with the vibration data collected using vibration sensors. Once the simulation model is verified under normal operating condition, simulated data for the fault mode for which minimal or no evidence is available in historical plant process data is developed. This simulated data is used to develop fault signatures to achieve robust predictive models. This paper presents modeling details and verification of the model that can used to generate data for fault modes that are not available at a plant site for condition monitoring purpose.

46 - INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AN↗

Developing a complete AI-accelerated workflow for superconductor discovery

The quest to identify new superconducting materials with enhanced properties is hindered by the prohibitive cost of computing electron-phonon spectral functions, severely limiting the materials space that can be explored. Here, we introduce a Bootstrapped Ensemble of Equivariant Graph Neural Networks (BEE-NET), a machine-learning model trained to predict the Eliashberg spectral function and superconducting critical temperature with a mean-absolute-error of 0.87 K relative to DFT-based Allen-Dynes calculations. Intriguingly, BEE-NET achieves a true-negative-rate of 99.4%, enabling highly efficient screening for the rare property of superconductivity. Integrated into a multi-stage, AI-accelerated discovery pipeline that incorporates elemental-substitution strategies and machine-learned interatomic potentials, our workflow reduced over 1.3 million candidate structures to 741 dynamically and thermodynamically stable compounds with DFT-confirmed T c > 5 K. We report the successful synthesis and experimental confirmation of superconductivity in two of these previously unreported compounds. This study establishes a data-driven framework that integrates machine learning, quantum calculations, and experiments to systematically accelerate superconductor discovery.

Gibson, Jason B. [Quantum Formatics, Cambridge, MA↗

Design Choices in Anomaly Detection for Industrial Control Systems: Insights from Gas Pipeline Data

Industrial control systems (ICS) remain vulnerable to increasingly sophisticated cyberattacks, yet evaluating anomaly detection models in these environments is challenging due to temporal dependencies, missing-not-at-random patterns, and extremely imbalanced datasets. These factors make common practices—especially random data splits and naïve imputation—prone to severe temporal leakage, which can inflate reported performance and obscure real-world limitations. In this work, we systematically examine classical machine learning models, temporal deep learning architecture, and tensor-decomposition–based methods on a gas-pipeline dataset using a fully temporally separated evaluation pipeline designed to mimic realistic deployment conditions. Our findings show that proper temporal handling and MNAR-aware preprocessing significantly alter the relative performance of popular anomaly-detection methods, providing practical guidance for designing reliable, leakage-resistant ICS intrusion-detection systems.

97 MATHEMATICS AND COMPUTING↗

Precision beam diagnostics at the NuMI facility using muon monitor observations

The Neutrinos at the Main Injector (NuMI) facility at Fermilab delivers an intense neutrino beam for multiple experiments by producing pions that decay into neutrinos, muons, and other particles. Magnetic horns—the primary pion focusing elements in the NuMI beamline—exhibit predominantly linear optics, enabling a predictable relationship between the proton beam and the resulting pion and muon phase spaces. This study has two primary objectives: first, to evaluate and confirm the linearity of the horn focusing mechanism using analytical models and numerical simulations; and second, to demonstrate that key beam parameters—such as proton beam intensity, beam position on target, and horn current—can be extracted from muon monitor observations within this linear optics framework. Using a machine learning model trained on spill-by-spill muon monitor data, we infer the horn current with a precision of ±0.05%, the beam intensity with ±0.1%, and the beam position on target with ±0.018⁢ mm horizontally and ±0.013⁢ mm vertically. This approach provides a reliable cross-check of beam parameters, helping to reduce systematic uncertainties that are critical for future experiments such as the Deep Underground Neutrino Experiment, which will rely on the neutrino beam produced by the Long-Baseline Neutrino Facility.

Beam control↗

Machine learning force field model for kinetic Monte Carlo simulations of itinerant Ising magnets

Here, we present a scalable machine learning (ML) framework for large-scale kinetic Monte Carlo (kMC) simulations of itinerant electron Ising systems. As the effective interactions between Ising spins in such itinerant magnets are mediated by conducting electrons, the calculation of energy change due to a local spin update requires solving an electronic structure problem. Such repeated electronic structure calculations could be overwhelmingly prohibitive for large systems. Assuming the locality principle, a convolutional neural network (CNN) model is developed to directly predict the effective local field and the corresponding energy change associated with a given spin update based on Ising configuration in a finite neighborhood. As the kernel size of the CNN is fixed at a constant, the model can be directly scalable to kMC simulations of large lattices. Our approach is reminiscent of the ML force field models widely used in first-principles molecular dynamics simulations. Applying our ML framework to a square-lattice double-exchange Ising model, we uncover unusual coarsening of ferromagnetic domains at low temperatures. Our work highlights the potential of ML methods for large-scale modeling of similar itinerant systems with discrete dynamical variables.

machine learning↗

HGQ: High Granularity Quantization for Real-time Neural Networks on FPGAs

Neural networks with sub-microsecond inference latency are required by many critical applications. Targeting such applications deployed on FPGAs, we present High Granularity Quantization (HGQ), a quantization-aware training framework that optimizes parameter bit-widths through gradient descent. Unlike conventional methods, HGQ determines the optimal bit-width for each parameter independently, making it suitable for hardware platforms supporting heterogeneous arbitrary precision arithmetic. In our experiments, HGQ shows superior performance compared to existing network compression methods, achieving orders of magnitude reduction in resource consumption and latency while maintaining the accuracy on several benchmark tasks. These improvements enable the deployment of complex models previously infeasible due to resource or latency constraints. HGQ is open-source and is used for developing next-generation trigger systems at the CERN ATLAS and CMS experiments for particle physics, enabling the use of advanced machine learning models for real-time data selection with sub-microsecond latency.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Establishing an acoustic-property relationship in laser powder bed fusion with machine learning

Quality control of Laser Powder Bed Fusion (PBF-LB) additively manufactured parts is an important hurdle inhibiting the technology’s use structural applications. Acoustic monitoring of the laser powder bed fusion process can detect defects in-situ that are known to degrade mechanical properties. However, processing-structure-property (PSP) relationships are required to extrapolate from detected defects to part performance. Here, this study explores how acoustics may be a suitable signature linking processing conditions to properties, thus effectively substituting for structure in the PSP relationship. Establishing such a relationship would enable a part’s mechanical performance to be directly predicted from its acoustic signature, reducing the need for destructive testing or microstructural analysis to ensure a part will meet performance requirements. One hundred CoCrFeMnNi high entropy alloy tensile bars were printed across 13 process conditions in a series of 6 prints. The acoustic signatures of these tensile bars were used to train machine learning models to predict each part’s mechanical properties. By using both process information and acoustic information to predict mechanical properties, yield strength was predicted 18% more accurately and ductility to failure was predicted 10% more accurately than is achieved when using duplicate parts to predict part performance. Finally, individual acoustic frequencies were investigated to determine why acoustic signatures improve mechanical property predictions and the potential physical origins of these signatures. This work demonstrates how blending acoustics, process information, and machine learning can provide in-situ diagnostics of mechanical properties and improve the reliability of the PBF-LB process.

Acoustic emission↗

The Impact of Time-Aware Design Choices in ICS Anomaly Detection

Industrial control systems (ICS) remain vulnerable to increasingly sophisticated cyberattacks, yet evaluating anomaly detection models in these environments is challenging due to temporal dependencies, missing-not-at-random patterns, and extremely imbalanced datasets. These factors make common practices—especially random data splits and na¨ıve imputation— prone to severe temporal leakage, which can inflate reported performance and obscure real-world limitations. In this work, we systematically examine classical machine learning models, temporal deep learning architecture, and tensordecomposition– based methods on a gas-pipeline dataset using a fully temporally separated evaluation pipeline designed to mimic realistic deployment conditions. Our findings show that proper temporal handling and MNAR-aware preprocessing significantly alter the relative performance of popular anomaly-detection methods, providing practical guidance for designing reliable, leakage-resistant ICS intrusion-detection systems.

97 MATHEMATICS AND COMPUTING↗

Aging heat treatment design for Haynes 282 made by wire-feed additive manufacturing using high-throughput experiments and interpretable machine learning

Wire-feed additive manufacturing (WFAM) produces superalloys with complex thermal cycles and unique microstructures, often requiring optimized heat treatments. To address this challenge, we present a hybrid approach that combines high-throughput experiments, precipitation simulation, and machine learning to design effective aging conditions for the WFAM Haynes 282 superalloy. Our results demonstrate that the γ’ radius is the critical microstructural feature for strengthening Haynes 282 during post-heat treatment compared with the matrix composition and γ’ volume fraction. New aging conditions at 770°C for 50 hours and 730°C for 200 hours were discovered based on the machine learning model and were applied to enhance yield strength, bringing it on par with the wrought counterpart. This approach has significant implications for future AM alloy production, enabling more efficient and effective heat treatment design to achieve desired properties.

CALPHAD↗

Targeted Biomining and Machine Learning Approaches in Critical Minerals Revealed by a Biogeochemical Survey of a Coal Mine Drainage Remediation System

Abandoned coal mine drainage (AMD) remediation systems in Pennsylvania can concentrate critical minerals and materials (CMM) at levels comparable to mining-grade ores. Remediation systems have varying engineering features and are open to the environment, resulting in diverse microbial colonization and seasonal climate influences that may impact CMM speciation. The location of CMMs, the types of bacterial communities tolerant of these pollutant conditions, and the influence of localized climate on CMM rich remediation systems are not well characterized. Through a one-year spatiotemporal survey of biogeochemistry at a remediation system, we have initiated the process to address these questions. Rare Earth Elements (REE) ranged 180-1,200 ppm and greater than 1,500 bacterial ASVs were classified via 16S sequencing. Analyses indicate biogeochemical differences are heavily influenced by engineering features. Additionally, REE precipitants correlate strongly with the elements Al, Cu, Zn, Be, and U. Unearthing these trends has refined our line of inquiry to explore biological mining opportunities more closely with these metals. Furthermore, we created a Machine Learning Model for predicting AMD REE content, with 89% accuracy, using the data from this study and several others. Further training data is required to create a more reputable model. Recently, global research efforts have prioritized modeling work or the use of the few historical surveys to design experiments. Through our data, we challenge this approach, emphasizing the importance of expanding fundamental survey efforts prior to advanced product design and experimentation.

critical minerals↗

Autonomous Utility Pole Identification

The implementation of small unmanned aerial systems (sUAS) for the purpose of powerline inspection is an emerging concept among utility companies. Such operations can produce a significant amount of useful data for the purpose of training different machine learning models or keep track of the integrity of electrical infrastructure. As such, this system is developed for the purpose of properly cataloging and leveraging this collected data.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Optimization of simulated high-field side lower hybrid current drive coupling using machine learning predictions of scrape-off layer density

Lower hybrid current drive (LHCD) is a potential source of non-inductive off-axis current drive (CD) for tokamaks. Although LHCD has been successfully deployed on a number of tokamaks, it is highly sensitive to the scrape-off layer (SOL) conditions local to the LHCD launcher. Large gaps between the launcher and plasma core, SOL turbulence, or edge density perturbations due to edge-localized modes can hamper CD or cause large reflected power. These coupling issues in part motivated the installation of an LHCD launcher on the high-field side (HFS) of DIII-D. On the HFS, the SOL is less turbulent and more controllable compared to the low-field side. This quiescence may result in more predictable edge conditions and thus a more predictable CD. Here, in this work, HFS SOL reflectometry measurements are predicted from global plasma parameters using machine learning models. The SOL predictions coupled with the full-wave simulation of the LHCD launcher allow for the prediction of reflected power, directivity, and arcing risk before the discharge. Launcher performance is then optimized using multi-objective Bayesian optimization, finding the shot parameters that result in an optimal SOL density that maximizes CD while minimizing the risk of arcing. The predictions and optimizations of LHCD performance are then accelerated using a surrogate model of the full-wave LHCD simulation.

Bayesian optimization↗

Large-Scale Visualization of 3D Unstructured Groundwater Model Using Cave Automated Virtual Environment

The immersive three-dimensional (3D) virtual reality (VR) visualization of groundwater models allows us to deepen our understanding of aquifer systems and provide better solutions to present groundwater-related problems, such as groundwater recharge, water quality, and sustainability. Visualization assists in accurately developing groundwater models and revealing important subsurface features, including faulting, folding, and unconformity. However, assessing model accuracy poses challenges due to the complexity of geology and groundwater systems. This research demonstrates a workflow to visualize and analyze raw 3D unstructured groundwater model data using an immersive Cave Automated Virtual Environment (CAVE). To visualize the unstructured groundwater model data, the raw dataset is converted into interactive CAVE-compatible formats utilizing a set of tools: ParaView, Blender, and Unity. This enables researchers to immerse themselves in the data, identifying influential patterns and relationships. e resulting insights can inform the development of sophisticated machine-learning models for groundwater level prediction. The CAVE’s immersive capabilities allow intuitive exploration from various perspectives, providing a more holistic understanding of the factors affecting groundwater levels. These insights are crucial to improve predictive models. The CAVE results also facilitate collaborative analysis and have potential applications in training and education. is research demonstrates the value of immersive VR tools such as the CAVE for unraveling intricacies within high-dimensional scientific data to drive real-world forecasting and modeling applications.

54 ENVIRONMENTAL SCIENCES↗

CovTransformer: A transformer model for SARS-CoV-2 lineage frequency forecasting

With hundreds of SARS-CoV-2 lineages circulating in the global population, there is an ongoing need for predicting and forecasting lineage frequencies and thus identifying rapidly expanding lineages. Accurate prediction would allow for more focused experimental efforts to understand pathogenicity of future dominating lineages and characterize the extent of their immune escape. Here, we first show that the inherent noise and biases in lineage frequency data make a commonly-used regression-based approach unreliable. To address this weakness, we constructed a machine learning model for SARS-CoV-2 lineage frequency forecasting, called CovTransformer, based on the transformer architecture. We designed our model to navigate challenges such as a limited amount of data with high levels of noise and bias. We first trained and tested the model using data from the UK and the USA, and then tested the generalization ability of the model to many other countries and US states. Remarkably, the trained model makes accurate predictions two months into the future with high levels of accuracy both globally (in 31 countries with high levels of sequencing effort) and at the US-state level. Our model performed substantially better than a widely used forecasting tool, the multinomial regression model implemented in Nextstrain, demonstrating its utility in SARS-CoV-2 monitoring. Assuming a newly emerged lineage is identified and assigned, our test using retrospective data shows that our model is able to identify the dominating lineages 7 weeks in advance on average before they became dominant. Overall, our work demonstrates that transformer models represent a promising approach for SARS-CoV-2 forecasting and pandemic monitoring.

60 APPLIED LIFE SCIENCES↗

Fast and Accurate Pixel Calibration of Tof Neutron Diffractometers with Machine Learning

At a spallation neutron source, neutron pulses of varying energies are generated, and the detection of neutrons by instrument detectors is recorded as time-of-flight from the emission of the neutron pulse to its arrival at specific detector pixels with high time resolution. The flight path of neutrons from the moderator to the sample and then to the detector must be precisely calibrated at the detector-pixel level using standard powders, so the neutron events from all pixels can be time-focused to produce high-resolution diffraction patterns. Modern time-of-flight neutron diffractometers at spallation neutron sources are equipped with two-dimensional detectors with millimeter-scale pixelations. The number of pixels in a diffraction instrument can reach millions, which makes a single-pixel-level calibration process time-consuming or even impossible with conventional refinement or fitting approaches. Here we present a machine-learning-aided calibration process using a train-and-predict approach, in which machine learning models are trained on the relationship between an individual pixel time-of-flight diffraction pattern and its diffraction constant. These models use a portion of the available pixels for training, and a good model then predicts the diffraction constants precisely and rapidly for large sets of pixel diffraction patterns.

detector pixel calibration↗

Quantifying mean, variability, and uncertainty in indoor radon exposure in Pennsylvania using random forest and quantile regression forest models

Radon is a naturally occurring radioactive gas that poses a serious health risk as the primary cause of lung cancer in non-smokers. Despite the well-known adverse association with health outcomes, current radon exposure assessments are limited to county-level or average-level estimates, which fail to capture regional variability. This study uses Machine Learning models, including Random Forest (RF) and Quantile Regression Forest (QRF), to estimate the indoor radon concentrations at the ZCTA (Zip code tabulation area)-level and characterize uncertainties in model estimates. Incorporating geological, meteorological, and building-specific data, the models aim to improve radon risk assessment by capturing mean exposure, variability, and extreme concentration levels. Processed radon test data (n = 718,111) were analyzed using average, variability, and quantile prediction methods. Models that estimate the average radon exposure at the ZCTA-level can yield promising model-fit results, but they do not capture the underlying variability of indoor radon exposure within a ZCTA. We utilize volatility analyses to identify characteristics indicative of high variability of indoor radon exposure. We also show that a QRF model can be used to estimate upper quantiles of residential radon exposure, thereby uncovering localized areas of elevated exposure that were not apparent in mean estimates. The results highlighted the need for a deep characterization of exposure risk and show that regions with moderate average exposure levels could still harbor extreme outliers with implications for evaluating health risks. Utilizing multiple radon exposure models allows for a deeper characterization of radon risk within a geographic area and can better identify high-risk areas. The results from this study provide a foundation for developing mitigation strategies and examining associations between radon exposure and health outcomes at fine scales. Future research should extend the geographic scope and incorporate additional environmental risk factors to establish a comprehensive framework for risk assessment.

Lee, Heechan [ORNL]↗

Two-level overlapping additive Schwarz preconditioner for training scientific machine learning applications

In this work we introduce a novel two-level overlapping additive Schwarz preconditioner for accelerating the training of scientific machine learning applications. The design of the proposed preconditioner is motivated by the nonlinear two-level overlapping additive Schwarz preconditioner. The neural network parameters are decomposed into groups (subdomains) with overlapping regions. In addition, the network’s feed-forward structure is indirectly imposed through a novel subdomain-wise synchronization strategy and a coarse-level training step. Through a series of numerical experiments, which consider physicsinformed neural networks and operator learning approaches, we demonstrate that the proposed two-level preconditioner significantly speeds up the convergence of the standard (LBFGS) optimizer while also yielding more accurate machine learning models. Moreover, the devised preconditioner is designed to take advantage of model-parallel computations, which can further reduce the training time.

97 MATHEMATICS AND COMPUTING↗