Search NASA⌕ Search

SEARCH · Search NASA

Results for “Convolutional neural network”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 307 records · Page 17

OPEN ALPHADIFFRACT

Open-source release of the AlphaDiffract data generation and training system. Includes only the public Materials Project dataset retrievers.AlphaDiffract is a deep learning framework that achieves state-of-the-art performance in predicting the crystal system, space group, and lattice parameters directly from PXRD patterns. AlphaDiffract utilizes a 1D adaptation of the ConvNeXt architecture, a modern convolutional neural network that integrates key design principles from transformers, coupledwith dedicated prediction heads for each crystallographic property.

Prince, Michael [Argonne National Laboratory (ANL)↗

OpenCRUMS USA: An Open Machine Learning Framework for Characterizing Variability in Aerosol Reanalysis Data

Advances in artificial intelligence (AI) have called for exploring how these techniques can be used for exploring patterns in large climate datasets. To that regard, the U.S. Department of Energy AI for Earth System Predictability (AI4ESP) supported a pilot initiative called the Open Classification of Regimes in the Southeast USA (OpenCRUMS USA) project to explore how AI can be used to characterize modes of spatial variability in large climate datasets. For this study, we focus on comparing two methods for characterizing the modes of spatial variability of surface aerosol concentration over the Houston region: empirical orthogonal functions (EOFs) and layerwise relevance propagation (LRP) applied to a convolutional neural network (CNN) classifier. We show that EOF analysis typically attributes spatial variability modes that span all of southeast Texas, prohibiting the attribution of spatial variability to localized regions. However, using LRP on the CNN classifier resolves the explanatory parameters at a finer spatial resolution than EOFs. This allows for the attribution of the spatial variability of surface aerosols to local regions of organic carbon which was not possible using EOFs. In addition, the LRP analysis also suggests that synoptic-scale transport of dust is most prevalent during anticyclonic and pretrough synoptic conditions as categorized by self-organizing maps.

54 ENVIRONMENTAL SCIENCES↗

Variation in forest root image annotation by experts, novices, and AI

Abstract Background The manual study of root dynamics using images requires huge investments of time and resources and is prone to previously poorly quantified annotator bias. Artificial intelligence (AI) image-processing tools have been successful in overcoming limitations of manual annotation in homogeneous soils, but their efficiency and accuracy is yet to be widely tested on less homogenous, non-agricultural soil profiles, e.g., that of forests, from which data on root dynamics are key to understanding the carbon cycle. Here, we quantify variance in root length measured by human annotators with varying experience levels. We evaluate the application of a convolutional neural network (CNN) model, trained on a software accessible to researchers without a machine learning background, on a heterogeneous minirhizotron image dataset taken in a multispecies, mature, deciduous temperate forest. Results Less experienced annotators consistently identified more root length than experienced annotators. Root length annotation also varied between experienced annotators. The CNN root length results were neither precise nor accurate, taking ~ 10% of the time but significantly overestimating root length compared to expert manual annotation ( p = 0.01). The CNN net root length change results were closer to manual ( p = 0.08) but there remained substantial variation. Conclusions Manual root length annotation is contingent on the individual annotator. The only accessible CNN model cannot yet produce root data of sufficient accuracy and precision for ecological applications when applied to a complex, heterogeneous forest image dataset. A continuing evaluation and development of accessible CNNs for natural ecosystems is required.

Handy, Grace↗

Application of deep learning to single-shot gas-phase laser-induced breakdown spectroscopy

Single-shot fs laser-induced breakdown spectroscopy (LIBS) has the potential to capture ns-scale electrode desorption phenomena in pulsed power fusion drivers. However, the successful implementation of the diagnostic for this purpose is challenging, as it requires interpreting single-shot measurements collected from low-density gas mixtures. In this work, we demonstrate the efficacy of a Bayesian-optimized convolutional neural network (CNN) to interpret these measurements. We generated 256 distinct measurement conditions at relevant gas pressures ranging from 80–530 mTorr by mixing 100–250 sccm H 2 and 50–200 sccm CH 4 in increments of 10 sccm. Despite the considerable overlap between signals separated by 20 sccm, the CNN is able to predict the H 2 flow rate with a root-mean-square error (RMSE) of 15.9 sccm and the CH 4 flow rate with an RMSE of 12.0 sccm. The average relative prediction error is <9% for each gas and largely remains below or near 10%.

Brown, Nathan Parnell [Sandia National Lab. (SNL-N↗

UNNT: A novel Utility for comparing Neural Net and Tree-based models

The use of deep learning (DL) is steadily gaining traction in scientific challenges such as cancer research. Advances in enhanced data generation, machine learning algorithms, and compute infrastructure have led to an acceleration in the use of deep learning in various domains of cancer research such as drug response problems. In our study, we explored tree-based models to improve the accuracy of a single drug response model and demonstrate that tree-based models such as XGBoost (eXtreme Gradient Boosting) have advantages over deep learning models, such as a convolutional neural network (CNN), for single drug response problems. However, comparing models is not a trivial task. To make training and comparing CNNs and XGBoost more accessible to users, we developed an open-source library called UNNT (A novel Utility for comparing Neural Net and Tree-based models). The case studies, in this manuscript, focus on cancer drug response datasets however the application can be used on datasets from other domains, such as chemistry.

59 BASIC BIOLOGICAL SCIENCES↗

Identifying the quantum properties of hadronic resonances using machine learning

With the great promise of deep learning, discoveries of new particles at the Large Hadron Collider (LHC) may be imminent. Following the discovery of a new Beyond the Standard model particle in an all-hadronic channel, deep learning can also be used to identify its quantum numbers. Convolutional neural networks (CNNs) using jet-images can significantly improve upon existing techniques to identify the quantum chromodynamic (QCD) (‘color’) as well as the spin of a two-prong resonance using its substructure. Additionally, jet-images are useful in determining what information in the jet radiation pattern is useful for classification, which could inspire future taggers. These techniques improve the categorization of new particles and are an important addition to the growing jet substructure toolkit, for searches and measurements at the LHC now and in the future.

Filipek, Jakub↗

Status of the Muon Neutrino Charged-Current Mesonless Cross Section Measurement in the NOvA Near Detector

NOvA is a long-baseline accelerator neutrino experiment at Fermilab whose physics goals include precision neutrino oscillation as well as cross-section measurements. We present the status of the measurement of a muon neutrino charged-current cross section with zero mesons in the final state at the NOvA near detector. This measurement is being made with respect to the kinematics of the final state muon. The chosen interaction channel is especially sensitive to quasielastic and meson exchange current interactions and aims to provide experimental constraints for the development of models of neutrino interactions. It will also provide a handle for constraining cross section systematic uncertainties in oscillation analyses in current and future experiments. For particle identification, we use a convolutional neural network (CNN) trained on individual particles simulated in the NOvA near detector that allows us to select the desired signal while reducing the potential bias from neutrino interaction modeling. Charged pion background constraining is further improved via Michel electron tagging.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Status of the Muon Neutrino Charged-Current Zero Mesons Cross Section at the NOvA Near Detector

NOvA is a long-baseline accelerator neutrino experiment at Fermilab whose physics goals include precision neutrino oscillation as well as cross-section measurements. We present the status of the measurement of a muon neutrino charged-current cross section with zero mesons in the final state at the NOvA near detector. This measurement is being made with respect to the kinematics of the final state muon. The chosen interaction channel is especially sensitive to quasielastic and meson exchange current interactions and aims to provide experimental constraints for the development of models of neutrino interactions. It will also provide a handle for constraining cross section systematic uncertainties in oscillation analyses in current and future experiments. For particle identification, we use a convolutional neural network (CNN) trained on individual particles simulated in the NOvA near detector that allows us to select the desired signal while reducing the potential bias from neutrino interaction modeling. Charged pion background constraining is further improved via Michel electron tagging.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Muon Neutrino Reconstruction at ICARUS with Machine Learning

The ICARUS T600 LArTPC detector successfully ran for three years at the underground LNGS laboratories, providing a first sensitive search for LSND-like anomalous electron neutrino appearance in the CNGS beam. After a significant overhauling at CERN, the T600 detector has been placed in its experimental hall at Fermilab, fully commissioned, and the first events observed with full detector readout. Regular data-taking began in May 2021 with neutrinos from the Booster Neutrino Beam (BNB) and neutrinos six degrees off-axis from the Neutrinos at the Main Injector (NuMI). Modern developments in machine learning have allowed for the development of an end-to-end machine learning-based event reconstruction for ICARUS data. This reconstruction folds in 3D voxel-level feature extraction using sparse convolutional neural networks and particle clustering using graph neural networks to produce outputs suitable for physics analyses. This poster will summarize the performance of a high-purity and high-efficiency end-to-end machine learning-based selection of muon neutrinos from the BNB and highlight studies of electromagnetic shower reconstruction from a neutral pion selection.

43 PARTICLE ACCELERATORS↗

A First Search for Argon-Bound Neutron-Antineutron Oscillation using the MicroBooNE LArTPC

The use of Liquid Argon Time Projection Chambers (LArTPCs) as a detector technology in neutrino experiments has grown considerably over the past two decades. The excellent spatial and calorimetric resolution offered by LArTPCs enable precise neutrino oscillation measurements as well as beyond-Standard Model searches. One such search, which is the focus of this note, is the search for nucleus-bound neutron-antineutron (n ₋ n̄) oscillation. The n ₋ n̄ oscillation process is a baryon number violating process that produces a unique, star-like topology as a result of multiple final state pions. This unique signature is a key feature that may be used to search for this signal process. This note describes a machine learning-based analysis of MicroBooNE data, making use of a sparse convolutional neural network to search for n ₋ n̄ oscillation-like signals in MicroBooNE. While the future DUNE LArTPC can search for this signature with high sensitivity, existing MicroBooNE data can be used to demonstrate and validate methodologies that can be used as part of the DUNE search. This document presents the first-ever search for n ₋ n̄ oscillation in a LArTPC, using MicroBooNE off-beam data (data collected when the neutrino beam was not running).

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

A first search for argon-bound neutron-antineutron oscillation using the MicroBooNE LArTPC

The use of Liquid Argon Time Projection Chambers (LArTPCs) as a detector technology in neutrino experiments has grown considerably over the past two decades. The excellent spatial and calorimetric resolution offered by LArTPCs enable precise neutrino oscillation measurements as well as beyond-Standard Model searches. One such search, which is the focus of this note, is the search for nucleus-bound neutron-antineutron (n – n̄) oscillation. The n – n̄ oscillation process is a baryon number violating process that produces a unique, star-like topology as a result of multiple final state pions. This unique signature is a key feature that may be used to search for this signal process. This note describes a machine learning-based analysis of MicroBooNE data, making use of a sparse convolutional neural network to search for n – n̄ oscillation-like signals in MicroBooNE. While the future DUNE LArTPC can search for this signature with high sensitivity, existing MicroBooNE data can be used to demonstrate and validate methodologies that can be used as part of the DUNE search. This document presents the first-ever search for n – n̄ oscillation in a LArTPC, using MicroBooNE off-beam data (data collected when the neutrino beam was not running).

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Understanding Twinning and Deformation in High Entropy Alloys

A combination of high strength and high ductility has been observed in multi-principal element alloys due to twin formation attributed to low stacking fault energy (SFE). In the pursuit of low SFE alloys, a key bottleneck is the lack of understanding of the composition–SFE cor- relations that would guide tailoring SFE via alloy composition. Using density functional theory (DFT), we show that dopant radius, which have been postulated as a key descriptor for SFE in dilute alloys, does not fully explain SFE trends across different host metals. Instead, charge density is a much more central descriptor. It allows us to (1) explain contrasting SFE trends in Ni and Cu host metals due to various dopants in dilute concentrations, (2) explain the large SFE variations observed in the literature even within a given alloy composition due to the nearest neighbor environments in “model” concentrated alloys, and (3) develop a machine learning model that can be used to predict SFEs in multi-elemental alloys. This model opens a possibility to use charge density as a descriptor for predicting SFE in alloys. Furthermore, a descriptor-less machine learning (ML) model based only on charge density images extracted from density functional theory (DFT) is developed to predict stacking fault energies (SFE) in concentrated alloys. The model is based on convolutional neural networks (CNNs) as one of the promising ML techniques for dealing with complex images and data. Identification of correct descriptors is a key bottleneck to develop ML models for predicting materials properties. Often, in most ML models, textbook physical descriptors such as atomic radius, valence charge and electronegativity are used as descriptors which have limitations because these properties change in concentrated alloys when multiple elements are mixed to form a solid solution. We illustrate that, within the scope of DFT, the search for descriptors can be circumvented by electronic charge density, which is the backbone of the Kohn-Sham DFT and describes the system completely. The performance of our model is demonstrated by predicting SFE of concentrated alloys with an RMSE and R2 of 6.18 mJ/m2 and 0.87, respectively, validating the accuracy of the proposed approach.

36 MATERIALS SCIENCE↗

In-Situ Process Monitoring Evaluation and Demonstration using Advanced Characterization with Laser Powder Bed Systems

Oak Ridge National Laboratory’s (ORNL) Manufacturing Demonstration Facility (MDF) worked with EOS Group to evaluate the current in-situ sensor capabilities of an EOS M290 Laser Powder Bed Fusion machine. The M290 was fitted with a 1 Mega-Pixel (MP) grayscale visible-light camera and a 5 MP temporally integrated (TI) near-infrared (NIR) camera. One print from stainless steel (SS) 316 and two from Inconel 625 (IN625) were performed where data including in-situ imaging and a machine log file were captured. These data were subsequently analyzed using a Dynamic Multi-Scale Segmentation Convolutional Neural Network (DMSCNN) trained on user defined classes and correlated to as-printed flaws, in the form of porosity, discovered in X-Ray Computed Tomography (XCT). In Phase I, two indications were detected in-situ and spatially correlated to stochastic lack-of-fusion flaws discovered using XCT. In Phase II, using these links from in-situ signatures to XCT flaw populations, a second neural network (NN) was trained to create a Voxelized Property Prediction Model (VPPM) to predict porosity percentages within the part using only features garnered from the in-situ data from two IN625 complex geometries. The VPPM was able to accurately predict porosity values for IN625 parts with an R 2 value of 0.764.

36 MATERIALS SCIENCE↗

wa-hls4ml: A Benchmark and Surrogate Models for hls4ml Resource and Latency Estimation

As machine learning (ML) is increasingly implemented in hardware to address real-time challenges in scientific applications, the development of advanced toolchains has significantly reduced the time required to iterate on various designs. These advancements have solved major obstacles, but also exposed new challenges. For example, processes that were not previously considered bottlenecks, such as hardware synthesis, are becoming limiting factors in the rapid iteration of designs. To mitigate these emerging constraints, multipleefforts have been undertaken to develop an ML-based surrogate model that estimates resource usage of synthesized ML accelerator architectures. We introduce wa-hls4ml, a benchmark for ML accelerator resource and latency estimation, and its corresponding initial dataset of over 680 000 fully connected and convolutional neural networks, all synthesized using hls4ml and targeting Xilinx FPGAs. The benchmark evaluates the performance of resource and latency predictors against several common ML model architectures, primarily originating from scientific domains, as exemplar models, and the average performance across a subset of the dataset. Additionally, we introduce GNN- and transformer-based surrogate models that predict latency and resources for ML accelerators. We present the architecture and performance of the models and find that the models generally predict latency and resources for the 75% percentile within several percent of the synthesized resources on the synthetic test dataset.

Hawks, Benjamin G. [Fermilab]↗

Web-Based Tools for Data-Informed Remedy Optimization: Software Theory and User Guide

This report documents the development and application of two web-based decision-support tools for pump-and-treat (P&T) groundwater remediation systems: PTOLEMY (Pump-and-Treat Optimized Location Evaluation to Maximize Yields) and OPTIMA (Optimization for Pump-and-Treat Implementation, Management, & Assessment). These tools enhance remedy design and management by leveraging advanced computational methods – specifically deep learning and multi-objective optimization – within a user-friendly platform. By integrating data-driven models with established hydrogeological knowledge, PTOLEMY and OPTIMA enable more efficient evaluation of well placement and operational strategies, helping site managers balance multiple remediation objectives under complex conditions. Both tools are implemented as modules within the SOCRATES (Suite Of Comprehensive Rapid Analysis Tools for Environmental Sites) web platform, which provides data access, visualization, and analytics to support remedy optimization across sites in the U.S. Department of Energy Office of Environmental Management complex. PTOLEMY is a rapid screening module designed to identify promising locations for new extraction wells. It employs a multi-channel three-dimensional convolutional neural network (MC3D-CNN) trained on high-fidelity simulation data to predict the relative performance (in terms of contaminant mass recovery) of potential well sites. Through an interactive web interface, PTOLEMY visualizes the probability of high performance across a site, highlighting areas where an extraction well is likely to yield above-threshold contaminant removal over a multi-year period. PTOLEMY’s map-based displays and exportable results support transparent communication of screening analyses. By focusing attention on the most favorable candidate locations, the tool augments traditional engineering judgment and physics-based modeling, providing a data informed basis for subsequent detailed evaluations. OPTIMA is a multi objective optimization module designed to find wellfield layouts and operating schedules that meet various cleanup goals. It quickly evaluates thousands of candidate setups – combinations of well locations, timing, and rates – and returns a small set of best trade-off options for comparison. At its core, OPTIMA uses a U-Net-based surrogate model – a deep-learning emulator of a groundwater flow and transport simulator – to dramatically accelerate scenario evaluations. Coupling this fast surrogate with the NSGA-II (Non-dominated Sorting Genetic Algorithm II) evolutionary algorithm, OPTIMA explores a wide decision space of well locations and schedules to identify Pareto-optimal solutions that trade off key objectives (e.g., minimizing cleanup time, maximizing contaminant mass removal, and minimizing plume extent). The tool outputs a family of optimal configurations and visualizes their trade-offs (Pareto frontiers of cleanup metrics and maps of optimized well placements). Site managers can use these results to understand the range of viable strategies and to select candidate designs for more detailed verification. OPTIMA is currently under active development and not yet fully released; this guide provides early documentation to support planning and gather user feedback.

54 ENVIRONMENTAL SCIENCES↗

Describing Point Defect Topology in 2D Energy Materials Through Computer Vision

Point defects such as vacancies and impurity atoms strongly impact the performance of 2D materials. Traditional efforts often rely on manual detection, a process that is time-intensive, prone to human error, and challenging to scale. Here we leverage machine learning (ML) methods to identify and quantify vacancies within 2D transition metal carbides (Ti3C2, MXenes), aiming to expedite detection while improving accuracy. MXenes exhibit valuable defect-defined electrochemical properties, but we currently lack statistical understanding of defect topology needed to fully harness these materials. We employ a convolutional neural network for semantic segmentation of experimental MXene images, opening an opportunity to conduct a rigorous statistical study on defect hierarchy while investigating local relaxation in the lattice. We show how the integration of ML can yield fundamental insight into point defects, providing a powerful tool that will play an increasingly crucial role in the future of materials science.

2D materials↗

Describing Point Defect Topology in 2D Energy Materials Through Computer Vision

Point defects such as vacancies and impurity atoms strongly impact the performance of 2D materials. Traditional efforts often rely on manual detection, a process that is time-intensive, prone to human error, and challenging to scale. Here we leverage machine learning (ML) methods to identify and quantify vacancies within 2D transition metal carbides (Ti3C2, MXenes), aiming to expedite detection while improving accuracy. MXenes exhibit valuable defect-defined electrochemical properties, but we currently lack statistical understanding of defect topology needed to fully harness these materials. Here we employ a convolutional neural network for semantic segmentation of experimental MXene images, opening an opportunity to conduct a rigorous statistical study on defect hierarchy while investigating local relaxation in the lattice. We show how the integration of ML can yield fundamental insight into point defects, providing a powerful tool that will play an increasingly crucial role in the future of materials science. ML is often not just a matter of straightforward application, and pretrained models proved ineffective in this case. Instead, we trained our own neural network (NN) and applied data augmentation techniques and fine-tuning to the training dataset. Since labeled microscopy data is often scarce, we developed training data from a previously published wide-frame MXene image, using customized Gaussian fitting to locate atomic positions. Our trained model was then applied to a large dataset of experimental images, enabling a statistical study of defect configurations across three samples prepared with different HF etchant concentrations (5%, 9.1%, and 12.5%), as shown in Fig. 1. This also allowed us to investigate local strain around vacancies, though we find that we are limited by the precision of measurements using high-angle annular dark field (HAADF) images, as shown in Fig. 2. This study demonstrates how ML enables large-scale, quantitative analysis of atomic defects - an otherwise infeasible task with traditional methods. While our NN was specialized for Ti3C2 MXenes, the pipeline we developed provides a foundation for future ML models tailored to other materials. Ultimately, we envision embedding the NN onto the microscope to give real-time feedback to the user. To make this a reality, continued work is necessary to fully understand the NN's capabilities and limitations. This study gets one step closer to our goals of automated experimentation moving away from traditional methods of manual labeling. As ML capabilities advance, we hope to continue adapting and applying these techniques in microscopy.

2D materials↗