Search NASA⌕ Search

SEARCH · Search NASA

Results for “convolutional neural net”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Measurement of hybrid rocket solid fuel regression rate for a slab burner using deep learning

This study presents an imaging-based deep learning tool to measure the fuel regression rate in a 2D slab burner experiment for hybrid rocket fuels. The slab burner experiment is designed to verify mechanistic models of reacting boundary layer combustion in hybrid rockets by the measurement of fuel regression rates. A DSLR camera with a high intensity flash is used to capture images throughout the burn and the images are then used to find the fuel boundary to calculate the regression rate. A U-net convolutional neural network architecture is explored to segment the fuel from the experimental images. Here, a Monte-Carlo Dropout process is used to quantify the regression rate uncertainty produced from the network. The U-net computed regression rates are compared with values from other techniques from literature and show error less than 10%. An oxidizer flux dependency study is performed and shows the U-net predictions of regression rates are accurate and independent of the oxidizer flux, when the images in the training set are not over-saturated. Training with monochrome images is explored and is not successful at predicting the fuel regression rate from images with high noise. The network is superior at filtering out noise introduced by soot, pitting, and wax deposition on the chamber glass as well as the flame when compared to traditional image processing techniques, such as threshold binary conversion and spatial filtering. U-net consistently provides low error image segmentations to allow accurate computation of the regression rate of the fuel.

42 ENGINEERING↗

A machine-learning approach to measure 3D sample properties from 2D Transmission Electron Microscopy images

Transmission Electron Microscopy (TEM) is a powerful tool for the characterization of materials at the nanoscale; however, its inherent two-dimensional (2D) nature poses significant challenges to accurately measure three-dimensional (3D) properties. We introduce a supervised machine-learning model that predicts 3D structural information, such as sample thickness and curvature, from a series of conventional 2D TEM images. The model, a U-Net convolutional neural network, is trained on a large synthetic dataset generated from dynamical diffraction simulations that model TEM’s complex, nonlinear image formation, accounting for sample thickness and curvature. This physically realistic framework enables exploration of a broad parameter space impractical to sample experimentally. We demonstrate that the trained model has accurate predictions for experimental single-crystal silicon samples, achieving performance comparable to established measurement techniques. This work highlights the critical role of robust, simulation-based training in overcoming the limitations of real-world imaging artifacts and inconsistent sample geometries. By integrating machine learning with numerical simulations, we offer an efficient and scalable framework for quantitative TEM analysis, paving the way for more sophisticated 3D characterization of complex materials.

Dynamical diffraction↗

Prospects and Limitations of Predicting Fuel Ignition Properties from Low-Temperature Speciation Data

Using chemical kinetic modeling and statistical analysis, we investigate the possibility of correlating key chemical “markers”–typically small molecules–formed during very lean (φ ~ 0.001) oxidation experiments with near-stoichiometric (φ ~ 1) fuel ignition properties. One goal of this work is to evaluate the feasibility of designing a fuel-screening platform, based on small laboratory reactors that operate at low temperatures and use minimal fuel volume. Buras et al. [Combust. Flame2020,216, 472–484] have shown that convolutional neural net (CNN) fitting can be used to correlate first-stage ignition delay times (IDTs) with OH/HO2 measurements during very lean oxidation in low-T flow reactors with better than factor-of-2 accuracy. In this work, we test the limits of applying this correlation-based approach to predict the low-temperature heat release (LTHR) and total IDT, including the sensitivity of total IDT to the equivalence ratio, φ. We demonstrate that first-stage IDT can be reliably correlated with very lean oxidation measurements using compressed sensing (CS), which is simpler to implement than CNN fitting. LTHR can also be predicted via CS analysis, although the correlation quality is somewhat lower than for first-stage IDT. In contrast, the accuracy of total IDT prediction at φ = 1 is significantly lower (within a factor of 4 or worse). Furthermore, these results can be rationalized by the fact that the first-stage IDT and LTHR are primarily determined by low-temperature chemistry, whereas total IDT depends on low-, intermediate-, and high-temperature chemistry. Oxidation reactions are most important at low temperatures, and therefore, measurements of universal molecular markers of oxidation do not capture the full chemical complexity required to accurately predict the total IDT even at a single equivalence ratio. As a result, we find that φ-sensitivity of ignition delay cannot be predicted at all using solely correlation with lean low-T chemical speciation measurements.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Deep learning unresolved lensed light curves

ABSTRACT Gravitationally lensed sources may have unresolved or blended multiple images, and for time varying sources, the light curves from individual images can overlap. We use convolutional neural nets to both classify the light curves as due to unlensed, double, or quad lensed sources and fit for the time delays. Focusing on lensed supernova systems with time delays Δt ≳ 6 d, we achieve 100 per cent precision and recall in identifying the number of images and then estimating the time delays to σΔt ≈ 1 d, with a 1000× speedup relative to our previous Monte Carlo technique. This also succeeds for flux noise levels $\sim 10{{\ \rm per\ cent}}$. For Δt ∈ [2, 6] d, we obtain 94–98 per cent accuracy, depending on image configuration. We also explore using partial light curves where observations only start near maximum light, without the rise time data, and quantify the success.

79 ASTRONOMY AND ASTROPHYSICS↗

Improving Variational Autoencoders for New Physics Detection at the LHC With Normalizing Flows

We investigate how to improve new physics detection strategies exploiting variational autoencoders and normalizing flows for anomaly detection at the Large Hadron Collider. As a working example, we consider the DarkMachines challenge dataset. We show how different design choices (e.g., event representations, anomaly score definitions, network architectures) affect the result on specific benchmark new physics models. Once a baseline is established, we discuss how to improve the anomaly detection accuracy by exploiting normalizing flow layers in the latent space of the variational autoencoder.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

A comparison of deep-learning-based inpainting techniques for experimental X-ray scattering

The implementation is proposed of image inpainting techniques for the reconstruction of gaps in experimental X-ray scattering data. The proposed methods use deep learning neural network architectures, such as convolutional autoencoders, tunable U-Nets, partial convolution neural networks and mixed-scale dense networks, to reconstruct the missing information in experimental scattering images. In particular, the recovered pixel intensities are evaluated against their corresponding ground-truth values using the mean absolute error and the correlation coefficient metrics. The results demonstrate that the proposed methods achieve better performance than traditional inpainting algorithms such as biharmonic functions. Overall, tunable U-Net and mixed-scale dense network architectures achieved the best reconstruction performance among all the tested algorithms, with correlation coefficient scores greater than 0.9980.

97 MATHEMATICS AND COMPUTING↗

Development of a Full-Scale Connected U-Net for Reflectivity Inpainting in Spaceborne Radar Blind Zones

CloudSat’s Cloud Profiling Radar is a valuable tool for remotely monitoring high-latitude snowfall, but its ability to observe hydrometeor activity near the Earth’s surface is limited by a radar blind zone caused by ground clutter contamination. This study presents the development of a deeply supervised U-Net-style convolutional neural network to predict cold season reflectivity profiles within the blind zone at two Arctic locations. The network learns to predict the presence and intensity of near-surface hydrometeors by coupling latent features encoded in blind zone-aloft clouds with additional context from collocated atmospheric state variables (i.e., temperature, specific humidity, and wind speed). Results show that the U-Net predictions outperform traditional linear extrapolation methods, with low mean absolute error, a 38% higher Sørensen–Dice coefficient, and vertical reflectivity distributions 60% closer to observed values. The U-Net is also able to detect the presence of near-surface cloud with a critical success index (CSI) of 72% and cases of shallow cumuliform snowfall and virga with 18% higher CSI values compared to linear methods. An explainability analysis shows that reflectivity information throughout the scene, especially at cloud edges and at the 1.2-km blind zone threshold, along with atmospheric state variables near the tropopause, are the most significant contributors to model skill. This surface-trained generative inpainting technique has the potential to enhance current and future remote sensing precipitation missions by providing a better understanding of the nonlinear relationship between blind zone reflectivity values and the surrounding atmospheric state.

54 ENVIRONMENTAL SCIENCES↗

Automated bubble analysis of high-speed subcooled flow boiling images using U-net transfer learning and global optical flow

Capturing and analyzing the bubble dynamics is crucial to improving the understanding of boiling heat transfer mechanisms and predicting boiling heat transfer coefficient and boiling crisis. High speed video (HSV) imaging has been used for decades towards this end. Still, there is no universal approach to quantitatively analyze bubble dynamics from HSV images. In this study, we propose a data-driven post-processing approach to segment, track, and identify wall-attached vapor bubbles from HSV images of the boiling process in subcooled flow conditions. Firstly, we employ a transfer learning framework with a U-Net-based convolution neural network (CNN) architecture to detect and segment bubbles in HSV images of diverse contrast and surface texture using very little data (e.g., 10 images) for training. Then, we evaluate the trained CNN model with 100 ground-truth images, and the validation results show that the model accuracy and precision in detecting the optical footprint of bubbles are higher than 90%. Finally, we suggest a criterion to identify a condensing bubble based on the divergence of the bubble displacement, which is calculated from sequential segmented bubble images using a global optical flow code. Using this combination of machine learning and optical flow, we can identify nucleation sites and track the growth of bubbles nucleating at each site to quantify nucleation site density, nucleation frequency, and other fundamental boiling parameters. The proposed system is validated using results obtained on a special heater, which enables both infrared (IR) thermometry and HSV imaging on a metallic surface. We compare the fundamental boiling parameters obtained by the two different diagnostics. The results show good agreement. In conclusion, the difference between the measurements of nucleation site density, averaged nucleation frequency, and averaged growth time performed with the two techniques is always within ± 20% and mostly ± 10% of the values measured with IR thermometry.

42 ENGINEERING↗

A Survey: Handling Irregularities in Neural Network Acceleration with FPGAs

In the last decade, Artificial Intelligence (AI) through Deep Neural Networks (DNNs) has penetrated virtually every aspect of science, technology, and business. Many types of DNNs have been and continue to be developed, including Convolutional Neural Networks (CNNs), Recurrent Neural Net- works (RNNs), and Graph Neural Networks (GNNs). The overall problem for all of these Neural Networks (NNs) is that their target applications generally pose stringent constraints on latency and throughput, while also having strict accuracy requirements. There have been many previous efforts in creating hardware to accelerate NNs. The problem designers face is that optimal NN models typically have significant irregularities, making them hardware-unfriendly. In this paper, we first define the problems in NN acceleration by characterizing common irregularities in NN processing into 4 types; then we summarize the existing works that handle the four types of irregularities efficiently using hardware, especially FPGAs; finally, we provide a new vision of next-generation FPGA-based NN acceleration: that the emerging heterogeneity in the next-generation FPGAs is the key to achieving higher performance.

Geng, Tong↗

Towards physics-inspired data-driven weather forecasting: integrating data assimilation with a deep spatial-transformer-based U-NET in a case study with ERA5

Abstract. There is growing interest in data-driven weather prediction (DDWP), e.g., using convolutional neural networks such as U-NET that are trained on data from models or reanalysis. Here, we propose three components, inspired by physics, to integrate with commonly used DDWP models in order to improve their forecast accuracy. These components are (1) a deep spatial transformer added to the latent space of U-NET to capture rotation and scaling transformation in the latent space for spatiotemporal data, (2) a data-assimilation (DA) algorithm to ingest noisy observations and improve the initial conditions for next forecasts, and (3) a multi-time-step algorithm, which combines forecasts from DDWP models with different time steps through DA, improving the accuracy of forecasts at short intervals. To show the benefit and feasibility of each component, we use geopotential height at 500 hPa (Z500) from ERA5 reanalysis and examine the short-term forecast accuracy of specific setups of the DDWP framework. Results show that the spatial-transformer-based U-NET (U-STN) clearly outperforms the U-NET, e.g., improving the forecast skill by 45 %. Using a sigma-point ensemble Kalman (SPEnKF) algorithm for DA and U-STN as the forward model, we show that stable, accurate DA cycles are achieved even with high observation noise. This DDWP+DA framework substantially benefits from large (O(1000)) ensembles that are inexpensively generated with the data-driven forward model in each DA cycle. The multi-time-step DDWP+DA framework also shows promise; for example, it reduces the average error by factors of 2–3. These results show the benefits and feasibility of these three components, which are flexible and can be used in a variety of DDWP setups. Furthermore, while here we focus on weather forecasting, the three components can be readily adopted for other parts of the Earth system, such as ocean and land, for which there is a rapid growth of data and need for forecast and assimilation.

54 ENVIRONMENTAL SCIENCES↗

A Physics-Constrained Deep Learning Model for Simulating Multiphase Flow in 3D Heterogeneous Porous Media

Physics-based simulators for multiphase flow in porous media emulate nonlinear processes with coupled physics, and usually require extensive computational resources for software development, maintenance and simulation execution. As a result, a huge demand exists for fast modeling of coupled processes in a wide range of subsurface applications including geological sequestration, hydrocarbon recovery and geothermal energy extraction. In this work, an efficient physics-constrained deep learning model is developed for solving multiphase flow in 3-Dimensional (3D) heterogeneous porous media. The model fully leverages the spatial topology predictive capability of convolutional neural networks, specifically U-Net with successive contracting and expansive steps, and is coupled with an efficient continuity-based smoother to predict flow responses that need spatial continuity. Furthermore, the transient regions are penalized to steer the training process such that the model can accurately capture flow in these regions. The model takes inputs including properties of porous media, fluid properties and well controls, and predicts the temporal-spatial evolution of the state variables (pressure and saturation). While maintaining the continuity of fluid flow, the 3D spatial domain is decomposed into 2D images for reducing training cost, and the decomposition results in an increased number of training data samples and better training efficiency. Additionally, a surrogate model is separately constructed as a postprocessor to calculate well flow rate based on the predictions of state variables from the deep learning model. We use the example of CO 2 injection into saline aquifers, and apply the physics-constrained deep learning model that is trained from physics-based simulation data and emulates the physics process. The model performs prediction with a speedup of ~ 1400 times compared to physics-based simulations, and the average temporal errors of predicted pressure and saturation plumes are 0.27% and 0.099% respectively. Furthermore, water production rate is efficiently predicted by a surrogate model for well flow rate, with a mean error less than 5%. Therefore, with its unique scheme to cope with the fidelity in fluid flow in porous media, the physics-constrained deep learning model can become an efficient predictive model for computationally demanding inverse problems or other coupled processes.

58 GEOSCIENCES↗

Toward ultra-efficient high-fidelity predictions of wind turbine wakes: Augmenting the accuracy of engineering models with machine learning

This study proposes a novel machine learning (ML) methodology for the efficient and cost-effective prediction of high-fidelity three-dimensional velocity fields in the wake of utility-scale turbines. The model consists of an autoencoder convolutional neural network with U-Net skipped connections, fine-tuned using high-fidelity data from large-eddy simulations (LES). The trained model takes the low-fidelity velocity field cost-effectively generated from the analytical engineering wake model as input and produces the high-fidelity velocity fields. The accuracy of the proposed ML model is demonstrated in a utility-scale wind farm for which datasets of wake flow fields were previously generated using LES under various wind speeds, wind directions, and yaw angles. Comparing the ML model results with those of LES, the ML model was shown to reduce the error in the prediction from 20% obtained from the Gauss Curl hybrid (GCH) model to less than 5%. In addition, the ML model captured the non-symmetric wake deflection observed for opposing yaw angles for wake steering cases, demonstrating a greater accuracy than the GCH model. The computational cost of the ML model is on par with that of the analytical wake model while generating numerical outcomes nearly as accurate as those of the high-fidelity LES.

Mechanics↗

Structural Health Monitoring of Microreactor Safety Systems Using Convolutional Neural Networks

Microreactors, a class of modular reactors with net power output of less than 20 MWth, have innovative applications in nuclear and nonnuclear industries due to their portability, reliability, resilience, and high capacity factors. In order to operate microreactors on a wider scale, it is essential to bring down maintenance life-cycle costs while ensuring the integrity of operating such systems. Autonomous operations in microreactors using augmented digital-twin (DT) technology can serve as a cost-effective solution by increasing awareness about the system’s health. Structural health monitoring (SHM) is a key component of nuclear DT frameworks. Artificial neural networks can be beneficial to detect degradation in the nuclear safety systems, such as piping equipment systems, by monitoring the sensor data obtained from the plant and its corresponding structures, systems and components. In this report, an SHM methodology is presented which uses convolutional neural networks to determine degraded locations and their corresponding degradation-severity levels at various locations of nuclear piping equipment systems. A simple pipe system, subjected to seismic loads, is selected to design the post-hazard SHM framework. The effectiveness of the proposed SHM methodology is demonstrated by obtaining high accuracy in detecting degraded locations as well as the severity levels.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

Image-Based Fracture Surface Defect Characterization Methods for Additively Manufactured Ti-6Al-4V Tested in Fatigue

Abstract Fatigue initiation in additively manufactured samples/parts often occurs at processed-induced defects such as lack-of-fusion (LoF), keyhole, or other morphological/microstructural defects that have unique characteristics and measurable qualities. Attempts at identifying and minimizing such defects have utilized optimized processing conditions along with in situ and ex situ characterization that includes metallography and/or X-ray computed tomography (XCT). This paper highlights the benefits of using fracture surface analyses to detect and quantify defects that may not be detected by metallography/XCT due to sectioning and resolution limits. In addition to using manual quantification of fatigue initiating LoF and keyhole defects on fracture surfaces, image-based machine learning using convolutional neural networks such as U-Net were also used to automate the process. Statistical analyses were used to identify the extreme cases of defects that initiated and accelerated fatigue and to model the distribution of defect size and shape characteristics to distinguish the type of defect. Initial results show agreement between trained machine learning models and ground truth data in defect segmentation, and the distributions of defect characteristics are distinguishable to particular process-induced defect types.

Materials Science↗

Mesoscale Cellular Convection Detection and Classification Using Convolutional Neural Networks: Insights From Long-Term Observations at ARM Eastern North Atlantic Site

Marine boundary layer clouds are crucial in Earth's climate system. They frequently manifest as closed or open cell mesoscale cellular convection (MCC). MCC clouds are challenging to represent accurately in current climate models, highlighting the need for detailed observational data sets and in-depth analyses. This study utilizes over 8 years of observations from the U.S. Department of Energy (DOE) Atmospheric Radiation Measurement (ARM) User Facility Eastern North Atlantic (ENA) site at Graciosa Island, Azores, to investigate these clouds. We first apply a convolutional neural network with a U-Net architecture to classify open and closed cells, marking the first application of such an approach for automatically detecting MCC patterns from ground-based radar measurements. This method addresses some observational gaps in satellite data related to low temporal resolution, nighttime challenges, and limited vertical structure capture. The analysis of the MCC cases shows clear differences between closed and open MCCs: Closed MCC clouds are characterized by lower cloud tops and bases, shallower cloud geometrical depth, weaker horizontal wind speeds, stronger atmospheric stability, and a more homogeneous liquid water path than open MCCs. Finally, we demonstrate two potential applications of our radar-based MCC classifications: (a) facilitating the investigation of aerosol-cloud interactions and (b) exploring meteorological factors along with MCC's evolution by integrating satellite imagery and back-trajectory analysis. The identified MCC cases offer a valuable resource for the scientific community to study MCC processes further and improve climate model accuracy.

54 ENVIRONMENTAL SCIENCES↗

Deep Learning Image Segmentation for Atmospheric Rivers

Abstract The identification of atmospheric rivers (ARs) is crucial for weather and climate predictions as they are often associated with severe storm systems and extreme precipitation, which can cause large impacts on society. This study presents a deep learning model, termed ARDetect, for image segmentation of ARs using ERA5 data from 1960 to 2020 with labels obtained from the TempestExtremes tracking algorithm. ARDetect is a convolutional neural network (CNN)-based U-Net model, with its structure having been optimized using automatic hyperparameter tuning. Inputs to ARDetect were selected to be the integrated water vapor transport (IVT) and total column water (TCW) fields, as well as the AR mask from TempestExtremes from the previous time step to the one being considered. ARDetect achieved a mean intersection-over-union (mIoU) rate of 89.04% for ARs, indicating its high accuracy in identifying these weather patterns and a superior performance than most deep learning–based models for AR detection. In addition, ARDetect can be executed faster than the TempestExtremes method (seconds vs minutes) for the same period. This provides a significant benefit for online AR detection, especially for high-resolution global models. An ensemble of 10 models, each trained on the same dataset but having different starting weights, was used to further improve on the performance produced by ARDetect, thus demonstrating the importance of model diversity in improving performance. ARDetect provides an effective and fast deep learning–based model for researchers and weather forecasters to better detect and understand ARs, which have significant impacts on weather-related events such as floods and droughts.

Galea, Daniel↗

Reconstructing the exit wave of 2D materials in high-resolution transmission electron microscopy using machine learning

Reconstruction of the exit wave function is an important route to interpreting high-resolution transmission electron microscopy (HRTEM) images. Here we demonstrate that convolutional neural networks can be used to reconstruct the exit wave from a short focal series of HRTEM images, with a fidelity comparable to conventional exit wave reconstruction. We use a fully convolutional neural network based on the U-Net architecture, and demonstrate that we can train it on simulated exit waves and simulated HRTEM images of graphene-supported molybdenum disulphide (an industrial desulfurization catalyst). We then apply the trained network to analyse experimentally obtained images from similar samples, and obtain exit waves that clearly show the atomically resolved structure of both the MoS 2 nanoparticles and the graphene support. We also show that it is possible to successfully train the neural networks to reconstruct exit waves for 3400 different two-dimensional materials taken from the Computational 2D Materials Database of known and proposed two-dimensional materials.

2D materials↗