Search NASA⌕ Search

SEARCH · Search NASA

Results for “Neural network models”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

Database and deep-learning scalability of anharmonic phonon properties by automated brute-force first-principles calculations

Understanding the anharmonic phonon properties of crystal compounds—such as phonon lifetimes and thermal conductivities—is essential for investigating and optimizing their thermal transport behaviors. These properties also impact optical, electronic, and magnetic characteristics through interactions between phonons and other quasiparticles and fields. In this study, we develop an automated first-principles workflow to calculate anharmonic phonon properties and build a comprehensive database encompassing more than 6500 inorganic compounds. Utilizing this dataset, we train a graph neural network model to predict thermal conductivity values and spectra from structural parameters, demonstrating a scaling law in which prediction accuracy improves with increasing training data size. High-throughput screening with the model enables the identification of materials exhibiting extreme thermal conductivities—both high and low. The resulting database offers valuable insights into the anharmonic behavior of phonons, thereby accelerating the design and development of advanced functional materials.

Ohnishi, Masato [University of Tokyo (Japan); Inst↗

Deep learning forecasts the spatiotemporal evolution of fluid-induced microearthquakes

Microearthquakes generated by subsurface fluid injection record the evolving stress state and permeability of reservoirs. Forecasting their spatiotemporal evolution is therefore critical for applications such as enhanced geothermal systems, carbon dioxide sequestration and other geoengineering applications. Here we propose a transformer neural network model that ingests hydraulic stimulation history and prior microearthquake observations to forecast four key quantities: cumulative microearthquake count, cumulative logarithmic seismic moment, and the 50th- and 95th-percentile extents of the microearthquake cloud. Applied to the EGS Collab Experiment 1 dataset, the model achieves R2 > 0.98 for the 1-s forecast horizon and R2 > 0.88 for the 15-s forecast horizon across all targets, and supplies uncertainty estimates through a learned standard deviation term. These accurate, uncertainty-quantified forecasts enable real-time inference of fracture propagation and permeability evolution, demonstrating the strong potential of deep-learning approaches to improve seismic-risk assessment and guide mitigation strategies in future fluid-injection operations.

Chung, Jaehong↗

Deep probabilistic direction prediction in 3D with applications to directional dark matter detectors

Abstract We present the first method to probabilistically predict 3D direction in a deep neural network model. The probabilistic predictions are modeled as a heteroscedastic von Mises-Fisher distribution on the sphere S 2 , giving a simple way to quantify aleatoric uncertainty. This approach generalizes the cosine distance loss which is a special case of our loss function when the uncertainty is assumed to be uniform across samples. We develop approximations required to make the likelihood function and gradient calculations stable. The method is applied to the task of predicting the 3D directions of electrons, the most complex signal in a class of experimental particle physics detectors designed to demonstrate the particle nature of dark matter and study solar neutrinos. Using simulated Monte Carlo data, the initial direction of recoiling electrons is inferred from their tortuous trajectories, as captured by the 3D detectors. For 40 keV electrons in a 70% He 30% CO 2 gas mixture at STP, the new approach achieves a mean cosine distance of 0.104 (26 ∘ ) compared to 0.556 (64 ∘ ) achieved by a non-machine learning algorithm. We show that the model is well-calibrated and accuracy can be increased further by removing samples with high predicted uncertainty. This advancement in probabilistic 3D directional learning could increase the sensitivity of directional dark matter detectors.

Computer Science↗

Digital Twin-Enabled Adaptive Control for Hydroelectric Systems: Turbine Governor and Voltage Regulation

This paper presents a comprehensive digital twin (DT) framework for hydroelectric systems that enables adaptive control of turbine governors and excitation systems without requiring detailed manufacturer specifications. The proposed framework integrates neural network-based system identification with stabilizing adaptive control laws for the installed turbine controller and middle-branch adaptive tuning for the installed voltage regulator. Using real operational data from Unit C-8 at Rocky Reach Dam (1,349 MW capacity), highfidelity neural network models are developed to capture turbine and generator dynamics without requiring detailed manufacturer specifications. The DT enables safe controller synthesis and validation in simulation before deployment. For turbine control, the proposed method achieves a 79.9% mean square error (MSE) reduction compared with that of an optimal controller. For voltage regulation, the adaptive excitation controller achieves approximately 42.6% MSE reduction while preserving installed protection logic. The results demonstrate that DT technology provides a practical pathway for modernizing hydropower control systems with minimal operational disruption.

Gui, Yonghao [ORNL] (ORCID:0000000250435534)↗

Three and Two Phase Rotating Field Inductive Couplers for Wireless Power Transfer with One Phase per Layer Windings

Multiphase inductive wireless charging coils have been proposed recently to improve coupler surface power density, reduce component stress and size, and provide near-constant power delivery to charge mobile electric systems. Several aspects for the fundamental characterization of multiphase coils are explored up to six phases including approximate mutual inductance with size and turn variation, induced voltage, and output power estimation. The relative component stress and size of passive components for resonant operation are compared between the multiphase variants. A combination of an experimentally validated 3D electromagnetic finite element analysis (FEA) and power electronic co-simulations are used to validate the estimated quantities approximated with a mixture of analytical equations and an artificial neural network model for mutual inductance. A novel three-phase transmitter, two-phase receiver coil pair is also proposed for electric vehicle charging to reduce the number of connections and compensation complexity on the vehicle-side with improved power output compared to a two-phase configuration.

Lewis, Donovin D. [University of Kentucky]↗

Evaluation of GlassNet for physics-informed machine learning of glass stability and glass-forming ability

Glassy materials form the basis of many modern applications, including nuclear waste immobilization, touch-screen displays, and optical fibers, and also hold great potential for future medical and environmental applications. However, their structural complexity and large composition space make design and optimization challenging for certain applications. Of particular importance for glass processing and design is an estimate of a given composition's glass-forming ability (GFA). However, there remain many open questions regarding the underlying physical mechanisms of glass formation, especially in oxide glasses. It is apparent that a proxy for GFA would be highly useful in glass processing and design, but identifying such a surrogate property has proven itself to be difficult. While glass stability (GS) parameters have historically been used as a GFA surrogate, recent research has demonstrated that most of these parameters are not accurate predictors of the GFA of oxide glasses. Here, in this work, we explore the application of an open-source pre-trained neural network model, GlassNet, that can predict the characteristic temperatures necessary to compute GS with reasonable performance and assess the feasibility of using these physics-informed machine learning (PIML)-predicted GS parameters to estimate GFA. In doing so, we track the uncertainties at each step of the computation—from the original ML prediction errors to the compounding of errors during GS estimation, and finally to the final estimation of GFA. While GlassNet exhibits reasonable accuracy on all individual properties, we observe a large compounding of error in the combination of these individual predictions for the PIML prediction of GS, finding that random forest models offer similar accuracy to GlassNet. We also break down the performance of GlassNet on different glass families and find that the error in GS prediction is correlated with the error in crystallization peak temperature prediction. Lastly, we utilize this finding to assess the relationship between top-performing GS parameters and GFA for two ternary glass systems: sodium borosilicate and sodium iron phosphate glasses. We conclude that to obtain true ML predictive capability of GFA, significantly more data needs to be collected.

36 MATERIALS SCIENCE↗

On the Training and Generalization of Deep Operator Networks

Here, we present a novel training method for deep operator networks (DeepONets), one of the most popular neural network models for operators. DeepONets are constructed by two subnetworks, namely the branch and trunk networks. Typically, the two subnetworks are trained simultaneously, which amounts to solving a complex optimization problem in a high dimensional space. In addition, the nonconvex and nonlinear nature makes training very challenging. To tackle such a challenge, we propose a two-step training method that trains the trunk network first and then sequentially trains the branch network. The core mechanism is motivated by the divide-and-conquer paradigm and is the decomposition of the entire complex training task into two subtasks with reduced complexity. Therein the Gram–Schmidt orthonormalization process is introduced which significantly improves stability and generalization ability. On the theoretical side, we establish a generalization error estimate in terms of the number of training data, the width of DeepONets, and the number of input and output sensors. Numerical examples are presented to demonstrate the effectiveness of the two-step training method, including Darcy flow in heterogeneous porous media.

deep operator networks↗

ORNL-Chi-Geometry

Library for benchmarking neural network models on classification tasks for chirality detection in atomistic structures of organic compounds.

Weaver, Rylie [Oak Ridge National Laboratory (ORNL↗

fast3

Code to train neural network models for binding affinity prediction

Kim, Hyojin [Lawrence Livermore National Laborator↗

Nonlinear behavior of urban flood peaks in the U.S. Mid-Atlantic region

Urbanization, i.e., increasing urban development areas in a watershed, is well known as a major cause of increasing flood magnitudes. This study analyzes the observed flood peaks at 262 watersheds in the U.S. Mid-Atlantic region with varying levels of urban development and free from reservoir impacts. Our analysis reveals an interesting, V-shaped nonlinear behavior: flood peaks first decrease and then increase with increasing percentage of urban development area at the watershed scale (PDAW), with the shift occurring at a PDAW threshold of around 10%. Regression analyses suggest that the V-shaped pattern primarily results from complex interactions among climate conditions (e.g., storm-event rainfall) and landscape properties (e.g., elevation, distance to the coast). A neural network model was then developed to capture such interactions, satisfactorily reproducing the V-shaped pattern with an R-squared value of 0.58, RMSE of 6.72 mm/day, and NSE of 0.55. These findings highlight the need to account for nonlinear dynamics in flood prediction and management in the coastal environment.

flood peaks↗

CO2 Plume Imaging with Accelerated Deep Learning-based Data Assimilation Considering Multiple Realizations: Application to the Illinois Basin-Decatur Carbon Sequestration Project

We propose a fast and efficient deep learning workflow for near real-time data assimilation, forecasting and visualization of CO2 plume evolution in saline aquifer and demonstrate its application at a field site. Unlike the previous work, this study incorporates the impact of spatial heterogeneity using multiple realizations. In the proposed workflow, a neural network model utilizes available monitoring data such as downhole pressure measurements as input and predicts the propagating pressure ‘front’ using the diffusive time of flight (DTOF) map which is considered as representative reservoir image of the flow field. The DTOF is the arrival time of pressure front propagation, which can be computed by the Fast Marching Method rapidly without flow simulations. Reservoir model calibration can be implemented by selecting the training data samples that describe the predicted DTOF map based on observed data. The power and efficacy of our workflow is demonstrated by application to the Illinois Basin-Decatur Project.

CO2 plume imaging↗

Machine Learning Vacancy Formation Energy in Nickel-Based Superalloys

Thermal vacancies play a critical role in high-temperature Ni-based superalloys and influence various properties such as creep resistance, oxidation, etc. This study systematically investigates the impact of commonly used transition metals (Cr, Co, Fe), refractory metals (Nb, Ta, Mo, W) and other elements (Al, Cu, Ti, Mn) on the thermodynamic stability of 36 binary, 20 ternary, 11 quaternary, 9 quinary, and 3 senary FCC Ni-based alloys covering various elemental combinations. Density functional theory-based studies on Ni-X binary alloys show that higher concentrations of Cr, Nb, Ta, Al, and Ti introduce significant lattice distortions and broaden the distribution of vacancy formation energies (standard deviation up to 0.15 eV). These elements partially donate electrons, reducing their self-consistent chemical potentials relative to single-element reference values and lowering vacancy formation energies, while Co, Fe, Mo, and W show lower charge localization. These trends extend from 3-6 element alloys, where Cr, Nb, and Ta-rich compositions have low-energy states (~0.5 eV) that increase vacancy concentrations. Finally, graph neural network models are developed to screen over 5000 virtual alloys. Eleven leading compositions are identified with mean vacancy formation energy higher than 1.75 eV and vacancy concentration ~2 orders of magnitude lower than pure Ni at 1000 K. These results provide valuable guidelines to achieve controlled defect engineering in structural alloys.

DFT↗

Neural nets on the MPP

The Massively Parallel Processor (MPP) is an ideal machine for computer experiments with simulated neural nets as well as more general cellular automata. Experiments using the MPP with a formal model neural network are described. The results on problem mapping and computational efficiency apply equally well to the neural nets of Hopfield, Hinton et al., and Geman and Geman.

Hastings, Harold M.↗

Electronic Neural Networks

Memory based on neural network models content-addressable and fault-tolerant. System includes electronic equivalent of synaptic network; particular, matrix of programmable binary switching elements over which data distributed. Switches programmed in parallel by outputs of serial-input/parallel-output shift registers. Input and output terminals of bank of high-gain nonlinear amplifiers connected in nonlinear-feedback configuration by switches and by memory-prompting shift registers.

Lambe, John↗

Electronic hardware implementations of neutral networks

This paper examines some of the present work on the development of electronic neural network hardware. In particular, the investigations currently under way at JPL on neural network hardware implementations based on custom VLSI technology, novel thin film materials, and an analog-digital hybrid architecture are reviewed. The availability of such hardware will greatly benefit and enhance the present intense research effort on the potential computational capabilities of highly parallel systems based on neural network models.

Thakoor, A. P.↗

Binary synaptic connections based on memory switching in a-Si:H for artificial neural networks

A scheme for nonvolatile associative electronic memory storage with high information storage density is proposed which is based on neural network models and which uses a matrix of two-terminal passive interconnections (synapses). It is noted that the massive parallelism in the architecture would require the ON state of a synaptic connection to be unusually weak (highly resistive). Memory switching using a-Si:H along with ballast resistors patterned from amorphous Ge-metal alloys is investigated for a binary programmable read only memory matrix. The fabrication of a 1600 synapse test array of uniform connection strengths and a-Si:H switching elements is discussed.

Thakoor, A. P.↗