Search NASA⌕ Search

SEARCH · Search NASA

Results for “predictive”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 217 records · Page 12

Predicting Partial Atomic Charges in Metal–Organic Frameworks: An Extension to Ionic MOFs

Molecular simulation is an invaluable tool to predict and understand the usage of metal–organic frameworks (MOFs) for gas storage and separation applications. Accurate partial atomic charges, commonly obtained from density functional theory (DFT) calculations, are often required to model the electrostatic interactions between the MOF and adsorbates, especially when the adsorbates have dipole or quadrupole moments, such as water and CO 2 . Machine learning (ML) models have been previously employed to predict partial charges and avoid the computational cost associated with DFT calculations. However, previous ML models suffer from small training data sets, which limit their scope of application. In this work, we introduce two novel machine learning models, PACMOF2-neutral and PACMOF2-ionic, aimed at predicting the density-derived electrostatic and chemical (DDEC6) partial atomic charges for both neutral and ionic MOFs. These models not only yield DFT-level accuracy at a fraction of the computational cost but also demonstrate a remarkable improvement in prediction of adsorption, as validated with grand canonical Monte Carlo simulations. Furthermore, the robustness and fast computational time of the PACMOF2 models, along with their transferability to other porous materials such as covalent organic frameworks and zeolites, underscores their potential in high-throughput screening of MOFs for diverse applications.

36 MATERIALS SCIENCE↗

Chemistry Informed Machine Learning-Based Heat Capacity Prediction of Solid Mixed Oxides

Knowing heat capacity is crucial for modeling temperature changes with the absorption and release of heat and for calculating the thermal energy storage capacity of oxide mixtures with energy applications. The current prediction methods (ab initio simulations, computational thermodynamics, and the Neumann–Kopp rule) are computationally expensive, not fully generalizable, or inaccurate. Machine learning has the potential of being fast, accurate, and generalizable, but it has been scarcely used to predict mixture properties, particularly for mixed oxides. Here, we demonstrate a method for the generalizable prediction of heat capacity of solid oxide pseudobinary mixtures using heat capacity data obtained from computational thermodynamics and descriptors from ab initio databases. Further, models trained through this workflow achieved an error (mean absolute error of 0.43 J mol –1 K –1 ) lower than the uncertainty in differential scanning calorimetry measurements, and the workflow can be extended to predict other properties derived from the Gibbs free energy and for higher-order oxide mixtures.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Prediction of Specificity of α-Conotoxins to Subtypes of Human Nicotinic Acetylcholine Receptors with Semi-supervised Machine Learning

Conotoxins are a family of highly toxic neurotoxins composed of cysteine-rich peptides produced by marine cone snails. The most lethal cone snail species to humans is Conus geographus, with fatality rates of up to ∼65% from a single sting, which is caused mostly by the activity of α-conotoxins against human nicotinic acetylcholine receptors (nAChRs). While sequence-based machine learning (ML) classifiers have been trained to identify targets of conotoxins binding voltage-gated ion channels, no ML model has been built to predict the subtype-specific nAChR targets of α-conotoxins. Here, we trained an ML model in a semi-supervised manner to predict the specificity of α-conotoxin binding toward different human nAChR subtypes to overcome the challenge of limited data in subtype-specific nAChR targets of α-conotoxins and the issue that one α-conotoxin can bind multiple nAChR subtypes with high selectivity. We considered additional features of sequences of α-conotoxins in training our ML model, including the secondary structure propensities and electrostatic properties, which resulted in better prediction capability for the ML model. Notably, we identify that most α-conotoxins bind to α3β2, α1γδ, and α7 subtypes of human nAChRs. Our findings from this study provide a framework for predicting targets of various kinds of toxins.

59 BASIC BIOLOGICAL SCIENCES↗

Development of Data-Driven Models for Performance Prediction and Chemical Dosing of a Full-Scale Controlled Phosphorus Precipitation Reactor

This study evaluated the use of data-driven models to improve control of a struvite precipitation reactor that removes phosphorus from wastewater while producing a fertilizer product. The researchers developed predictive models for influent orthophosphate concentration, effluent orthophosphate concentration, and phosphorus removal using operational data from a full-scale MagPrex™ reactor at a water resource recovery facility in Denver, Colorado. Model predictions were used to recommend magnesium chloride dosing adjustments needed to achieve a target effluent phosphorus concentration. Several machine learning approaches were tested, with ridge regression providing the best predictions for influent orthophosphate concentration and phosphorus removal, and XGBoost providing the best predictions for effluent orthophosphate concentration. Simulation results indicated that the decision-support approach could correctly identify dosing adjustments in most cases and reduce chemical use. Full-scale implementation achieved lower accuracy due to changing operating conditions and limited historical data in some operating ranges. Here, the results demonstrate the potential of data-driven tools to support phosphorus recovery process control while also identifying practical limitations that affect deployment in full-scale systems.

42 ENGINEERING↗

High-Throughput Screening and Accurate Prediction of Ionic Liquid Viscosities Using Interpretable Machine Learning

Ionic liquids (ILs) are a novel group of green solvents with great promise for various industrial applications, including carbon capture and lignocellulosic biomass deconstruction. However, the use of ILs at the industrial scale remains challenging due to their high viscosities at ambient temperatures. To develop ILs with lower viscosities, a systematic study of their quantitative structure–property relationship (QSPR) is desirable. Here, we developed four machine learning (ML) models to predict viscosity at various temperature and pressure ranges, trained over a wide range of ILs consisting of various cationic and anionic families. ML methods including two-factor polynomial regression (two-factor PR), support vector regression (SVR), feed-forward neural networks (FFNN), and categorical boosting (CATBoost) were developed based on features that have proven useful in previous ML studies: COSMO-RS (conductor-like screening model for real solvents)-derived surface screening charge densities (sigma profiles). FFNN and CATBoost were the most accurate in predicting IL viscosities with lower average absolute relative deviation and higher R2 values on the test set. Tanimoto similarity scores were calculated to characterize the chemical space and structural similarity of the investigated ions. Furthermore, SHapley Additive exPlanation (SHAP) analysis was employed to interpret the ML results. Temperature, the polar area of ILs, and the nonpolar regions of ions are key features that influence the viscosity predictions. Importantly, the IL viscosity prediction here is the most accurate reported to date.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Theoretical Prediction and Experimental Verification of IrO x Supported on Titanium Nitride for Acidic Oxygen Evolution Reaction

Reducing iridium (Ir) catalyst loading for acidic oxygen evolution reaction (OER) is a critical strategy for large-scale hydrogen production via proton exchange membrane (PEM) water electrolysis. However, simultaneously achieving high activity, long-term stability, and reduced material cost remains challenging. To address this challenge, we develop a frame-work by combining density functional theory (DFT) prediction using model surfaces and proof-of-concept experimental ver-ification using thin films and nanoparticles. DFT results predict that oxidized Ir monolayers over titanium nitride (IrO x /TiN) should display higher OER activity than IrO x while reducing Ir loading. Further, this prediction is verified by depositing Ir monolayers over TiN thin films via physical vapor deposition. The promising thin film results are then extended to commercially viable powder IrO x /TiN catalysts, which demonstrate a lower overpotential and higher mass activity than commercial IrO 2 , and a long-term stability of 250 hours to maintain a current density of 10 mA cm -2 . The superior OER performance of IrO x /TiN is further confirmed using proton exchange membrane water electrolyzer (PEMWE), which shows a lower cell voltage than commercial IrO 2 to achieve a current density of 1 A cm -2 . Both DFT and in situ X-ray absorption spectroscopy reveal that the high OER performance of IrO x /TiN strongly depends on the IrO x - TiN interaction via direct Ir-Ti bonding. This study highlights the importance of close interaction between theoretical prediction based on mechanistic understanding and experimental verification based on thin film model catalysts to facilitate the development of more practical powder IrO x /TiN catalysts with high activity and stability for acidic OER.

08 HYDROGEN↗

Advancing the Prediction of MS/MS Spectra Using Machine Learning

Tandem mass spectrometry (MS/MS) is an important tool for the identification of small molecules and metabolites where resultant spectra are most commonly identified by matching them with spectra in MS/MS reference libraries. While popular, this strategy is limited by the contents of existing reference libraries. In response to this limitation, various methods are being developed for the in silico generation of spectra to augment existing libraries. Recently, machine learning and deep learning techniques have been applied to predict spectra with greater speed and accuracy. Here, in this work, we investigate the challenges these algorithms face in achieving fast and accurate predictions on a wide range of small molecules. The challenges are often amplified by the use of generic machine learning benchmarking tactics, which lead to misleading accuracy scores. Curating data sets, only predicting spectra for sufficiently high collision energies, and working more closely with experimental mass spectrometrists are recommended strategies to improve overall prediction accuracy in this nuanced field.

47 OTHER INSTRUMENTATION↗

Exploring the Relative Importance of the MJO and ENSO to North Pacific Subseasonal Predictability

Abstract Here we explore the relative contribution of the Madden‐Julian Oscillation (MJO) and El Niño Southern Oscillation (ENSO) to midlatitude subseasonal predictive skill of upper atmospheric circulation over the North Pacific, using an inherently interpretable neural network applied to pre‐industrial control runs of the Community Earth System Model version 2. We find that this interpretable network generally favors the state of ENSO, rather than the MJO, to make correct predictions on a range of subseasonal lead times and predictand averaging windows. Moreover, the predictability of positive circulation anomalies over the North Pacific is comparatively lower than that of their negative counterparts, especially evident when the ENSO state is important. However, when ENSO is in a neutral state, our findings indicate that the MJO provides some predictive information, particularly for positive anomalies. We identify three distinct evolutions of these MJO states, offering fresh insights into opportune forecasting windows for MJO teleconnections.

58 GEOSCIENCES↗

Convolutional Neural Networks Trained on Internal Variability Predict Forced Response of TOA Radiation by Learning the Pattern Effect

Abstract Predicting forced, long‐term radiative feedbacks from internal climate variability has been a decades‐long quest in climate science. We train a convolutional neural network (CNN) to predict annual‐ and global‐mean top of the atmosphere radiation anomalies from time‐varying maps of near‐surface temperature in climate models. Trained on internal variability alone, the nonlinear CNN can predict radiation under strong climate change, outperforms a regularized linear regression approach, and works within and across different climate models. We show with explainable artificial intelligence methods that the CNN draws predictive skill from physically meaningful regions but at much smaller spatial scales than currently assumed.

Rugenstein, Maria [Colorado State University Fort ↗

Data‐Driven Predictions of Peak Warming Under Rapid Decarbonization

Abstract The severe impacts associated with recent record‐setting annual global temperatures elevate the need to accurately predict the hottest conditions that could occur even if the most ambitious decarbonization goals are achieved. We use convolutional neural networks (CNNs) to predict peak global warming from recent observed temperature maps and future cumulative CO 2 emissions. For the SSP1‐1.9 decarbonization scenario there is >99% probability that mean global warming exceeds 1.5°C, approximately even odds that it reaches 2°C, and ∼90% probability that the hottest year globally exceeds 2023 by at least 0.5°C. Further, for the SSP2‐4.5 decarbonization scenario, there is >90% probability that the hottest annual global temperature anomaly is twice the 2023 anomaly. That our framework makes highly accurate out‐of‐sample predictions of the hottest historical year provides confidence in the predicted future probabilities, suggesting substantial risks from the extreme local conditions that are likely to result from globally hot years during rapid decarbonization.

Diffenbaugh, Noah S. [Doerr School of Sustainabili↗

Photoinduced hydrogen dissociation in thymine predicted by coupled cluster theory

The fate of thymine upon excitation by ultraviolet radiation has been the subject of intense debate. Today, it is widely believed that its ultrafast excited state gas phase decay stems from a radiationless transition from the bright ππ* state to a dark nπ* state. However, conflicting theoretical predictions have made the experimental data difficult to interpret. Here we simulate the early gas phase ultrafast dynamics in thymine at the highest level of theory to date. This is made possible by performing wavepacket dynamics with a recently developed coupled cluster method. Our simulation confirms an ultrafast ππ* to nπ* transition (τ = 41 ± 14 fs). Furthermore, the predicted oxygen-edge X-ray absorption spectra agree quantitatively with experiment. We also predict an as-yet uncharacterized πσ* channel that leads to hydrogen dissociation at one of the two N-H bonds. Similar behavior has been identified in other heteroaromatic compounds, including adenine, and several authors have speculated that a similar pathway may exist in thymine. However, this was never confirmed theoretically or experimentally. This prediction calls for renewed efforts to experimentally identify or exclude the presence of this channel.

Kjønstad, Eirik F.↗

Automatic speech recognition predicts contemporaneous earthquake fault displacement

Abstract Significant progress has been made in probing the state of an earthquake fault by applying machine learning to continuous seismic waveforms. The breakthroughs were originally obtained from laboratory shear experiments and numerical simulations of fault shear, then successfully extended to slow-slipping faults. Here we apply the Wav2Vec-2.0 self-supervised framework for automatic speech recognition to continuous seismic signals emanating from a sequence of moderate magnitude earthquakes during the 2018 caldera collapse at the Kīlauea volcano on the island of Hawai’i. We pre-train the Wav2Vec-2.0 model using caldera seismic waveforms and augment the model architecture to predict contemporaneous surface displacement during the caldera collapse sequence, a proxy for fault displacement. We find the model displacement predictions to be excellent. The model is adapted for near-future prediction information and found hints of prediction capability, but the results are not robust. The results demonstrate that earthquake faults emit seismic signatures in a similar manner to laboratory and numerical simulation faults, and artificial intelligence models developed for encoding audio of speech may have important applications in studying active fault zones.

58 GEOSCIENCES↗

Cross-scale covariance for material property prediction

A simulation can stand its ground against an experiment only if its prediction uncertainty is known. The unknown accuracy of interatomic potentials (IPs) is a major source of prediction uncertainty, severely limiting the use of large-scale classical atomistic simulations in a wide range of scientific and engineering applications. Here we explore covariance between predictions of metal plasticity, from 178 large-scale (~10 8 atoms) molecular dynamics (MD) simulations, and a variety of indicator properties computed at small-scales (≤10 2 atoms). All simulations use the same 178 IPs. In a manner similar to statistical studies in public health, we analyze correlations of strength with indicators, identify the best predictor properties, and build a cross-scale “strength-on-predictors” regression model. This model is then used to estimate regression error over the statistical pool of IPs. Small-scale predictors found to be highly covariant with strength are computed using expensive quantum-accurate calculations and used to predict flow strength, within the statistical error bounds established in our study.

36 MATERIALS SCIENCE↗

SA-GAT-SR: self-adaptable graph attention networks with symbolic regression for high-fidelity material property prediction

Recent advances in machine learning have demonstrated an enormous utility of deep learning approaches, particularly Graph Neural Networks (GNNs) for materials science. These methods have emerged as powerful tools for high-throughput prediction of material properties, offering a compelling enhancement and alternative to traditional first-principles calculations. While the community has predominantly focused on developing increasingly complex and universal models to enhance predictive accuracy, such approaches often lack physical interpretability and insights into materials behavior. Here, we introduce a novel computational paradigm—Self-Adaptable Graph Attention Networks integrated with Symbolic Regression (SA-GAT-SR)—that synergistically combines the predictive capability of GNNs with the interpretative power of symbolic regression. Our framework employs a self-adaptable encoding algorithm that automatically identifies and adjust attention weights so as to screen critical features from an expansive 180-dimensional feature space while maintaining O(n) computational scaling. The integrated SR module subsequently distills these features into compact analytical expressions that explicitly reveal quantum-mechanically meaningful relationships, achieving 23 × acceleration compared to conventional SR implementations that heavily rely on first-principle calculations-derived features as input. This work suggests a new framework in computational materials science, bridging the gap between predictive accuracy and physical interpretability, offering valuable physical insights into material behavior.

36 MATERIALS SCIENCE↗

RNA-Puzzles Round V: blind predictions of 23 RNA structures

RNA-Puzzles is a collective endeavor dedicated to the advancement and improvement of RNA three-dimensional structure prediction. With agreement from structural biologists, RNA structures are predicted by modeling groups before publication of the experimental structures. We report a large-scale set of predictions by 18 groups for 23 RNA-Puzzles: 4 RNA elements, 2 Aptamers, 4 Viral elements, 5 Ribozymes and 8 Riboswitches. We describe automatic assessment protocols for comparisons between prediction and experiment. Our analyses reveal some critical steps to be overcome to achieve good accuracy in modeling RNA structures: identification of helix-forming pairs and of non-Watson–Crick modules, correct coaxial stacking between helices and avoidance of entanglements. Three of the top four modeling groups in this round also ranked among the top four in the CASP15 contest.

59 BASIC BIOLOGICAL SCIENCES↗

Prediction of BiS2-type pnictogen dichalcogenide monolayers for optoelectronics

Abstract In this work, we introduce a 2D materials family with chemical formula MX 2 (M={As, Sb, Bi} and X={S, Se, Te}) having a rectangular 2D lattice. This materials family has been predicted by systematic ab-initio structure search calculations in two dimensions. Using density-functional theory and many-body perturbation theory, we study the structural, vibrational, electronic, optical, and excitonic properties of the predicted MX 2 family. Our calculations reveal that the predicted SbX 2 and BiX 2 monolayers are stable while the AsX 2 layers exhibit an in-plane ferroelectric instability. All materials display strong excitonic effects and good optical absorption within the infrared-to-visible range. Hence, these monolayers can harvest solar energy and serve in optoelectronics applications. Furthermore, our results indicate that exfoliation of the predicted MX 2 monolayers from their bulk counterparts is experimentally viable.

Materials Science↗

A human-in-the-loop explanation framework for morphologically transparent AI predictions from whole-slide images

Deep learning models enable the prediction of clinical endpoints from whole-slide images (WSIs), but many such models function as “black boxes”, lacking transparency about whether and which histomorphological patterns drive their predictions, hindering interpretability and clinical adoption. Here we propose a human-in-the-loop explanation framework, MorphoXAI, which provides both local and global interpretability for deep learning models by incorporating human-expert interpretations. At the global level, it reveals the histomorphological patterns on which the model consistently relies to distinguish between classes of WSIs, as well as the patterns associated with confusion between classes. At the local level, it indicates which of these patterns are used in the prediction of an individual WSI and which regions within the slide correspond to such patterns. We validated our method across multiple deep learning–based WSI analysis tasks spanning different tissue types. The results show that our framework generates explanations that accurately reflect the histomorphology underlying the model’s predictions at both global and local levels. For interpretability and clinical utility in diagnostic contexts, human evaluation results showed that our explanations were easy to interpret, rich in diagnostic features, and directly helpful for diagnostic decision-making, thereby enhancing pathologist-AI collaboration. Our work highlights that unifying global and local explanations and grounding them in expert-interpreted morphology enhances the interpretability and verifiability of deep learning models, thereby facilitating the transparent deployment of such models in clinical practice.

Lou, Peiliang↗

CryoTRANS: predicting high-resolution maps of rare conformations from self-supervised trajectories in cryo-EM

Cryogenic electron microscopy (cryo-EM) has revolutionized structural biology, enabling efficient determination of structures at near-atomic resolutions. However, a common challenge arises from the severe imbalance among various conformations of vitrified particles, leading to low-resolution reconstructions in rare conformations due to a lack of particle images in these quasi-stable states. We introduce CryoTRANS, a method that predicts high-resolution maps of rare conformations by constructing a self-supervised pseudo-trajectory between density maps of varying resolutions. This trajectory is represented by an ordinary differential equation parameterized by a deep neural network, ensuring retention of detailed structures from high-resolution density maps. By leveraging a single high-resolution density map, CryoTRANS significantly improves the reconstruction of rare conformations and has been validated on four real-world datasets: alpha-2-macroglobulin, actin-binding protein complexes, SARS-CoV-2 spike glycoprotein, and the 70S ribosome. CryoTRANS can also predict high-resolution structures in cryogenic electron tomography maps using a high-resolution cryo-EM map.Cryogenic electron microscopy (cryo-EM) has revolutionized structural biology, enabling efficient determination of structures at near-atomic resolutions. However, a common challenge arises from the severe imbalance among various conformations of vitrified particles, leading to low-resolution reconstructions in rare conformations due to a lack of particle images in these quasi-stable states. We introduce CryoTRANS, a method that predicts high-resolution maps of rare conformations by constructing a self-supervised pseudo-trajectory between density maps of varying resolutions. This trajectory is represented by an ordinary differential equation parameterized by a deep neural network, ensuring retention of detailed structures from high-resolution density maps. By leveraging a single high-resolution density map, CryoTRANS significantly improves the reconstruction of rare conformations and has been validated on four real-world datasets: alpha-2-macroglobulin, actin-binding protein complexes, SARS-CoV-2 spike glycoprotein, and the 70S ribosome. CryoTRANS can also predict high-resolution structures in cryogenic electron tomography maps using a high-resolution cryo-EM map.

47 OTHER INSTRUMENTATION↗