Search NASA⌕ Search

SEARCH · Search NASA

Results for “Machine learning models”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 433 records · Page 24

A high-throughput experimentation platform for data-driven discovery in electrochemistry

Automating electrochemical analyses combined with artificial intelligence is poised to accelerate discoveries in renewable energy sciences and technologies. This study presents an automated high-throughput electrochemical characterization (AHTech) platform as a cost-effective and versatile tool for rapidly assessing liquid analytes. The Python-controlled platform combines a liquid handling robot, potentiostat, and customizable microelectrode bundles for diverse, reproducible electrochemical measurements in microtiter plates, minimizing chemical consumption and manual effort. To showcase the capability of AHTech, we screened a library of 180 small molecules as electrolyte additives for aqueous zinc metal batteries, generating data for training machine learning models to predict Coulombic efficiencies. Key molecular features governing additive performance were elucidated using Shapley Additive exPlanations and Spearman’s correlation, pinpointing high-performance candidates like cis-4-hydroxy-d-proline, which achieved an average Coulombic efficiency of 99.52% over 200 cycles. The workflow established herein is highly adaptable, offering a powerful framework for accelerating the exploration and optimization of extensive chemical spaces across diverse energy storage and conversion fields.

Lin, Dian-Zhao [Johns Hopkins University, Baltimor↗

Predicting fusion ignition at the National Ignition Facility with physics-informed deep learning

Here, an inertial confinement fusion experiment, carried out at the National Ignition Facility, has achieved ignition by generating fusion energy exceeding the laser energy that drove the experiment. Prior to the experiment, a generative machine learning model that combines radiation hydrodynamics simulations, deep learning, experimental data, and Bayesian statistics was used to predict, with a probability greater than 70%, that ignition was the most likely outcome for this shot.

Spears, Brian K. [Lawrence Livermore National Labo↗

Microbial vitamin biosynthesis links gut microbiota dynamics to chemotherapy toxicity

ABSTRACT Dose-limiting toxicities pose a major barrier to cancer treatment. While preclinical studies show that the gut microbiota influences and is influenced by anticancer drugs, data from patients paired with careful side effect monitoring remains limited. Here, we investigate capecitabine (CAP)-microbiome interactions through longitudinal metagenomic sequencing of stool from 56 advanced colorectal cancer patients. CAP significantly altered the gut microbiome, enriching for menaquinol (vitamin K2) biosynthesis genes. Transposon library screens, targeted gene deletions, and media supplementation revealed that menaquinol biosynthesis protectsEscherichia colifrom drug toxicity. Stool menaquinol gene and metabolite levels were associated with decreased peripheral sensory neuropathy. Machine learning models trained in this cohort predicted toxicities in an independent cohort. Taken together, these results suggest treatment-associated increases in microbial vitamin biosynthesis serve a chemoprotective role for bacterial and host cells. Further, our findings provide a foundation for in-depth mechanistic dissection, human intervention studies, and extension to other cancer treatments. IMPORTANCE Side effects are common during the treatment of cancer. The trillions of microbes found within the human gut are sensitive to anticancer drugs, but the effects of treatment-induced shifts in gut microbes for side effects remain poorly understood. We profiled gut microbes in colorectal cancer patients treated with capecitabine and carefully monitored side effects. We observed a marked expansion in genes for producing vitamin K2 (menaquinone). Vitamin K2 rescued gut bacterial growth and was associated with decreased side effects in patients. We then used information about gut microbes to develop a predictive model of drug toxicity that was validated in an independent cohort. These results suggest that treatment-associated increases in bacterial vitamin production protect both bacteria and host cells from drug toxicity, providing new opportunities for intervention and motivating the need to better understand how dietary intake and bacterial production of micronutrients like vitamin K2 influence cancer treatment outcomes.

Microbiology↗

Reconstruction of atmospheric neutrinos in DUNE’s horizontal-drift far-detector module

This paper reports on the capabilities in reconstructing and identifying atmospheric neutrino interactions in one of the Deep Underground Neutrino Experiment’s (DUNE) far detector modules, a liquid argon time projection chamber (LArTPC) with horizontal drift (FD-HD) of ionization electrons. The reconstruction is based upon the workflow developed for DUNE’s long-baseline oscillation analysis, with some necessary machine-learning models’ retraining and the addition of features relevant only to atmospheric neutrinos such as the neutrino direction reconstruction. Where relevant, the impact of the detection of the charged particles of the hadronic system is emphasized, and comparisons are carried out between the case when lepton-only information is considered in the reconstruction (as is the case for many neutrino oscillation experiments), versus when all particles identified in the LArTPC were included. Three neutrino direction reconstruction methods have been developed and studied for the atmospheric analyses: using lepton-only information, using all reconstructed particles, and using only correlations from reconstructed hits. The results indicate that incorporating more than just lepton information significantly improves the resolution of both neutrino direction and energy reconstruction. The angle reconstruction algorithms developed in this work result in no strong dependence on particle direction for reconstruction efficiencies or neutrino flavor identification. This comprehensive review of the reconstruction of atmospheric neutrinos in DUNE’s FD-HD LArTPC is the first step towards developing a first neutrino oscillation sensitivity analysis, which will ready DUNE for its first measurements.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Emerging Technologies for Privacy Preservation in Energy Systems

This study explores the intersection of digitalization and privacy within the energy sector, focusing on the emerging challenges and opportunities presented by integrating Distributed Energy Resources (DERs) and advanced metering infrastructure. The need for robust digital privacy measures has become crucial as the energy industry evolves towards a more decentralized, digitalized, and decarbonized future. This study delves into four cutting-edge privacy-preserving technologies—Homomorphic Encryption (HE), Secure Multiparty Computation (SMPC), Differential Privacy (DP), and Federated Learning (FL)—each offering unique solutions to safeguard consumer data by increasing digital connectivity and data exchange. Through a detailed examination of these methods, the study explains how each technology operates, its applications within the energy sector, and the specific privacy challenges it addresses. Homomorphic Encryption allows for secure computations on encrypted data, enabling data analysis without compromising privacy. Secure Multiparty Computation enables collaborative data analysis across different entities while protecting the confidentiality of the inputs. Differential Privacy introduces randomness into the assembled data set, preventing the identification of individual records in statistical databases. Lastly, Federated Learning offers a paradigm shift in data analysis, where machine learning models are trained at the edge, minimizing the centralization of sensitive data. The research underscores the significance of implementing these privacy-enhancing technologies to comply with strict data protection regulations, foster consumer trust, and enhance the security of the energy infrastructure. By providing a comprehensive overview of these methodologies and their practical implications for the energy sector, this study aims to contribute to the ongoing discourse on digital privacy, offering insights into how the energy industry can navigate the complexities of data privacy in the digital age.

Cali, Umit↗

Convolution Neural Network for Fault Identification in Distribution Feeder with High Penetration Solar PV

Identification and zonal classification of the faults is a decisive factor in the relay’s decision to trip or not. Different types of fault like three-phase, line-to-line-to-ground and single-line-to-ground can occur at various locations in the feeder. These faults are seen as the variation in the instantaneous values of three-phase voltages and currents, i.e., waveforms, that are measured at the relay location. The objective of this work is to develop a machine learning model that can identify a fault and classify it to various protection zones based on measured waveforms. In this work, a data-driven relay based on Convolutional Neural Network (CNN) is proposed for fault identification in distribution feeders with high penetration solar PV. The proposed CNN model takes local current and voltage waveforms as input and classify it into fault, no-fault or a capacitor switching. Further, the CNN also attempts to identify fault zones based on the images of waveforms. The overall testing accuracy of the trained model exceeds 95%.

Ramesh, Meghana↗

Dashboard for Visualizing Molecular Property Prediction Machine Learning Results

This is a dashboard for exploring the results of machine learning models for predicting molecular properties from molecular structure. It includes tools for: 1. Modifying molecules to observe the change in predicted properties 2. Exploring the relationship between molecular structure and predicted properties 3. Recommending structurally similar molecules with improved properties 4. Exploring the impact of data subsampling on model performance metrics

Xu, Audrey↗

Code Description for "Brief Communication: Monitoring snow depth using small, cheap, and easy-to-deploy ground surface temperature sensors"

Temporally continuous snow depth estimates are vital for understanding changing snow patterns and impacts on permafrost in the Arctic. We train a random forest machine learning model to predict snow depth from variability in ground surface temperature. To our knowledge, this is the first time that small ground surface temperature sensors have been used to estimate snow depth. The model performs well at sites where the model was trained and at pan-arctic evaluation sites (RMSE <= 0.15 m). Small temperature sensors are cheap and easy-to-deploy, so this technique enables spatially distributed and temporally continuous snowpack monitoring to an extent previously infeasible. The model is flexible and can be applied to datasets retroactively to retrieve snow depth estimates at additional sites. This code package includes a *.joblib file of the trained random forest model and a *.ipynb file showing how to clean input data, train the random forest model, and apply the model.

Bachand, Claire↗

TransPlatformer

We propose TransPlatformer for translating toxicogenomics from one platform to another. Transcriptomic profiling has evolved through multiple generations of technology, from microarrays (e.g., Affymetrix, CodeLink) to more recent high-throughput sequencing and targeted panels such as S1500+. Microarrays, which dominated gene expression studies in the early 2000s, provided affordable and high-throughput transcript quantification but suffered from cross-hybridization issues and limited dynamic range . RNA-Seq, introduced in the late 2000s, revolutionized transcriptomics by enabling unbiased and comprehensive gene expression analysis, albeit at higher costs and computational demands . Despite advances, many studies rely on historical microarray data, necessitating the translation of legacy data into modern platforms to ensure continuity and comparability. This translation is complicated by factors such as platform-specific probe design, differences in transcript coverage, and batch effects . Existing methods for cross-platform mapping include statistical normalization, machine learning models, and biological anchoring approaches. The ability to translate transcriptomic data between platforms has broad implications, including enhanced meta-analyses, improved toxicological modeling, and better integration of historical datasets with contemporary research. TransPlatformer seeks to contribute to this effort by evaluating translation methodologies and proposing novel strategies to improve cross-platform gene expression harmonization. In this repository there are code examples for TransPlatformer implementation

Cong, Guojing↗

listenr

SAND2024-13902O listenr is an R package containing code that allows users to fit echo state networks (a machine learning model) on data to obtain predictions. The code also allows users to compute spatio-temporal feature importance on the echo state network, as described in “Characterizing Climate Pathways Using Feature Importance on Echo State Networks.” The package also provides a function for computing principal components.

Ries, Daniel↗

Intelligent-Immunity

This code builds machine learning models for transcription and protein data generated for the purpose of classifying innate immune signatures.

Martinez, Kaitlyn (Katy) [@lanl]↗

DeepLensSBI: Deep inference of simulated strong lenses in ground-based surveys

This code is used to train and test machine learning models and generate results and plots presented in 2501.08524 [astro-ph.IM]. The code is written in python. The goal of this work is to train ML models trained on simulated images of strong gravitational lenses. The trained model can then quickly infer properties of the lensed objects with uncertainty quantification.

Poh, Jason [Univ. of Chicago, IL (United States)] ↗

Spatio-temporal Fourier Transformer for Long-term Dynamics Prediction (StFT) v1.0

We propose a novel machine learning model spatio-temporal Fourier transformer (StFT) to emulate long-term dynamics of multi-scale and multi-physics systems. Our method StFT overcomes the limitations of rapid error accumulation, particularly in long-term forecasting of systems characterized by complex and coupled dynamics. StFT achieves outstanding accuracy and computational efficiency by effectively capturing multi-scale interactions, and quantify the uncertainties inherent in the predictions. Our model leverages a structured hierarchy of StFT blocks, and explicitly captures dynamics across both macro- and micro- spatial scales. Evaluations conducted on three benchmark datasets (plasma, fluid, and atmospheric dynamics) demonstrate the advantages of our approach over state-of-the-art ML methods.

Bai, Zhe [Lawrence Berkeley National Laboratory (L↗

inverse-cnf

Inverse machine learning model using Conditional Normalizing Flow (CNF)

DeBardeleben, Nathan Andrew [Los Alamos National L↗

Machine learning methods for weather forecasting

SAND2025-14466O This repository contains code for developing, training, and evaluating machine learning models for weather and climate forecasting, including forecast skill assessment, feature importance analysis, and reproducible workflows for model comparison. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Holthuijzen, Maike [Sandia National Lab. (SNL-CA),↗

Iterative ML and Experiments for Emerging VOCs

SAND2026-17074O Iterative ML and Experiments for Emerging VOCs is a tool that analyzes and predicts the behaviors of SARS-CoV-2 variants. It processes experimental data on ACE2 (the receptor for the SARS-CoV-2 virus that allows it to infect the cell) and antibody binding using machine learning models, including neural networks, to forecast ACE2 interactions and variant expression. The tool employs transfer learning and global epistasis modeling, integrating public datasets with proprietary data to enhance prediction accuracy. Additionally, it fits concentration-response curves to determine dissociation constants and generates visualizations to support research findings, thereby aiding in the identification of new antibodies for emerging variants of concern. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Sheffield, Thomas [Sandia National Lab. (SNL-NM), ↗

PRIME: Protein Representation Inference for Mutation Evaluation

Protein language machine learning models built upon existing ESM-2 model developed by Evolutionary Scale (evolutionaryscale.ai) and an in-house protein language model based on the BERT model developed by Google. The code also includes model training scripts and saved checkpoints from our own training using publicly available SARS-CoV-2 protein sequences.

Gibson, Kaetlyn [Los Alamos National Lab]↗