Search NASA⌕ Search

SEARCH · Search NASA

Results for “generative machine learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 361 records · Page 20

Machine learning surrogates for ion energy–angle distributions in thermal and RF plasma sheaths

Ion energy–angle distributions (IEADs) at material surfaces are a critical input for plasma–material interaction (PMI) studies in fusion devices, yet they are computationally expensive to obtain using particle-in-cell (PIC) simulations. In this work, we develop a machine learning surrogate based on a deep deconvolutional neural network (DDeCNN) trained on large databases generated with the hPIC2 code. The surrogate is capable of reconstructing IEADs from sheath parameters for both thermal and radio-frequency (RF) plasmas, including cases with multiple ion species. Across thousands of test cases, the model achieves high accuracy, with over 97 % of predictions classified as good or average based on standard error metrics (MAE, MSE, L2). Even in the more challenging RF and multi-species regimes, the surrogate reliably captures the multi-peak structure of PIC results. Once trained, the surrogate produces IEADs in milliseconds on a common workstation, yielding speedups of six to seven orders of magnitude compared with running a full PIC simulation. This computational gain enables dense parameter scans and direct coupling of IEAD predictions with PMI and erosion models on whole-device scales in fusion-relevant conditions.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Nanoengineering of non-aqueous liquid electrolyte solutions for future lithium metal batteries

Research and development of non-aqueous electrolyte solutions are essential for practical advancement towards the production of high-energy lithium metal batteries (LMBs). An ideal LMB electrolyte solution should enable highly efficient, uniform and prolonged lithium metal plating and stripping, preserve the electrodes’ electro(chemo)mechanical properties and ensure compatibility with all cell components. However, despite extensive research efforts, scientists have yet to achieve an electrolyte design that meets these requirements simultaneously. Here, by examining the nanoengineering aspects of various non-aqueous electrolyte solution designs, we elucidate the understanding of the nanoscale physicochemical and electrochemical processes taking place in LMBs, which are mainly governed by the thermodynamic and kinetic properties of the electrolyte system. We also explore emerging research directions and propose an accelerated, iterative framework that integrates nanoengineering principles with machine learning, high-throughput computation and experimentation to facilitate the development of next-generation non-aqueous electrolyte solutions for practical LMBs.

Weintz, Dominik↗

Finite-element-based simulations of electrodes for CO 2 cascade reduction reactions

The multielectron reduction of CO 2 to liquid fuels could be a path to scalable energy storage, but reaching this goal requires major advances in catalysis and systems engineering. Cascade catalysis, which couples sequential reactions without isolating intermediates, has emerged as a promising route to enhance selectivity and efficiency in CO 2 reduction (CO 2 R). In this review, we examine how finite-element-based simulations of continuum model [finite element method (FEM)] approaches are being used to analyze and guide CO 2 R cascade systems. We first outline the fundamentals of cascade catalysis and recent advances in catalytic materials (metallic, molecular, and hybrid architectures). We then focus on FEM developments at the electrode and device scales, emphasizing how these models capture transport phenomena, local microenvironments, and geometry-dependent effects. To clarify design principles, we present case studies of cascade electrodes organized in systems without and with integrated semiconductors. We further emphasize the integration of FEM with multiscale frameworks (density functional theory, molecular dynamics, kinetic Monte Carlo) and its role in bridging atomic-level insights with device-level performance. Finally, we identify current limitations and future prospects, including improved boundary conditions, coupling with operando experiments, and machine learning-accelerated model development. Together, these insights provide design principles for next-generation CO 2 R cascade systems for efficient solar fuel production.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Resolving turbulent magnetohydrodynamics: a hybrid operator-diffusion framework

We present a hybrid machine learning framework that combines physics-informed neural operators (PINOs) with score-based generative diffusion models to simulate the full spatio-temporal evolution of two-dimensional, incompressible, resistive magnetohydrodynamic turbulence across a broad range of Reynolds numbers (Re). The framework leverages the equation-constrained generalization capabilities of PINOs to predict coherent, low-frequency dynamics, while a conditional diffusion model stochastically corrects high-frequency residuals, enabling accurate modeling of fully developed turbulence. Trained on a comprehensive ensemble of high-fidelity simulations with Re ϵ {100, 250, 500, 750, 1000, 3000, 10000}, the approach achieves state-of-the-art accuracy in regimes previously inaccessible to deterministic surrogates. At Re = 1000 and 3000, the model faithfully reconstructs the full spectral energy distributions of both velocity and magnetic fields late into the simulation, capturing non-Gaussian statistics, intermittent structures, and cross-field correlations with high fidelity. At extreme turbulence levels (Re = 10 000), it remains the first surrogate capable of recovering the high-wavenumber evolution of the magnetic field, preserving large-scale morphology and enabling statistically meaningful predictions.

Diffusion-Integrated Neural Operators↗

SVM-Based Synchronized Fault Detection for 100% Renewable Microgrids

Traditional protection schemes face significant challenges when applied to microgrids with high penetrations of renewables with inverter-based resources (IBRs). The proliferation of advanced sensing and communication technologies has generated copious data, offering an opportunity to overcome these limitations using data-driven machine learning approaches. This work proposes a novel approach based on a support vector machine (SVM) for detecting faults within a 100% renewable microgrid. The approach encompasses a systematic offline training stage for the development of a linear SVM-based fault detection algorithm. This process covers offline data collection from the microgrid under study, the extraction of features such as positive- and negative-sequence components and the total harmonic distortion of the voltage and current measurements of the relays, and the design of the linear SVM-based classifier. During the online implementation, however, different classifiers can exhibit asynchronicity in detecting the fault inception at different subcycle-to-cycle period-level delays. To circumvent this asynchronicity issue, a separate algorithm is developed for each relay to estimate the fault inception time as close to the real fault time. The performance of the proposed SVM-based synchronized fault detection method is evaluated using online time-domain simulation studies on a microgrid test system. The results corroborate the reliability of the fault detection scheme when tested under various fault cases (fault types, locations, and impedances) and non-fault cases during both grid-tied and islanded operation modes.

100% microgrid↗

NEUROSPF: A Tool For the Symbolic Analysis of Neural Networks

This paper presents NEUROSPF, a tool for the symbolic analysis of neural networks. Given a trained neural network model, the tool extracts the architecture and model parameters and translates them into a Java representation that is amenable for analysis using the Symbolic PathFinder symbolic execution tool. Notably, NEUROSPF encodes specialized peer classes for parsing the model’s parameters, thereby enabling efficient analysis. With NEUROSPF the user has the flexibility to specify either the inputs or the network internal parameters as symbolic, promoting the application of program analysis and testing approaches from software engineering to the field of machine learning. For instance, NEUROSPF can be used for coverage-based testing and test generation, finding adversarial examples and also constraint-based repair of neural networks, thus improving the reliability of neural networks and of the applications that use them.

neural networks↗

SVM-Based Synchronized Fault Detection for 100% Renewable Microgrids: Preprint

Traditional protection schemes face significant challenges when applied to microgrids with high penetrations of renewables with inverter-based resources (IBRs). The proliferation of advanced sensing and communication technologies has generated copious data, offering an opportunity to overcome these limitations using data-driven machine learning approaches. This work proposes a novel approach based on a support vector machine (SVM) for detecting faults within a 100% renewable microgrid. The approach encompasses a systematic offline training stage for the development of a linear SVM-based fault detection algorithm. This process covers offline data collection from the microgrid under study, the extraction of features such as positive- and negative-sequence components and the total harmonic distortion of the voltage and current measurements of the relays, and the design of the linear SVM-based classifier. During the online implementation, however, different classifiers can exhibit asynchronicity in detecting the fault inception at different subcycle-to-cycle period-level delays. To circumvent this asynchronicity issue, a separate algorithm is developed for each relay to estimate the fault inception time as close to the real fault time. The performance of the proposed SVM-based synchronized fault detection method is evaluated using online time-domain simulation studies on a microgrid test system. The results corroborate the reliability of the fault detection scheme when tested under various fault cases (fault types, locations, and impedances) and non-fault cases during both grid-tied and islanded operation modes.

100% microgrid↗

Generative Models for Crystalline Materials

Understanding structure-property relationships in materials is fundamental in condensed matter physics and materials science. Over the past few years, machine learning (ML) has emerged as a powerful tool for advancing this understanding and accelerating materials discovery. Early ML approaches primarily focused on constructing and screening large material spaces to identify promising candidates for various applications. More recently, research efforts have increasingly shifted toward generating crystal structures using end-to-end generative models. This review analyzes the current state of generative modeling for crystal structure prediction and de novo generation. It examines crystal representations, outlines the generative models used to design crystal structures, and evaluates their respective strengths and limitations. Furthermore, the review highlights experimental considerations for evaluating generated structures and provides recommendations for suitable existing software tools. Emerging topics, such as modeling disorder and defects, integration in advanced characterization, incorporating synthetic feasibility constraints, and model explainability are explored. Ultimately, this work aims to inform both experimental scientists looking to adapt suitable ML models to their specific circumstances and ML specialists seeking to understand the unique challenges related to inverse materials design and discovery.

Metni, Houssam [Karlsruhe Inst. of Technology (KIT↗

When more data hurts: Optimizing data coverage while mitigating diversity-induced underfitting in an ultrafast machine-learned potential

Machine-learned interatomic potentials (MLIPs) are becoming an essential tool in materials modeling. However, optimizing the generation of training data used to parametrize the MLIPs remains a significant challenge. This is because MLIPs can fail when encountering local environments too different from those present in the training data. The difficulty of determining a priori the environments that will be encountered during molecular dynamics simulation necessitates diverse, high-quality training data. Here, this study investigates how training data diversity affects the performance of MLIPs using the Ultra-Fast force field (UF 3 ) to model amorphous silicon nitride. We employ expert and autonomously generated data to create the training data and fit four force field variants to subsets of the data. Our findings reveal a critical balance in training data diversity: insufficient diversity hinders generalization, while excessive diversity can exceed the MLIP's learning capacity, reducing simulation accuracy. Specifically, we found that the UF 3 variant trained on a subset of the training data, in which nitrogen-rich structures were removed, offered vastly better prediction and simulation accuracy than any other variant. By comparing these UF 3 variants, we highlight the nuanced requirements for creating accurate MLIPs, emphasizing the importance of application-specific training data to achieve optimal performance in modeling complex material behaviors.

ab initio molecular dynamics↗

Using Artificial Intelligence to Improve Reliability and Operational Efficiency of Small-Scale Hydroelectric Distributed Generation

Reliability and resilience are critical concerns for distributed generation (DG) at the rural electric level. The integration of renewable energy sources, such as small-scale hydroelectric distributed generators (hydro DGs), introduces operational challenges, particularly regarding aging infrastructure and grid stability. Artificial Intelligence (AI)-driven Machine Learning (ML) models and applications of Large Language Models (LLMs) offer promising solutions for optimizing DG operations and enhancing resilience. This paper explores AI-based models for improving efficiency, fault resolution, and outage mitigation in small-scale hydro DGs. Furthermore, it highlights the development of a centralized, AI-powered information portal for rural electric cooperatives and municipalities. The research evaluates hydro DG plant models and discusses the applicability of AI-powered question-answering tools for real-time operations, focusing on statistical data, load flow, voltage regulation, and generation power. The findings demonstrate AI’s potential to transform DG management to ensure greater stability and resilience in rural electric grids.

Bhattacharyya, Arjun [ORNL] (ORCID:000900060976046↗

A generalizable machine learning approach to predict land surface temperature

Monitoring of land surface and atmospheric states is highly reliant on satellite data. Traditionally, data products are generated using carefully tuned and validated algorithms for low-earth orbit (LEO) sensors. However, the emerging constellation of geostationary (GEO) sensors contributes global, high temporal resolution observations which can better capture the diurnal variability of key observables like land surface temperature (LST). Using high performance computing and datasets from the NASA Earth Exchange, we exploit co-located, co-temporal observations from LEO and GEO satellites to develop a deep learning-based method for sensor-to-sensor algorithm emulation. Our model is trained on GOES-16 thermal bands to predict MODIS Terra LST and achieves a validation error <2K. Further, application of the model to unseen times of day and a second GEO sensor observing an unseen spatial domain demonstrate the generalization of the deep learning model across space, time and spectra. We anticipate that the synergies between a variety of active orbit configurations can be used to accelerate application of existing algorithms to new datasets.

Kate Marie Duffy↗

OmicsMLMentor: A Web Application for Guided Machine Learning Analysis of Omics Data

Expression-based omics technologies (e.g. proteomics, metabolomics, transcriptomics, etc.) increasingly rely on supervised and unsupervised machine learning (ML) models to find key biomolecules distinguishing conditions, identify natural groupings in biological data, or generate predictions for outcomes of interest. Fitting ML models to omics data presents several challenges, including handling missing data, selecting a normalization method, choosing a valid model, and optimizing hyperparameters, all requiring statistical programming skills to address these challenges. Thus, the open-source web application SLOPE was designed to lower the barrier to ML modeling for omics data. SLOPE supports the fitting of 15 ML models (10 supervised and 5 unsupervised) tailored to omics datasets, such as proteomics, metabolomics, lipidomics, and transcriptomics. SLOPE offers several omics-specific features, including methods for handling missingness (imputation, conversion, removal), normalization tests, ranking of models based on the structure of a user’s data and user input, and optimal hyperparameter selections using cross-validation splits. By streamlining ML workflows for omics analysis, SLOPE address critical gaps in existing online web tools, facilitating a broader adoption of these models for omics research. Here, SLOPE is applied to data from a lignin exposure study to highlight the workflow for fitting both supervised and unsupervised models to data.

lipidomics↗

An integrated approach to examine fuel-cladding chemical interaction in HT9/U-10Zr metallic fast reactor fuels: Coupling machine learning with electron microscopy and local mechanical properties analysis

The metallic U-Zr nuclear fuel alloy has garnered renewed interest as a promising candidate for next-generation sodium-cooled fast reactors. Recent studies and technology assessments have identified several areas requiring improvements, enhanced knowledge, and reliable data to strengthen the U-Zr fuel design basis for qualification and commercial applications. One of the most challenging phenomena impacting this fuel system’s performance is fuel-cladding chemical interaction (FCCI). This work aimed to harvest FCCI data by examining selected HT9/U-10Zr (wt. %) fuel samples of prototypic full-length fuel pins through an integrated approach. This approach integrated scanning electron microscopy (SEM) microstructure characterization with localized mechanical properties examination to deepen understanding of FCCI phenomenon in HT9/U-10Zr fuel system. Particularly, this study focused on MFF fuel pins irradiated at Fast Flux Test Facility (FFTF), which aimed to qualify metallic fuel as a driver fuel for FFTF and to assess its viability for larger-scale fast reactors. Electron microscopy provided high confidence in detecting and distinguishing the different FCCI layers, while small-scale mechanical testing (SSMT) probed the mechanical properties of these layers. SEM examination of a MFF-2 pin 192167, with a time averaged inner cladding temperature (TICT) slightly over 500°C, revealed minimal cladding-side FCCI (cladding wastage). In contrast, significantly thicker cladding wastage comprising two distinct sublayers was observed in samples from the thermally hot MFF-3 pin 193045 and MFF-5 pin 195011 where the TICT ranged from 610-635°C. SSMT indicated complete embrittlement in the sublayer adjacent to the fuel and a tendency toward embrittlement in the other sublayer. Additionally, a new machine learning method was developed, validated, and used to quantify cladding wastage thickness. The machine learning method reliably predicted the wastage thickness across various fuel pins and sample cross-sections. Furthermore, the available cladding wastage data from HT9/U-10Zr fuel system demonstrated a strong temperature dependency. However, the dataset remains small, and ongoing research activities are essential to further understand the FCCI phenomenon and develop a reliable FCCI model for enhanced fuel performance simulation under various conditions.

36 - MATERIALS SCIENCE↗

SPIDARman: System-Level Physics-Informed Detection of Anomalies in Reactor Collected Data Considering Human Errors

In nuclear power plants (NPPs), anomalies arising from sensors or human errors (HEs) can undermine the performance and reliability of plant operations. Anomaly detection models can be employed to detect sensor errors and HEs. Additionally, physics-informed machine learning models can utilize the known physics of the system, as described by mathematical equations, to ensure that sensor values are consistent with physical laws. Hence, we propose SPIDARman: System-level Physics-Informed Detection of Anomalies in Reactor Collected Data Considering Human Errors, a holistic physics-informed anomaly detection approach based on generative adversarial networks (GANs) to detect anomalies in both automatically collected sensor data and manually collected surveillance data. Here we test our approach on data collected from a flow loop testbed, showcasing its potential to detect anomalies. Results demonstrate that the proposed model performs better than the baseline GAN-based models in detecting sensor and surveillance anomalies, suggesting the potential of physics-informed anomaly detection GAN models in NPPs.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Intelligent Pixel Detectors: Towards a Radiation Hard ASIC with On-Chip Machine Learning in 28 nm CMOS

Detectors at future high energy colliders will face enormous technical challenges. Disentangling the unprecedented numbers of particles expected in each event will require highly granular silicon pixel detectors with billions of readout channels. With event rates as high as 40 MHz, these detectors will generate petabytes of data per second. To enable discovery within strict bandwidth and latency constraints, future trackers must be capable of fast, power efficient, and radiation hard data-reduction at the source. We are developing a radiation hard readout integrated circuit (ROIC) in 28nm CMOS with on-chip machine learning (ML) for future intelligent pixel detectors. We will show track parameter predictions using a neural network within a single layer of silicon and hardware tests on the first tape-outs produced with TSMC. Preliminary results indicate that reading out featurized clusters from particles above a modest momentum threshold could enable using pixel information at 40 MHz.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Artificial intelligence-driven approaches for materials design and discovery

Materials design is an important component of modern science and technology, yet traditional approaches rely heavily on trial and error and can be inefficient. Computational techniques, enhanced by modern artificial intelligence, have reshaped the landscape of designing new materials. Among these approaches, inverse design has shown great promise in designing materials that meet specific property requirements. Here, in this Review, we present key computational advances in materials design over the past few decades. We follow the evolution of relevant materials design techniques, from high-throughput forward machine learning methods and evolutionary algorithms, to advanced artificial intelligence strategies such as reinforcement learning and deep generative models. We highlight the paradigm shift from conventional screening approaches to inverse generation driven by deep generative models. Finally, we discuss current challenges and future perspectives of materials inverse design. This Review may serve as a brief guide to the approaches, progress and outlook of designing future functional materials with technological relevance.

computational methods↗

Single photon emitters in van der Waals solids for quantum photonics: materials, theory and molecular-scale characterization probes

Strong light–matter interactions in two-dimensional layered materials (2D materials) have attracted the interest of researchers from interdisciplinary fields for more than a decade now. A unique phenomenon in some 2D materials is their large exciton binding energies (BEs), increasing the likelihood of exciton survival at room temperature. It is this large BE that mediates the intense light–matter interactions of many of the 2D materials, particularly in their monolayer limit, where the interplay of excitonic phenomena poses a wealth of opportunities for high-performance optoelectronics and quantum photonics. Within quantum photonics, quantum information science (QIS) is growing rapidly, where photons are a promising platform for information processing due to their low-noise properties, excellent modal control, and long-distance propagation. A central element for QIS applications is a single photon emitter (SPE) source, where an ideal on-demand SPE emits exactly one photon at a time into a given spatiotemporal mode. Recently, 2D materials have shown practical appeal for QIS which is directly driven from their unique layered crystalline structure. This structural attribute of 2D materials facilitates their integration with optical elements more easily than the SPEs in conventional three-dimensional solid state materials, such as diamond and SiC. In this review article, we will discuss recent advances made with 2D materials towards their use as quantum emitters, where the SPE emission properties maybe modulated deterministically. Here, the use of unique scanning tunneling microscopy tools for the in-situ generation and characterization of defects is presented, along with theoretical first-principles frameworks and machine learning approaches to model the structure-property relationship of exciton–defect interactions within the lattice towards SPEs. Given the rapid progress made in this area, the SPEs in 2D materials are emerging as promising sources of nonclassical light emitters, well-poised to advance quantum photonics in the future.

2D layered materials↗

Enhancing Electron Microscopy Image Classification Using Data Augmentation

Manual labeling for machine learning tasks such as image classification is tedious and labor-intensive; as a result, scientific datasets suitable for deep learning applications are scarce and limited. While data augmentation techniques have shown promise for extending image datasets, very little work has been done to understand the impact of combining multiple augmentation methods sequentially or the limits of their effectiveness when combined. Our work addresses this gap by examining how standard and combinatorial data augmentation affects the performance of machine learning models when trained on small datasets for label classification tasks. For our analysis, we generate single, double and quadruple-augmented datasets for a microscopy image classification task using six standard augmentation methods, and compare the resultant improvements observed in binary classification accuracy with three standard image classification models (DenseNet169, MobileNetV2, ResNet101V2). Our experiments show a non-monotonic relationship between the number of simultaneous augmentation methods and classification accuracy, indicating that there is a trade-off between the degree of augmentation and the model performance. These findings suggest that the optimal number of augmentation methods will vary by domain and use case. We also find that the order in which augmentation methods are applied to a limited dataset matters when combining augmentation schemes, with our use case showing performance differences up to 2.6% when the augmentation order is reversed for double-augmented datasets. Our work offers insights to the limits of data augmentation when working on image classification tasks with limited datasets.

Welsman, Jordan A↗