Search NASA⌕ Search

SEARCH · Search NASA

Results for “adaptive machine learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Adaptive Quantum Generative Training using an Unbounded Loss Function

We propose a generative quantum learning algorithm using the Adaptive Derivative-Assembled Problem Tailored ansatz (ADAPT) framework in which the loss function to be minimized is the maximal quantum Rényi divergence of order two, an unbounded function that mitigates barren plateaus which inhibit training variational circuits. We benchmark this method against other state-of-the-art adaptive algorithms by learning random two-local thermal states. We perform numerical experiments of up to 12 qubits comparing our method learning algorithms that use linear objective functions and show that Rényi-ADAPT is capable of constructing shallow quantum circuits competitive with existing methods, while the gradients remain favorable resulting from the maximal Rényi divergence loss function.

quantum algorithms, quantum machine learning, quan↗

Real-time plasma monitoring framework for advanced plasma control and ML-research in DIII-D

Real-time and adaptive plasma control is crucial for robust tokamak operation, requiring sensitivity and tolerance measurements of the plasma state. This paper presents the implementation of an integrated real-time plasma monitoring framework on the DIII-D tokamak to support advanced control approaches, including machine-learning (ML) methods. The system is built on the SHIELD framework, a high-performance modular architecture that provides a unified pipeline for integrating diverse diagnostics. The framework leverages high-bandwidth digitizers, fast numerical processing, and deterministic, low-latency interconnects to stream high-fidelity data from diagnostics such as electron cyclotron emission (ECE), beam emission spectroscopy (BES), CO interferometers, and a visible tangential divertor camera (TangTV). The system’s validity is demonstrated through direct comparisons of real-time and offline data. Furthermore, we present two key applications of the developed plasma monitoring system with ML-based plasma control strategies, including real-time divertor detachment and active Alfvén Eigenmode control. As a result, this work presents a robust and scalable approach for integrating high-frequency, multidimensional diagnostics into advanced control algorithms for future fusion devices.

AI/ML↗

Predictive models of the genetic bases underlying budding yeast fitness in multiple environments

Abstract The ability of organisms to adapt and survive depends on the effects of genes and the environment on fitness. However, the multigenic nature of fitness and genotype-by-environment interactions hinder our understanding of the genetic basis of fitness. Here, we established fitness prediction models for 35 environments using machine learning and existing fitness data and different genetic variant types for a Saccharomyces cerevisiae population. Models revealed that the predictive ability of genetic variants varied across environments, with copy number variants explaining the majority of fitness variation in most cases. Model interpretation showed that different variant types identified distinct gene sets associated with predictive variants. These gene sets were significantly enriched in experimentally validated genes affecting fitness in only a subset of environments, indicating that many genes influencing fitness remain unexplored. Notably, non-experimentally validated genes were more important than validated ones for fitness predictions. Gene contributions to predictions were both isolate- and environment-dependent, pointing to gene-by-gene and gene-by-environment interactions. Furthermore, models uncovered experimentally validated and novel candidate genetic interactions for a well-characterized stress, the fungicide benomyl. These findings highlight the feasibility of identifying the genetic basis of fitness by using different genetic variant types and offer novel targets for future functional analysis.

DNA copy number variations↗

Safe and Robust Binary Classification and Fault Detection Using Reinforcement Learning

In this paper, we propose a learning-based method utilizing the Soft Actor-Critic (SAC) algorithm to train a binary Support Vector Machine (SVM) classifier. This classifier is designed to identify valid input spaces in high-dimensional, highly constrained systems while minimizing the total runtime of offline simulations. The simulations adapt their runtime based on the likelihood that a given training input will be informative to the classifier. Furthermore, we introduce a method for using the trained SAC model to predict whether a desired system input is likely to violate constraints, along with a technique to adjust the input as necessary. Additionally, we explore the potential of this model to detect faults or adversarial attacks within the system. The effectiveness of our approach is demonstrated through various simulations of challenging classification problems and a constrained quadrotor model.

Netter, Josh [Georgia Institute of Technology, Atl↗

A Molecular View of Methane Activation on Ni(111) through Enhanced Sampling and Machine Learning

A combination of machine learned interatomic potentials (MLIPs) and enhanced sampling simulations is used to investigate the activation of methane on a Ni(111) surface. The work entails the development and iterative refinement of MLIPs, initially trained on a dataset constructed via ab initio molecular dynamics (AIMD) simulations, supplemented by adaptive biasing forces, to enrich the sampling of catalytically relevant configurations. Our results reveal that by incorporating collective variables that capture the behavior of the reactant molecule, as well as additional frames that describe the dynamic response of the catalytic surface, it is possible to enhance considerably the accuracy of predicted energies and forces. By employing enhanced sampling schemes in the refinement of the MLIP, we systematically explore the potential energy surface, leading to a refined MLIP capable of predicting DFT-level energies and forces and replicating key geometric characteristics of the catalytic system. The resulting free energy landscapes at several temperatures provide a detailed view of the thermodynamics and dynamics of methane activation. Specifically, as methane approaches and dissociates on the catalytic surface, the process involves the dynamic interplay of CH 4 and the Ni catalyst that includes both enthalpic and entropic contributions. The progression towards the transition state involves an CH 4 moiety that is increasingly restrained in its ability to rotate or translate, while the stage following the transition state is characterized by a notable rise of the Ni atom that interacts with the cleaved C–H bond. Furthermore, this leads to an increase in the mobility of the adsorbed species, a feature that becomes more pronounced at higher temperatures.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

MULTI-LEADER: MULTI-source LEarning-Accelerated Design of high-Efficiency multi-stage compRessor (Final Technical Report)

The objective of MULTI-LEADER is to cut design costs by 80% while generating more energy-efficient designs of multi-stage compressors by developing and implementing novel machine learning (ML) techniques, which enable faster and fewer design iterations, improved solver performance, and concurrent multi-disciplinary design. Current industrial practices for the design of multi-stage compressors involve simulation-based design optimization with successive levels of model fidelity, iteratively evaluated between distinct disciplines, one stage at a time to tackle the high dimensional design variations. This project addresses these key design challenges: (1) concurrent optimization of multiple stages under many non-linear constraints; (2) multitude of evaluation of high-fidelity and expensive solvers and their gradients during optimization convergence in high-dimensional design; (3) multi-disciplinary design to maximize aerodynamic performance while guaranteeing structural integrity and additive manufacturability; (4) utilization of multiple fidelity of solvers with disparate parameterization and modeling assumptions. MULTI-LEADER achieved more than 5x speed up in detailed design of more energy-efficient compressors via these machine learning (ML) innovations: (i) rapid design surrogates by multi-source learning from diverse fidelities across multiple disciplines, (ii) physics-constrained data-augmented modeling for improved empiricism, (iii) generative manifold embedding for high dimensional concurrent design without gradient information; (iv) budget-constrained fidelity-adaptive sampling towards fewer design iterations.

33 ADVANCED PROPULSION SYSTEMS↗

Final report- UFL - RAPIDS2: A SciDAC Institute for Computer Science, Data, and Artificial Intelligence

The research initiatives supported by the U.S. Department of Energy (DOE) Grant DE-SC0022265 are fundamentally aimed at pioneering advanced machine learning (ML) techniques for scientific data compression within high-performance computing (HPC) environments. This comprehensive body of work addresses the critical challenge posed by the exponential growth of data generated by scientific simulations in domains such as fusion energy, climate modeling, and computational fluid dynamics (CFD). A core objective is to develop compression algorithms that achieve substantial data reduction—often by orders of magnitude—while rigorously ensuring the fidelity of both the primary data (PD) and scientifically crucial derived quantities of interest (QoI). The methodologies deployed under this grant integrate sophisticated deep learning architectures, prominently featuring autoencoders, advanced generative models like conditional diffusion, and hybrid learning techniques. Key innovations include the development of Guaranteed Autoencoders (GAE) and the Guaranteed Conditional Diffusion with Tensor Correction (GCDTC) framework, which provide explicit, instance-level error bounds on reconstructed data. Furthermore, specialized strategies such as nonlinear constraint satisfaction are employed to preserve the integrity of QoI, a vital requirement for the trustworthiness of downstream scientific analyses. This research also focuses on the design and implementation of scalable, GPU-accelerated software pipelines that seamlessly integrate into existing HPC workflows, ensuring both computational efficiency and practical applicability. The CAESAR framework, for example, unifies foundation and generative models to create an adaptive and efficient compression solution for spatio-temporal scientific data. Collectively, these efforts represent a significant advancement in mitigating the scientific data deluge, enabling more effective data management, accelerated scientific discovery, and optimized utilization of HPC resources.

97 MATHEMATICS AND COMPUTING↗

Adaptive Algebraic Derivative Estimation for Battery Electric Buses Energy Consumption Forecasting

The limited service life of onboard batteries for EVs is a challenge, underscoring the need for real-time battery usage prediction. This paper proposes an adaptive Algebraic Derivative Estimation (ADE) approach for forecasting the energy consumption of battery electric buses. By dynamically adjusting the sliding window length, the adaptive ADE retains the fixed-length ADE’s key advantage—namely, operating online without reliance on extensive historical datasets—while substantially bolstering forecast accuracy by actively trading estimation bias off estimation variance. Comparative experiments against both the conventional ADE with a fixed length and a representative machine learning algorithm, XGBoost, were conducted, with performance evaluated via root mean square error, mean absolute error, and the coefficient of determination. The results demonstrate that the proposed approach significantly outperforms baseline methods.

Cui, Tianyang [The University of Texas at Dallas]↗

Mathematical Morphological Filtering with a Self-Adaptive Reconstruction Technique and Application to Local Seismic Data

Recorded seismic data are generally contaminated by noise from different sources, which masks the signals of interest. In the seismology community, frequency filtering (FF) is the standard method for noise suppression. However, when the signal of interest and noise share the same frequency band, the latter cannot be filtered out without infringing on the former. We implemented a noise suppression approach based on the mathematical morphology theorem. The method involves compound operations of dilation and erosion using structuring elements of varying lengths and decomposes an input noisy waveform into several time functions with differing characteristics. Further, the filtered waveform is constructed from the time functions using a self-adaptive reconstruction technique. Application to a data set of >4700 local waveforms suggests that the implemented mathematical morphological filtering (MMF) approach is efficient for data with low signal-to-noise ratio (SNR) and significantly outperforms FF in that SNR range. For most of the dataset, FF, machine learning (ML) denoising, and continuous wavelet transform (CWT) thresholding result in higher SNR values compared with the MMF method. However, for ~42% of the waveforms, MMF outperforms FF, and the SNR gain achieved with MMF is as large as ~23 dB. Compared to ML denoising and CWT thresholding, this proportion drops to only ~10%–14%. Our results suggests that in an operational setting, MMF cannot replace the other noise suppression methods; however, signal detection can be improved if MMF is used to supplement them in some scenarios. MMF could help detect signals in problematic low-SNR data, which are currently being missed particularly when using FF alone.

58 GEOSCIENCES↗

Multi-Omics Reveals Temporal Scales of Carbon Metabolism in Synechococcus Elongatus PCC 7942 Under Light Disturbance

Central carbon metabolism in model cyanobacteria involves multiple pathways to adapt to energy-light limitations across diel cycles. However, the success in mechanistic modeling for phenotypic prediction of the protein regulators in the metabolic state depends on capturing the vast possibilities emerging from multiple regulatory pathways in complex biological processes. Here, we developed a physics-informed machine learning approach based on energy-landscape concepts to predict regulatory proteins responding to cyclic circadian and unforeseen light perturbations in cyanobacterial metabolic networks. Our approach provides interpretable de novo models for inferring gene expression dynamics from Synechococcus elongatus over diel cycles and using redox proteome analysis to distinguish immediate light-responsive elements from circadian-regulated processes in carbon metabolism pathways. We identified distinct temporal signatures with the analysis of the redox proteome: there was an immediate shift in cysteine redox states accompanied by a limited change in protein abundance under constant illumination and after 2 hours of darkness. This discovery indicates that the generation of reductants coordinates photoinduced electron transport with redox metabolic pathways in two discernable molecular mechanisms: fast redox-based protein modifications occur immediately after the light disturbance, followed by slow transcriptional regulations across networks. This temporal regulation reveals how metabolic networks integrate rapid light responses with programmed circadian rhythms to maintain cellular homeostasis under the light-energy limitations over the diel cycle.

Biomolecular & subcellular processes↗

Dynamical Sketching for Enhanced Communication Efficiency in Federated Learning

Federated learning (FL) has revolutionized distributed machine learning by enabling collaborative model training without sharing local data. However, communication efficiency and privacy guarantees remain significant challenges. This paper introduces a dynamic sketching mechanism in FL, optimizing the trade-off between communication efficiency and model accuracy. By dynamically selecting the sketch matrix size, our approach adapts to the evolving characteristics of the data and the model, ensuring optimal performance across diverse scenarios. We leverage Bayesian optimization to systematically tune the sketch parameters, achieving an effective balance between resource efficiency and model performance. Experimental results on the MNIST dataset using a convolutional neural network (CNN) architecture validate the proposed method's efficiency and scalability. Our dynamic sketching approach significantly outperforms fixed-size sketching techniques, achieving higher compression ratios (up to 62x) and providing better privacy guarantees while maintaining high model accuracy. These findings highlight the robustness and versatility of our approach and make it a valuable solution for privacy-preserving, communication-efficient federated learning.

Afrose, Sharmin [ORNL]↗

Uncertainty guided online ensemble for non-stationary data streams in fusion science

Machine Learning (ML) is poised to play a pivotal role in the development and operation of next-generation fusion devices. Fusion data shows non-stationary behavior with distribution drifts, resulted by both experimental evolution and machine wear-and-tear. ML models assume stationary distribution and fail to maintain performance when encountered with such non-stationary data streams. Online learning techniques have been leveraged in other domains, however it has been largely unexplored for fusion applications. In this paper, we investigate online learning for continuous adaptation to drifting data streams in the prediction of Toroidal Field (TF) coils deflection at the DIII-D fusion facility. We further address the short-term performance degradation inherent to standard online learning, which arises because ground truth is unavailable at prediction time. To mitigate this issue, we propose an uncertainty-guided online ensemble framework. The method leverages the Deep Gaussian Process Approximation (DGPA) for calibrated uncertainty estimation and uses these uncertainty measures to guide a meta-algorithm that aggregates predictions from learners trained over different historical horizons. Our results show that online learning reduces prediction error by 80% compared to a static model. The online ensemble and the proposed uncertainty-guided ensemble further reduce error by approximately 6%, and 10% respectively, relative to standard single-model online learning, while also providing calibrated uncertainty estimates to support operational decision-making.

AI↗

Computational Modeling to Advance Novel Medical Isotopes for Radiotheranostics: A DOE-NIH Joint Workshop Executive Summary

The DOE-NIH Joint Workshop on Computational Modeling to Advance Novel Medical Isotopes for Radiotheranostics, held on September 27, 2024, brought together experts from government, academia, and industry to address critical challenges in radionuclide production and clinical translation. Here, the workshop emphasized interdisciplinary collaboration, particularly between the Department of Energy (DOE) and the National Institutes of Health (NIH), to strengthen the domestic isotope supply, streamline regulatory pathways, and further integrate computational tools into radiopharmaceutical therapy (RPT). Key discussions explored the role of AI-driven modeling, machine learning, and digital twin technologies in optimizing dosimetry, dynamically personalizing treatments, and reducing time to clinical adoption. Advances in predictive computational modeling were highlighted as essential for improving radionuclide yield, purity, and synthesis efficiency. Regulatory considerations and equitable access were central themes, with participants advocating for harmonized global standards, adaptive trial designs, and expanded infrastructure for clinical implementation. DOE computational and production infrastructure was emphasized. Future priorities identified include increased investment in radionuclide production infrastructure, expanded workforce development in radiopharmaceutical sciences and computational modeling, and the creation of robust public-private partnerships. The workshop concluded that continued strategic collaboration and sustained resources will be vital for advancing next-generation radiotheranostics, ensuring safe and effective therapies accessible to all patients.

digital twins↗

A knowledge-informed large language model framework for U.S. nuclear power plant shutdown initiating event classification for probabilistic risk assessment

Identifying and classifying shutdown initiating events (SDIEs) is critical for developing shutdown probabilistic risk assessment for nuclear power plants. Existing computational approaches cannot achieve satisfactory performance due to the challenges of unavailable large, labeled datasets, imbalanced event types, and label noise. To address these challenges, we propose a hybrid pipeline that integrates a knowledge-informed machine learning model to prescreen non-SDIEs and a large language model (LLM) to classify SDIEs into four types. In the prescreening stage, we proposed a set of 44 SDIE text patterns that consist of the most salient keywords and phrases from six SDIE types. Text vectorization based on the SDIE patterns generates feature vectors that are highly separable by using a simple binary classifier. The second stage builds Bidirectional Encoder Representations from Transformers (BERT)-based LLM, which learns generic English language representations from self-supervised pretraining on a large dataset and adapts to SDIE classification by fine-tuning it on an SDIE dataset. The proposed approaches are evaluated on a dataset with 10,928 events using precision, recall ratio, F 1 score, and average accuracy. In conclusion, the results demonstrate that the prescreening stage can exclude more than 97% non-SDIEs, and the LLM achieves an average accuracy of 95.1% for SDIE classification.

99 - GENERAL AND MISCELLANEOUS↗

ELM‐MOSART‐DOC: A Large‐Scale Riverine Dissolved Organic Carbon Model and Its Application Over the United States

Riverine dissolved organic carbon (DOC), primarily sourced from soil organic carbon (SOC), plays a crucial role in regional and global carbon cycles. However, the complexities of the underlying mechanisms and limited observations present significant challenges for predictive understanding of DOC at regional or larger scales. Recently, we developed a machine learning‐based (ML) map of DOC transformation rates, bridging the gap between SOC and DOC leaching flux and simplifying terrestrial DOC representation. Building on this advancement, we introduce ELM‐MOSART‐DOC, a DOC module integrated into the riverine component of the Energy Exascale Earth System Model (E3SM)—the Model for Scale Adaptive River Transport (MOSART). ELM‐MOSART‐DOC simulates DOC transport and transformation across both headwater streams and river networks, including those managed. Model validation demonstrates the ability of ELM‐MOSART‐DOC to accurately capture long‐term average DOC concentrations, with Kling‐Gupta Efficiency (KGE) scores of 0.58 and 0.76 at large and local stations, respectively. We further assess the impact of reservoirs through different simulation schemes, revealing that reservoirs significantly alter DOC fluxes by regulating streamflow patterns and promoting DOC mineralization. Model simulations indicate that reservoirs reduce total DOC flux from the Mississippi River into the ocean by 7.5%, with the long‐term average annual export decreasing from 3.34 to 3.14 teragrams (Tg) per year. ELM‐MOSART‐DOC integrates process‐based modeling with ML parameterization to enhance the predictive understanding of riverine biogeochemical processes. This approach reduces uncertainties in modeling regional and global carbon cycle ESMs and provides new insights into carbon cycling and its implications for global environmental change.

Li, Lingbo [Univ. of Houston, TX (United States); ↗

A high-throughput experimentation platform for data-driven discovery in electrochemistry

Automating electrochemical analyses combined with artificial intelligence is poised to accelerate discoveries in renewable energy sciences and technologies. This study presents an automated high-throughput electrochemical characterization (AHTech) platform as a cost-effective and versatile tool for rapidly assessing liquid analytes. The Python-controlled platform combines a liquid handling robot, potentiostat, and customizable microelectrode bundles for diverse, reproducible electrochemical measurements in microtiter plates, minimizing chemical consumption and manual effort. To showcase the capability of AHTech, we screened a library of 180 small molecules as electrolyte additives for aqueous zinc metal batteries, generating data for training machine learning models to predict Coulombic efficiencies. Key molecular features governing additive performance were elucidated using Shapley Additive exPlanations and Spearman’s correlation, pinpointing high-performance candidates like cis-4-hydroxy-d-proline, which achieved an average Coulombic efficiency of 99.52% over 200 cycles. The workflow established herein is highly adaptable, offering a powerful framework for accelerating the exploration and optimization of extensive chemical spaces across diverse energy storage and conversion fields.

Lin, Dian-Zhao [Johns Hopkins University, Baltimor↗

ML based control systems for nuclear physics experiments

The Experimental Physics Software and Computing Infrastructure (EPSCI) group at Jefferson Lab is leading the use of machine learning (ML) to enhance control systems in nuclear physics experiments. Collaborating closely with domain experts and data scientists, we have developed an ML-based control system that uses a Gaussian process to dynamically adjust the high voltage of the GlueX Central Drift Chamber. This results in stable detector performance by adapting to environmental changes, thereby reducing the offline calibration effort. Furthermore, we are developing ML-driven systems for optimizing the polarization of photon beams and polarized cryotargets. These systems will maintain the optimal microwave frequency in cryogenic targets and make real-time adjustments to diamond radiators for polarized photon sources, tasks traditionally handled by human operators. By automating these functions, we aim to optimize the polarization, reduce downtime, and minimize human error. This talk will highlight the development of reliable ML-based control systems and the policies to ensure they are both effective and trustworthy.

Jeske, Torri↗

Statistically Resolved Planetary Boundary Layer Height Diurnal Variability Using Spaceborne Lidar Data

The Planetary Boundary Layer Height (PBLH) significantly impacts weather, climate, and air quality. Understanding the global diurnal variation of the PBLH is particularly challenging due to the necessity of extensive observations and suitable retrieval algorithms that can adapt to diverse thermodynamic and dynamic conditions. This study utilized data from the Cloud-Aerosol Transport System (CATS) to analyze the diurnal variation of PBLH in both continental and marine regions. By leveraging CATS data and a modified version of the Different Thermo-Dynamics Stability (DTDS) algorithm, along with machine learning denoising, the study determined the diurnal variation of the PBLH in continental mid-latitude and marine regions. The CATS DTDS-PBLH closely matches ground-based lidar and radiosonde measurements at the continental sites, with correlation coefficients above 0.6 and well-aligned diurnal variability, although slightly overestimated at nighttime. In contrast, PBLH at the marine site was consistently overestimated due to the viewing geometry of CATS and complex cloud structures. The study emphasizes the importance of integrating meteorological data with lidar signals for accurate and robust PBLH estimations, which are essential for effective boundary layer assessment from satellite observations.

54 ENVIRONMENTAL SCIENCES↗