Search NASA⌕ Search

SEARCH · Search NASA

Results for “continual learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

deadtrees.earth — An open-access and interactive database for centimeter-scale aerial imagery to uncover global tree mortality dynamics

Excessive tree mortality is a global concern and remains poorly understood as it is a complex phenomenon. We lack global and temporally continuous coverage on tree mortality data. Ground-based observations on tree mortality, e.g., derived from national inventories, are very sparse, and may not be standardized or spatially explicit. Earth observation data, combined with supervised machine learning, offer a promising approach to map overstory tree mortality in a consistent manner over space and time. However, global-scale machine learning requires broad training data covering a wide range of environmental settings and forest types. Low altitude observation platforms (e.g., drones or airplanes) provide a cost-effective source of training data by capturing high-resolution orthophotos of overstory tree mortality events at centimeter-scale resolution. Here, we introduce deadtrees.earth, an open-access platform hosting more than two thousand centimeter-resolution orthophotos, covering more than 1,000,000 ha, of which more than 58,000 ha are manually annotated with live/dead tree classifications. This community-sourced and rigorously curated dataset can serve as a comprehensive reference dataset to uncover tree mortality patterns from local to global scales using space-based Earth observation data and machine learning models. This will provide the basis to attribute tree mortality patterns to environmental changes or project tree mortality dynamics to the future. The open nature of deadtrees.earth, together with its curation of high-quality, spatially representative, and ecologically diverse data will continuously increase our capacity to uncover and understand tree mortality dynamics.

Citizen science↗

Utility and Industry Perceptions of Control Room Modernization Over the Last 10 Years

Within the current United States (U.S.) nuclear power fleet, main control room modernization (CRM) is an important step towards cost savings. In recent decades, plants have been engaged in upgrades to varying degrees. This process requires a nuanced, balanced, and timely approach that ensures continued safety and long-term sustainability. In 2012 a survey was issued to individuals from the nuclear industry to learn their perspectives on a range of CRM issues. The survey targeted the benefits and challenges for utilities undertaking this process, including the main drivers and barriers to technology upgrades, regulatory compliance, and the effects these factors have on concepts of operations, strategic approaches, and staffing. In 2022, the survey was issued again to understand whether CRM perceptions had changed in the last 10 years. Our findings identify changes in industry thinking from a decade ago. Here we reveal perspective shifts that represent increased optimism and, in some instances, increased doubt regarding the opportunities and challenges inherent in CRM and implementation. We also report nuanced differences in CRM perspectives between utility and surrounding nuclear industry respondents.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

STag. II. Classification of Serendipitous Supernovae Observed by Galaxy Redshift Surveys

With the number of supernovae observed expected to drastically increase thanks to large-scale surveys like the Dark Energy Spectroscopic Instrument (DESI), it is necessary that the tools we use to classify these objects keep up with this increase. We previously created Supernova Tagging and Classification (STag) to address this problem by employing machine learning techniques alongside logistic regression in order to assign “tags” to spectra based on spectral features. STag II is a continuation of this work, which now makes use of model supernova spectra combined with real DESI spectra in order to train STag to better deal with realistic data. Furthermore, we also make use of the rlap score as a trustworthiness cut, making for a more robust and accurate supernova classifier than before.

Astrostatistics techniques↗

A Digital Twin Framework Utilizing Machine Learning for Robust Predictive Maintenance: Enhancing Tire Health Monitoring

We introduce a novel digital twin (DT) framework for the predictive maintenance of long-term physical systems. Using monitoring tire health as an application, we show how the DT framework can be used to enhance automotive safety and efficiency, and how the technical challenges can be overcome using a three-step approach. First, to manage the data complexity over a long operation span, we employ data reduction techniques to concisely represent physical tires using historical performance and usage data. Relying on these data, for fast real-time prediction, we train a transformer-based model offline on our concise dataset to predict future tire health over time, represented as remaining casing potential (RCP). Based on our architecture, our model quantifies both epistemic and aleatoric uncertainties, providing reliable confidence intervals around predicted RCP. Second, to incorporate real-time data, we update the predictive model in the DT framework, ensuring its accuracy throughout its lifespan with the aid of hybrid modeling and the use of the discrepancy function. Third, to assist decision-making in predictive maintenance, we implement a tire state decision algorithm, which strategically determines the optimal timing for tire replacement based on RCP forecasted by our transformer model. This approach ensures that our DT accurately predicts system health, continually refines its digital representation, and supports predictive maintenance decisions. Furthermore, our framework effectively embodies a physical system, leveraging big data and machine learning (ML) for predictive maintenance, model updates, and decision-making.

advanced computing infrastructure↗

In situ Detection of Plasma Induced Surface Interaction based on Deep Learning based Visual Diagnostics (Technical Report)

It is characteristic for many plasma devices to undergo plasma-material interaction leading to surface erosion. These processes, often not easily detectable, lead to changes in device performance and lifespan. State-of-the-art lifetime tests and wear experiments require over 1000s hours. A self-consistent model for accurately predicting the erosion's effects is not available. In situ detection of these processes is not a trivial task since the surface variations at the early stages have a micron scale. Such limitations not only restrict testing and prediction capabilities but also slow the development of new thrusters and limit mission duration. To address these challenges, an in-situ diagnostic for real-time erosion assessment has been developed, aiming to expedite lifetime testing and broaden experimental campaigns. Several works were dedicated to real-time and in situ monitoring of material erosion during plasma exposure using laser holography, microscopy, and with telemicroscopes. However, the applicability of these approaches is limited due to complexity, cost and less flexibility as they often require placing diagnostic equipment inside the vacuum chamber. In collaboration with Princeton Collaborative Research Facility (PCRF), Princeton Plasma Physics Laboratory (PPPL), a new diagnostic approach is developed, where geometry modifications to the ceramic channel walls were introduced that would result in accelerated channel erosion. We employed Long-distance microscope (LDM) imagery, combined with Deep-Learning based Shape from focus or depth from focus (DFF or SFF) approach, that provides an accessible and cost-effective solution. LDM employs focus variation techniques to continuously capture multiple images of the target object at distinct focal planes. DFF, an optical focus variation method, generates a 3D topographical surface depth map from a sequence of variably focused images. Combined with the developed diagnostic, this approach offers a controllable means to study erosion under accelerated conditions. In this work, we develop Neural Network-based DFF algorithm applicable for LDM data to quantitatively evaluate plasma induced surface modification from LDM data. Next, we develop Deep Learning-based super-resolution depth map image reconstruction technique to increase the resolution of depth maps obtained from DFF algorithm to improve the accuracy of erosion measurements. Thirdly, we develop several image processing techniques to remove noise and improve the quality of depth map image. Here we report the results of initial tests for this approach. An experimental setup designed and built in PPPL was employed that consists of a 3-cm gridded ion source that produces a neutralized argon beam with energies up to 600 eV. A hexagonal boron nitride (h-BN) ceramic target, designed based on computational predictions, was used. Tests were conducted to reconstruct the complex geometry of the target under the lighting conditions of the operated ion source.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Unraveling the Dynamics of Nucleosome Arrays

The organization of genomic DNA into chromatin is a fundamental determinant of genome stability, regulation, and cellular function. Nucleosomes, the basic repeating units of chromatin, assemble into higher-order structures whose organization and heterogeneity remain difficult to characterize using conventional ensemble-averaged techniques. A key need in the field is the development of experimental approaches capable of directly visualizing nucleosome assemblies and their structural variability at the single-molecule level. This LDRD Lab-Wide project focused on establishing and evaluating atomic force microscopy (AFM)–based approaches for the characterization of nucleosome assemblies. The work emphasized experimental workflows for preparing, imaging, and assessing multi-nucleosome systems, rather than isolated single nucleosomes. Through method development and exploratory measurements, the project demonstrated the feasibility of applying scanning probe microscopy to investigate chromatin-relevant assemblies and provided preliminary insight into the strengths and limitations of this approach for future quantitative studies. Results and lessons learned from this effort were disseminated to the broader scientific community through multiple national conference presentations, helping to position LLNL for continued work in chromatin and genome organization research.

59 BASIC BIOLOGICAL SCIENCES↗

Predicting Drug Effects from High-dimensional Asymmetric Drug Data Sets using Graph Neural Networks: A Comprehensive Analysis of Multi-target Drug Effect Prediction

Graph neural networks (GNNs) have emerged as one of the most effective Machine learning (ML) techniques for drug effect prediction from drug molecular graphs. Despite having immense potential, GNN models lack performance when using data sets that contain high dimensional asymmetrically co-occurrent drug effects as targets with complex correlations between them. Training individual learning models for each drug effect and incorporating every prediction result for a wide spectrum of drug effects is beyond practicality. Such an implication provides a testbed to address this challenge as multi-target prediction problems, aiming to predict all drug effects at a time. We develop standard and hybrid graph neural networks (GNNs)to perform two separate tasks that are multi-regression for continuous values and multi-label classification for categorical values contained in our data sets. Since this step makes the target data even more sparse and introduces asymmetric label co-occurrence, the learning of multi-label classification models becomes difficult and heavily impacts the GNN's performance. To address these challenges, we propose a new data oversampling technique to improve multi-label classification performances on all the given imbalanced molecular graph data sets. Using the technique, we improve the data imbalance ratio of the drug effects better than before while protecting the data set's integrity. Finally, we evaluate multi-label classification performance using the best-performant hybrid GNN model on all the oversampled data sets obtained from the proposed oversampling technique. These results outperform those of other ML models including GNN models when they are trained on the original data sets or oversampled data sets using MLSMOTE (a well-known oversampling technique) in all evaluation metrics precision, recall, and F1 score by a significant margin.

Bose, Avishek [ORNL]↗

Operating advanced scientific instruments with AI agents that learn on the job

Advanced scientific user facilities, such as next generation X-ray light sources and self-driving laboratories, are revolutionizing scientific discovery by automating routine tasks and enabling rapid experimentation and characterizations. However, these facilities must continuously evolve to support new experimental workflows, adapt to diverse user projects, and meet growing demands for more intricate instruments and experiments. This continuous development introduces significant operational complexity, necessitating a focus on usability, reproducibility, and intuitive human-instrument interaction. In this work, we explore the integration of agentic AI, powered by Large Language Models (LLMs), as a transformative tool to achieve this goal. We present our approach to developing a human-in-the-loop pipeline for operating advanced instruments including an X-ray nanoprobe beamline and an autonomous robotic station dedicated to the design and characterization of materials. Specifically, we evaluate the potential of various LLMs as trainable scientific assistants for orchestrating complex, multi-task workflows, which also include multimodal data, optimizing their performance through optional human input and iterative learning. We demonstrate the ability of AI agents to bridge the gap between advanced automation and user-friendly operation, paving the way for more adaptable and intelligent scientific facilities.

Large Language Models↗

Learning stochastic dynamics and predicting emergent behavior using transformers

We show that a neural network originally designed for language processing can learn the dynamical rules of a stochastic system by observation of a single dynamical trajectory of the system, and can accurately predict its emergent behavior under conditions not observed during training. We consider a lattice model of active matter undergoing continuous-time Monte Carlo dynamics, simulated at a density at which its steady state comprises small, dispersed clusters. We train a neural network called a transformer on a single trajectory of the model. The transformer, which we show has the capacity to represent dynamical rules that are numerous and nonlocal, learns that the dynamics of this model consists of a small number of processes. Forward-propagated trajectories of the trained transformer, at densities not encountered during training, exhibit motility-induced phase separation and so predict the existence of a nonequilibrium phase transition. Transformers have the flexibility to learn dynamical rules from observation without explicit enumeration of rates or coarse-graining of configuration space, and so the procedure used here can be applied to a wide range of physical systems, including those with large and complex dynamical generators.

97 MATHEMATICS AND COMPUTING↗

Machine Learning-Based Technique for Automated Sensor Characterization

The development of novel instrumentation requires an iterative cycle with three stages: design, prototyping, and testing. Recent advancements in simulation and nanofabrication techniques have significantly accelerated the design and prototyping phases. Nonetheless, detector characterization continues to be a major bottleneck in device development. During the testing phase, a significant time investment is required to characterize the device in different operating conditions and find optimal operating parameters. The total effort spent on characterization and parameter optimization can occupy a year or more of an expert s time. In this work, we present a novel technique for automated sensor calibration that aims to accelerate the testing stage of the development cycle. This technique leverages closed-loop Bayesian optimization (BO), using real-time measurements to guide parameter selection and identify optimal operating states. We demonstrate the method with a novel low-noise CCD, showing that the machine learning-driven tool can efficiently characterize and optimize operation of the sensor in a couple of days without supervision of a device expert.

Zepeda, Cuevas [Chicago U., KICP]↗

Advances in in situ/operando techniques for catalysis research: enhancing insights and discoveries

Abstract Catalysis research has witnessed remarkable progress with the advent of in situ and operando techniques. These methods enable the study of catalysts under actual operating conditions, providing unprecedented insights into catalytic mechanisms and dynamic catalyst behavior. This review discusses key in situ techniques and their applications in catalysis research. Advances in in situ electron microscopy allow direct visualization of catalysts at the atomic scale under reaction conditions. In situ spectroscopy techniques like X-ray absorption spectroscopy and nuclear magnetic resonance spectroscopy can track chemical states and reveal transient intermediates. Synchrotron-based techniques offer enhanced capabilities for in situ studies. The integration of in situ methods with machine learning and computational modeling provides a powerful approach to accelerate catalyst optimization. However, challenges remain regarding radiation damage, instrumentation limitations, and data interpretation. Overall, continued development of multi-modal in situ techniques is pivotal for addressing emerging challenges and opportunities in catalysis research and technology.

Chen, Linfeng↗

Evaluation of Hardware and Software Bill of Materials (HBOMs/SBOMs) Extraction Methods

Hardware and software bills of materials (HBOMs and SBOMs) provide important visibility into the components, dependencies, and supply chain relationships within programmable digital devices. This visibility is critical for advanced nuclear reactor applications, where use of common or shared hardware components, software libraries, suppliers, or manufacturing processes may create common cause failure (CCF) vulnerabilities despite apparent diversity. This paper evaluates current approaches for obtaining and analyzing HBOMs and SBOMs in support of CCF, diversity and defense-in-depth (D3) assessments, and begins to explore potential methods for artificial intelligence/machine learning-based analysis. The availability of BOM information from advanced reactor manufacturers and vendors, representative hardware and software categories found in advanced reactor systems continues to limit research [13]. This paper compares commonly used BOM formats, including CycloneDX, SPDX, and SWID. It also surveys publicly available tools for generating BOMs from source code, compiled binaries, and hardware-related information, noting limitations in language coverage, system age, and format interoperability. Finally, this paper evaluates methods for correlating BOM data with vulnerability and exploitability information, including VEX, CVE, and CWE resources. The findings indicate that publicly available nuclear-vendor BOMs are limited, making third-party extraction and research into novel analysis techniques necessary.

Cybersecurity↗

Structural and mechanical properties of monolayer amorphous carbon and boron nitride

Amorphous materials exhibit various characteristics that are not featured by crystals and can sometimes be tuned by their degree of disorder (DOD). Here, we report results on the mechanical properties of monolayer amorphous carbon (MAC) and monolayer amorphous boron nitride (maBN) with different DOD. The pertinent structures are obtained by kinetic-Monte-Carlo (kMC) simulations using machine-learning potentials (MLP) with density-functional-theory (DFT)-level accuracy. An intuitive order parameter, namely the areal fraction F x occupied by crystallites within the continuous random network, is proposed to describe the DOD. We find that F x captures the essence of the DOD: Samples with the same F x but different sizes and arrangements of crystallites, obtained using two distinct kMC procedures, have virtually identical radial distributions functions as well as bond-length and bond-angle distributions. Furthermore, by simulating the fracture process with molecular dynamics, we found that the mechanical responses of MAC and maBN before fracture are mainly determined by F x and are insensitive to the sizes and specific arrangements and to some extent the numbers and area distributions of the crystallites. The behavior of cracks in the two materials is analyzed and found to mainly propagate in meandering paths in the CRN region and to be influenced by crystallites in distinct ways that toughen the material. Furthermore, the present results reveal the relation between structure and mechanical properties in amorphous monolayers and may provide a universal toughening strategy for 2D materials.

2-dimensional systems↗

Bifunctional Electrocatalysts with High-Entropy Alloys: Bridging Hydrogen Evolution and Oxygen Reduction

High-entropy alloys (HEAs) have emerged as a promising class of bifunctional electrocatalysts capable of simultaneously driving the hydrogen evolution reaction (HER) and the oxygen reduction reaction (ORR) with high activity and durability. Their near-equiatomic multicomponent compositions give rise to unique physicochemical characteristics, including lattice distortion, sluggish diffusion, high-entropy stabilization, and pronounced electronic heterogeneity, that collectively generate diverse and synergistic active sites inaccessible in conventional alloys. This review summarizes recent progress in HEA-based bifunctional electrocatalysis, with a focus on the fundamental mechanisms governing HER and ORR activity, stability, and selectivity. We discuss advances in synthesis strategies, ranging from confined growth and step-alloying to scalable continuous-flow methods, that enable precise control over composition, size, and surface structure. Complementary computational and data-driven approaches, including density functional theory, machine-learning-assisted screening, and descriptor development, are highlighted as essential tools for navigating the vast HEA design space and establishing structure−property relationships. Particular attention is paid to adsorption-energy distributions, multisite cooperativity, and environmental effects under realistic electrochemical conditions. Finally, we outline current challenges and future opportunities for integrating mechanistic understanding with AI-guided, closed-loop design frameworks to accelerate the discovery of next-generation HEA bifunctional electrocatalysts for sustainable energy conversion.

Alloys↗

Building Datasets and Training Methods for ML Based Magnet Quench Detection

Detecting quenches in superconducting (SC) magnets during training is a challenging process that involves capturing physical events that occur at different frequencies and appear as various signal features. These events may be correlated across instrumentation type, thermal cycle, and ramp. These events together build a more complete picture of continuous processes occurring in the magnet, and may allow us to flag potential precursors for quench detection. We present our work on building an automatic machine learning (ML) based quench detection system. We build upon our existing work on unsupervised auto-encoders for acoustic sensors and quench antenna (QA) by first establishing a supervised ML training pipeline. We show the results of an event tagging, analysis, and simulation framework on our QA and acoustic data which are used concurrently to build a training dataset for a supervised implementation. We then show how this supervised training can be used as a prior in a semi-supervised framework and compare this to the unsupervised neural network auto-encoder performance.This allows us to have a more concrete understanding of the performance of our algorithms relative to physical events occurring in the magnet, and also provides a baseline software tool to generically evaluate our quench prediction autoencoders under completely unsupervised, supervised, and semi-supervised training conditions.

Khan, Maira [Fermilab]↗

MLCommons Science Benchmarks

Benchmarks are a cornerstone of modern machine learning practice, providing standardized eval- uations that enable reproducibility, comparison, and scientific progress. Yet, as AI systems particularly deep learning models become increasingly dynamic, traditional static benchmarking approaches are losing their relevance. Models rapidly evolve in architecture, scale, and capability; datasets shift; and deployment contexts continuously change, creating a moving target for evaluation. Without adaptive benchmarking frame- works, both scientific assessment and real-world de- ployment risk becoming misaligned with actual system behavior. Drawing on our experience from MLCommons, educa- tional initiatives, and government programs such as the DOE s Million Parameter Consortium, we identify key barriers that hinder the broader adoption and utility of benchmarking in AI. These include substantial resource demands, limited access to specialized hardware, lack of expertise in benchmark design, and uncertainty among practitioners about how to relate benchmark results to their own application domains. Moreover, current benchmarks often emphasize peak performance on leadership-class hardware, offering limited guidance for more diverse, real-world deployment scenarios. We argue that benchmarking itself must become dy- namic in order to incorporate evolving models, updated data, and heterogeneous computational platforms while maintaining transparency, reproducibility, and inter- pretability. Democratizing this process requires not only technical innovation, but also systematic educational efforts spanning undergraduate to professional levels to develop sustained expertise in benchmark design and use. Finally, benchmarks should be framed and com- municated to support application-relevant comparisons, enabling both developers and users to make informed, context-sensitive decisions. Advancing dynamic and inclusive benchmarking practices will be essential to ensure that evaluation keeps pace with the evolving AI landscape and supports responsible, reproducible, and accessible AI deployment.

Hawks, Benjamin G. [Fermilab]↗

Monitoring river flow status using low-cost wildlife camera and image segmentation artificial intelligence

Continuous measurement and monitoring of surface water coverage in non-perennial streams are essential for understanding the exchange fluxes between surface and subsurface waters under both inundated and non-inundated conditions. In this study, a wildlife camera photo-based framework was developed to monitor small stream water inundation, depth, discharge, and velocity. Two advanced machine learning models, YOLOv8 and Mask2Former, were utilized to efficiently analyze images captured by wildlife cameras. The accuracy of the framework was validated against on-site depth measurements at six sites in the Yakima River Basin, along with the gage height, discharge, and velocity data from four USGS sites. This approach facilitates long-term, continuous monitoring and quantification of river intermittency and water availability with high precision and low cost, thereby advancing river ecosystem research and management.

machine learning↗

Best of both worlds: Enforcing detailed balance in machine learning models of transition rates

The slow microstructural evolution of materials often plays a key role in determining material properties. When the unit steps of the evolution process are slow, direct simulation approaches such as molecular dynamics become prohibitive and Kinetic Monte-Carlo (kMC) algorithms, where the state-to-state evolution of the system is represented in terms of a continuous-time Markov chain, are instead frequently relied upon to efficiently predict long-time evolution. The accuracy of kMC simulations however relies on the complete and accurate knowledge of reaction pathways and corresponding kinetics. This requirement becomes extremely stringent in complex systems such as concentrated alloys where the astronomical number of local atomic configurations makes the a priori tabulation of all possible transitions impractical. Machine learning models of transition kinetics have been used to mitigate this problem by enabling the efficient on-the-fly prediction of kinetic parameters. While conventional KMC methods based on transition state theory naturally yield reversible dynamics that exactly obey the detailed balance criterion, providing strong guarantees on the properties of the stationary distribution, many recently-proposed ML-based approaches to barrier predictions provide no such guarantees. In this study, we derive conditions under which physics-informed ML architectures exactly enforce the detailed balance condition by construction, even when relying on non-extensive descriptions of states in terms of local environments around mobile defects. In conclusion, using the diffusion of a vacancy in a concentrated alloy as an example, we show that such ML architectures also exhibit superior performance in terms of prediction accuracy, demonstrating that the imposition of physical constraints can facilitate the accurate learning of barriers at no increase in computational cost.

36 MATERIALS SCIENCE↗