Search NASA⌕ Search

SEARCH · Search NASA

Results for “Generative learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 343 records · Page 19

Using a Large Language Model for Accurate Technical Language Generation in the Predictive Maintenance of Circulating Water Systems in Nuclear Power Plants

Machine learning (ML) methods for predictive maintenance (PdM) are emerging as effective proactive strategies for diagnosing equipment degradation and enabling effective decision-making. However, explainability and trustworthiness of artificial intelligence are two salient challenges that need to be addressed for wider deployment of these technologies in nuclear power plants (NPPs). Large language models (LLMs) offer a unique approach to tackle these challenges by explaining PdM, work orders, diagnosis results, and ML algorithms to users, who may not be familiar with ML and PdM in general. Moreover, by dynamically retrieving relevant information from technical documents and evaluating factuality of LLM generation, the accuracy and relevance of LLM generations can be improved. This work demonstrates using LLMs to explain the causes and consequences of circulating water system failures based on multiyear NPP work orders. This work tests the capability of multimodal LLM approaches in explaining the differences in the circulating water system from both the Salem and Hope Creek NPPs using both text and image resources. This work also demonstrates the use of multimodal LLMs in describing the diagnosis tab of a predictive maintenance software named VIsualization for PrEdictive maintenance Recommendation (VIPER) to users.

22 - GENERAL STUDIES OF NUCLEAR REACTORS↗

Lessons Learned for Transmission Cost Allocation in U.S. Regional Markets

Expanding electric transmission can facilitate generator interconnection and improve grid reliability. Assigning costs for new transmission infrastructure is highly contentious because these costs can have a direct impact on energy prices and ratepayer bills. In this report, we evaluate what factors influence successful transmission cost allocation agreements. Through a review of legal disputes, existing cost allocation practices, and regional case studies, we identify potential strategies to minimize cost allocation disputes for future projects. The report also highlights the processes by which regions can update their cost allocation methods. While we do not consider cost allocation methods currently under development for compliance with FERC Order 1920, the trends and lessons learned identified in this report can inform discussions on effective cost allocation methods to reduce barriers for transmission development.

29 ENERGY PLANNING, POLICY, AND ECONOMY↗

Reduced-order modeling for efficient cross section library development in high-temperature gas reactor pebble-bed depletion analysis

Accurate modeling of running-in and equilibrium conditions in pebble-bed reactors (PBRs) requires precise microscopic multigroup neutron cross sections. In Griffin, deterministic neutronics calculations rely on multivariate interpolation over large cross section libraries, resulting in significant memory usage and performance bottlenecks. This work, together with a companion paper on Griffin integration, explores reduced-order models (ROMs) to replace interpolation with lightweight surrogates. Several ROM techniques are benchmarked, with deep neural networks (DNNs) demonstrating superior memory efficiency, scalability, and predictive accuracy. A total of 295 DNNs were trained to build a comprehensive isotope library, integrated into Griffin through a custom LibTorch interface for depletion analysis. Initial results demonstrate that DNN-based ROMs drastically reduce memory demands while preserving accuracy, enabling finer tabulations and additional state variables without overhead. In conclusion, the framework also supports online cross section generation and real-time DNN updates through transfer learning, improving fidelity by capturing self-shielding and evolving nuclide compositions during burnup.

22 - GENERAL STUDIES OF NUCLEAR REACTORS↗

In the Mix : A Workshop Merging Computational Chemistry and Electrochemistry Alongside Data Science

As chemistry expands to more complex and interdisciplinary areas, a new generation of diverse researchers must engage with science and learn effective cross-disciplinary collaboration and communication. To these ends, we designed and implemented In the Mix, a graduate student-led, two-day workshop for undergraduate students promoting collaborative science in the context of energy storage innovations. Here, the interactive workshop was designed for future and emerging researchers to gain hands-on experience with data science, computational chemistry, and electrochemistry techniques that are critical for developing materials for battery technologies. Participants also visited commercial renewable energy facilities to help them connect discovery-based research with industry and broader societal considerations. The workshop content and structure ensured that participants experienced the interrelatedness of the fields and understood the importance of collaborative research to yield scientific advances with real-world applications. An external team evaluated the workshop and participants’ perceptions of their experiences. While our research context was energy storage, the workshop goals and outcomes are applicable to other contexts. Interdisciplinary, experiential workshops are a key avenue to broadening participation in science and research, and the ideas presented here can be readily modified for other scientific contexts and/or incorporated as broader impact activities.

25 ENERGY STORAGE↗

Convergent Protocols for Computing Protein–Ligand Interaction Energies Using Fragment-Based Quantum Chemistry

Fragment-based quantum chemistry methods offer a way to sidestep the steep nonlinear scaling of electronic structure calculations so that large molecular systems can be investigated using high-level methods. Here, we use fragmentation to compute protein–ligand interaction energies in systems with several thousand atoms, using a new software platform for managing fragment-based calculations that implements a screened many-body expansion. Convergence tests using a minimal-basis semiempirical method (HF-3c) indicate that two-body calculations, with single-residue fragments and simple hydrogen caps, are sufficient to reproduce interaction energies obtained using conventional supramolecular electronic structure calculations, to within 1 kcal/mol at about 1% of the computational cost. We also demonstrate that the HF-3c results are illustrative of trends obtained with density functional theory in basis sets up to augmented quadruple-ζ quality. Strategic deployment of fragmentation facilitates the use of converged biomolecular model systems alongside high-quality electronic structure methods and basis sets, bringing ab initio quantum chemistry to systems of hitherto unimaginable size. This will be useful for generation of high-quality training data for machine learning applications.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Moment extraction using an unfolding protocol without binning

Deconvolving (“unfolding”) detector distortions is a critical step in the comparison of cross-section measurements with theoretical predictions in particle and nuclear physics. However, most existing approaches require histogram binning while many theoretical predictions are at the level of statistical moments. We develop a new approach to directly unfold distribution moments as a function of another observable without having to first discretize the data. Our moment unfolding technique uses machine learning and is inspired by Boltzmann weight factors and generative adversarial networks (GANs). We demonstrate the performance of this approach using jet substructure measurements in collider physics. With this illustrative example, we find that our moment unfolding protocol is more precise than bin-based approaches and is as or more precise than completely unbinned methods.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Learning genetic perturbation effects with variational causal inference

Advances in sequencing technologies have enhanced the understanding of gene regulation in cells. In particular, Perturb-seq has enabled high-resolution profiling of the transcriptomic response to genetic perturbations at the single-cell level. This understanding has implications in functional genomics and potentially for identifying therapeutic targets. Various computational models have been developed to predict perturbational effects. While deep learning models excel at interpolating observed perturbational data, they tend to overfit in the lack of enough data and may not generalize well to unseen perturbations. In contrast, mechanistic models, such as linear causal models based on gene regulatory networks, hold greater potential for extrapolation, as they encapsulate regulatory information that can predict responses to unseen perturbations. However, their application has been limited to small studies due to overly simplistic assumptions, making them less effective in handling noisy, large-scale single-cell data. We propose a hybrid approach that combines a mechanistic causal model with variational deep learning, termed Single Cell Causal Variational Autoencoder (SCCVAE). The mechanistic model employs a learned regulatory network to represent perturbational changes as shift interventions that propagate through the learned network. SCCVAE integrates this mechanistic causal model into a variational autoencoder, generating rich, comprehensive transcriptomic responses. Our results indicate that SCCVAE exhibits superior performance over current state-of-the-art baselines for extrapolating to predict unseen perturbational responses. Additionally, for the observed perturbations, the latent space learned by SCCVAE allows for the identification of functional perturbation modules and simulation of single-gene knockdown experiments of varying penetrance, presenting a robust tool for interpreting and interpolating perturbational responses at the single-cell level.

59 BASIC BIOLOGICAL SCIENCES↗

Assessing shellfish water exposure to fecal bacteria pollution in Salish Sea: three-dimensional modeling and implications for monitoring

Fecal bacteria (FB) contamination poses significant risks to shellfish safety and management in coastal and estuarine waters. Despite extensive pollution identification and correction efforts, FB contamination in shellfish-growing areas persists in the Salish Sea, highlighting the need to identify overlooked sources and better understand FB transport from riverine and shoreline inputs to shellfish beds. To address this, a high-resolution three-dimensional hydrodynamic model coupled with FB kinetics was developed and applied to a case study site in Salish Sea—Portage Bay—to simulate freshwater plume circulation, flushing dynamics, and bacterial transport. Daily FB loading from the major freshwater inflow—Nooksack River was generated by both linear interpolation and integrating a machine learning approach (XGBoost), trained on historical hydrological and meteorological data. The model successfully reproduced both the magnitude and seasonal variation of FB concentrations in Portage Bay for the year of 2021, demonstrating that simplified FB kinetics with first-order decay due to mortality was effective in this dynamic coastal environment with short flushing time. Model results identified the Nooksack River as the dominant far-field FB source, while scenario simulations showed that near-field coastal stormwater outfalls elevated local FB levels following rainfall, particularly under low-flow conditions. The XGBoost prediction provided comparable or superior accuracy to linear interpolation, particularly during periods of missing observational data, by capturing short-term variability and event-driven loading more effectively. Integrating data-driven riverine FB inputs with mechanistic coastal numerical modeling provides a robust framework for operational forecasting of shellfish bed exposure risk and supports adaptive monitoring and management of shellfish growing areas in the Salish Sea and similar coastal systems.

Salish Sea↗

Updates and Lessons Learned from NuMI Beamline at Fermilab

The Neutrinos at the Main Injector (NuMI) beamline at Fermilab generates an intense muon neutrino beam for the NOvA (NuMI Off-axis $\nu_e$ Appearance) long-baseline neutrino experiment. Over the years, the NuMI beamline has been pivotal in advancing neutrino physics, providing invaluable data and insights. This proceeding paper discusses updates and the lessons learned from recent experiences during the beam operations, maintenance, and monitoring of the NuMI beamline. Key topics include the optimization of beam performance and challenges in maintaining beamline stability. The paper aims to share best practices and provide a road-map for future beamline projects, including the Long-Baseline Neutrino Facility (LBNF).

43 PARTICLE ACCELERATORS↗

Updates and Lessons Learned from NuMI Beamline at Fermilab

The Neutrinos at the Main Injector (NuMI) beamline at Fermilab generates an intense muon neutrino beam for the NOvA (NuMI Off-axis 𝜈𝑒 Appearance) long baseline neutrino experiment. Over the years, the NuMI beamline has been pivotal in advancing neutrino physics, providing invaluable data and insights. This presentation offers updates and a comprehensive review of the lessons learned from the operation, maintenance, and monitoring of the NuMI beamline. Key topics include the optimization of beam performance, challenges in maintaining beamline stability, and proposed Machine Learning implementations to enhance monitoring. The talk aims to share best practices and provide a roadmap for future beamline projects, including the Long-Baseline Neutrino Facility (LBNF).

Wickremasinghe, Athula↗

Machine Learning on Heterogeneous, Edge, and Quantum Hardware for Particle Physics (ML-HEQUPP)

The next generation of particle physics experiments will face a new era of challenges in data acquisition, due to unprecedented data rates and volumes along with extreme environments and operational constraints. Harnessing this data for scientific discovery demands real-time inference and decision-making, intelligent data reduction, and efficient processing architectures beyond current capabilities. Crucial to the success of this experimental paradigm are several emerging technologies, such as artificial intelligence and machine learning (AI/ML) and silicon microelectronics, and the advent of quantum algorithms and processing. Their intersection includes areas of research such as low-power and low-latency devices for edge computing, heterogeneous accelerator systems, reconfigurable hardware, novel codesign and synthesis strategies, readout for cryogenic or high-radiation environments, and analog computing. This white paper presents a community-driven vision to identify and prioritize research and development opportunities in hardware-based ML systems and corresponding physics applications, contributing towards a successful transition to the new data frontier of fundamental science.

Gonski, Julia [SLAC]↗

Polymer Additive Manufacturing for Marine Renewable Energy Applications: Best Practices, Research Trends, and Current Challenges

Additive manufacturing (AM) is a rapidly growing technology space, not only for prototyping, but is also becoming more feasible at larger scales and increasing component quantities. There are a large variety of AM processes and materials available to users and effectively applying those processes and materials to a specific use case can be challenging. One specific area where AM could be particularly beneficial is marine renewable energy (MRE). Not only is MRE a relatively nascent industry with a near-term need for rapid deployments and prototype testing, but developers could also see long-term benefits from the broad variety of environmentally resistant materials available and the ability to manufacture complex geometries that AM technologies offer. Over the past 4 years, AM materials have played an increasing role in the Advanced Materials project; a multi-year, multi-laboratory research project funded by the U.S. Department of Energy's Water Power Technologies Office, with the main goal of reducing barriers to the adoption of complex materials in the MRE industry. The primary focus of this project is to develop test methods and generate datasets to understand the long-term performance of advanced materials in marine environmental and address specific material challenges as they arise. This report provides an extensive overview of the research that has been performed specific to AM polymers as part of the Advanced Materials project. The intention of this document is to provide recommendations of best practices with regards to material selection, mechanical test method development, and design practices, lessons learned along the way, current research trends, and ongoing challenges with regards to AM polymers in marine environments. In particular, this report focuses on several key aspects: Material and process selection, Environmental conditioning and subsequent degradation quantification through mechanical characterization, Composite reinforcements on AM polymer substrates, Adhesion of instrumentation for mechanical characterization and loads measurements, Protective coatings for preventing biofouling and water ingress, Other MRE case studies where AM has proved particularly useful. Ultimately, we hope that the test methods that have been developed, data generated, and lessons learned from this research will be valuable to the MRE community (researchers and developers alike), as well as other industries, and can be used as a reference point as the respective MRE and AM industries continue to grow and mature.

16 TIDAL AND WAVE POWER↗

Monitoring Fracture Hydromechanical Evolution in the Lab and Field Using Unsupervised Metric Learning

Fractures evolve in time through thermal‐hydraulic‐mechanical‐chemical (THMC) processes that alter their long‐range hydraulic transport properties and modify subsurface behavior and activities. The location of subsurface fractures makes it necessary to use remote sensing techniques such as passive or active seismic monitoring for fracture characterization. In this paper, we develop a machine learning approach to monitor the evolution of fracture properties using passive seismic sources in a laboratory setting and using active seismic monitoring from the Sanford Underground Research Facility in Lead, South Dakota, at a depth of 1.25 km in amphibolite rock during stimulation of natural fractures as well as during induced fracturing. The unsupervised metric learning technique applies tandem neural networks (twin (Siamese) or triplet) with contrastive loss and adaptive margins to track slowly varying systems for which class or similarity labels are not available. The approach adopts locality‐sensitive hashing to divide time‐ordered contiguous data into an arbitrary number of pseudo‐classes. Contrastive‐loss training with many hash bins generates an evolving latent‐space trajectory. This approach enables unsupervised metric learning for seismic data stacks under the condition of contiguous state sampling and slowly varying fracture properties. The displacement discontinuity theory provides a mechanistic foundation for the fracture‐dependent trajectories that are related to relaxation of fractures with time‐dependent specific stiffness responding to changes in stress or fluid saturation.

02 PETROLEUM↗

RLMolLM: Reinforcement Learning-Enhanced Language Model Framework for Inverse Molecular Design

Inverse molecular design faces significant challenges due to vast chemical space and complex property requirements. While language models show promise for molecular generation, they struggle with validity, multi-property optimization, and structural constraints. This work presents RLMolLM, a reinforcement learning framework combining Proximal Policy Optimization (PPO) with genetic algorithms to address these limitations. Our approach optimizes multiple user-specified properties including quantitative estimates of drug-likeness (QED), synthetic accessibility (SA), and ADMET (absorption, distribution, metabolism, excretion, and toxicity) endpoints without requiring complete model retraining, while maintaining capability for scaffold-constrained generation where specific substructures must be preserved. We outperform state-of-the-art methods for molecular optimization, achieving best QED scores across GDB13, Moses, and Zinc datasets with up to 31% improvement over previous methods while maintaining excellent validity, uniqueness, and novelty metrics. For simultaneous multi-property optimization, our framework achieves substantial improvements in ADMET properties including 4.5-fold reduction in hERG toxicity and enhanced Caco-2 permeability compared to Moses dataset. Under structural constraints, the framework significantly improves molecular validity while preserving scaffolds and effectively optimizing properties. In conclusion, this versatile solution advances pharmaceutical and materials molecular design through effective integration of reinforcement learning and genetic algorithms with multi-property optimization and scaffold preservation.

Genetic algorithms↗

Enhancing generative molecular design via uncertainty-guided fine-tuning of variational autoencoders

In recent years, deep generative models have been successfully applied to various molecular design tasks, particularly in the life and materials sciences. One critical challenge for pre-trained generative molecular design (GMD) models is to fine-tune them to be better suited for downstream design tasks that aim at optimizing specific molecular properties. However, redesigning and training an existing effective generative model from scratch for each new design task are impractical. Furthermore, the black-box nature of typical downstream tasks that involve property prediction makes it nontrivial to optimize the generative model in a task-specific manner. In this work, we propose an uncertainty-guided fine-tuning strategy that can effectively enhance a pre-trained variational autoencoder (VAE) for GMD through performance feedback in an active learning setting. The strategy begins by quantifying the model uncertainty of the generative model using an efficient active subspace-based UQ (uncertainty quantification) scheme. Next, the decoder diversity within the characterized model uncertainty class is explored to expand the viable space of molecular generation. The low-dimensionality of the active subspace makes this exploration tractable using a black-box optimization scheme, which in turn enables us to identify and leverage a diverse set of high-performing models to generate enhanced molecules. Empirical results across six target molecular properties using multiple VAE-based generative models demonstrate that our uncertainty-guided fine-tuning strategy consistently leads to improved models that outperform the original pre-trained models.

97 MATHEMATICS AND COMPUTING↗

Delamination-informed lifecycle decisions: A dielectric and machine learning framework for composite sorting and recycling

Composite materials are widely used in aerospace, marine, and automotive sectors due to their high strength-to-weight ratio and durability. However, their long-term reliability can be compromised by damage accumulation. Specifically, delamination initiation serves as a precursor to structural failure, which is often difficult to detect during damage inspection. Identifying and sorting delamination initiation in samples not only increases operational safety while providing critical information for end-of-life decisions, which influences both the service life extension value and the efficiency of fiber extraction during recycling. This research addresses two challenges: (1) developing a nondestructive, ex-situ framework to sort composite materials based on damage severity, particularly delamination, and (2) understanding how damage in composites influences resin removal during pyrolysis. Both experimental work and finite element analysis were performed to predict critical stress levels that are associated with delamination onset. Based on these results, three loading levels 50 %, 75 %, and 90 % of maximum stress, were selected for controlled experiments, generating composite samples with varying extents of damage for machine learning model training. Microscopic imaging of these samples confirmed the damage progression from matrix cracking to delamination, validating the computational predictions. We explored supervised machine learning using dielectric measurements to classify damage states. Preliminary results show an artificial neural network can identify early delamination which is a potential precursor to failure, with 94.44 % accuracy on our dataset. A parallel investigation into the effect of damage severity on pyrolysis recycling showed that heavily delaminated samples required significantly less energy for comparable matrix removal than undamaged samples.

dielectric variables↗

Portable, heterogeneous ensemble workflows at scale using libEnsemble

libEnsemble is a Python-based toolkit for running dynamic ensembles, developed as part of the DOE Exascale Computing Project. The toolkit utilizes a unique generator–simulator–allocator paradigm, where generators produce input for simulators, simulators evaluate those inputs, and allocators decide whether and when a simulator or generator should be called. The generator steers the ensemble based on simulation results. Generators may, for example, apply methods for numerical optimization, machine learning, or statistical calibration. libEnsemble communicates between a manager and workers. Flexibility is provided through multiple manager–worker communication substrates each of which has different benefits. These include Python’s multiprocessing, mpi4py, and TCP. Multisite ensembles are supported using Balsam or Globus Compute. We overview the unique characteristics of libEnsemble as well as current and potential interoperability with other packages in the workflow ecosystem. We highlight libEnsemble’s dynamic resource features: libEnsemble can detect system resources, such as available nodes, cores, and GPUs, and assign these in a portable way. These features allow users to specify the number of processors and GPUs required for each simulation; and resources will be automatically assigned on a wide range of systems, including Frontier, Aurora, and Perlmutter. Such ensembles can include multiple simulation types, some using GPUs and others using only CPUs, sharing nodes for maximum efficiency. We also describe the benefits of libEnsemble’s generator–simulator coupling, which easily exposes to the user the ability to cancel, and portably kill, running simulations based on models that are updated with intermediate simulation output. We demonstrate libEnsemble’s capabilities, scalability, and scientific impact via a Gaussian process surrogate training problem for the longitudinal density profile at the exit of a plasma accelerator stage. In conclusion, the study uses gpCAM for the surrogate model and employs either Wake-T or WarpX simulations, highlighting efficient use of resources that can easily extend to exascale.

Dynamic ensembles↗

Women in Nuclear Security: The United Arab Emirates Story

Historically in the United Arab Emirates, security personnel stationed at nuclear sites were typically men with military or policing backgrounds. Although these groups bring important, transferrable skills and experience to the nuclear security mission, some Emirati stakeholders recognize the increased benefit of a greater diversity of voices and backgrounds within the nuclear security team hierarchy. By recruiting from a wider range of disciplines and encouraging gender diversity, the security team could develop broader insight into security challenges, and a greater variety of perspectives could bring forth more innovative solutions. When aligning a diverse, multinational workforce with a common mission, all members of a security organization must fundamentally change their individual approaches to the work; the entire staff must be on the same page. This can be particularly challenging for some people who find it difficult to transition from a framework of beliefs because of past experiences based on prescriptive, exclusionary, hierarchical thought. Also, national cultural values are learned early, held deeply, and change slowly over the course of generations, and inherent biases rooted in these values can create barriers to the collective progress of a team. Creating a more open, diverse, and transparent culture for the team by accepting a unique, egalitarian, professional nuclear security culture that all members adhere to can be a challenging task for some. In this article, I share my experiences related to professional roadblocks, personal challenges, and lessons learned, particularly those regarding the empowerment of women in science and engineering, while developing and establishing the United Arab Emirates nuclear security program. I also discuss the support (and underrepresentation) of women, including those currently in critical leadership positions, and their careers in nuclear security.

Barakah Nuclear Power Plant↗