Search NASASearch

SEARCH · Search NASA

Results for “Conditional diffusion models”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Final report- UFL - RAPIDS2: A SciDAC Institute for Computer Science, Data, and Artificial Intelligence

The research initiatives supported by the U.S. Department of Energy (DOE) Grant DE-SC0022265 are fundamentally aimed at pioneering advanced machine learning (ML) techniques for scientific data compression within high-performance computing (HPC) environments. This comprehensive body of work addresses the critical challenge posed by the exponential growth of data generated by scientific simulations in domains such as fusion energy, climate modeling, and computational fluid dynamics (CFD). A core objective is to develop compression algorithms that achieve substantial data reduction—often by orders of magnitude—while rigorously ensuring the fidelity of both the primary data (PD) and scientifically crucial derived quantities of interest (QoI). The methodologies deployed under this grant integrate sophisticated deep learning architectures, prominently featuring autoencoders, advanced generative models like conditional diffusion, and hybrid learning techniques. Key innovations include the development of Guaranteed Autoencoders (GAE) and the Guaranteed Conditional Diffusion with Tensor Correction (GCDTC) framework, which provide explicit, instance-level error bounds on reconstructed data. Furthermore, specialized strategies such as nonlinear constraint satisfaction are employed to preserve the integrity of QoI, a vital requirement for the trustworthiness of downstream scientific analyses. This research also focuses on the design and implementation of scalable, GPU-accelerated software pipelines that seamlessly integrate into existing HPC workflows, ensuring both computational efficiency and practical applicability. The CAESAR framework, for example, unifies foundation and generative models to create an adaptive and efficient compression solution for spatio-temporal scientific data. Collectively, these efforts represent a significant advancement in mitigating the scientific data deluge, enabling more effective data management, accelerated scientific discovery, and optimized utilization of HPC resources.

97 MATHEMATICS AND COMPUTING

Report on the AAPM grand challenge on deep generative modeling for learning medical image statistics

Abstract Background The findings of the 2023 AAPM Grand Challenge on Deep Generative Modeling for Learning Medical Image Statistics are reported in this Special Report. Purpose The goal of this challenge was to promote the development of deep generative models for medical imaging and to emphasize the need for their domain‐relevant assessments via the analysis of relevant image statistics. Methods As part of this Grand Challenge, a common training dataset and an evaluation procedure was developed for benchmarking deep generative models for medical image synthesis. To create the training dataset, an established 3D virtual breast phantom was adapted. The resulting dataset comprised about 108 000 images of size 512 512. For the evaluation of submissions to the Challenge, an ensemble of 10 000 DGM‐generated images from each submission was employed. The evaluation procedure consisted of two stages. In the first stage, a preliminary check for memorization and image quality (via the Fréchet Inception Distance [FID]) was performed. Submissions that passed the first stage were then evaluated for the reproducibility of image statistics corresponding to several feature families including texture, morphology, image moments, fractal statistics, and skeleton statistics. A summary measure in this feature space was employed to rank the submissions. Additional analyses of submissions was performed to assess DGM performance specific to individual feature families, the four classes in the training data, and also to identify various artifacts. Results Fifty‐eight submissions from 12 unique users were received for this Challenge. Out of these 12 submissions, 9 submissions passed the first stage of evaluation and were eligible for ranking. The top‐ranked submission employed a conditional latent diffusion model, whereas the joint runners‐up employed a generative adversarial network, followed by another network for image superresolution. In general, we observed that the overall ranking of the top 9 submissions according to our evaluation method (i) did not match the FID‐based ranking, and (ii) differed with respect to individual feature families. Another important finding from our additional analyses was that different DGMs demonstrated similar kinds of artifacts. Conclusions This Grand Challenge highlighted the need for domain‐specific evaluation to further DGM design as well as deployment. It also demonstrated that the specification of a DGM may differ depending on its intended use.

Radiology, Nuclear Medicine & Medical Imaging

Latent diffusion can map beam loss to two-dimensional phase-space projections

Beam loss monitors (BLMs) and beam current monitors (BCMs) are ubiquitous at particle accelerators around the world. These simple devices provide noninvasive high-level beam measurements but give no insight into the detailed 6D (𝑥,𝑦,𝑧,𝑝 𝑥 ,𝑝 𝑦 ,𝑝 𝑧 ) beam phase-space distributions or dynamics. We show that generative conditional latent diffusion models can learn intricate patterns to solve the extreme inverse problem of mapping waveforms of tens of BLMs or BCMs along an accelerator to detailed 2D projections of a charged particle beam’s 6D phase-space density. This transformational method can be used at any particle accelerator to transform simple noninvasive devices into detailed beam phase-space diagnostics. We demonstrate this concept via multiparticle simulations of the high-intensity beam in the kilometer-long Los Alamos Neutron Science Center linear proton accelerator.

43 PARTICLE ACCELERATORS

A comparison of probabilistic generative frameworks for molecular simulations

Generative artificial intelligence is now a widely used tool in molecular science. Despite the popularity of probabilistic generative models, numerical experiments benchmarking their performance on molecular data are lacking. Here, in this work, we introduce and explain several classes of generative models, broadly sorted into two categories: flow-based models and diffusion models. We select three representative models: neural spline flows, conditional flow matching, and denoising diffusion probabilistic models, and examine their accuracy, computational cost, and generation speed across datasets with tunable dimensionality, complexity, and modal asymmetry. Our findings are varied, with no one framework being the best for all purposes. In a nutshell, (i) neural spline flows do best at capturing mode asymmetry present in low-dimensional data, (ii) conditional flow matching outperforms other models for high-dimensional data with low complexity, and (iii) denoising diffusion probabilistic models appear the best for low-dimensional data with high complexity. Our datasets include a Gaussian mixture model and the dihedral torsion angle distribution of the Aib9 peptide, generated via a molecular dynamics simulation. We hope our taxonomy of probabilistic generative frameworks and numerical results may guide model selection for a wide range of molecular tasks.

Artificial intelligence

ClimGen: Learning the Forcing-Response Relationship in Climate System

Solar Radiation Management (SRM) is emerging as a potential geoengineering strategy to address the anthropogenic impact on climate, but its effective implementation requires an iterative and large ensemble of highly accurate and efficient climate projections. Traditional climate projections rely on executing computationally demanding and time-consuming numerical climate models. Recent advances in machine learning (ML) aim to enhance these approaches by emulating traditional methods. In this work, we propose a novel framework for directly learning the relationship between solar radiation flux at the top of the atmosphere and the corresponding surface temperature response. To evaluate the feasibility of this direct ML-based projection, we developed a dataset using an intermediate complexity model, incorporating a comprehensive suite of different forcing patterns and evaluation metrics to rigorously assess the ML model’s performance. We introduce a Conditional Denoising Diffusion Probabilistic Model (cDDPM) for this task, which demonstrates encouraging skill in representing climate statistics under previously unseen forcing patterns. This approach provides a promising pathway for direct climate projections by accurately learning the forcing-response relationship, with a wide range of applications in impact mitigation, emissions policy design, and SRM strategies.

Chen, Tse-Chun [BATTELLE (PACIFIC NW LAB)] (ORCID:

Towards universal unfolding of detector effects in high-energy physics using denoising diffusion probabilistic models

Correcting for detector effects in experimental data, particularly through unfolding, is critical for enabling precision measurements in high-energy physics. However, traditional unfolding methods face challenges in scalability, flexibility, and dependence on simulations. We introduce a novel approach to multidimensional object-wise unfolding using conditional Denoising Diffusion Probabilistic Models (cDDPM). Our method utilizes the cDDPM for a non-iterative, flexible posterior sampling approach, incorporating distribution moments as conditioning information, which exhibits a strong inductive bias that allows it to generalize to unseen physics processes without explicitly assuming the underlying distribution. Our results highlight the potential of this method as a step towards a "universal" unfolding tool that reduces dependence on truth-level assumptions, while enabling the unfolding of a wide range of measured distributions with improved adaptability and accuracy.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC

Phase-field modeling of diffusion bonding in 316H stainless steel: Impact of processing conditions on grain morphology and bonding quality

A novel multi-phase, multi-component phase‐field model is presented to study the diffusion bonding of 316H stainless steel. Combined with targeted experimental investigations, this model simulates the bond-growth process and predicts the bonding quality. Unlike previous models, our approach captures the simultaneous evolution of voids and grain structures, while quantifying bonding quality using defined bonding ratio. A comprehensive analysis of bond process control is performed by changing temperature, pressure and surface roughness observing the resulting bond structure, which is consistent with experimental observations and analytical predictions. Temperature is determined to be the dominant factor, with the transition from a flat to a robust bond occurring between 1000 °C and 1050 °C. At the ideal bonding temperature of 1050 °C, a surface roughness exceeding 0.6 μm or an applied stress below 4 MPa results in poor bonding quality. Beyond this, higher pressures and smoother surfaces reduce void size, accelerate void shrinkage, and lead to improved bond integrity. This diffuse-interface model can be extended to other material systems if supplied with appropriate thermodynamic and kinetic data. In conclusion, this makes it an effective modeling platform for optimizing high-temperature diffusion bonding and developing reliable bonded components such as compact heat exchangers.

Diffusion bonding

Simulation of plasma and neutral transport in PISCES-RF using SOLPS-ITER

In this research, the fluid plasma transport code SOLPS-ITER is applied and validated against experimental data from the plasma interaction surface component experimental station (PISCES)-RF linear plasma device to establish a physics basis for plasma and neutral transport in its two magnetic field (B-field) geometry setups-(1)the cusp and (2) non-cusp or linear B-field. The main focus of this study is to understand (1) radial plasma transport (2) heat and particle loads on the upstream dump and downstream target plate, and (3) the physics of plasma-neutral interactions in PISCES-RF. The simulation setup adheres to typical PISCES-RF experimental conditions, with a 2D helicon power deposition profile as an input heating source. SOLPS-ITER simulations reproduce experimental conditions with the Bohm diffusion model for both B-field configurations of the PISCES-RF experiment. Major energy loss channels include neutral radiation and power deposited on the wall and dump plate, with only 1% of the input power reaching the target. The ionization front is well confined near the dump plate due to the heating and puffing regions. Additionally, SOLPS-ITER simulation results are also found to be in very good agreement with the particle-in-cell calculations using the code-PICOS++ which supports the validity of SOLPS in low collisionality regime.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

CaloChallenge 2022: a community challenge for fast calorimeter simulation

Here, we present the results of the ‘Fast Calorimeter Simulation Challenge 2022’—the CaloChallenge. We study state-of-the-art generative models on four calorimeter shower datasets of increasing dimensionality, ranging from a few hundred voxels to a few tens of thousand voxels. The 31 individual submissions span a wide range of current popular generative architectures, including variational autoencoders (VAEs), generative adversarial networks (GANs), normalizing flows, diffusion models, and models based on conditional flow matching. We compare all submissions in terms of quality of generated calorimeter showers, as well as shower generation time and model size. To assess the quality we use a broad range of different metrics including differences in one-dimensional histograms of observables, KPD/FPD scores, AUCs of binary classifiers, and the log-posterior of a multiclass classifier. The results of the CaloChallenge provide the most complete and comprehensive survey of cutting-edge approaches to calorimeter fast simulation to date. In addition, our work provides a uniquely detailed perspective on the important problem of how to evaluate generative models. As such, the results presented here should be applicable for other domains that use generative AI and require fast and faithful generation of samples in a large phase space.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND

DiffESM: Conditional Emulation of Temperature and Precipitation in Earth System Models With 3D Diffusion Models

Earth system models (ESMs) are essential for understanding the interaction between human activities and the Earth's climate. However, the computational demands of ESMs often limit the number of simulations that can be run, hindering the robust analysis of risks associated with extreme weather events. While low-cost climate emulators have emerged as an alternative to emulate ESMs and enable rapid analysis of future climate, many of these emulators only provide output on at most a monthly frequency. This temporal resolution is insufficient for analyzing events that require daily characterization, such as heat waves or heavy precipitation. We propose using diffusion models, a class of generative deep learning models, to effectively downscale ESM output from a monthly to a daily frequency. Trained on a handful of ESM realizations, reflecting a wide range of radiative forcings, our DiffESM model takes monthly mean precipitation or temperature as input, and is capable of producing daily values with statistical characteristics close to ESM output. Combined with a low-cost emulator providing monthly means, this approach requires only a small fraction of the computational resources needed to run a large ensemble. We evaluate model behavior using a number of extreme metrics, showing that DiffESM closely matches the spatio-temporal behavior of the ESM output it emulates in terms of the frequency and spatial characteristics of phenomena such as heat waves, dry spells, or rainfall intensity.

54 ENVIRONMENTAL SCIENCES

A novel conditional generative model for efficient ensemble forecasts of state variables in large-scale geological carbon storage

Integrating monitoring data to efficiently update reservoir pressure and CO 2 plume distribution forecasts presents a significant challenge in geological carbon storage (GCS) applications. Inverse modeling techniques are commonly used to fuse observational data and refine reservoir model parameters, thereby improving state variable forecasts. However, these techniques often rely on linear or Gaussian assumptions, which can limit their effectiveness in accurately predicting state variables. Moreover, simulating large-scale three-dimensional (3D) GCS problems is computationally expensive, making iterative runs in inverse problems prohibitive. To address these challenges, we propose a conditional generative model utilizing the score-based diffusion method for real-time 3D pressure and saturation field distribution predictions. Our approach involves solving the score function with a mini-batch-based Monte Carlo estimator to generate labeled data. This data is subsequently employed to train a fully connected neural network, enabling it to learn the conditional sample generator within a supervised learning framework. This method enables the rapid generation of a large ensemble of predictions, facilitating comprehensive uncertainty quantification of state variables. Here we applied our method to forecast the dynamic 3D distributions of pressure and saturation fields over a 30-year injection period. The statistical assessment with low root mean square error (RMSE) values demonstrates that our method can accurately predict the spatiotemporal distributions of both pressure and saturation fields. Moreover, the developed conditional generative model shows high computational efficiency by generating 100 ensemble forecasts of 3D state variables in less than 10 min. The consistency between ensemble averages and ground truth values further illustrates the model’s capability to capture state variable dynamics during the CO 2 plume injection process. Notably, the ground truth values fall within the ensemble forecasts, indicating that our uncertainty quantification effectively captures variability and potential noise in the observations. Thus, the developed conditional generative model proves to be a more efficient, accurate, and practical tool for GCS applications, facilitating timely risk analysis and informed decision-making.

58 GEOSCIENCES

Computational materials reliability assessment of hydrogen fueled gas turbine power generation engines

The use of blended fuel sources in land based gas turbine engines drives variations in the resulting operational profile (temperatures and pressures) which can impact engine reliability. Furthermore, variability in the manufacture of components affects the resulting microstructure which directly impacts material performance and reliability. Currently, data-driven models are typically used for maintaining and inspecting fleets of engines. Without explicitly capturing material and operational sources of variability conservatism must be used in developing component-level reliability models. Therefore, there exists an opportunity to use information from materials-scale physics models to better inform reliability modeling and reduce conservatism; the impact is more cost-efficient operation and maintenance of current and future fleets. Specifically, this work establishes a computational framework for evaluating the probabilistic high temperature creep performance of hot-section Ni-based superalloys where uncertainty comes from both microstructural and operational variability. A novel high-fidelity physics model which phenomenologically captures grain-boundary sensitive phenomena has been established. A probabilistic calibration procedure was used to calibrate the model and capture uncertainty in the parameterized model coefficients. A design of experiments methodology was established for identifying informative microstructural digital representations for suitable for forward model evaluation. Results show that training a machine-learning surrogate using this design criteria outperforms random selection of microstructural representations. Finally, two surrogate models were developed: (1) a deterministic surrogate model which predicts the local field response given microstructure, constitutive model parameters, and operating conditions (stress, temperature) and (2) a probabilistic model, where uncertainty comes from constitutive law uncertainty, built using denoising diffusion probabilistic models which samples responses given (1) microstructure and (2) operating conditions. These surrogate models enable partner Siemens Energy to rapidly perform UQ analysis specific to creep deformation across a range of microstructures and operating conditions. The impact is that these ML and physics codes can be used to establish more advanced reliability models for the inspection, servicing, and maintenance of land based gas turbine engines.

36 MATERIALS SCIENCE

Atomic layer deposition of nanofilms on porous polymer substrates: Strategies for success

Atomic layer deposition (ALD) is a versatile technique for engineering the surfaces of porous polymers, imbuing the flexible, high-surface-area substrates with inorganic and hybrid material properties. Previously reported enhancements include fouling resistance, electrical conductance, thermal stability, photocatalytic activity, hydrophilicity, and oleophilicity. However, there are many poorly understood phenomena that introduce challenges in applying ALD to porous polymers. In this paper, we address five common challenges and ways to overcome them: (1) entrapped precursor, (2) embrittlement, (3) film fracture, (4) deformation, and (5) pore collapse. These challenges are often interrelated and can exacerbate one another. To investigate these phenomena, we applied various ALD chemistries to porous polymers including polyethersulfone, polysulfone, polyvinylidene fluoride, and polycarbonate track-etched membranes. Reaction-diffusion modeling revealed why certain precursors and processing conditions result in embrittling subsurface material growth, entrapment of unreacted precursors, and nongrowth. We quantify the limits of ALD processing temperatures that are dictated by thermal expansion mismatch and can lead to fractured ALD films. The results herein allow us to make recommendations to avoid, mitigate, or overcome the difficulties encountered when performing ALD and plasma-enhanced ALD on porous polymers. We intend this article to serve as a “lessons learned” guide informed by previous experience to provide a better understanding of the difficulties and limitations of ALD on porous polymers and knowledge-based guidelines for successful depositions. This knowledge can accelerate future research and help experimentalists navigate and troubleshoot as they expose porous polymers to reactive precursor vapors.

36 MATERIALS SCIENCE

Field-level reconstruction from foreground-contaminated 21-cm maps

Current and upcoming 21-cm experiments will soon be able to map 21-cm spatial fluctuations in three dimensions for a wide range of redshifts. However, bright foreground contamination and the nature of radio interferometry create significant challenges, making it difficult to access rich cosmological information from the Fourier modes that lie within the “foreground wedge”. Here, in this work, we introduce two approaches aiming to reconstruct the full 21-cm density field, including the missing modes in the wedge: (a) a field-level inference under an effective field theory (EFT) framework; (b) a diffusion-based deep generative model trained on simulations. Under the EFT framework, we implement a fully differentiable forward model that maps the initial conditions of matter fluctuations to the observed, foreground-filtered 21-cm maps. This enables a gradient-based sampler to simultaneously sample the initial conditions and bias parameters, allowing a physically motivated mode reconstruction. Alternatively, we apply a variational diffusion model to perform 21-cm density reconstruction at the map level. Our model is trained on semi-numerical simulations over a wide range of astrophysical parameters. Our results from both approaches should provide improved cosmological constraints from the field level and also enable cross-correlation between experiments that have little or no overlapping modes.

cosmological perturbation theory

GenAI4UQ: A software for forward and inverse uncertainty quantification using conditional generative AI

We introduce GenAI4UQ, a software package for forward and inverse uncertainty quantification in model calibration, parameter estimation, and ensemble forecasting. GenAI4UQ leverages a generative AI-based conditional modeling framework to address limitations of traditional inverse modeling techniques, such as Markov Chain Monte Carlo (MCMC) methods. By replacing computationally intensive iterative processes with a direct, learned mapping, GenAI4UQ enables efficient calibration of input parameters and generation of predictions directly from observations. The software supports rapid ensemble forecasting with robust uncertainty quantification while maintaining computational and storage efficiency. Built-in auto-tuning of hyperparameters simplifies model training, ensuring accessibility for users with varying expertise. Its versatile conditional generative framework is applicable across diverse scientific domains. While GenAI4UQ offers significant advantages in flexibility and efficiency, users should interpret its uncertainty estimates with caution in data-sparse scenarios, as the model may overestimate uncertainty—an effect common to all surrogate-based approaches including MCMC with surrogate models. Despite this, GenAI4UQ transforms inverse modeling by providing a fast, reliable, and user-friendly solution. It empowers researchers and practitioners to quickly estimate parameter distributions and generate model predictions for new observations, facilitating efficient decision-making and advancing the state of uncertainty quantification in computational modeling.

97 MATHEMATICS AND COMPUTING

Conditional deep generative models for simultaneous simulation and reconstruction of entire events

We extend the particle-flow neural assisted simulations (arnassus) framework of fast simulation and reconstruction to entire collider events. In particular, we use two generative artificial intelligence tools, continuous normalizing flows and diffusion models, to create a set of reconstructed particle-flow objects conditioned on truth-level particles from CMS Open Simulations. While previous work focused on jets, our updated methods now can accommodate all particle-flow objects in an event along with particle-level attributes like particle type and production vertex coordinates. This approach is fully automated, entirely written in Python, and GPU-compatible. Using a variety of physics processes at the LHC, we show that the extended arnassus is able to generalize beyond the training dataset and outperforms the standard, public tool elphes.

Dreyer, Etienne [Weizmann Institute of Science, Re

The accuracy of multi-group models for nonlocal electron transport in magnetized plasmas

In the extreme conditions of inertial confinement fusion experiments, heat flow plays a vital role, but local diffusive models frequently break down and overestimate the heat flow. The situation becomes more complicated again in the significant magnetic fields generated during laser–plasma interactions or in magnetized fusion schemes. Accurate non-local and magnetized heat flow computations can be carried out using Vlasov–Fokker–Planck (VFP) simulations, but these are computationally expensive. There is, therefore, significant interest in using faster multi-group models to accurately calculate the non-local heat flow in magnetized plasmas. We benchmark two such multi-group models for calculating the heat flow, M1 and hybrid-AWBS-BGK, against diffusive models and full VFP simulations, before applying the models to realistic example test cases, both magnetized and unmagnetized. We find that the multi-group models generally perform very well for moderate non-localities up to kλmfp∼0.01, but the computational cost increases dramatically. hybrid-AWBS-BGK performs more effectively than M1 at high non-localities, up to kλmfp∼1, due to its adaptive solver and robust P1 closure, but tends to fail in very strong magnetic fields. Both codes are much faster than VFP simulations but are still slow in steep temperature gradients.

Arran, C. (ORCID:0000000286448118)