Search NASA⌕ Search

SEARCH · Search NASA

Results for “Surrogate Optimization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 181 records · Page 10

Digital twin framework for PIP-II linac: AI-driven multi-scale modeling from ion source to 800 MeV

The PIP-II superconducting linac at Fermilab is designed to deliver multi-megawatt proton beams for neutrino physics and other high-intensity applications. To expedite commissioning and enhance operational reliability, we have developed an EPICS-based data flow framework that seamlessly integrates digital twins (DT) with physical twins (PT). These digital twins comprise high-fidelity beam dynamics models or data-driven surrogate models connected to their physical counterparts through real-time diagnostics and advanced machine-learning algorithms.Central to this framework is Linac_Gen, an accelerated simulation tool that incorporates convolutional neural networks, random forests, and genetic algorithms to provide up to a tenfold speedup in optimizing the accelerator geometry model. An EPICS translator layer ensures interoperability by efficiently mapping lattice parameters across diverse simulation platforms.Our EPICS-based framework supports multiple operational modes—monitoring, passive learning, closed-loop control, and online learning—covering the entire machine lifecycle. By leveraging HPC resources and multi-objective optimization techniques, the digital twin enables adaptive trajectory correction, real-time fault detection, and predictive modeling of beam stability. This comprehensive approach paves the way for robust, high-intensity operation and data-driven accelerator R&D at Fermilab.

Pathak, Abhishek [Fermilab]↗

TRANSFER LEARNING FOR FIELD EMISSION MITIGATION IN CEBAF SRF CAVITIES

The Continuous Electron Beam Accelerator Facility (CEBAF) at Jefferson Lab operates hundreds of super-conducting radio frequency (SRF) cavities in its two linear accelerators (linacs). Field emission (FE) is an ongoing operational challenge in higher gradient SRF cavities. FE generates high levels of neutron and gamma radiation leading to damaged accelerator hardware and a radiation hazard environment. During machine development periods, we performed gradient scans to record data capturing the relationship between cavity gradients and radiation levels measured throughout the linacs. However, the field emission environment at CEBAF varies considerably over time as the configuration of the radio frequency (RF) gradients changes and due to the changing behaviour of field emitters. An artificial intelligence/machine learning (AI/ML) approach with transfer learning could be a valuable tool to mitigate FE and lower the radiation levels. In this work, we mainly focus on leveraging the RF trip data gathered during CEBAF operations. We develop a transfer learning-based surrogate model for radiation detector readings given RF cavity gradients to track the CEBAF?s changing configuration and environment. Then, we could use the developed model as an optimization process for redistributing the RF gradients within a linac to minimize radiation levels.

Ahammed, K.↗

TRANSFER LEARNING FOR FIELD EMISSION MITIGATION IN CEBAF SRF CAVITIES

The Continuous Electron Beam Accelerator Facility (CEBAF) at Jefferson Lab operates hundreds of super-conducting radio frequency (SRF) cavities in its two linear accelerators (linacs). Field emission (FE) is an ongoing operational challenge in higher gradient SRF cavities. FE generates high levels of neutron and gamma radiation leading to damaged accelerator hardware and a radiation hazard environment. During machine development periods, we performed gradient scans to record data capturing the relationship between cavity gradients and radiation levels measured throughout the linacs. However, the field emission environment at CEBAF varies considerably over time as the configuration of the radio frequency (RF) gradients changes and due to the changing behaviour of field emitters. An artificial intelligence/machine learning (AI/ML) approach with transfer learning could be a valuable tool to mitigate FE and lower the radiation levels. In this work, we mainly focus on leveraging the RF trip data gathered during CEBAF operations. We develop a transfer learning-based surrogate model for radiation detector readings given RF cavity gradients to track the CEBAF?s changing configuration and environment. Then, we could use the developed model as an optimization process for redistributing the RF gradients within a linac to minimize radiation levels.

Ahammed, K.↗

A 3D helical filament surrogate model for 3D tokamak equilibria

A novel approach for efficient representation of three-dimensional (3D) tokamak equilibria is investigated, where a set of helical current filaments occupying the plasma region are employed to resolve deviations from the two-dimensional (2D) axi-symmetric state. A discrete set of 3D filaments, located at rational surfaces for a given toroidal mode number n and following the 2D equilibrium field lines (thus forming closed current loops), are found to provide a surrogate model of 3D equilibria with reasonable accuracy. Specifically, application of the filament model to 3D perturbed equilibria, due to the resonant magnetic perturbation (RMP) in DIII-D and MAST-U discharges, reveals that (1) a single helical filament per rational surface is sufficient; (2) 21 such helical filaments are capable of representing the n = 2 3D response field in MAST-U with less than 10% relative error as compared to that computed by a full magnetohydrodynamic code; (3) optimizing currents (both amplitude and phase) flowing in 3D filaments with fixed geometry, the highest accuracy fitting is found to depend on the characteristics of the 3D equilibria such as the coil current phasing of the RMP coils in our case studies. Here, whis filament approach is also applicable for generating surrogate models of other type of 3D tokamak equilibria, including those during the initial phase of the plasma disruption.

MARS-F↗

TorbeamNN: machine learning-based steering of ECH mirrors on KSTAR

We have developed TorbeamNN: a machine learning surrogate model for the TORBEAM ray tracing code to predict electron cyclotron heating (ECH) and current drive locations in tokamak plasmas. TorbeamNN provides more than a 100 times speed-up compared to the highly optimized and simplified real-time implementation of TORBEAM without any reduction in accuracy compared to the offline, full fidelity TORBEAM code. The model was trained using KSTAR ECH mirror geometries and works for both O-mode and X-mode absorption. The TorbeamNN predictions have been validated both offline and real-time in experiment. TorbeamNN has been utilized to track an ECH absorption vertical position target in dynamic KSTAR plasmas as well as under varying toroidal mirror angles and with a minimal average tracking error of 0.5 cm.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Harnessing the power of gradient-based simulations for multi-objective optimization in particle accelerators

Abstract Particle accelerator operation requires simultaneous optimization of multiple objectives. Multi-objective optimization (MOO) is particularly challenging due to trade-offs between the objectives. Evolutionary algorithms, such as genetic algorithms (GAs), have been leveraged for many optimization problems, however, they do not apply to complex control problems by design. This paper demonstrates the power of differentiability for solving MOO problems in particle accelerators using a deep differentiable reinforcement learning (DDRL) algorithm. We compare the DDRL algorithm with model-free reinforcement learning (MFRL), GA, and Bayesian optimization (BO) for simultaneous optimization of heat load and trip rates in the continuous electron beam accelerator facility. The underlying problem enforces strict constraints on both individual states and actions as well as cumulative (global) constraints on energy requirements of the beam. Using historical accelerator data, we develop a physics-based surrogate model which is differentiable and allows for back-propagation of gradients. The results are evaluated in the form of a Pareto-front with two objectives. We show that the DDRL outperforms MFRL, BO, and GA on high dimensional problems.

43 PARTICLE ACCELERATORS↗

Surrogate models for development of unconventional shale reservoirs by an integrated numerical approach of hydraulic fracturing, flow and geomechanics, and machine learning

We develop well-completion surrogate models by taking an integrated workflow of hydraulic fracturing, flow, geomechanics, and machine learning simulation. There are three steps in the proposed workflow. First, history-matching processes are conducted with the field data including pumping and production data for characterization. Second, full-physics simulation is performed with various parameters of the field development (e.g., cluster spacing, clusters per stage, pumping rates and times, amount of proppant, and well spacing) to generate multiple simulation results by changing the parameters of the completion design with well-known hydraulic fracturing, reservoir, geomechanics simulators to calculate fracture geometry, reservoir depressurization, induced stress changes. The workflow is demonstrated over a field in the Southern Midland Basin. Here, we take two completion scenarios: a single well case followed by a multi-well case. Finally, a Long Short-Term Memory (LSTM) machine learning algorithm is employed to create surrogate models that can replicate the full-physics simulation results. Furthermore, results show that the trained models applied in the single well and multi-well cases for a particular geological system can provide good accuracy close to those provided by full-physics simulations. Specifically, the site-specific surrogate models can predict fracture parameters (length, height, and surface area) and cumulative production accurately with computational efficiency, suggesting our proposed workflow can be used as a pragmatic tool for expediting the well completion optimization process.

Geomechanics↗

Batch Extraction Studies to Evaluate Trace Element Behavior in PUREX Conditions

The multilab Intentional Forensics Venture is working to identify which stable elements (i.e., taggants) at trace concentrations relative to U would persist throughout the nuclear fuel cycle in a voluntary fuel tagging scheme. A taggant would provide the nuclear forensics community with a “barcode” to help identify nuclear materials found outside of regulatory control. A portion of this project was focused on reprocessing effects and determining which, if any, elements would coextract with U(VI) in standard Pu–U reduction extraction (PUREX) conditions. Elements with a propensity to coextract could, in theory, be used as taggants from a PUREX perspective. Although retention is not a performance requirement, the taggant signature would need to partition predictably from the U stream after the PUREX process to maintain forensic utility. This report documents results from several batch extraction studies with numerous trace elements from HNO 3 (1.5–5 M), with and without U(VI), into 30% tri-n-butyl phosphate (TBP) in kerosene. Extraction and back-extraction tests were used to evaluate nearly 60 elements in surrogate conditions for PUREX, and distribution coefficients (i.e., D-values) for most species were <0.1, indicating few species are likely to co-extract with U through PUREX. Additional studies are needed to optimize sample volumes and dilutions to dial in these low D-values. The D-values (D) were determined for several of the more promising elements, including Re and Se. Ultimately, we conclude that only a limited number of the ~ 60 elements investigated are extractable in the U stream of PUREX, based on measured D values, meaning most candidate elemental taggants would likely be lost at this stage of the nuclear fuel cycle, even when considering a range of acid concentrations.

38 RADIATION CHEMISTRY, RADIOCHEMISTRY, AND NUCLEA↗

Digital Twin Framework for PIP-II Linac: AI-Driven Multi-Scale Modeling from Ion Source to 800 MeV

The PIP-II linac will enable >1.2 MW beam power for DUNE, requiring unprecedented operational reliability across its warm front-end (RFQ, MEBT) and five distinct SRF sections operating at 162.5/325/650 MHz. We present a comprehensive digital twin framework uniquely combining a fully differentiable fast beam transport code with neural network surrogates trained on high-fidelity PIC simulations, capturing space charge and nonlinear dynamics beyond traditional envelope codes while achieving 10⁴ speedup at <1% accuracy. End-to-end differentiability enables gradient-based optimization across 500+ parameters simultaneously previously impossible with conventional tools while the model incorporates static/dynamic errors and serves as a virtual commissioning platform for diverse hardware integration. The framework facilitates reinforcement learning for pulsed/CW mode transitions, predictive maintenance through anomaly detection, and autonomous tuning algorithm development with real-time execution capability. Validation against physics simulations shows excellent agreement for the front-end, with initial results demonstrating potential for 30% commissioning time reduction and proactive fault mitigation, providing a scalable blueprint for operating next-generation high-intensity accelerators.

Pathak, Abhishek [Fermilab] (ORCID:000000021704208↗

Digital Twin Framework for PIP-II Linac: AI-Driven Multi-Scale Modeling from Ion Source to 800 MeV

The PIP-II linac will enable >1.2 MW beam power for DUNE, requiring unprecedented operational reliability across its warm front-end (RFQ, MEBT) and five distinct SRF sections operating at 162.5/325/650 MHz. We present a comprehensive digital twin framework uniquely combining a fully differentiable fast beam transport code with neural network surrogates trained on high-fidelity PIC simulations, capturing space charge and nonlinear dynamics beyond traditional envelope codes while achieving 10⁴× speedup at <1% accuracy. End-to-end differentiability enables gradient-based optimization across 500+ parameters simultaneously—previously impossible with conventional tools—while the model incorporates static/dynamic errors and serves as a virtual commissioning platform for diverse hardware integration. The framework facilitates reinforcement learning for pulsed/CW mode transitions, predictive maintenance through anomaly detection, and autonomous tuning algorithm development with real-time execution capability. Validation against physics simulations shows excellent agreement for the front-end, with initial results demonstrating potential for 30% commissioning time reduction and proactive fault mitigation, providing a scalable blueprint for operating next-generation high-intensity accelerators.

Pathak, Abhishek [Fermilab] (ORCID:000000021704208↗

Energy-efficient, Large-scale Molecular Dynamics Simulations via Hardware- and Algorithm-level Optimization

This work aims to develop a framework for energy-efficient computing that will enable molecular dynamics (MD) simulations of large-scale phenomena with atomic precision and simultaneously remove computational bottlenecks limiting the speed of MD simulations. We seek to implement such an approach through the development of surrogate models for the interatomic force calculation combined with the use of mixed numerical precision formats. For a model system of neutral atoms (only pairwise interactions), significant force calculation efficiency improvements were achieved, without detrimental effects on atomic structures or average energies, using single precision, by developing a surrogate model (deep neural network), and by quantizing this surrogate model. For a model system of charged atoms, the reciprocal-space calculation of electrostatic interactions was identified as the main bottleneck, and the development of a surrogate model should be pursued to achieve an estimated one-order-of-magnitude additional speedup.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Comparing Control Performance Between Simulation and Experiment using the Microreactor Automated Control System Testbed

In the advanced reactor domain, a flexible and scalable software/hardware infrastructure is crucial for integrating and validating various control technologies. This study used the Microreactor Automated Control System (MACS) hardware platform as a testbed. MACS was originally designed to mirror Idaho National Laboratory (INL)'s Microreactor Applications Research Validation and Evaluation (MARVEL), a 85-kW thermal fission microreactor. It features control drums for simulated reactivity control; lights that function as a surrogate reactor core, with the brightness being proportional to the reactor power; and light sensors that emulate neutron detectors. To transform MACS into a physical twin of MARVEL for evaluating control methods, the Control and Optimization Modular Modeling Application for Nuclear Deployment (COMMAND) software was employed. This software integrated the hardware with two models of the MARVEL core, based on Reactor Excursion and Leak Analysis Program (RELAP5-3D) and Monte Carlo N-Particle (MCNP) models. The study aimed to demonstrate the gap between control theory and actual practice—a gap that often necessitates empirical adjustments such as control gain retuning, filters, time discretization, and integrator anti-windup measures. Controllers were developed based on increasingly complex simulations without hardware, starting from the base MARVEL model and then introducing actuator saturation constraints and sensor noise. The final control strategy was then tested using MACS, and a comparative performance analysis was conducted.

46 - INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AN↗

Fast Adaptive Neural Control of Resonant Extraction at Fermilab

We present the development of a machine learning (ML) based regulation system for third-order resonant beam extraction in the Mu2e experiment at Fermilab. Classical and ML-based controllers have been optimized using semi-analytic simulations and evaluated in terms of regulation performance and training efficiency. We compare several controller architectures and discuss the integration of neural control into an adaptive framework. We also present progress on surrogate models that predict the controller response given a spill intensity and controller action history. To enable real-time deployment, we report progress on implementing low-latency, edge-based inference suitable for hardware-constrained environments. Our results demonstrate the feasibility and advantages of ML-based control in managing complex, time-varying physical systems, with broader implications for accelerator operations and other domains requiring fast, adaptive regulation.

Berlioz, Jose Rene [Fermilab]↗

Adaptive Computing for Scale-Up Problems

Adaptive Computing is an application-agnostic outer loop framework to strategically deploy simulations and experiments to guide decision making for scale-up analysis. Resources are allocated over successive batches, which makes the allocation adaptive to some objective such as optimization or model training. The framework enables the characterization and management of uncertainties associated with predictive models of complex systems when scale-up questions lead to significant model extrapolation. A key advancement of this framework is its integration of multi-fidelity surrogate modeling, uncertainty management, and automated orchestration of various computing and experimentation resources into a single integrated software package. This enables efficient multi-fidelity modeling across multiple computing resources by incorporating real-world constraints such as relative queue times and throughput on individual machines into the multi-fidelity sampling decision. We discuss applications of this framework to problems in the renewable energy space, including biofuels production, material synthesis, perovskite crystal growth, and building electrical loads.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Analysis of semivolatile organics in liquid radioactive residue using sorptive stir bar and solvent back-extraction

Current methods for semivolatiles analysis in radioactive samples can produce large volumes of radioactive solvent residue. A method utilizing stir-bar sorptive extraction has been explored in this work for its applicability to radioactive waste samples. This low solvent analytical method may accelerate remediation, minimize hazardous solvent waste, and reduce exposure risk to workers. Organic compounds (polyaromatic hydrocarbons, chlorinated aromatics, and phenolic compounds) were chosen as surrogates for common Liquid Waste System (LWS) contaminants at Savannah River Site (Aiken, SC). Stir-bar extraction parameters (extraction time, matrix modification, and effective pH range) and solvent back extraction parameters (solvent type, volume, and extraction time) were optimized experimentally for the chosen compounds. Affinity of the stir-bar extraction polymer to radionuclides Cs-137 and Am-241 was observed to determine radionuclide concentration effects. The stir-bar method achieved mean recovery of 100 ± 0.7% (1σ), relative to 114 ± 7% using solvent extraction, while reducing weekly method hands-on time by 93.4% and solvent volume consumption by 99.3%. Sensitivity was improved by 378% in simulated tank waste and 278% in real-world LWS matrix, relative to solvent extraction. This work has produced a safe and optimized method for the low solvent analysis of organics in legacy radioactive tank waste by stir-bar sorptive extraction.

GC-MS↗

Divertor Plasma Detachment Control Neural Network

DivControlNN is a state-of-the-art software tool that leverages advanced machine learning techniques to predict and control divertor plasma behavior in fusion reactors. Plasma, a highly energetic and electrically charged gas, requires meticulous management to protect reactor components and maintain optimal energy production. Conventional simulation methods, although extremely detailed, typically demand extensive computational time-making them unsuitable for real-time control scenarios. DivControlNN addresses this challenge by learning from tens of thousands of high-fidelity simulations, thereby creating a rapid surrogate model that can deliver near-instantaneous predictions. At the core of its functionality is a sophisticated technique known as latent space mapping, which condenses complex, high-dimensional plasma data into a compact, lower-dimensional representation. This streamlined representation enables the system to quickly forecast essential plasma properties and determine the precise conditions required for effective detachment. Detachment is a crucial process in which the plasma is cooled before reaching the divertor plates, thereby reducing heat loads and mitigating material erosion. In recent experiments conducted on the KSTAR tokamak in South Korea, DivControlNN successfully guided the detachment process without any fine-tuning-even when applied to a new tungsten divertor configuration. By achieving a computational speed-up of over one hundred million times compared to traditional simulation methods while maintaining low prediction errors, DivControlNN stands to significantly enhance real-time control and diagnostic capabilities in future fusion reactors. This breakthrough paves the way for safer, more reliable reactor operation and represents a major advancement toward realizing fusion energy as a practical, sustainable, and clean power source.

Xu, Xueqiao [Lawrence Livermore National Laborator↗

Transient anisotropic kernel for probabilistic learning on manifolds

PLoM (Probabilistic Learning on Manifolds) is a method introduced in 2016 for handling small training datasets by projecting an Itô equation from a stochastic dissipative Hamiltonian dynamical system, acting as the MCMC generator, for which the KDE-estimated probability measure with the training dataset is the invariant measure. PLoM performs a projection on a reduced-order vector basis related to the training dataset, using the diffusion maps (DMAPS) basis constructed with a time-independent isotropic kernel. In this paper, we propose a new ISDE projection vector basis built from a transient anisotropic kernel, providing an alternative to the DMAPS basis to improve statistical surrogates for stochastic manifolds with heterogeneous data. The construction ensures that for times near the initial time, the DMAPS basis coincides with the transient basis. For larger times, the differences between the two bases are characterized by the angle of their spanned vector subspaces. The optimal instant yielding the optimal transient basis is determined using an estimation of mutual information from Information Theory, which is normalized by the entropy estimation to account for the effects of the number of realizations used in the estimations. Consequently, this new vector basis better represents statistical dependencies in the learned probability measure for any dimension. Three applications with varying levels of statistical complexity and data heterogeneity validate the proposed theory, showing that the transient anisotropic kernel improves the learned probability measure.

Diffusion maps↗

Active learning of ternary alloy structures and energies

Abstract Machine learning models with uncertainty quantification have recently emerged as attractive tools to accelerate the navigation of catalyst design spaces in a data-efficient manner. Here, we combine active learning with a dropout graph convolutional network (dGCN) as a surrogate model to explore the complex materials space of high-entropy alloys (HEAs). We train the dGCN on the formation energies of disordered binary alloy structures in the Pd-Pt-Sn ternary alloy system and improve predictions on ternary structures by performing reduced optimization of the formation free energy, the target property that determines HEA stability, over ensembles of ternary structures constructed based on two coordinate systems: (a) a physics-informed ternary composition space, and (b) data-driven coordinates discovered by the Diffusion Maps manifold learning scheme. Both reduced optimization techniques improve predictions of the formation free energy in the ternary alloy space with a significantly reduced number of DFT calculations compared to a high-fidelity model. The physics-based scheme converges to the target property in a manner akin to a depth-first strategy, whereas the data-driven scheme appears more akin to a breadth-first approach. Both sampling schemes, coupled with our acquisition function, successfully exploit a database of DFT-calculated binary alloy structures and energies, augmented with a relatively small number of ternary alloy calculations, to identify stable ternary HEA compositions and structures. This generalized framework can be extended to incorporate more complex bulk and surface structural motifs, and the results demonstrate that significant dimensionality reduction is possible in thermodynamic sampling problems when suitable active learning schemes are employed.

Chemistry↗