Search NASA⌕ Search

SEARCH · Search NASA

Results for “Linear accelerators”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

Accelerating GNNs on GPU Sparse Tensor Cores through N:M Sparsity-Oriented Graph Reordering

Recent advancements in GPU hardware support have introduced the capability to leverage N:M sparse patterns for substantial performance gains. Graphs in Graph Neural Networks (GNNs) are typically sparse, but the sparsity is often irregular, not conforming to such sparse patterns. In this paper, we propose a novel graph reordering algorithm, the first of its kind, to reshape irregular graph data into the N:M structured sparse pattern at the tile level, allowing linear-algebra-based graph operations in GNNs to benefit from the N:M sparse hardware. The optimization is lossless, maintaining the accuracy of GNN. It can remove 98-100\% violations of the N:M sparse patterns at the vector level, and increase the proportion of conforming graphs in SuiteSparse collection from 5-9\% to 88.7-93.5\%. On A100 GPUs, the optimization accelerates Sparse Matrix Matrix (SpMM) by up to 43X (2.3X -- 7.5X on average) and speeds up the key graph operations in GNNs on real graphs by as much as 8.6X (3.5X on average).

artificial intelligence, graph neural networks↗

Simulation of the thermoelectric effect in a multi-metallic superconducting cavity

Superconducting radio-frequency accelerating cavities made with different material layers, such as copper, Nb or Nb₃Sn, are susceptible to thermoelectric effects due to differences in Seebeck coefficients between the metals. A temperature gradient across the surfaces can drive thermoelectric currents, which may impact the cavity performance. A layered Cu/Nb/Nb3Sn single-cell cavity was tested with cryocoolers in 2022. Three heaters were mounted on the cavity surface at different locations and three single-axis cryogenic fluxgate magnetometers were attached close to the cavity equator. A linear increase in the magnetic field was measured while increasing the heaters' power. The cavity setup was analyzed with COMSOL and the results showed a trend similar to that observed in the experiment. This contribution details the approach chosen for the simulation and some of the challenges encountered.

Accelerator Physics↗

Flat Spectra of Energetic Particles in Interplanetary Shock Precursors

The observed energy spectra of accelerated particles at interplanetary shocks often do not match the diffusive shock acceleration (DSA) theory predictions. In some cases, the particle flux forms a plateau over a wide range of energies, extending upstream of the shock for up to seven flux e-folds before submerging into the background spectrum. Remarkably, at and downstream of the shock we have studied in detail, the flux falls off in energy as ϵ -1 , consistent with the DSA prediction for a strong shock. The upstream plateau suggests a particle transport mechanism different from those traditionally employed in DSA models. We show that a standard (linear) DSA solution based on a widely accepted diffusive particle transport with an underlying resonant wave–particle interaction is inconsistent with the plateau in the particle flux. To resolve this contradiction, we modify the DSA theory in two ways. First, we include a dependence of the particle diffusivity κ on the particle flux F (nonlinear particle transport). Second, we invoke short-scale magnetic perturbations that are self-consistently generated by, but not resonant with, accelerated particles. They lead to the particle diffusivity increasing with the particle energy as ∝ϵ 3/2 that simultaneously decreases with the particle flux as 1/F. The combination of these two trends results in the flat spectrum upstream. We speculate that nonmonotonic spatial variations of the upstream spectrum, apart from being time-dependent, may also result from non-DSA acceleration mechanisms at work upstream, such as stochastic Fermi or magnetic pumping acceleration.

79 ASTRONOMY AND ASTROPHYSICS↗

Shock-induced bubble jets: a dual perspective of bubble collapse and interfacial instability theory

Interactions between shock waves and gas bubbles in a liquid can lead to bubble collapse and high-speed liquid jet formation, relevant to biomedical applications such as shock wave lithotripsy and targeted drug delivery. This study reveals a complex interplay between acceleration-induced instabilities that drive jet formation and radial accelerations causing overall bubble collapse under shock wave pressure. Using high-speed synchrotron X-ray phase contrast imaging, the dynamics of micrometre-sized air bubbles interacting with laser-induced underwater shock waves are visualised. These images offer full optical access to phase discontinuities along the X-ray path, including jet formation, its propagation inside the bubble, and penetration through the distal side. Jet formation from laser-induced shock waves is suggested to be an acceleration-driven process. A model predicting jet speed based on the perturbation growth rate of a single-mode Richtmyer–Meshkov instability shows good agreement with experimental data, despite uncertainties in the jet-driving mechanisms. The jet initially follows a linear growth phase, transitioning into a nonlinear regime as it evolves. To capture this transition, a heuristic model bridging the linear and nonlinear growth phases is introduced, also approximating jet shape as a single-mode instability, again matching experimental observations. Upon piercing the distal bubble surface, jets can entrain gas and form a toroidal secondary bubble. Linear scaling laws are identified for the pinch-off time and volume of the ejected bubble relative to the jet’s Weber number, characterising the balance of inertia and surface tension. At low speeds, jets destabilise due to capillary effects, resulting in ligament pinch-off.

Drops and Bubbles: Bubble dynamics↗

Electric Drive Technologies Consortium (EDTC)/ Cost competitive, high-Performance, highly Reliable (CPR) Power Devices on 4H-SiC (Final Report)

4H-Silicon carbide (4H-SiC) is a wide bandgap semiconductor that offers superior material properties over silicon, including higher critical electric field, thermal conductivity, and electron saturation velocity. These advantages make 4H-SiC highly attractive for high-voltage, high-efficiency power electronics. However, realizing the full potential of SiC requires device technologies that are not only high-performing but also manufacturable and reliable under real-world operating conditions. This report summarizes the outcomes of a five-year R&D effort funded by the U.S. Department of Energy (DOE) under the Electric Drive Technologies Consortium (EDTC), focused on developing cost-competitive, high-performance, and highly reliable (CPR) power devices on 4H-SiC substrates. The program targeted scalable and manufacturable 1.2 kV-class SiC MOSFETs optimized for next-generation electric vehicles, renewable energy systems, and industrial power conversion. The project delivered transformative advancements in SiC power device performance and ruggedness. Particularly, Specific on-resistance (R on,sp ) was reduced by up to 37%, from ~4.0 m$\Omega \cdot$cm 2 in earlier designs to an industry-leading 2.40 m$\Omega \cdot$cm 2 , driven by optimized doping, refined JFET widths, and layout engineering. Breakdown voltages (BV) exceeded 1600 V, marking improvement over legacy baselines, and demonstrating the robustness of newly implemented junction profiles and edge terminations. Short-circuit withstand time (SCWT) saw a remarkable 4$\times$ increase, from ~2 $\mu$s to over 8 $\mu$s, achieved through the successful deployment of deep P-well structures (~1.8–2.0 $\mu$m) via channeling implantation. This innovative process breakthrough enabled precise junction formation without MeV-class implantation tools, reduced leakage under high field stress, and allowed even the shortest-channel devices (down to 0.3 $\mu$m) to achieve both high BV and excellent ruggedness—breaking the traditional trade-off between conduction efficiency and blocking capability. Several novel architectures pushed the performance envelope further. JBSFETs—featuring embedded Schottky portions—eliminated bipolar degradation and drastically reduced third-quadrant leakage, while Ladder MOSFETs introduced a clever orthogonal conduction path that achieved a 15.4% reduction in R on,sp over standard linear designs. Switching performance reached new benchmarks: short-channel devices showed a 31% reduction in total switching energy compared to 0.5 $\mu$m counterparts, while maintaining manageable gate drive requirements. Layout-optimized structures not only improved transconductance but also accelerated switching transitions, pointing to real-world benefits in converter-level efficiency. The devices also passed rigorous reliability validation. Stress-tested across TDDB, HTGB, HTRB, HVP, and burn-in, the devices screened under 30 V/10 hr and 43 V/1 s protocols consistently exhibited tighter lifetime distributions and long-term oxide robustness. These screening techniques proved effective in identifying latent defects and ensuring deployment-grade reliability. Meanwhile, advanced 3D TCAD simulations revealed and resolved electric field hotspots—particularly in HEXFET corners—where fields exceeding 4.8 MV/cm were mitigated through geometry-aware layout corrections. Overall, the results of this project demonstrate a manufacturable and scalable SiC power device platform that addresses key DOE performance targets for efficient, robust, and reliable 1.2kV 4H-SiC Power Devices. The developed technologies represent a meaningful step forward in the commercial readiness of high-voltage SiC solutions and provide a strong foundation for continued advancement in wide bandgap power electronics.

42 ENGINEERING↗

Extracting symplectic maps for space-charge dominated beams

Symplecticity of transfer maps is important for reliable evaluation of space-charge dominated beams in accelerators. Unfortunately, most simulation codes that include collective effects, such as space charge, do not use canonical phase-space variables and therefore are not symplectic in the presence of electromagnetic fields. In this paper, we present a numerical method to extract local linear symplectic transfer maps using particle tracking simulation code for space-charge dominated beams. We demonstrate this method for the photoinjector (113 MHz SRF gun) section of the Coherent electron Cooling (CeC) Proof of Principle (POP) experiment.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Integral Kernel Methods for Nonlinear Parabolic-Elliptic Systems

Nonlinear parabolic-elliptic systems arise in many physical, biological, and chemical phenomena such as chemotaxis, ion transport, self-gravitating particles, and Brownian vortices. Existing methods struggle with the strong coupling and high nonlinearity and nonlocality of some of these systems, especially the ill-conditioned, convection-dominated problems. To overcome numerical difficulties, current approaches rely on initial guesses, preconditioning, or iterative techniques with no convergence guarantees. They might suffer from poor scalability, large memory usage, and difficulty to parallelize. Inspired by the connection of parabolic-elliptic systems to stochastic processes, we introduce a novel meshless, monolithic, and fully explicit method that naturally encapsulates the elliptic and parabolic operators into a single step which updates each node deterministically with global information. By being fully quadrature-based, it avoids solving systems of discretized equations and does not utilize initial guesses or preconditioning, while requiring little memory and being easy to parallelize. We first derive the method in an integral kernel formulation with quadratic complexity in the number of integration nodes and then leverage kernel-independent fast multipole methods (FMM) to present a scalable algorithm with linear complexity. We provide numerical examples for the Poisson-Nernst-Planck equations in one, two, and three dimensions, together with the derivation of the integral kernel for each case. Furthermore, the examples demonstrate the fast convergence and scalability of the FMM-accelerated algorithm, as well as its suitability for convection-dominated problems, making it competitive against traditional PDE solvers.

PDE systems↗

Simplifying activations with linear approximations in neural networks

A key step in Neural Networks is activation. Among the different types of activation functions, sigmoid, tanh, and others involve the usage of exponents for calculation. From a hardware perspective, exponential implementation implies the usage of Taylor series or repeated methods involving many addition, multiplication, and division steps, and as a result are power-hungry and consume many clock cycles. We implement a piecewise linear approximation of the sigmoid function as a replacement for standard sigmoid activation libraries. This approach provides a practical alternative by leveraging piecewise segmentation, which simplifies hardware implementation and improves computational efficiency. In this paper, we detail piecewise functions that can be implemented using linear approximations and their implications for overall model accuracy and performance gain. Our results show that for the DenseNet, ResNet, and GoogLeNet architectures, the piecewise linear approximation of the sigmoid function provides faster execution times compared to the standard TensorFlow sigmoid implementation while maintaining comparable accuracy. Specifically, for MNIST with DenseNet, accuracy reaches 99.91% (Piecewise) vs. 99.97% (Base) with up to 1.31x speedup in execution time. For CIFAR-10 with DenseNet, accuracy improves to 98.97% (Piecewise) vs. 99.40% (Base) while achieving 1.24x faster execution. Similarly, for CIFAR-100 with DenseNet, the accuracy is 97.93% (Piecewise) vs. 98.39% (Base), with a 1.18x execution time reduction. These results confirm the proposed method’s capability to efficiently process large-scale datasets and computationally demanding tasks, offering a practical means to accelerate deep learning models, including LSTMs, without compromising accuracy.

Activation function↗

AI-Enabled Operations at Fermi Complex: Multivariate Time Series Prediction for Outage Prediction and Diagnosis

The Main Control Room of the Fermilab accelerator complex continuously gathers extensive time-series data from thousands of sensors monitoring the beam. However, unplanned events such as trips or voltage fluctuations often result in beam outages, causing operational downtime. This downtime not only consumes operator effort in diagnosing and addressing the issue but also leads to unnecessary energy consumption by idle machines awaiting beam restoration. The current threshold-based alarm system is reactive and faces challenges including frequent false alarms and inconsistent outage-cause labeling. To address these limitations, we propose an AI-enabled framework that leverages predictive analytics and automated labeling. Using data from $2,703$ Linac devices and $80$ operator-labeled outages, we evaluate state-of-the-art deep learning architectures, including recurrent, attention-based, and linear models, for beam outage prediction. Additionally, we assess a Random Forest-based labeling system for providing consistent, confidence-scored outage annotations. Our findings highlight the strengths and weaknesses of these architectures for beam outage prediction and identify critical gaps that must be addressed to fully harness AI for transitioning downtime handling from reactive to predictive, ultimately reducing downtime and improving decision-making in accelerator management.

Jain, Milan [PNL, Richland] (ORCID:000000021676111↗

GPU-Accelerated Analytic Simulation of Sparse Ionization Signal Formation in Pixelated Projection Detector

This paper presents a GPU-accelerated simulation package, TRED, for next-generation neutrino detectors with pixelated charge readout, leveraging community-driven software ecosystems to ensure adaptability and extensibility. We introduce two generic contributions: (i) an effective-charge representation based on Gaussian quadrature rules, in which the linear- interpolation factors for the field response inside each voxel are absorbed into the effective charge, and (ii) a sparse, block- binned tensor representation that enables efficient FFT-based computation of induced signals on readout electrodes for sparsely activated detector volumes. The former captures structure inside a voxel without dense sampling, while the latter achieves low memory usage and scalable runtime, as demonstrated in bench- mark studies. The underlying data representation is applicable to large-scale detectors and to other computational problems involving sparse activity.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Synergistic Integration of Multiple Wave Energy Converters with Adaptive Resonance and Offshore Floating Wind Turbines through Bayesian Optimization

We developed a synergistic ocean renewable system where an array of Wave Energy Converters (WEC) with adaptive resonance was collocated with a Floating Offshore Wind Turbine (FOWT) such that the WECs, capturing wave energy through the resonance adapting to varying irregular waves, consequently reduced FOWFT loads and turbine motions. Combining Surface-Riding WECs (SR-WEC) individually designed to feasibly relocate their natural frequency at the peak of the wave excitation spectrum for each sea state, and to obtain the highest capture width ratio at one of the frequent sea states for annual average power in a tens of kilowatts scale with a 15 MW FOWT based on a semi-submersible, Bayesian Optimization is implemented to determine the arrangement of WECs that minimize the annual representation of FOWT’s wave excitation spectra. The time-domain simulation of the system in the optimized arrangement is performed, including two sets of interactions: one set is the wind turbine dynamics, mooring lines, and floating body dynamics for FOWT, and the other set is the nonlinear power-take-off dynamics, linear mooring, and individual WECs’ floating body dynamics. Those two sets of interactions are further coupled through the hydrodynamics of diffraction and radiation. For sea states comprising Annual Energy Production, we investigate the capture width ratio of WECs, wave excitation on FOWT, and nacelle acceleration of the turbine compared to their single unit operations. We find that the optimally arranged SR-WECs reduce the wave excitation spectral area of FOWT by up to 60% and lower the turbine’s peak nacelle acceleration by nearly 44% in highly occurring sea states, while multiple WECs often produce more than the single operation, achieving adaptive resonance with a larger wave excitation spectra for those sea states. The synergistic system improves the total Annual Energy Production (AEP) by 1440 MWh, and we address which costs of Levelized Cost Of Energy (LCOE) can be reduced by the collocation.

Engineering↗

Accelerating phase field simulations through a hybrid adaptive Fourier neural operator with U-net backbone

Prolonged contact between a corrosive liquid and metal alloys can cause progressive dealloying. For one such process as liquid-metal dealloying (LMD), phase field models have been developed to understand the mechanisms leading to complex morphologies. However, the LMD governing equations in these models often involve coupled non-linear partial differential equations (PDE), which are challenging to solve numerically. In particular, numerical stiffness in the PDEs requires an extremely refined time step size (on the order of 10 -12 s or smaller). This computational bottleneck is especially problematic when running LMD simulation until a late time horizon is required. This motivates the development of surrogate models capable of leaping forward in time, by skipping several consecutive time steps at-once. In this paper, we propose a U-shaped adaptive Fourier neural operator (U-AFNO), a machine learning (ML) based model inspired by recent advances in neural operator learning. U-AFNO employs U-Nets for extracting and reconstructing local features within the physical fields, and passes the latent space through a vision transformer (ViT) implemented in the Fourier space (AFNO). We use U-AFNOs to learn the dynamics of mapping the field at a current time step into a later time step. We also identify global quantities of interest (QoI) describing the corrosion process (e.g., the deformation of the liquid-metal interface, lost metal, etc.) and show that our proposed U-AFNO model is able to accurately predict the field dynamics, in spite of the chaotic nature of LMD. Most notably, our model reproduces the key microstructure statistics and QoIs with a level of accuracy on par with the high-fidelity numerical solver, while achieving a significant 11, 200 × speed-up on a high-resolution grid when comparing the computational expense per time step. Finally, we also investigate the opportunity of using hybrid simulations, in which we alternate forward leaps in time using the U-AFNO with high-fidelity time stepping. We demonstrate that while advantageous for some surrogate model design choices, our proposed U-AFNO model in fully auto-regressive settings consistently outperforms hybrid schemes.

36 MATERIALS SCIENCE↗

Kinetic modeling of hot tail runaway electron generation during plasma disruptions using the JOREK code

The generation of runaway electrons (REs) during disruptions poses a significant challenge for the operation of tokamaks. The production of these high-energy electrons can cause substantial damage, particularly when the plasma current is high, making it a critical concern for ITER. For the high-temperature plasmas anticipated in ITER, the primary generation of REs may be dominated by the hot tail mechanism, which consists of the acceleration of hot electrons from the pre-disruption population which have not yet thermalized with the bulk following the rapid cooling of the plasma. To account for the significant 3D effects on RE production, a hot tail modeling framework has been developed within the non-linear 3D extended MHD code JOREK. This paper presents the structure of this framework, which is based on test electrons evolving in MHD fields. The verification of the method shows good agreement with the reference DREAM code for 0D test cases, as well as for axisymmetric simulations of 15 MA ITER H-mode disruption scenarios. Furthermore, a proof-of-principle application to a DIII-D case demonstrates the framework’s capability to capture for the first time the hot tail generation in 3D MHD simulations in realistic geometry. Preliminary results suggest that the production of REs is significantly reduced by stochastic losses.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Accelerating multigrid with streaming chiral SVD for Wilson fermions in lattice QCD

A modification to the setup algorithm for the multigrid preconditioner of Wilson fermions in lattice QCD is presented. A larger basis of test vectors than that used in regular multigrid is calculated by the smoother and truncated by singular value decomposition on the chiral components of the test vectors. The truncated basis is used to form the prolongation and restriction matrices of the multigrid hierarchy. This modification of the setup method is demonstrated to increase the convergence of linear solvers on an anisotropic lattice with m π ≈ 239 MeV from the Hadron Spectrum Collaboration and an isotropic lattice with m π ≈ 220 MeV from the MILC Collaboration. The lattice volume dependence of the method is also examined. Increasing the number of test vectors improves speedup up to a point, but storing these vectors becomes impossible in limited memory resources such as GPUs. To address storage cost, we implement a streaming singular value decomposition of the basis of test vectors on the chiral components and demonstrate a decrease in the number of fine level iterations by a factor of 1.7 for m q ≈ m crit

Iterative methods↗

Chromatic Transport Models for ImpactX

This note describes the basic levels of Hamiltonian approximation (for straight-axis elements) that are used in the symplectic beam dynamics modeling code ImpactX. These include: purely linear models, models based on a paraxial (chromatic) expansion of the Hamiltonian, and models based on the exact nonlinear Hamiltonian. The focus here is on paraxial (chromatic) models.

43 PARTICLE ACCELERATORS↗

Accelerating uncertainty quantification in incremental dynamic analysis using dimension reduction-based surrogate modeling

We propose a surrogate modeling framework based on dimension reduction to facilitate the quantification of seismic risk of structural systems in performance-based earthquake engineering. The framework adopts incremental dynamic analysis (IDA) for addressing hazard variability, and promotes significant computational efficiency improvement for propagating epistemic uncertainties associated with the structural models. It utilizes both linear and nonlinear dimension reduction approaches, equipped with inverse mappings, to learn a functional between the input parameter space (e.g., the epistemic uncertainties of the structure) to the high-dimensional output space created through the IDA implementation across different ground motions and seismic intensity levels. Polynomial chaos expansion is adopted as the surrogate model to learn this functional in the reduced space. A nine-story steel moment-resisting frame with uncertain structural properties is used as a testbed. Furthermore, we select the seismic fragility curves as a measure of the structure’s seismic performance, since it provides an estimate of the probability of entering specified damage states for given levels of ground shaking.

42 ENGINEERING↗

Dynamics of McMillan mappings II. axially symmetric map

Here, in this article, we investigate the transverse dynamics of a single particle in a model integrable accelerator lattice, based on a McMillan axially-symmetric electron lens. Although the McMillan e-lens has been considered as a device potentially capable of mitigating collective space charge forces, some of its fundamental properties have not been described yet. The main goal of our work is to close this gap and understand the limitations and potentials of this device. It is worth mentioning that the McMillan axially symmetric map provides the first-order approximations of dynamics for a general linear lattice plus an arbitrary thin lens with motion separable in polar coordinates. Therefore, advancements in its understanding should give us a better picture of more generic and not necessarily integrable round beams. In the first part of the article, we classify all possible regimes with stable trajectories and find the canonical action-angle variables. This provides an evaluation of the dynamical aperture, Poincaré rotation numbers as functions of amplitudes, and thus determines the spread in nonlinear tunes. Also, we provide a parameterization of invariant curves, allowing for the immediate determination of the map image forward and backward in time. The second part investigates the particle dynamics as a function of system parameters. We show that there are three fundamentally different configurations of the accelerator optics causing different regimes of nonlinear oscillations. Each regime is considered in great detail, including the limiting cases of large and small amplitudes. In addition, we analyze the dynamics in Cartesian coordinates and provide a description of observable variables and corresponding spectra.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Unifying thermochemistry concepts in computational heterogeneous catalysis

Thermophysical properties of adsorbates and gas-phase species define the free energy landscape of heterogeneously catalyzed processes and are pivotal for an atomistic understanding of the catalyst performance. These thermophysical properties, such as the free energy or the enthalpy, are typically derived from density functional theory (DFT) calculations. Enthalpies are species-interdependent properties that are only meaningful when referenced to other species. The widespread use of DFT has led to a proliferation of new energetic data in the literature and databases. However, there is a lack of consistency in how DFT data is referenced and how the associated enthalpies or free energies are stored and reported, leading to challenges in reproducing or utilizing the results of prior work. Additionally, DFT suffers from exchange–correlation errors that often require corrections to align the data with other global thermochemical networks, which are not always clearly documented or explained. In this review, we introduce a set of consistent terminology and definitions, review existing approaches, and unify the techniques using the framework of linear algebra. This set of terminology and tools facilitates the correction and alignment of energies between different data formats and sources, promoting the sharing and reuse of ab initio data. Standardization of thermochemistry concepts in computational heterogeneous catalysis reduces computational cost and enhances fundamental understanding of catalytic processes, which will accelerate the computational design of optimally performing catalysts.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗