Search NASA⌕ Search

SEARCH · Search NASA

Results for “Accelerators”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 523 records · Page 29

Designing and Utilizing Material Acceleration Platforms: Need for Workforce Development

In the quest to accelerate scientific discovery, the materials science field is rapidly moving toward the implementation of robotics and artificial intelligence driven workflows. Our recent summer school “Future Labs: Robotic Synthesis Coupled with Machine Learning for Energy Materials” provided learning opportunities for students, researchers, and educators in the materials science community. We describe this experience and provide our perspective on which new directions could be pursued to enable the future workforce to acquire cross-disciplinary skills.

Educational policy↗

Accelerating resonant spectroscopy simulations using multishifted biconjugate gradient

Resonant spectroscopies, which involve intermediate states with finite lifetimes, provide important insights into collective excitations in quantum materials that are otherwise inaccessible. However, theoretical understanding in this area is often limited by the numerical challenges of solving Kramers-Heisenberg-type response functions for large-scale systems. To address this, we introduce a multishifted biconjugate gradient algorithm that exploits the shared structure of Krylov subspaces across spectra with varying incident energies, effectively reducing the computational complexity to that of linear spectroscopies. Both mathematical proofs and numerical benchmarks confirm that this algorithm substantially accelerates spectral simulations, achieving constant complexity independent of the number of incident energies, while ensuring accuracy and stability. This development provides a scalable, versatile framework for simulating advanced spectroscopies in quantum materials.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Ancilla-entangling Floquet kicks for accelerating quantum algorithms

Quantum simulation with adiabatic annealing can provide insight into difficult problems that are impossible to study with classical computers. However, it deteriorates when the systems scale up due to the shrinkage of the excitation gap and thus places an annealing rate bottleneck for high success probability. Here, in this study, we accelerate quantum simulation using digital multiqubit gates that entangle primary system qubits with ancillary qubits. The practical benefits originate from tuning the ancillary gauge degrees of freedom to enhance the quantum algorithm's original functionality in the system registry. For simple but nontrivial short-ranged, infinite long-ranged transverse-field Ising models, and the hydrogen molecule model after qubit encoding, we show improvement in the time to solution by one hundred percent but with higher accuracy through exact state-vector numerical simulation in a digital-analog setting. The findings are further supported by time-averaged Hamiltonian theory.

97 MATHEMATICS AND COMPUTING↗

Ion irradiation induced crystalline disorder accelerates interfacial phonon conversion and reduces thermal boundary resistance

Traditional understanding of the thermal boundary resistance (TBR) across solid-solid interfaces posits that the vibrational densities of states overlap between materials dictates interfacial energy transport, with phonon scattering occurring at the interface. Using atomistic simulations, we show a mechanism for control of TBR; point defects near an interface can lead to both short- and midrange disorder, accelerating the conversion of vibrational energy between bulk and interfacial modes, ultimately reducing the TBR. In conclusion, we experimentally demonstrate this reduction through ion irradiation of gallium nitride and subsequently measuring the TBR across Al/GaN interfaces.

Pfeifer, Thomas W.↗

Investigation of post-breakup Coulomb acceleration using a trajectory model

Intermediate mass fragments ejected during the deexcitation of excited projectilelike fragments may promptly decay following ejection; the daughter particles that are subsequently produced are subject to interactions with the residual nucleus that affect final-state observables, a process herein referred to as post-breakup Coulomb acceleration. A simple classical Coulomb interaction model was used to study modification of 8 Be (2 + ), 5 Li (3/2 – ), 7 Li (7/2 – ), 7 Be (7/2 – ), and states in 12 B emitted following heavy-ion collisions of 28 Si + 12 C at 35 MeV/nucleon. Here, in contrast to previous work studying 8 Be (2 + ), excellent agreement between simulation and experiment was obtained using only Coulomb forces when either a Lorentzian or R-matrix line shape was used to describe the initial relative energy rather than a Gaussian. In consideration of the obtained results, improvements to the model and evaluation of experimental data are discussed as future directions, but it was concluded that the effects observed in the present data can be accurately described using only elements of classical mechanics and that the process is largely understood for a wide range of state lifetimes and mass and charge (a)symmetries.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Accelerating particle-in-cell kinetic plasma simulations via reduced-order modeling of space-charge dynamics using dynamic mode decomposition

We present a data-driven reduced-order modeling of the space-charge dynamics for electromagnetic particle-in-cell (EMPIC) plasma simulations based on dynamic mode decomposition (DMD). The dynamics of the charged particles in kinetic plasma simulations such as EMPIC is manifested through the plasma current density defined along the edges of the spatial mesh. We showcase the efficacy of DMD in modeling the time evolution of current density through a low-dimensional feature space. Not only do such DMD based predictive reduced-order models help accelerate EMPIC simulations, they also have the potential to facilitate investigative analysis and control applications. Here, we demonstrate the proposed DMD-EMPIC scheme for reduced-order modeling of current density and speedup in EMPIC simulations involving electron beam under the influence of magnetic field, virtual cathode oscillations, and backward wave oscillator.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Complete quasilinear model for the acceleration-driven lower hybrid drift instability and a computational assessment of its validity

A complete quasilinear model is derived for the electrostatic acceleration-driven lower hybrid drift instability in a uniform two-species low-beta plasma in which current is perpendicular to the background magnetic field. The model consists of coupled nonlinear velocity space diffusion equations for the volume-averaged ion and electron distribution functions. Each species' diffusion coefficient depends on a time-evolving spectral density of the electric-field energy per unit volume and a time-evolving dispersion relation. The dispersion relation is expressed analytically in integral form without the use of asymptotic limits and applies to arbitrary distribution functions, so long as they can be expressed as a function of one velocity coordinate, e.g., f⁡(vy) or f⁡(v⊥). The quasilinear model conserves energy and is complete in that it fully describes the evolution of the distribution functions, including resonant and nonresonant particle-wave interactions, while accounting for distribution-function-dependent mixed-complex frequencies. Further, the quasilinear diffusion model is solved numerically and self-consistently using a Crank-Nicolson temporal discretization and a second-order finite-volume velocity-space discretization. Numerical solutions are compared to nonlinear fourth-order accurate continuum kinetic Vlasov-Poisson simulations. Evolution of electric-field energy, growth rates, distribution functions, and diffusion coefficients are shown to be in agreement with Vlasov simulations. The quasilinear model is shown to predict anomalous transport terms, like resistivity and heating, to within a factor of order unity. Discrepancies between the quasilinear model and Vlasov simulations are assessed and attributed primarily to lack of damping in the quasilinear description and to the use of unperturbed-orbit susceptibilities in the linear theory dispersion relation. The results illuminate the predictive accuracy of the quasilinear model, place approximate bounds on its validity, and provide much needed vetting of quasilinear theory's ability to predict the nonlinear state of a microturbulent plasma.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Enhanced Isomer Population via Direct Irradiation of Solid-Density Targets Using a Compact Laser-Plasma Accelerator

Excitation of long-lived states in bromine nuclei using a tabletop laser-plasma accelerator providing pulsed (<100 fs) electron beams provided a sensitive probe of γ strength and level densities in the nuclear quasicontinuum and may indicate angular momentum coupling through electron-nuclear interactions. Solid-density active $LaBr_{3}$ targets absorb real and virtual photons up to 35 ± 2.5 MeV and deexcite through γ cascade into different states. Here, a factor of 4.354 ± 0.932 enhancement of the $^{80}Br^{m}/^{80}Br^{g}$ isomeric ratio was observed following electron irradiation, as compared to bremsstrahlung. Additional angular momentum transfer could possibly occur through nuclear-plasma or electron-nuclear interactions enabled by the ultrashort electron beam. Further investigation of these mechanisms could have far-reaching impact including decreased storage of long-term nuclear waste and an improved understanding of heavy element formation in astrophysical settings.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Machine learning-accelerated discovery of iron cobalt phosphides as rare-earth-free magnets

Here, the discovery of rare-earth-free permanent magnets has been a goal of scientists for decades. The absence of rare-earth elements will alleviate a pressing concern about the availability of rare-earth elements used in permanent magnets. These magnets are crucial for applications such as wind turbines, electric cars, and memory devices. Rare-earth magnets are special owing to a large magnetic anisotropy energy (K 1 ). In contrast, iron cobalt phosphides hold promise since doping P into cubic FeCo can induce anisotropy, leading to a large coercivity, without introducing rare-earth elements. We present a comprehensive search over the Fe-Co-P ternary space for magnets, utilizing recently developed adaptive machine learning feedback to efficiently screen over 850 000 structures. We focus on machine learning acceleration as a paradigm for materials design. Further adaptive genetic algorithm searches and first-principles calculations aid in the identification of 16 new structures below the known convex hull. Five of them possess high magnetic polarization (J s > 1 T). The structures with desirable magnetic properties center on (Fe,Co) 2⁢ P. This supports conventional wisdom, which focuses on the mixture of the two known end compounds: Fe 2 ⁢P and Co 2 ⁢P. Our work provides guidance for synthesis. We find Fe 7 ⁢CoP 4 shows the most promise (J s = 1.03T and K 1 = 0.83MJ/m 3 ).

36 MATERIALS SCIENCE↗

From Cell to System: Accelerated hpc Simulations of BESS Aging under Frequency Regulation and Arbitrage use cases

Lithium-ion battery energy storage systems (BESS) packs have emerged as a leading solution for grid-scale energy storage, enhancing resiliency and balancing load fluctuations. Yet, experimental characterization of large-format LIB packs-particularly to assess performance and degradation over hundreds of cycles - demands substantial hardware investment and multi-year testing campaigns. In this work, we couple a hierarchical, physics-based modeling framework agnostic to electrode chemistries with high-performance computing to accelerate systems level evaluation by upto two orders of magnitude. Building on the open-source liionpack platform, we implement cell, module, and pack-scale electrochemical models enriched with mechanistic aging mechanisms and deploy them on an HPC cluster to simulate 150−200kWh systems over 500 - 1,000 cycles with in days. We subject these virtual B ESS to both constant-current cycling and realistic grid service profiles spanning frequency regulation, ramp-rate support, and energy arbitrage-and quantify the resulting degradation patterns. Our results reveal that localized cell aging can induce substantial nonuniformity at module and pack levels, with service-specific cycling protocols driving distinct aging modes. This rapid, multiscale modeling approach provides a powerful design-space exploration tool for optimizing electrical architecture, control strategies, and operational schedules to prolong pack lifetime and lower total cost of ownership.

Ayalasomayajula, Surya [ORNL] (ORCID:0009000860788↗

FitCache: A Transparent Drop-In Framework for Multi-Tier Caching to Accelerate Distributed Deep Learning Workloads

Training in Deep learning (DL) remains highly compute- and data-intensive, with I/O becoming a critical bottleneck as models and datasets scale. Recent studies report that data loading can dominate training time, especially on large-scale HPC systems with shared parallel file systems (PFS). Existing caching approaches either rely on single-tier designs or require intrusive modifications to training pipelines, limiting their portability and effectiveness. In this work, we present FitCache, a transparent drop-in framework for multi-tier caching to accelerate distributed DL training by coordinating fast local memory (e.g., DRAM, Persistent Memory (PMem)) and NVMe as hierarchical caches atop PFS. Our design adapts to hardware diversity, i.e., if NVMe is missing, memory transparently acts as a caching tier, ensuring stable performance. FitCache transparently intercepts I/O requests and issues concurrent fetches across all tiers, returning data from the fastest responder without centralized metadata or static redirection paths. FitCache adapts to dynamic workloads and heterogeneous clusters while maintaining POSIX compatibility. Experiments on Frontier (2048 GPUs) and smaller research clusters show that FitCache reduces training time by up to 40% and per-batch I/O latency by up to 71.6% compared to Lustre Orion PFS, offering a drop-in solution for scalable DL training.

Hu, Guangxing [ORNL] (ORCID:0009000283203614)↗

Accelerating Traction Motor Optimization Design with AI Surrogate Models

The advancement of artificial intelligence systems enables the use of data-driven physics-based surrogate models to explore design spaces rapidly and deeply for engineering projects. This work presents a surrogate model workflow that accelerates electric traction motor design optimization by replacing finite element analysis (FEA) with an artificial neural network (ANN) and using this model in a genetic algorithm for design optimization. A baseline interior permanent-magnet motor is parameterized and sampled to generate FEA-labeled training data, after which a feed-forward ANN predicts key outputs (e.g., loss components and weight). The validated surrogate enables genetic-algorithm optimization and deep search over the design space without new FEA runs, producing Pareto-optimal trade-offs between weight and losses and set of optimized designs for rapid downselection of manufacturable motor designs.

Ribeiro, Pedro [ORNL] (ORCID:0009000921026641)↗

XaaS: Acceleration as a Service to Enable Productive High-Performance Cloud Computing

High-performance computing (HPC) and the cloud have evolved independently, specializing their innovations into performance or productivity. Acceleration as a Service (XaaS) is a recipe to empower both fields with a shared execution platform that provides transparent access to computing resources, regardless of the underlying cloud or HPC service provider. Bridging HPC and cloud advancements, XaaS presents a unified architecture built on performance-portable containers. Here, our converged model concentrates on low-overhead, high-performance communication and computing, targeting resource-intensive workloads from climate simulations to machine learning. XaaS lifts the restricted allocation model of Function as a Service (FaaS), allowing users to benefit from the flexibility and efficient resource utilization of serverless computing while supporting long-running and performance-sensitive workloads from HPC.

97 MATHEMATICS AND COMPUTING↗

Integrating Energy-Efficient Computing with Computational Research to Accelerate Energy Technology

NREL's computational sciences center hosts the largest high performance computing (HPC) capabilities dedicated to energy research while functioning as a living laboratory for energy-efficient computing. NREL's HPC capabilities support the research needs of the Department of Energy's Office of Energy Efficiency and Renewable Energy (EERE). In ten years of operation, HPC use in EERE-sponsored research has grown by a factor of 30, including work in electricity generation, energy efficiency, transportation, and energy system modeling. This paper analyzes this research portfolio, providing examples of individual use cases. The paper documents NREL's history of operating one of the world's most energy-efficient data centers while examining pathways to reduce economic and environmental impact beyond reduction of Power Usage Efficiency (PUE). This paper concludes by examining the unique opportunities created for accelerating improvements in data center efficiency created by combining an HPC system dedicated to energy research and a research program in energy-efficient computing.

97 MATHEMATICS AND COMPUTING↗

BCSR on GPU: A Way Forward Extreme-scale Graph Processing on Accelerator-enabled Frontier Supercomputer

Handling large graphs in a distributed environment requires effective partitioning across processors and efficient management of local partitions. In 2D partitioning, local graphs often become too sparse, making memory-efficient data structures crucial. Using the Compressed Sparse Row (CSR) format wastes space, especially for > 83% of vertices with empty edges for the sparse graphs. This study explores bit-CSR (BCSR), a modified CSR representation, on GPUs to reduce memory usage in graph computations. We achieved 16.67% memory savings on a sparse rmat dataset with 268 million vertices and 357 million edges, without performance degradation, supported by both theoretical and experimental storage savings of 33%. However, we observed a 1.7× slowdown in degree lookup times due to bitwise operations on AMD CPUs. This analysis highlights the potential of BCSR on GPUs for improving Graph500 benchmark performance on GPU-accelerated systems, such as the Frontier supercomputer.

Sattar, Naw Safrin↗

FPGA-Accelerated Range-Limited Molecular Dynamics

Long timescale Molecular Dynamics (MD) simulation of small molecules is crucial in drug design and basic science. To accelerate a small data set that is executed for a large number of iterations, high-efficiency is required. Recent work in this domain has demonstrated that among COTS devices only FPGA-centric clusters can scale beyond a few processors. The problem addressed here is that, as the number of on-chip processors has increased from fewer than 10 into the hundreds, previous intra-chip routing solutions are no longer viable. We find, however, that through various design innovations, high efficiency can be maintained. These include replacing the previous broadcast networks with ring-routing and then augmenting the rings with out-of-order and caching mechanisms. Others are adding a level of hierarchical filtering and memory recycling. Two novel optimized architectures emerge, together with a number of variations. These are validated, analyzed, and evaluated. We find that in the domain of interest speed-ups over GPUs are achieved. Finally, the potential impact is that this system promises to be the basis for scalable long timescale MD with commodity clusters.

97 MATHEMATICS AND COMPUTING↗

GPU-Accelerated Analytic Simulation of Sparse Ionization Signal Formation in Pixelated Projection Detector

This paper presents a GPU-accelerated simulation package, TRED, for next-generation neutrino detectors with pixelated charge readout, leveraging community-driven software ecosystems to ensure adaptability and extensibility. We introduce two generic contributions: (i) an effective-charge representation based on Gaussian quadrature rules, in which the linear- interpolation factors for the field response inside each voxel are absorbed into the effective charge, and (ii) a sparse, block- binned tensor representation that enables efficient FFT-based computation of induced signals on readout electrodes for sparsely activated detector volumes. The former captures structure inside a voxel without dense sampling, while the latter achieves low memory usage and scalable runtime, as demonstrated in bench- mark studies. The underlying data representation is applicable to large-scale detectors and to other computational problems involving sparse activity.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Fostering Peat Moss Feedbacks to Accelerate Peatland Restoration

Extensive knowledge exists on plant-species traits and functions, but we understand less about how population- or community-level emergent traits influence ecosystem functioning. This knowledge gap is important for ecosystems like peatlands, arid drylands, salt marshes, seagrass meadows and mangroves, where emergent traits of plant communities can create plant-environment feedbacks that amplify or dampen ecosystem processes. Recent insights from restoration ecology suggest that these feedbacks can critically influence restoration success. Despite growing recognition of emergent trait-driven feedbacks in other ecosystems, they remain underexplored in peatland restoration, world’s most carbon-dense ecosystem. Here, we review emergent self-amplifying and self-dampening feedbacks with net positive effects for peat moss-dominated systems. We show how these feedbacks can promote key physical, chemical, and biological processes that enhance peat moss growth, increase water retention, and reduce microbial decomposition of organic matter. Understanding and fostering these feedbacks offers a promising framework to accelerate peatland restoration across diverse degradation states.

Sphagnum↗