Search NASA⌕ Search

SEARCH · Search NASA

Results for “high performance computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 577 records · Page 32

Performance of the CMS high-level trigger during LHC Run 2

The CERN LHC provided proton and heavy ion collisions during its Run 2 operation period from 2015 to 2018. Proton-proton collisions reached a peak instantaneous luminosity of 2.1 $\times$ 10$^{34}$ cm$^{-2}$s$^{-1}$, twice the initial design value, at $\sqrt{s}$ = 13 TeV. The CMS experiment records a subset of the collisions for further processing as part of its online selection of data for physics analyses, using a two-level trigger system: the Level-1 trigger, implemented in custom-designed electronics, and the high-level trigger, a streamlined version of the offline reconstruction software running on a large computer farm. This paper presents the performance of the CMS high-level trigger system during LHC Run 2 for physics objects, such as leptons, jets, and missing transverse momentum, which meet the broad needs of the CMS physics program and the challenge of the evolving LHC and detector conditions. Sophisticated algorithms that were originally used in offline reconstruction were deployed online. Highlights include a machine-learning b tagging algorithm and a reconstruction algorithm for tau leptons that decay hadronically.

high energy physics↗

Design Overview of a High-Pressure Helium Flow Visualization Apparatus for Blanket Cooling Studies

Cooling of the fusion blanket first wall remains a significant challenge given the adverse conditions of heat and particle flux encountered near the plasma. Helium emerges as an attractive cooling candidate because of its chemical and neutronic inertness and separability from hydrogenic species (e.g. tritium). Because of the low thermal mass of helium, optimization of these coolant channels is warranted to provide high heat transfer performance at low pumping costs. Increasingly, computational fluid dynamics (CFD) simulations are employed to model and optimize these flow channels, and accompanying experimental data are needed to validate the predictions of these models. To provide the aforementioned experimental data, a high-pressure helium flow visualization upgrade has been designed for the Helium Flow Loop Experiment facility. This apparatus was built to American Society of Mechanical Engineers boiler and pressure vessel standards to withstand operating pressure of 4 MPa and mated to high-pressure glass windows. Seedless flow visualization is performed via high-speed background oriented schlieren (BOS), with image correlation used for time-resolved two-dimensional velocimetry at frequencies in excess of 60 kHz. Rectangular flow channel test articles are additively manufactured via laser powder bed fusion and installed into this visualization apparatus, with one-sided heating supplied by resistive heaters. In conclusion, the chosen test geometries were informed by prior CFD simulations, and the helium flow structures observed via BOS (detachment, recirculation, etc.) will be used for the validation of these accompanying models, in support of the design and optimization of blanket cooling channel configurations.

Helium flow↗

Latent heat thermal energy storage performance maps enabling fast & accurate building energy simulations

Thermal energy storage (TES) using phase change materials (PCMs) has gained attention as an effective approach to manage energy demand fluctuations and shift peak building loads. PCM embedded heat exchangers (PCM-HXs) offer high energy storage density and low temperature variation during phase change, being suitable for load-shifting applications. However, this component is typically evaluated using computationally expensive methods, which present significant challenges when the ultimate goal is to assess the performance of PCM-HX integrated thermal energy storage systems in the full building context. In this paper, we present a methodology to generate highly accurate and computationally efficient PCM-HX performance maps which can be easily integrated into building energy simulation tools to analyze the feasibility of space conditioning systems with latent heat PCM-based TES. The performance maps are generated using a computationally efficient PCM-HX simulation tool based on a Generalized Resistance-Capacitance Model (GRCM) which can simulate arbitrary PCM-HXs with high accuracy and significantly less computational effort compared to full CFD simulations. The methodology was verified for a case study considering a 5-ton (~17.5 kW) air-to-water heat pump-thermal energy storage system (HP-TES), which was co-simulated in Modelica for a DOE prototype small-office building in Vienna, Austria, using Spawn of EnergyPlus™. The TES performance maps provided accurate predictions of PCM-HX behavior when used as Modelica component, with deviations within 2-4% while also achieving at least 103 computational time reduction. Leveraging this faster prediction capability, four PCMs with different melting temperatures for cooling (12°C, 16°C) and heating (31°C, 36°C) were assessed to investigate their impact on system performance. This work highlights the importance of robust PCM-HX models for efficient and high-fidelity building-level simulations, presenting new opportunities for advanced control strategy development and parametric analysis of TES configurations in a computationally efficient manner

Modelica Building Simulations↗

Finding MIDDLE Ground: Scalable and Secure Distributed Learning

Edge computing methods allow devices to efficiently train a high-performing, robust, and personalized model for predictive tasks. However, these methods succumb to privacy and scalability concerns such as adversarial data recovery and expensive model communication. Furthermore, edge computing methods unrealistically assume that all devices train an identical model. In practice, edge devices have varying computational and memory constraints which may not allow certain devices to have the space or speed to train a specific model. To overcome these issues, we propose MIDDLE: a model independent distributed learning algorithm which allows heterogeneous edge devices to assist each other’s training while communicating only non-sensitive information. MIDDLE unlocks the ability for edge devices, regardless of computational or memory constraints, to assist each other even with completely different model architectures. Furthermore, MIDDLE does not require model or gradient communication which greatly reduces communication size and time. We prove that MIDDLE attains the optimal convergence rate O(1/sqrt(TM)) of stochastic gradient descent for convex and non-convex smooth optimization (for total iterations T and batch size M). Finally, our experimental results demonstrate that MIDDLE (even in non-IID data settings) attains robust and high-performing models without model or gradient communication.

Bornstein, Marc I.↗

ZERNIPAX: A fast and accurate Zernike polynomial calculator in Python

Zernike polynomials serve as an orthogonal basis on the unit disc, and have proven to be effective in optics simulations, astrophysics, and more recently in plasma simulations. Unlike Bessel functions, Zernike polynomials are inherently finite and smooth at the disc center (r=0), ensuring continuous differentiability along the axis. This property makes them particularly suitable for simulations, requiring no additional handling at the origin. We developed ZERNIPAX, an open-source Python package capable of utilizing CPU/GPUs, leveraging Google's JAX package and available on GitHub as well as the Python software repository PyPI. Furthermore, our implementation of the recursion relation between Jacobi polynomials significantly improves computation time compared to alternative methods by use of parallel computing while still performing more accurately for high-mode numbers.

Astrophysics↗

MBX V1.2: Accelerating Data-Driven Many-Body Molecular Dynamics Simulations

The MBX software provides an advanced platform for molecular dynamics simulations, leveraging state-of-the-art MB-pol and MB-nrg data-driven many-body potential energy functions. Developed over the past decade, these potential energy functions integrate physics-based and machine-learned many-body terms trained on electronic structure data calculated at the "gold standard" coupled-cluster level of theory. Recent advancements in MBX have focused on optimizing its performance, resulting in the release of MBX v1.2. While the inherently many-body nature of MB-pol and MB-nrg ensures high accuracy, it poses computational challenges. MBX v1.2 addresses these challenges with significant performance improvements, including enhanced parallelism that fully harnesses the power of modern multicore CPUs. In conclusion, these advancements enable simulations on nanosecond time scales for condensed-phase systems, significantly expanding the scope of high-accuracy, predictive simulations of complex molecular systems powered by data-driven many-body potential energy functions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

buhito

buhito is a Python library for graph analysis and machine learning. Graphs can represent networks with objects as nodes and their relationships as edges. buhito focuses on graphlet methods that study graphs through enumerating their component subgraphs to enable interpretable and fast models of complex systems. The package provides tools for different algorithmic designs for computing, analyzing, and applying graphlets to research problems such as machine learning, data compression, and anomaly detection in graph-structured data. A central feature is performing decomposition data analysis on graphs for machine learning models. Implemented in Python and built upon open-source scientific libraries such as NetworkX, NumPy, and SciPy, buhito provides high-performance methods for researchers exploring the mathematical and computational foundations of graphlet analysis applicable to systems of different sizes.

Pimonova, Yulia↗

Comparative S/TEM study of superconducting Ta quantum resonators by wet and dry etching types

Superconducting resonators play a pivotal role in various quantum technology applications, such as quantum computing and high-frequency communication systems. The performance of these resonators is closely tied to the properties of the superconducting films used in their fabrication. Here, in this study, we investigated the impact of wet and dry etching on tantalum (Ta) films leveraging advanced scanning transmission electron microscopy-based characterization methods and examined the morphological, chemical, and strain changes caused by the etching processes. Consequently, we report the significant differences between the two etching methods, with dry etching resulting in straight slanted sidewalls and a thinner oxidized layer, while wet etching produced curved sidewalls and undercuts. Both methods led to the formation of a residual Ta wedge at the lower part of the sidewall, causing lattice deformation, which could adversely influence the homogeneous operations of superconducting devices. These insights enhance our understanding of how etching influences superconducting films, offering valuable guidance for optimizing resonators and related devices. Our findings mark a significant stride in advancing quantum technologies and high-frequency communication by enhancing our practical understanding of superconducting material fabrication.

36 MATERIALS SCIENCE↗

Computational design of high entropy alloy coating for hydrogen turbine applications

This project aims to develop novel high entropy alloy (HEA)-based coatings to protect critical components in hydrogen-fueled turbine power system. The HEA-coatings will demonstrate superior performance in hydrogen combustion environment to commercial NiCoCrAlY coating in current natural gas turbine system. The HEA coating facilitates the formation of a protective scale of alpha-alumina to slow down the inward diffusion of oxidizing species and the outward diffusion of metal elements, and possesses ultrahigh corrosion and spallation resistance to prolong the service lifetime of critical components in hydrogen turbine power system. Aimed to accelerate the discovery of novel HEA coating compositions, high throughput computational modeling including CALPHAD and density functional theory and machine learning are performed to predict phase stability, oxygen permeability, oxidation rate constant, coefficient of thermal expansion, and mechanical properties. Based on the modeling and machine learning prediction, experimental validation is performed. Preliminary results will be presented and approaches to minimize oxidation will be discussed.

alloy design↗

Computational materials reliability assessment of hydrogen fueled gas turbine power generation engines

The use of blended fuel sources in land based gas turbine engines drives variations in the resulting operational profile (temperatures and pressures) which can impact engine reliability. Furthermore, variability in the manufacture of components affects the resulting microstructure which directly impacts material performance and reliability. Currently, data-driven models are typically used for maintaining and inspecting fleets of engines. Without explicitly capturing material and operational sources of variability conservatism must be used in developing component-level reliability models. Therefore, there exists an opportunity to use information from materials-scale physics models to better inform reliability modeling and reduce conservatism; the impact is more cost-efficient operation and maintenance of current and future fleets. Specifically, this work establishes a computational framework for evaluating the probabilistic high temperature creep performance of hot-section Ni-based superalloys where uncertainty comes from both microstructural and operational variability. A novel high-fidelity physics model which phenomenologically captures grain-boundary sensitive phenomena has been established. A probabilistic calibration procedure was used to calibrate the model and capture uncertainty in the parameterized model coefficients. A design of experiments methodology was established for identifying informative microstructural digital representations for suitable for forward model evaluation. Results show that training a machine-learning surrogate using this design criteria outperforms random selection of microstructural representations. Finally, two surrogate models were developed: (1) a deterministic surrogate model which predicts the local field response given microstructure, constitutive model parameters, and operating conditions (stress, temperature) and (2) a probabilistic model, where uncertainty comes from constitutive law uncertainty, built using denoising diffusion probabilistic models which samples responses given (1) microstructure and (2) operating conditions. These surrogate models enable partner Siemens Energy to rapidly perform UQ analysis specific to creep deformation across a range of microstructures and operating conditions. The impact is that these ML and physics codes can be used to establish more advanced reliability models for the inspection, servicing, and maintenance of land based gas turbine engines.

36 MATERIALS SCIENCE↗

Ginkgo - A math library designed to accelerate Exascale Computing Project science applications

Large-scale simulations require efficient computation across the entire computing hierarchy. A challenge of the Exascale Computing Project (ECP) was to reconcile highly heterogeneous hardware with the myriad of applications that were required to run on these supercomputers. Mathematical software forms the backbone of almost all scientific applications, providing efficient abstractions and operations that are crucial to harness the performance of computing systems. Ginkgo is one such mathematical software library, nurtured by ECP, providing high-performance, user-friendly, and performance portable interfaces for applications in ECP and beyond. In this paper, we elaborate on Ginkgo’s philosophy of high-performance software that is sustainable, reproducible, and easy to use. We showcase the wide feature set of solvers and preconditioners available in Ginkgo and the central concepts involved in their design. We elaborate on four different ECP software integrations: MFEM, PeleLM + SUNDIALS, XGC, and ExaSGD that use Ginkgo to accelerate their science runs. Performance studies of different problems from these applications highlight the effectiveness of Ginkgo and the benefits incurred by these ECP applications.

Cojean, Terry↗

Portable Parallel Algorithms and Frameworks for Exascale Graph Analytics

Graphs (or networks) are a tool used to model the interactions among various entities. Efficiently processing large graphs has recently attracted significant attention due to the applications of graphs in various domains, such as biology, chemistry, and cyber-security. Analyzing the structure and properties of these graphs is an important component of many scientific computing pipelines. With the explosion in the volume of data, graphs have become very large and can contain hundreds of billions of vertices and trillions of edges. Therefore, it is crucial to develop high-performance methods to enable graph analysis to be done quickly and energy-efficiently. Furthermore, these solutions should be highly parallel in order to take advantage of modern parallel machines. However, designing efficient solutions is not enough. With the wide variety of computing environments available, each with different programmability and performance characteristics, it is necessary to develop solutions that are portable in terms of both performance (i.e., provide theoretical guarantees) and programmability (i.e., provide high level abstractions).

97 MATHEMATICS AND COMPUTING↗

Photoionization of seeded combustion products as a method of enhancing the efficiency of magnetohydrodynamic power generators

Here, in this study, we performed an experimental and computational investigation into the feasibility of utilizing photoionization to enhance the electrical conductivity of seeded oxy-fuel combustion products and improve the performance of magnetohydrodynamic (MHD) power generators. We applied a variety of optical and microwave diagnostics to study the ionization and recombination processes of potassium excited by an excimer laser in a high-velocity oxy-fuel free jet. Computational fluid dynamic (CFD) simulations were performed to model the thermophysical properties and species densities of the free jet. The CFD results were validated with position-dependent potassium concentration measurements. Electron recombination exponential lifetimes were measured through time-resolved microwave transmission. The experimental electron lifetimes were compared with lifetimes calculated from CFD-predicted species densities and literature recombination rates. It was determined that K + or O 2 are the most likely recombination partners for photoionized electrons. Time-resolved fluorescence measurements provided evidence of an ionization pathway involving a two-photon ionization of KOH . Finally, a zero-dimensional chemical kinetic model was developed to assess the fundamental viability of inducing a non-equilibrium electron population to provide a net energy return in combustion-driven MHD power generators. We determined that a high energy return is feasible for targeting electrode boundary layers with ultraviolet photoionization. We also found that photoionization could potentially lower the required temperature of the bulk gas flow.

20 FOSSIL-FUELED POWER PLANTS↗

HPDR: High-Performance Portable Scientific Data Reduction Framework

The rapid growth in scientific data generation is outpacing advancements in computing systems necessary for efficient storage, transfer, and analysis, particularly in the context of exascale computing. With the deployment of first-generation exascale computing systems and next-generation experimental facilities, this gap is widening and necessitates effective data reduction techniques to manage enormous data volumes. Over the past decade, various data reduction methods, including lossless compression, error-controlled lossy compression, and data refactoring, have been developed to accelerate I/O in scientific workflows. Despite significant reductions in data volume, these methods introduce considerable computational overhead, which can become the new bottleneck in data processing. To mitigate this, GPU-accelerated data reduction algorithms have been introduced. However, challenges remain in their integration into exascale workflows, including limited portability across different GPU architectures, substantial memory transfer overhead, and reduced scalability on dense multi-GPU systems. To address these challenges, we propose HPDR, a high-performance and portable data reduction framework. HPDR is designed to enable the execution of state-of-the-art reduction algorithms across diverse processor architectures while reducing memory transfer overhead to 2.3 % of the original, resulting in up to 3.5× faster throughput compared to existing solutions. It also achieves up to 96% of the theoretical speedup in multi-GPU settings. In addition, evaluations on accelerating I/O operations at scale up to 1,024 nodes of the Frontier supercomputer demonstrate that HPDR can achieve up to 103 TB/s reduction throughput, providing up to 4× acceleration in parallel I/O performance compared to existing data reduction routines. This work highlights the potential of HPDR to significantly enhance data reduction efficiency in exascale computing environments.

Chen, Jieyang [University of Oregon]↗

Machine Learning Thermodynamics And Kinetics of Defects For Accelerated Materials Discovery

Atomistic defects play a pivotal role in functional and structural materials’ performance across a myriad of technology applications. Quantitative prediction of the thermodynamics and kinetics of defect formation and migration, respectively, typically requires accurate but expensive first-principles approaches, such as density functional theory (DFT). Their computational expense limits the throughput needed to perform high-throughput materials discovery/screening exercises or to perform materials modeling tasks relying on extensive sampling techniques. Therefore, in this Sandia National Laboratories Laboratory Directed Research and Development (LDRD) project (Project #229366), we developed a variety of machine learning techniques, trained on density functional theory calculations, to accelerate the discovery and modeling of materials in which vacancy and interstitial defects primarily dictate material performance. These include applications such as metal oxides for water-splitting or mixed ionic-electronic conduction, metal hydrides for hydrogen storage, and transition metal dichalcogenides for electronics, and the approaches developed herein can further be applied to many other domains that similarly depend on materials’ thermodynamic and kinetic defect properties for their desired functionality.

36 MATERIALS SCIENCE↗

Development of a data-driven neural network model for electron thermal transport in NSTX

A data-driven electron thermal transport neural network (ETT-NN) model, trained on TRANSP interpretative analysis results of National Spherical Torus Experiment (NSTX), was developed to enable faster and more accurate ETT computation for spherical tokamaks (STs). The model incorporates both convolutional NNs and recurrent NNs, allowing it to simultaneously account for the spatial and temporal non-localities and multi-scale features of turbulent transport, which have been considered only in a limited manner in conventional models. The model was validated through interpretative analysis and predictive simulations using Tokamak Reactor Integrated Automated Suite for Simulation and Computation, demonstrating relatively high accuracy. Additionally, parameter scans were performed on test discharges known to exhibit specific turbulent modes, such as microtearing mode, trapped electron mode, kinetic ballooning mode, and electron temperature gradient mode. The scanning results revealed that the ETT-NN model exhibits the same trends as those observed in conventional gyrokinetic simulations or theories, while also capturing the global nature of turbulent transport, indicating that the data-driven model accurately reflects the underlying physical characteristics. Furthermore, due to the dimensionless nature of the model, we can feasibly expand its applicability by incorporating data from other devices and uncovering the characteristics of ETT in STs in the future.

NSTX↗

Reinforcement learning pulses for transmon qubit entangling gates

The utility of a quantum computer is highly dependent on the ability to reliably perform accurate quantum logic operations. For finding optimal control solutions, it is of particular interest to explore model-free approaches, since their quality is not constrained by the limited accuracy of theoretical models for the quantum processor—in contrast to many established gate implementation strategies. In this work, we utilize a continuous control reinforcement learning algorithm to design entangling two-qubit gates for superconducting qubits; specifically, our agent constructs cross-resonance and CNOT gates without any prior information about the physical system. Using a simulated environment of fixed-frequency fixed-coupling transmon qubits, we demonstrate the capability to generate novel pulse sequences that outperform the standard cross-resonance gates in both fidelity and gate duration, while maintaining a comparable susceptibility to stochastic unitary noise. We further showcase an augmentation in training and input information that allows our agent to adapt its pulse design abilities to drifting hardware characteristics, importantly, with little to no additional optimization. Our results exhibit clearly the advantages of unbiased adaptive-feedback learning-based optimization methods for transmon gate design.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

NEML2: An efficient and modular multiphysics constitutive modeling library for hybrid computing environments

This paper presents NEML2, an open-source, high-performance library developed for constitutive material modeling, designed to support the flexible and modular development of models for complex material behavior. Building on the foundational structure of its predecessor, NEML, the NEML2 library introduces significant improvements, including enhanced vectorization, automatic differentiation, and seamless integration with PyTorch, facilitating the application of machine learning techniques in material simulations. NEML2 provides a C++ backend with Python bindings, enabling users to create custom material models that can be executed efficiently on both CPU and GPU platforms. The library also supports coupling with Multiphysics simulation frameworks like MOOSE, making it suitable for realistic simulations involving coupled physical processes. Rigorous quality assurance through unit and regression testing ensures the reliability of results, while the extensible, user-friendly design encourages collaboration and reproducibility across the scientific community. This paper provides an overview of NEML2’s architecture, core features, and applications, highlighting its impact on accelerating material qualification and advancing computational methods in materials science.

GPU↗