Search NASASearch

SEARCH · Search NASA

Results for “Computer systems performance”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

The role of quantum computing in advancing scientific high-performance computing: A perspective from the ADAC institute

Quantum computing (QC) has gained significant attention over the past two decades due to its potential for speeding up classically demanding tasks. This transition from an academic focus to a thriving commercial sector is reflected in substantial global investments. While advancements in qubit counts and functionalities continue at a rapid pace, current quantum systems still lack the scalability for practical applications, facing challenges such as too high error rates and limited coherence times. Here, this perspective paper examines the relationship between QC and high-performance computing (HPC), highlighting their complementary roles in enhancing computational efficiency. It is widely acknowledged that even fully error-corrected QC will not be suited for all computational tasks. Rather, future compute infrastructures are anticipated to employ quantum acceleration within hybrid systems that integrate HPC and QC. While QC can enhance classical computing, traditional HPC remains essential for maximizing quantum acceleration. This integration is a priority for supercomputing centers and companies, sparking innovation to address the challenges of merging these technologies. The novelty of this work lies in its unique perspective, reflecting the collective insights of the Accelerated Data Analytics and Computing (ADAC) Institute, a global consortium of over 20 leading HPC centers. Recognizing the growing importance of QC, ADAC established a Quantum Computing Working Group in 2023 to foster collaboration and knowledge-sharing among its members. This paper synthesizes insights from the group’s collaborative efforts and incorporates findings from a member survey that captures shared experiences, ongoing projects, and strategic directions. By outlining the current landscape and challenges of QC integration into HPC ecosystems, this work offers HPC specialists practical and forward-looking guidance on the opportunities and implications of QC in computationally intensive endeavors.

Accelerated Data Analytics and

HP-MDR: High-performance and Portable Data Refactoring and Progressive Retrieval with Advanced GPUs

Scientific applications produce vast amounts of data, posing grand challenges in the underlying data management and analytic tasks. Progressive compression is a promising way to address this problem, as it allows for on-demand data retrieval with significantly reduced data movement cost. However, most existing progressive methods are designed for CPUs, leaving a gap for them to unleash the power of today’s heterogeneous computing systems with GPUs.In this work, we propose HP-MDR, a high-performance and portable data refactoring and progressive retrieval framework for GPUs. Our contributions are four-fold: (1) We carefully optimize the bitplane encoding and lossless encoding, two key stages in progressive methods, to achieve high performance on GPUs; (2) We propose pipeline optimization and incorporate it with data refactoring and progressive retrieval workflows to further enhance the performance for large data process; (3) We leverage our framework to enable high-performance data retrieval with guaranteed error control for common Quantities of Interest; (4) We evaluate HP-MDR and compare it with state of the arts using five real-world datasets. Experimental results demonstrate that HP-MDR delivers an average 13.68 × and 6.31 × throughput in data refactoring and progressive retrieval tasks, respectively. It also leads to 11.22 × throughput for recomposing required data representations under Quantity-of-Interest error control and 6.04 × performance for the corresponding end-to-end data retrieval, when compared with state-of-the-art solutions.

Li, Yanliang [University of Oregon]

Comparison of a Full-Scale and a 1:10 Scale Low-Speed Two-Stroke Marine Engine Using Computational Fluid Dynamics

International marine shipping is a growing component of international trade; a vast majority of all the world’s goods are being transported on large ocean-going vessels. The International Maritime Organization (IMO) introduced the Energy Efficiency Design Index in 2013, a regulatory framework of associated metrics for reducing emissions of CO 2 per tonne-mile from shipping by approximately 10% each decade. Therefore, decarbonizing the maritime sector requires the development of new fuel sources. Because of the extremely large physical size of the internal combustion engines present in shipping vessels, experimental iterative development of the engine and fuel system is cost-prohibitive. Thus, the ability to perform combustion system development in a scaled platform that can be more easily operated and modeled computationally is of interest. To that end, scaling relationships are needed to translate the results from a smaller engine to a larger counterpart. Scaling studies to date have been restricted to low scaling ratios, four-stroke light-duty engines, and under-resolved computational fluid dynamic simulations that likely do not accurately capture the physics of scaling. In this work, computational models of a 1:10 scale and a full-scale two-stroke crosshead low-speed marine engine were created and validated against experiments obtained in a real 1:10 scale engine installed at Oak Ridge National Laboratory. Further, due to the large size of the full-scale engine, the model required large high-performance computing resources to be evaluated. The availability of high-performance computing resources at the Department of Energy’s Leadership Computing Facilities is an enabler of the current work. The results of the small- and large-scale engine simulations were compared to analyze the effectiveness of the appropriate scaling laws under these extreme scaling ratio conditions.

33 ADVANCED PROPULSION SYSTEMS

Towards Sustainable Post-Exascale Leadership Computing

As computing systems approach the limits of traditional silicon technology, the diminishing returns in performance per watt present a significant barrier to sustaining growth in HPC. From a large-scale scientific supercomputing facility point of view, we propose a multifaceted strategy toward specialized hardware and architectures that are optimized for energy efficiency in specific applications. We also emphasize the need for integrating energy-aware practices across all levels of HPC, from system design and software development to operational policies. We discuss strategic opportunities such as the adoption of application-specific accelerators, the development of energy-efficient algorithms, and the implementation of data-driven operational analytics. Our goal is to develop a comprehensive roadmap ensuring that future leadership systems at OLCF can meet scientific demands while operating within stringent energy budgets, thereby supporting sustainable computing growth.

Shin, Woong

An Evaluation and Qualification of U.S.-Based Research Reactors for Irradiation Capabilities Supporting Advanced Nuclear Systems

Irradiation experiments are a prerequisite for evaluating nuclear reactor system designs, analyzing the performance of these systems, and obtaining licenses. Likewise, irradiation facilities are necessary for producing the radioisotopes used in industrial and medical applications. Recent developments in modeling and simulation capabilities and advancements in computational resources have further enabled the design of irradiation experiments for evaluating radiation-induced phenomena and determining nuclear fuel, material, and system design and safety criteria pertaining to both normal and accident scenarios. These computational tools and models require comprehensive experimental datasets acquired under prototypic radiation conditions—for exploring material and system performance under the uniquely harsh environments found in nuclear reactors—to enable verification and validation for qualification and licensing purposes. However, qualification of irradiation experimental facilities, primarily research and test reactors (RTRs), necessitates that their performance be evaluated based on the irradiation environment (e.g. flux, power, testing capabilities) using an appropriate scoring matrix. Although many university campus RTRs are available for research and development (R&D) activities and initiatives, this study focuses on evaluating and qualifying the irradiation facilities (mostly RTRs) within the United States that are suitable for advanced nuclear fuel, material, and system irradiation experiments aimed at establishing operational-performance limits and informing component and fuel designs so as to improve operational efficiencies and mitigate proliferation vulnerabilities, as well as for radioisotope production aimed at multipurpose applications. As a result, the findings of the present study support the acceleration of nuclear fuel and material qualifications, thus hastening new and advanced nuclear energy system demonstrations and radioisotope production efforts by using extended R&D.

irradiation experiment

HPC I/O innovations in the exascale era

As high performance computing architecture evolves to deliver ever-increasing performance, the middleware tools also need to adapt in order for applications to better use these higher-performance features. Here, the Adaptable Input Output System (ADIOS), which provides scalable IO performance for exascale HPC applications is one such middleware. During the Exascale Computing Project (ECP), key portions of the ADIOS environment were adapted to respond to ongoing developments in exascale computing and the stresses and opportunities inherent in those changes. This paper examines those changes and where appropriate compares them to pre-exascale implementations.

ADIOS

Massively parallel and universal approximation of nonlinear functions using diffractive processors

Nonlinear computation is essential for a wide range of information processing tasks, yet implementing nonlinear functions using optical systems remains a challenge due to the weak and power-intensive nature of optical nonlinearities. Overcoming this limitation without relying on nonlinear optical materials could unlock unprecedented opportunities for ultrafast and parallel optical computing systems. Here, we demonstrate that large-scale nonlinear computation can be performed using linear optics through optimized diffractive processors composed of passive phase-only surfaces. In this framework, the input variables of nonlinear functions are encoded into the phase of an optical wavefront—e.g., via a spatial light modulator (SLM)—and transformed by an optimized diffractive structure with spatially varying point-spread functions to yield output intensities that approximate a large set of unique nonlinear functions–all in parallel. We provide proof establishing that this architecture serves as a universal function approximator for an arbitrary set of bandlimited nonlinear functions, also covering wavelength-multiplexed nonlinear functions as well as multi-variate and complex-valued functions that are all-optically cascadable. Our analysis also indicates the successful approximation of typical nonlinear activation functions commonly used in neural networks, including the sigmoid, tanh, ReLU (rectified linear unit), and softplus. We numerically demonstrate the parallel computation of one million distinct nonlinear functions, accurately executed at wavelength-scale spatial density at the output of a diffractive optical processor. Furthermore, we experimentally validated this framework using in situ optical learning and approximated 35 unique nonlinear functions in a single shot using a compact setup consisting of an SLM and an image sensor. These results establish diffractive optical processors as a scalable platform for massively parallel universal nonlinear function approximation, paving the way for new capabilities in analog optical computing based on linear materials.

Rahman, Md Sadman Sakib [University of California,

The Radical Atom: Mechanosynthetic 3D Printing of an Atomically Precise SPM Tip

This research effort sought to overcome current limitations in scanning probe-based atomic manipulation to enable atomically precise manufacturing (APM). Previous theoretical and experimental works on atom by atom and molecule by molecule fabrication of precise structures are limited to essentially to two-dimensions. APM will enable a paradigm shift in 21st century manufacturing practices in which every single atom in a electronic chip, device or machine can be placed in an exact and predefined position in three-dimensions. By providing a general method for generating reproducible SPM tip structure, this project will drive forward the entire field of atomically precise scanning probe microscopy, opening the door to positional control of nearly arbitrary covalent chemistry. Such control could, for example, be used in applications such as novel 2.5 or 3D microchip fabrication. The creation of a unique manufacturing method through APM has the potential to impact technologies at the theoretical limits of performance, weight, and utility including: solid-state quantum and spintronic computing systems, high efficiency optical antenna, solar power systems, defect engineered materials and extremely efficient catalysts. Although this experiment focused on pick-and-place non-scalable APM, the better understanding of the chemistry is crucial to the eventual goal of scalable APM. To place individual atoms into a specified location is a seminal aspiration of researchers and engineers in the many fields and may have early premium applications in medical devices and microelectronics.

77 NANOSCIENCE AND NANOTECHNOLOGY

OES CO 2 Pipeline FEED Project Design Basis Memorandum

The OES CO₂ Pipeline project will move captured carbon dioxide from two ethanol facilities near Gibson City, Illinois, roughly 7.8 miles southeast to three injection wells outside Anchor, where it will be permanently stored underground. The system is designed to handle up to 4.5 million metric tonnes per year of dense-phase CO₂ at pressures up to 2,500 psig, using 16-inch mainline pipe and 10.750-inch laterals made from API 5L X-60 and X-65 steel. Wall thicknesses vary depending on location, with thinner pipe in open country, heavier wall at road crossings, and the heaviest where the pipe passes under highways or railroads via horizontal directional drill. The pipe gets a fusion-bonded epoxy coating, with an added abrasion-resistant layer wherever it's bored or drilled. Major water crossings will use HDD rather than open trenching. The pipeline will be cathodically protected, equipped with SCADA-compatible pressure and temperature instrumentation, and monitored for leaks using a computational pipeline monitoring system per API RP 1130. Hydrostatic testing will be performed at 1.25 times design pressure, and an ILI caliper run will follow to catch any construction defects. Several items, including fracture toughness requirements, specific NDE methods, and ILI tool selection, are left for the detailed design phase. The whole system falls under 49 CFR Part 195 and ASME B31.4, and Gulf Interstate Engineering prepared this document as the FEED-level design basis under the CarbonSAFE Phase III program.

09 BIOMASS FUELS

Role of bath-induced many-body interactions in the dissipative phases of the Su-Schrieffer-Heeger model

The Su-Schrieffer-Heeger chain is a prototype example of a symmetry-protected topological insulator. Coupling it nonperturbatively to local thermal environments, either through the intercell or the intracell fermion tunneling elements, modifies the topological window. To understand this effect, we employ the recently developed reaction-coordinate polaron transform (RCPT) method, which allows treating system-bath interactions at arbitrary strengths. The effective system Hamiltonian, which is obtained via the RCPT, exposes the impact of the baths on the SSH chain through renormalization of tunneling elements and the generation of many-body interaction terms. By performing exact diagonalization and computing the ensemble geometric phase, a topological invariant, which is applicable even to systems at finite temperature, we distinguish the trivial band insulator (BI) from the topological insulator (TI) phases. Furthermore, through the RCPT mapping, we are able to pinpoint the main mechanism behind the extension of the parameter space for the TI or the BI phases (depending on the coupling scheme, intracell or intercell), which is the bath-induced, dimerized, many-body interaction. In conclusion, we also study the effect of on-site staggered potentials on the SSH phase diagram and discuss extensions of our method to higher dimensions.

1-dimensional systems

Parallel derivative-free optimization for simulation-based design of behind-the-meter energy systems

In this work, the integrated design and dispatch of behind-the-meter or distributed resources (e.g. stationary battery storage and solar PV generation) is considered. A simulation-based framework is employed, generating high-fidelity results with closed-loop predictive control at a fine resolution, at the expense of high computational cost (several minutes to a few hours per design point). To address this challenge, parallel derivative-free design methods are considered. Four methods are compared, including state-of-the-art surrogate-based methods (Radial-Basis Functions and Gaussian processes) and sampling strategies, an evolutionary-based method, and a simple sequential grid refinement method. As a case study, two types of design problem with increasing complexity are considered, namely, the design of behind-the-meter resources (three design variables) and the inclusion of grid capacity (four design variables). The second yields a constrained design problem for which violations can only be determined after solving the computationally expensive simulation. For the three-dimensional case, all methods present a good performance, achieving a solution within 1% of the optimum after the first iteration, with the sequential grid refinement exhibiting the fastest convergence and achieving the best final objective value. This indicates that the parallel evaluation of multiple sampling points may be more important than the choice of method for small decision spaces. For the four-dimensional constrained case, the Genetic Algorithm presents the best tradeoff between performance and computational effort, while the rough objective function terrain generated by constraint violation penalties reduces the performance of surrogate-based methods. Contour plots with flat regions indicate flexibility in the optimal design and highlight the importance of characterizing the solution space.

24 POWER TRANSMISSION AND DISTRIBUTION

A machine learning decision criterion for reducing scan time for hyperspectral neutron computed tomography systems

We present the first machine learning-based autonomous hyperspectral neutron computed tomography experiment performed at the Spallation Neutron Source. Hyperspectral neutron computed tomography allows the characterization of samples by enabling the reconstruction of crystallographic information and elemental/isotopic composition of objects relevant to materials science. High quality reconstructions using traditional algorithms such as the filtered back projection require a high signal-to-noise ratio across a wide wavelength range combined with a large number of projections. This results in scan times of several days to acquire hundreds of hyperspectral projections, during which end users have minimal feedback. To address these challenges, a golden ratio scanning protocol combined with model-based image reconstruction algorithms have been proposed. This novel approach enables high quality real-time reconstructions from streaming experimental data, thus providing feedback to users, while requiring fewer yet a fixed number of projections compared to the filtered back projection method. In this paper, we propose a novel machine learning criterion that can terminate a streaming neutron tomography scan once sufficient information is obtained based on the current set of measurements. Our decision criterion uses a quality score which combines a reference-free image quality metric computed using a pre-trained deep neural network with a metric that measures differences between consecutive reconstructions. The results show that our method can reduce the measurement time by approximately a factor of five compared to a baseline method based on filtered back projection for the samples we studied while automatically terminating the scans.

97 MATHEMATICS AND COMPUTING

Iterative methods in GPU-resident linear solvers for nonlinear constrained optimization

Linear solvers are major computational bottlenecks in a wide range of decision support and optimization computations. The challenges become even more pronounced on heterogeneous hardware, where traditional sparse numerical linear algebra methods are often inefficient. For example, methods for solving ill-conditioned linear systems have relied on conditional branching, which degrades performance on hardware accelerators such as graphical processing units (GPUs). To improve the efficiency of solving ill-conditioned systems, our computational strategy separates computations that are efficient on GPUs from those that need to run on traditional central processing units (CPUs). Our strategy maximizes the reuse of expensive CPU computations. Iterative methods, which thus far have not been broadly used for ill-conditioned linear systems, play an important role in our approach. In particular, we extend ideas from Arioli et al., (2007) to implement iterative refinement using inexact LU factors and flexible generalized minimal residual (FGMRES), with the aim of efficient performance on GPUs. In conclusion, we focus on solutions that are effective within broader application contexts, and discuss how early performance tests could be improved to be more predictive of the performance in a realistic environment.

97 MATHEMATICS AND COMPUTING

Liquid Phase Modeling in Porous Media: Adsorption of Methanol and Ethanol in H-MFI in Condensed Water

Zeolites are used in the chemical and separation industries for their exceptional selectivity, adsorption capacity, regenerability, and stability in gas and liquid phase processing. Here, we developed an explicit solvation method for predicting solvent/condensed phase effects on adsorption free energies in microporous media such as zeolites based on the hybrid quantum mechanical/molecular mechanical free energy perturbation (QM/MM-FEP) technique. Our explicit solvation method for zeolite systems, called eSZS, aims to capture site-specific interactions during the adsorption process at the Brønsted acid sites of H-MFI zeolite while still considering the diverse configuration space of the solvent molecules. This strategy is ideal for chemical reactions or adsorbates that interact with the microporous medium in few distinct adsorbate/transition state configurations, i.e., the harmonic or similar approximations are acceptable for the adsorbate/transition state while such approximations break down for the solvent molecules that require extensive configuration space sampling. In this way, our approach effectively overcomes the limitations of implicit solvation models and classical force field methods for describing solvation effects on chemical reactions within porous materials such as zeolites. Specifically, in this study, we investigated various aspects of our hybrid QM/MM approach, including QM cluster size dependencies in a periodic electrostatically embedded cluster model (PEECM), rules for link atoms at the QM/MM boundary, and functional and basis set considerations for converged and reasonably accurate gas and aqueous phase methanol and ethanol adsorption free energy predictions in H-MFI. For gas phase adsorption of methanol and ethanol in H-MFI at a Brønsted acid site in T12 position, we compute adsorption free energies at 298 K of −0.61 and −0.75 eV, respectively, using a PEECM containing 50 Si and 1 Al atom with ωB97x-D/def2-TZVP level of theory. For solvent effect calculations, we sample the aqueous phase using grand canonical Monte Carlo (GCMC) simulations to (1) obtain a mean field of electrostatic interactions in the reaction system and (2) perform a rigorous free energy perturbation calculation. Similar to the experimentally and computationally observed endergonic solvation effects observed for hydrocarbon adsorption on metal surfaces, we also observe that a condensed aqueous environment destabilizes methanol and ethanol at these acid sites in H-MFI at 298 K. Specifically, the computed solvation free energies of adsorption (ΔΔG solv ) for methanol and ethanol are +0.44 and +0.54 eV, respectively. From this study, it is evident that adsorbates (methanol and ethanol) are competing with water for adsorption space inside the H-MFI zeolite, leading to an endergonic solvation effect. Here, we expect that the endergonic, aqueous solvent effect during adsorption in microporous zeolites is highly tunable by changing the pore size and hydrophobicity of the microporous material as this will affect the water density inside the pore structure.

Adsorption

A Typology of Quantum-Classical Faults

This paper introduces an extended taxonomy of faults specific to hybrid quantum-classical systems, addressing the unique challenges that arise from integrating quantum accelerators into high-performance computing (HPC) infrastructures. Building on the foundational fault classification by Avizienis et al., we incorporate fault types unique to quantum computing-such as qubit decoherence, spontaneous gate errors, and photon loss-alongside traditional and human-induced faults including development errors, operational mistakes, and malicious attacks. Our taxonomy classifies faults by their origin (natural vs. human-made), intent (accidental, deliberate non-malicious, or malicious), system boundaries (internal vs. external), and persistence (transient to permanent). We also explore how different architectural integration patterns-ranging from tight coupling to loose on-premise and cloud-based configurations-shape the manifestation and propagation of faults. These scenarios are analyzed in terms of timing mismatches, interface inconsistencies, and security threats such as data tampering and denial-of-service attacks. Through this fault-centric lens, we aim to support the co-design of dependable quantum-classical systems and highlight the critical role that integration strategies play in ensuring reproducibility, resilience, and security across hybrid computing platforms.

Giusto, Edorado [University of Naples Federico II,

Computational design of high entropy alloy coating for hydrogen turbine applications

This project aims to develop novel high entropy alloy (HEA)-based coatings to protect critical components in hydrogen-fueled turbine power system. The HEA-coatings will demonstrate superior performance in hydrogen combustion environment to commercial NiCoCrAlY coating in current natural gas turbine system. The HEA coating facilitates the formation of a protective scale of alpha-alumina to slow down the inward diffusion of oxidizing species and the outward diffusion of metal elements, and possesses ultrahigh corrosion and spallation resistance to prolong the service lifetime of critical components in hydrogen turbine power system. Aimed to accelerate the discovery of novel HEA coating compositions, high throughput computational modeling including CALPHAD and density functional theory and machine learning are performed to predict phase stability, oxygen permeability, oxidation rate constant, coefficient of thermal expansion, and mechanical properties. Based on the modeling and machine learning prediction, experimental validation is performed. Preliminary results will be presented and approaches to minimize oxidation will be discussed.

alloy design

University Data Management Pilot Utilizing the Nuclear Research Data System

Background In 2022, the Office of Science and Technology Policy (OSTP) issued a memo that significantly reshaped the landscape of access to federally funded research. The memo mandated that all taxpayer-funded research be made available to the public without delay upon publication, without an embargo period, superseding the 2013 OSTP public access policy. This public access policy promotes transparency and the democratization of knowledge, ensuring that the fruits of scientific endeavors funded by federal agencies could be immediately accessed and built upon by scientists, educators, students, and the public at large. To implement the requirements of the OSTP guidance and DOE Public Access Plan, the Office of Nuclear Energy (NE) has implemented public access plan guidance and has identified several areas where better data management practices would further expand public access to important nuclear energy related scientific data, reports, and other technical products. Significant NE supported efforts are already underway for data management and public access to important nuclear energy related data.1 2 To address gaps in data management practices, and improve retention and accessibility of data, NE is actively exploring enhanced data management options utilizing its high-performance computing resources administered by its Nuclear Scientific User Facility Program. A newly piloted system, the Nuclear Research Data System (NRDS) acts as a portal for data collection and dissemination. Nuclear Energy University Program Research and Development Portfolio According to Web of Science, NEUP has produced 2,345 journal publication that have been cited more than 61,000 times3 and countless conference proceedings. These publications are publicly available through OSTI.gov and in the open literature. Additional scientific and technical products including project milestones that are not publications and NEUP project final reports are vetted through OSTI.gov and released once reviewed and approved by DOE. Since 2009, NEUP has awarded close to 1,000 different R&D projects in technical areas across the NE research programs. As of June 2023, 512 NEUP reports are publicly available on OSTI. The underlying data for projects is still held at universities, and data transfer, co-location, and dissemination has not occurred in a systematic way. NEUP data is currently accessible through myriad university-based data repositories, or through direct requests to PIs. The program identified this patchwork of repositories, or often lack of publicly available data, as a significant barrier to an organized, accessible, and comprehensive solution to sharing data with the larger nuclear energy community. Approach The goal of this pilot project is to establish a pathway to a consolidated long-term repository for NEUP project data. To accomplish this goal, the pilot strives to accomplish the following objectives: Establish data collection standards, including a standard set of required supplementary information to contextualize and support raw data files. Work with the HPC group collect and upload information and to modify the NRDS system, as needed, to support a standardized approach. Resolve potential barriers to successful roll out of an expanded data collection strategy, including modifying data management plan guidelines and establishing a document and data release process that accounts for potential intellectual property and/or export control concerns. Results Overall, the pilot was successful in collecting 8,982 raw and processes data files, 220 reports, 56 calibration files, and 5,931 other supplementary documents. Supplementary documents included experimental plans, methods, journal publications and conference proceedings, milestone reports, and final reports. Figure 2 shows the number of data sets and supplementary project information provided by each project. Projects has significantly different input, depending on experimental data produced and completeness of the datasets provided.

Data collection

A time-parallel method for scalable heat transfer simulations of additive manufacturing

Here, a major challenge in simulating the thermal behavior in additive manufacturing processes is the disparate length and time scales between transport phenomena occurring in the melt pool and the component. A common simulation approach relies on spatial decomposition for parallel computing, but due to the nature of heat transfer in AM, where most of the computational expenditure is localized near the melt pool, the computational speedup from spatial parallelization saturates quickly. Therefore, additional parallelism by means of time-domain decomposition is needed to fully take advantage of high-performance computing (HPC) resources. This work introduces a time-parallel method to improve the computational scalability of additive manufacturing simulations on HPC systems, while maintaining high temporal resolution of heat transfer near the melt pool. The method, inspired by the nonlinear paraexp formalism, performs an iterative superposition of nonlinear solutions to the initial value problem, integrating the heat equation across overlapping time-parallel intervals. For a single layer of the NIST AMB2018–01 L7 benchmark problem, the method achieves a 38.51x speedup in wall-clock time with a maximum error in the global temperature solution of 0.99%. This reduces the total solution time from 196.72 min to 5.11 min on 128 nodes of the ORNL Frontier supercomputer. The tradeoff between accuracy and total wall-clock time is investigated and recommendations for time-parallel deployment for AM problems are made.

Additive manufacturing