Search NASA⌕ Search

SEARCH · Search NASA

Results for “Parallelization”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 469 records · Page 26

Entanglement Renormalization for Quantum Field Theories with Discrete Wavelet Transforms

We propose an adaptation of Entanglement Renormalization for quantum field theories that, through the use of discrete wavelet transforms, strongly parallels the tensor network architecture of the Multiscale Entanglement Renormalization Ansatz (a.k.a. MERA). Our approach, called wMERA, has several advantages of over previous attempts to adapt MERA to continuum systems. In particular, (i) wMERA is formulated directly in position space, hence preserving the quasi-locality and sparsity of entanglers; and (ii) it enables a built-in RG flow in the implementation of real-time evolution and in computations of correlation functions, which is key for efficient numerical implementations. As examples, we describe in detail two concrete implementations of our wMERA algorithm for free scalar and fermionic theories in (1+1) spacetime dimensions. Possible avenues for constructing wMERAs for interacting field theories are also discussed.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Difference equations and integral families for Witten diagrams

We show that tree-level and one-loop Mellin space correlators in anti-de Sitter space obey certain difference equations, which are the direct analog to the differential equations for Feynman loop integrals in the flat space. Finite-difference relations, which we refer to as “summation-by-parts relations”, in parallel with the integration-by-parts relations for Feynman loop integrals, are derived to reduce the integrals to a basis. We illustrate the general methodology by explicitly deriving the difference equations and summation-by-parts relations for various tree-level and one-loop Witten diagrams up to the four-point bubble level.

AdS-CFT Correspondence↗

SPECTER: efficient evaluation of the spectral EMD

The Energy Mover’s Distance (EMD) has seen use in collider physics as a metric between events and as a geometric method of defining infrared and collinear safe observables. Recently, the Spectral Energy Mover’s Distance (SEMD) has been proposed as a more analytically tractable alternative to the EMD. In this work, we obtain a closed-form expression for the Riemannian-like p = 2 SEMD metric between events, eliminating the need to numerically solve an optimal transport problem. Additionally, we show how the SEMD can be used to define event and jet shape observables by minimizing the distance between events and parameterized energy flows (similar to the EMD), and we obtain closed-form expressions for several of these observables. We also present the Specter framework, an efficient and highly parallelized implementation of the SEMD metric and SEMD-derived shape observables as an analogue of the previously-introduced Shaper for EMD-based computations. We demonstrate that computing the SEMD with Specter can be up to a thousand times faster than computing the EMD with standard optimal transport libraries.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Wire arc additive manufactured A36 steel performance for marine renewable energy systems

Additive manufacturing has established itself to be advantageous beyond small-scale prototyping, now supporting full-scale production of components for a variety of applications. Despite its integration across industries, marine renewable energy technology is one largely untapped application with potential to bolster clean energy production on the global scale. Wave energy converters (WEC) are one specific facet within this realm that could benefit from AM. As such, wire arc additive manufacturing (WAAM) has been identified as a practical method to produce larger scale marine energy components by leveraging cost-effective and readily available A36 steel feedstock material. The flexibility associated with WAAM can benefit production of WEC by producing more complex structural geometries that are challenging to produce traditionally. Additionally, for large components where fine details are less critical, the high deposition rate of WAAM in comparison to traditional wrought techniques could reduce build times by an order of magnitude. In this context of building and supporting WEC, which experience harsh marine environments, an understanding of performance under large loads and corrosive environments must be understood. Hence, WAAM and wrought A36 steel tensile samples were manufactured, and mechanical properties compared under both dry and corroded conditions. Here, the unique microstructure created via the WAAM process was found to directly correlate to the increased ultimate tensile and yield strength compared to the wrought condition. Static corrosion testing in a simulated saltwater environment in parallel with electrochemical testing highlighted an outperformance of corroded WAAM A36 steel than wrought, despite having a slighter higher corrosion rate. Ultimately, this study shows how marine energy systems may benefit from additive manufacturing components and provides a foundation for future applications of WAAM A36 steel.

36 MATERIALS SCIENCE↗

Enriched immersed finite element and isogeometric analysis: algorithms and data structures

Immersed finite element methods provide a convenient analysis framework for problems involving geometrically complex domains, such as those found in topology optimization and microstructures for engineered materials. However, their implementation remains a major challenge due to, among other things, the need to apply nontrivial stabilization schemes and generate custom quadrature rules. This article introduces the robust and computationally efficient algorithms and data structures comprising an immersed finite element preprocessing framework. The input to the preprocessor consists of a background mesh and one or more geometries defined on its domain. The output is structured into groups of elements with custom quadrature rules formatted such that common finite element assembly routines may be used without or with only minimal modifications. The key to the preprocessing framework is the construction of material topology information, concurrently with the generation of a quadrature rule, which is then used to perform enrichment and generate stabilization rules. While the algorithmic framework applies to a wide range of immersed finite element methods using different types of meshes, integration, and stabilization schemes, the preprocessor is presented within the context of the extended isogeometric analysis. This method utilizes a structured B-spline mesh, a generalized Heaviside enrichment strategy considering the material layout within individual basis functions’ supports, and face-oriented ghost stabilization. Using a set of examples, the effectiveness of the enrichment and stabilization strategies is demonstrated alongside the preprocessor’s robustness in geometric edge cases. Additionally, the performance and parallel scalability of the implementation are evaluated.

Computer implementation↗

Coupled Thermo-Hydrological-Mechanical-Chemical Behavior of Anisotropic Granite for Geologic Disposal of High-Level Radioactive Waste: A Core-Scale Laboratory Investigation

The coupled thermo-hydrological-mechanical-chemical (THMC) behavior of rock within an Excavation Damaged Zone (EDZ) is critical for the safety and long-term performance of a geological repository for high-level radioactive wastes. While many laboratory experiments have been conducted to investigate EDZ rocks, the flow and deformation characteristics resulting from anisotropic rock textures and microcrack distribution under triaxial loading and elevated temperatures remain poorly understood. Particularly, cracks at various scales serve as fast paths for fluid flow and solute transport and present as focal points of mechanical weakness, which complicate the coupled THMC processes in anisotropic EDZ rocks and challenge modeling predictions. Here, in this study, a series of core-scale experiments was conducted on three granite samples under repository-relevant conditions. These rock samples were obtained from the Grimsel Underground Research Laboratory (URL), featured by anisotropic minerals (represented by bedding layers) and microcrack distributions and coarse cm-scale grain sizes. During the experiments, samples were subjected to an elevated temperature at 90 °C and different triaxial loading conditions either by radial (normal to bedding layers) or axial (parallel to bedding layers) compaction. Water was injected into the samples, and the rock permeability evolutions and effluent water chemistry were monitored closely. For intact samples, thermal expansion of minerals at 90 °C resulted in a large, 75% irreversible permeability reduction and rock strengthening under radial compaction, while thermal impact was limited to a 15% permeability reduction under axial compaction. In contrast, for a sample containing open cracks, the growth of fractures during the experiment resulted in an abrupt permeability increase and fast failure at 90 °C. The effluent water chemistry indicates much more considerable mineral dissolution from large shear sliding than that in rocks dominated by mechanical compaction. These results helped better understand the coupled THMC processes in anisotropic rocks containing cracks, evaluate the behaviors of EDZ rocks, and predict the long-term evolution of EDZ for the performance of the repositories.

Coupled THMC processes↗

A worldwide climatology of extreme air masses

Extreme temperature events are among the most damaging weather phenomena. In a warming world, more heat extremes and fewer cold extremes are expected in most regions in the future, a trade-off that warrants further understanding of such events. Here, we track and analyze large, persistent areas of hot and cold extreme temperatures in parallel, relative to the location and time of year, to quantify overall regional exposure to extreme temperatures. To accomplish this, we compare the frequencies, movements, trends, and sources/sinks in each type of extreme air mass, calling them extreme cold or extreme hot air masses (ECAMs/EHAMs). For most land regions, ECAMs occur more often than EHAMs, and ECAMs are more common in each hemisphere’s winter, when cold-air advection is strongest and most widespread. Average movement of ECAMs has a stronger equatorward component in winter than in summer, while movement of EHAMs is eastward all year, with less meridional movement than ECAMs. EHAMs have become more common almost everywhere on land, and the reverse is true for ECAMs, with the strongest trends in the northern hemisphere occurring in autumn, especially in the Arctic. The number of EHAMs around the world is increasing at a higher rate (+ 2.32 events/year) than ECAM numbers are decreasing (-1.27 events/year), causing a net increase in extremes, especially in some midlatitude regions. These results help to document the different processes driving hot and cold extremes, and thus the asymmetric and regionally varying trends in frequency for extreme air masses.

Ryan, James M↗

A rheological model for loose sands with insights from DEM

A rheological model for loose granular media is developed to capture both solid-like and fluid-like responses during shearing. The proposed model is built by following the mathematical structure of an extended Kelvin–Voigt model, where an elastic spring and plastic slider act in parallel to a viscous damper. This arrangement requires the partition of the total stress into rate-independent and rate-dependent stress components. To model the solid-like behavior, a simple frictional plasticity model is adopted without modifications, thus contributing to the rate-independent stress. Instead, the fluid-like or rate-dependent stress is further decomposed into deviatoric and volumetric parts, by proposing a new formulation based on a combination of the μ(I) relation, originally developed under pressure-controlled shear, with a pressure-shear rate relation derived under volume-controlled shear. The proposed formulation allows the model to capture both the increase in the friction coefficient and the enhanced dilation at high shear rates. High-fidelity simulation data, obtained from discrete element method and multiscale modelling, are used to evaluate the performance of the proposed constitutive model. The model provides accurate results under both drained and undrained simple shear paths across a wide range of shear rates. Furthermore, it successfully reproduces at much lower computational cost the flowslide mobility computed through multiscale simulations, which is primarily regulated by the shear rate dependence of the material properties during the dynamic runout stage.

Elasticity↗

Stage-local partitioned two-step runge-kutta methods for large systems of ordinary differential equations

We introduce stage-local partitioned two-step Runge-Kutta methods are an extension of standard two-step Runge-Kutta methods, which are an alternative to the standard additive two-step Runge-Kutta methods currently existing in the literature. Furthermore, these new schemes are designed with an eye towards truly N-partitioned systems and leverage local stage approximations to make several computationally interesting approximations viable. Specifically, the focus on local stage approximations makes possible the construction of truly asynchronous schemes, in the parallel sense, possible. In addition, we show that an implicit-explicit approach to these schemes can lead to methods that require the inversion of only local nonlinear systems.

Applied Dynamical Systems↗

Impact of Forest Canopy Structure on Buoyant Plume Dynamics During Wildland Fires

Heterogeneous forest canopies can generate complex turbulent structures, but in the presence of a fire plume, these interactions are not fully understood. This study investigates the influence of forest canopy heterogeneity on buoyant plume dynamics resulting from surface thermal anomalies representing wildland fires, utilizing Large Eddy Simulation (LES). The Parallelized Large-Eddy Simulation Model (PALM) was employed to simulate six canopy configurations: no canopy, homogeneous canopy, external plume-edge canopy, internal plume-edge canopy, 100 m gap canopy, and 200 m gap canopy. Each configuration was analyzed with and without a static surface heat flux patch of 5000 W ∙ m -2 , resulting in a resting buoyant plume. Simulations were conducted under three crosswind speeds: 0, 5, and 10 m ∙ s -1 . Results show that canopy structure significantly modifies plume behavior, mean flow, and turbulent kinetic energy (TKE) budgets. Plume updraft speed and tilt varied with canopy configuration and crosswind speed. Horizontal pressure gradients associated with plume-atmosphere interaction were modified based on the canopy configuration, resulting in varying crosswind speed reductions at the plume region. Strong momentum absorption was observed above the canopy for the crosswind cases, with the greatest enhancement in the gap canopies. Momentum injection from below the canopy due to the heat source was also observed, resulting in plume structure modulation based on canopy configuration. TKE was found to be the largest in the gap canopy configurations. TKE budget analysis revealed that buoyant production dominated over shear production. At the center of the heat patch, the gap canopy configurations showed enhanced buoyancy within the gap. These results improve our knowledge of fire-canopy-atmosphere interactions that can inform fire models on the impacts of canopy heterogeneity on plume dynamics and ember ejections.

54 ENVIRONMENTAL SCIENCES↗

Properties of Low Tc AlMn TES

Low Tc AlMn transition-edge sensors (TESs) have been developed as sensitive thermometers for the Q-Array, which will use superconducting targets to measure the coherent elastic neutrino nucleus scattering spectrum in the RICOCHET experiment. The TESs are made of manganese-doped aluminum with a titanium and gold antioxidation layer. A prototype TES thermometer consists of two TESs in parallel, an input gold pad in metallic contact with the TESs and an output gold pad and gold thermal link meanders, which are each designed to control the flow of heat through the TESs. We have fabricated and measured low TC AlMn TES chips with or without thermal flow control structures. We present TC measurements of the TESs after the initial fabrication and further TC tuning by re-heating and summarize the thermal property studies of the prototype TES thermometer by measuring I-V curves and complex impedance.

Coherent elastic neutrino nucleus scattering, Cryo↗

Hierarchical Phased-Array Antennas Coupled to Al KIDs: A Scalable Architecture for Multi-band Millimeter/Submillimeter Focal Planes

We present the optical characterization of two-scale hierarchical phased-array antenna kinetic inductance detectors (KIDs) for millimeter/submillimeter wavelengths. Our KIDs have a lumped-element architecture with parallel plate capacitors and aluminum inductors. The incoming light is received with a hierarchical phased array of slot dipole antennas, split into 4 frequency bands (between 125 GHz and 365 GHz) with on-chip lumped-element band-pass filters, and routed to different KIDs using microstriplines. Individual pixels detect light for the 3 higher-frequency bands (190–365 GHz), and the signals from four individual pixels are coherently summed to create a larger pixel detecting light for the lowest frequency band (125–175 GHz). The spectral response of the band-pass filters was measured using Fourier transform spectroscopy (FTS), the far-field beam pattern of the phased-array antennas was obtained using an infrared source mounted on a 2-axis translating stage, and the optical efficiency of the KIDs was characterized by observing loads at 294 K and 77 K. We report on the results of these three measurements.

47 OTHER INSTRUMENTATION↗

Randomized Preconditioned Solvers for Strong Constraint 4D-Var Data Assimilation

The Strong Constraint 4D Variational (SC-4DVAR) data assimilation method is widely used in climate and weather applications. SC-4DVAR involves solving a minimization problem to compute the maximum a posteriori estimate, which we tackle using the Gauss-Newton method. The computation of the descent direction is expensive since it involves the solution of a large-scale and potentially ill-conditioned linear system, solved using the preconditioned conjugate gradient (PCG) method. Here, to address this cost, we efficiently construct scalable preconditioners using three different randomization techniques, which all rely on a certain low-rank structure involving the Gauss-Newton Hessian. The proposed techniques come with theoretical guarantees on the condition number, and at the same time, are amenable to parallelization. We also develop an adaptive approach to estimate the sketch size and choose between the reuse or recomputation of the preconditioner. We demonstrate the performance and effectiveness of our methodology on two representative model problems—the Burgers and barotropic vorticity equation—showing a drastic reduction in both the number of PCG iterations and the number of Gauss-Newton Hessian products after including the preconditioner construction cost.

Gauss-Newton↗

Exploratory Investigation of Coal in Nonequilibrium Plasma

Coal is an abundant natural resource and there is motivation to find new uses for it that do not intrinsically involve combustion. One approach is to explore new ways of processing coal, and in this work, we focus on the transformation of coal in a nonequilibrium plasma generated from an equimolar mixture of nitrogen and hydrogen. The outcome of the nonequilibrium plasma reaction is fundamentally different than a thermal control reaction carried out using the same gas composition, pressure, and temperature range. The nonequilibrium plasma produces a gas mixture that is enriched in acetylene and its derivatives. Furthermore, when compared to the thermal control experiment, the solid char byproduct of the nonequilibrium plasma has a very reactive surface and is spontaneously combustible at ambient temperature. Experiments performed to characterize the reaction kinetics of coal in the plasma suggest that the mechanism proceeds through a sequential process by which the coal particle temperature rises to a point where devolatilization can occur, the devolatilization reaction happens, followed by parallel reactions of released organic vapors in the plasma phase and surface activation. In conclusion, the reaction rate appears to be limited by the time it takes for the coal particle temperature to rise, consistent with previous results reported for reactions of coal in thermal plasma.

Acetylene↗

Optimizing inference of segmentation on high-resolution images in MLExchange

MLExchange is a machine learning (ML) operations platform providing web user-interfaces (UIs) for data visualization and analysis pipelines at synchrotron facilities. Among these UIs is the segmentation app which helps synchrotron users utilize ML algorithms to automatically segment high-resolution scientific images with minimal manual annotation effort. In this work, we share code optimizations that significantly speed up the segmentation inference workflow of large data in short time. By optimizing the sequence of CPU-GPU data transfers and introducing CPU parallelization to key operations, we improve the per-device, per-image frame computational efficiency and observe close to 3×$$\times$$ speedup over the original segmentation inference workflow run time when utilizing a single GPU. Further adaptations enabling multi-GPU inference yield more than 40×$$\times$$ speedup with 100 GPUs compared to the optimized single GPU inference workflow. This acceleration of the segmentation inference workflow will provide MLExchange users with easy access to segmentation results with little wait time.

Lu, Shizhao↗

Variation of the Passive Film on Compositionally Concentrated Dual-Phase Al 0.3 Cr 0.5 Fe 2 Mn 0.25 Mo 0.15 Ni 1.5 Ti 0.3 and Implications for Corrosion

The passive film on a dual-phase Al 0.3 Cr 0.5 Fe 2 Mn 0.25 Mo 0.15 Ni 1.5 Ti 0.3 FCC + Heusler (L2 1 ) compositionally concentrated alloy formed during extended exposure to an applied potential in the passive range in dilute chloride solution was characterized. Each phase, with its own distinct composition of passivating elements, formed unique passive films separated by a heterophase interface. High-resolution, surface sensitive characterization enabled chemical analysis of the passive film formed over individual phases. The film formed over the L2 1 phase had a higher concentration of Al, Ni, and Ti, while the film formed over FCC phase was of similar thickness but contained comparatively higher Cr, Fe, and Mo concentrations, consistent with the differences in bulk microstructure composition. The passive film was continuous across phase boundaries and the distribution of passivating elements (Al, Cr, and Ti) indicated both phases were independently passivated. Spatially resolved analysis of the surface chemistry of the dual-phase CCA revealed that the cation with the highest composition in passive film formed on the FCC phase was Cr (52.4 at. pct) and for the L2 1 phase was Ti (53.1 at. pct) despite the bulk concentration of each element being below 20 at. pct in their respective phases. Al, Cr, and Ti were enriched in both phases within the passive film relative to their respective bulk compositions. In parallel studies, single-phase alloys with compositions representative of the FCC and L2 1 phases were synthesized to evaluate the corrosion behavior of each phase in isolation. The corrosion behavior of the dual-phase alloy showed passivity evidenced by a pitting potential of 0.615 V SCE in 0.01 M NaCl. The pitting potential and other electrochemical parameters suggested a combination of behaviors of both single-phase samples, suggesting that the global corrosion behavior may be represented by a composite theory applied to phases, their area fractions, and interphase length. However, the interphase in the dual-phase CCA was a local corrosion initiation site and may limit localized corrosion protectiveness. The alloy design implications for optimization of second phase structure and morphology are discussed.

36 MATERIALS SCIENCE↗

The Influence of Residual Stress on Fatigue Crack Growth Rates in Stainless Steel Processed by Different Additive Manufacturing Methods

The properties and microstructure of Type 304L stainless steel produced by two additive manufacturing (AM) methods—directed energy deposition (DED) and powder bed fusion (PBF)—are evaluated and compared. Localized heating and steep temperature gradients of AM processes lead to significant residual stress and distinctive microstructures, which may be process-specific and influence mechanical behavior. Test data show that materials produced by DED and PDF have small differences in tensile strengths but clear differences in residual stress and microstructural features. Measured fatigue crack growth rates (FCGRs) for cracks propagating parallel to and perpendicular to the build directions differ between the two AM materials. To separate the influences of residual stress and microstructure, K-control test procedures with decreasing and constant stress intensity factor ranges are used to measure FCGRs in the near-threshold regime (crack growth rates ≤ 1 × 10 −8 m/cycle). Residual stress is quantified by the residual stress intensity factor, K res , measured by the online crack compliance method. Correcting the FCGR data for differences in K res brings results for specimens of the two AM materials into agreement with each other and with results for wrought specimens, when the latter are corrected for crack closure. Differences in microstructure and tensile strength have an insignificant influence on FCGRs in these tests.

additive manufacturing↗

Expert and operator perspectives on barriers to energy efficiency in data centers

Abstract It was last estimated in 2016 that data centers (DCs) comprise approximately 2% of total US electricity consumption. However, this estimate is currently being updated to account for the massive increase in computing needs due to streaming, cryptocurrency, and artificial intelligence (AI). To prevent energy consumption that tracks with increasing computing needs, it is imperative we identify energy efficiency strategies and investments beyond the low-hanging fruit solutions. In a two-phased research approach, we ask: What non-technical barriers still impede energy efficiency (EE) practices and investments in the data center sector, and what can be done to overcome these barriers? In particular, we are focused on social and organizational barriers to EE. In Phase I, we performed a literature review and found that technical solutions are abundant in the literature, but fail to address the top-down cultural shifts that need to take place in order to adapt new energy efficiency strategies. In Phase II, reported here, we interviewed 16 data center operators/experts to ground-truth our literature findings. Our interview protocols focus on three aspects of DC decision-making: procurement practices, metrics and monitoring, and perceived barriers to energy efficiency. We find that vendors are the key drivers of procurement decisions, advanced efficiency metrics are facility-specific, and there is convergence in the design of advanced facilities due to the heat density of parallelized infrastructure. Our ultimate goals for our research are to design DC decarbonization policies that target organizational structure, empower individual staff, and foster a supportive external market.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗