Search NASA⌕ Search

SEARCH · Search NASA

Results for “Massively Parallel Implementation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

209 records · Page 12

A Performance Portable, Fully Implicit Landau Collision Operator with Batched Linear Solvers

Modern accelerators use hierarchical parallel programming models that enable massive multithreading within a processing element (PE), with multiple PEs per device driven by traditional processes. Batching is a technique for exposing PE-level parallelism in algorithms that have traditionally run on MPI processes or multiple threads within a single process. Opportunities for batching arise in, for example, kinetic discretizations of magnetized plasmas where collisions are advanced in velocity space at each spatial point independently. This paper builds on previous work on a high-performance, fully nonlinear, Landau collision operator by batching the linear solver, as well as batching the spatial point problems and adding new support for multiple grids for multiscale, multispecies problems. An anisotropic relaxation verification test that agrees well with previously published results and analytical models is presented. The performance results from NVIDIA A100 and AMD MI250X nodes are presented with hardware utilization analysis for each architecture. Finally, the entire implicit Landau operator time advance is implemented in Kokkos for performance portability, running entirely on the device and is available in the PETSc numerical library.

97 MATHEMATICS AND COMPUTING↗

Massively parallel algorithms for trace-driven cache simulations

Trace driven cache simulation is central to computer design. A trace is a very long sequence of reference lines from main memory. At the t(exp th) instant, reference x sub t is hashed into a set of cache locations, the contents of which are then compared with x sub t. If at the t sup th instant x sub t is not present in the cache, then it is said to be a miss, and is loaded into the cache set, possibly forcing the replacement of some other memory line, and making x sub t present for the (t+1) sup st instant. The problem of parallel simulation of a subtrace of N references directed to a C line cache set is considered, with the aim of determining which references are misses and related statistics. A simulation method is presented for the Least Recently Used (LRU) policy, which regradless of the set size C runs in time O(log N) using N processors on the exclusive read, exclusive write (EREW) parallel model. A simpler LRU simulation algorithm is given that runs in O(C log N) time using N/log N processors. Timings are presented of the second algorithm's implementation on the MasPar MP-1, a machine with 16384 processors. A broad class of reference based line replacement policies are considered, which includes LRU as well as the Least Frequently Used and Random replacement policies. A simulation method is presented for any such policy that on any trace of length N directed to a C line set runs in the O(C log N) time with high probability using N processors on the EREW model. The algorithms are simple, have very little space overhead, and are well suited for SIMD implementation.

Nicol, David M.↗

A closer look in the mirror: reflections on the matter/dark matter coincidence

We argue that the striking similarity between the cosmic abundances of baryons and dark matter, despite their very different astrophysical behavior, strongly motivates the scenario in which dark matter resides within a rich dark sector parallel in structure to that of the standard model. The near cosmic coincidence is then explained by an approximate ℤ$_{2}$ exchange symmetry between the two sectors, where dark matter consists of stable dark neutrons, with matter and dark matter asymmetries arising via parallel WIMP baryogenesis mechanisms. Taking a top-down perspective, we point out that an adequate ℤ$_{2}$ symmetry necessitates solving the electroweak hierarchy problem in each sector, without our committing to a specific implementation. A higher-dimensional realization in the far UV is presented, in which the hierarchical couplings of the two sectors and the requisite ℤ$_{2}$-breaking structure arise naturally from extra-dimensional localization and gauge symmetries. We trace the cosmic history, paying attention to potential pitfalls not fully considered in previous literature. Residual ℤ$_{2}$-breaking can very plausibly give rise to the asymmetric reheating of the two sectors, needed to keep the cosmological abundance of relativistic dark particles below tight bounds. We show that, despite the need to keep inter-sector couplings highly suppressed after asymmetric reheating, there can naturally be order-one couplings mediated by TeV scale particles which can allow experimental probes of the dark sector at high energy colliders. Massive mediators can also induce dark matter direct detection signals, but likely at or below the neutrino floor.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Updates on the Predictive Materials Modeling Software Tools

Updates on NASA‘s efforts to build a Predictive Material Modeling (PMM) framework from the micro-scale to the macro-scale are presented in this abstract. The PMM effort is part of the Entry Systems Modeling (ESM) project under NASA’s Game Changing Development (GCD) program. To reduce the need for extensive testing and accelerate the design cycle process, ESM is developing simulation and modeling tools that enable the characterization of the properties of thermal protection materials and their response to extremely hot plasma. The Porous Microstructure Analysis (PuMA) software has been developed to compute effective material properties and perform material response simulations on digitized microstructures of porous media. PuMA is able to import three-dimensional digital images obtained from X-ray microtomography or to generate artificial microstructures that mimic real materials. PuMA also provides a module for interactive 3D visualizations. Version 3, which was recently released as open-source, includes modules to compute simple morphological properties such as porosity, volume fractions, pore diameter, and specific surface area. Additional capabilities include the determination of effective thermal and electrical conductivity (both radiative and solid conduction - including the ability to simulate local anisotropy for the latter); effective diffusivity and tortuosity from the continuum to the rarefied regime; techniques to determine the local material orientation, as well as mechanical properties (elasticity coefficients), and permeability. Computed properties are then used to inform a macro-scale material response model, such as those implemented in the Porous material Analysis Toolbox based on OpenFOAM (PATO) software developed within ESM. The computational model in PATO is a generic heat and mass transfer model for porous reactive materials containing several solid phases and a single gas phase. The detailed chemical interactions occurring between the solid phases and the gas phase are modeled at the pore scale, assuming Local Thermal Equilibrium. Recent efforts include the development of a mechanical erosion model as well as a unified model allowing an intrinsic coupling between fluid and material. Comparison to flight data (Mars Science Laboratory [MSL] Entry Descent and Landing Instrument [MEDLI] and Mars 2020 MEDLI2) is critical in order to validate these computational tools. Examples of ablative material response using the code will be presented, including 3D simulations of the full-scale heatshield of the MSL capsule. The simulations demonstrated the ability of the modern material response code, PATO, to handle the material response of geometrically complex and large domains through the use of massively parallel computations.

material modeling↗

Research and test facilities for development of technologies and experiments with commercial applications

One of NASA'S agency-wide goals is the commercial development of space. To further this goal NASA is implementing a policy whereby U.S. firms are encouraged to utilize NASA facilities to develop and test concepts having commercial potential. Goddard, in keeping with this policy, will make the facilities and capabilities described in this document available to private entities at a reduced cost and on a noninterference basis with internal NASA programs. Some of these facilities include: (1) the Vibration Test Facility; (2) the Battery Test Facility; (3) the Large Area Pulsed Solar Simulator Facility; (4) the High Voltage Testing Facility; (5) the Magnetic Field Component Test Facility; (6) the Spacecraft Magnetic Test Facility; (7) the High Capacity Centrifuge Facility; (8) the Acoustic Test Facility; (9) the Electromagnetic Interference Test Facility; (10) the Space Simulation Test Facility; (11) the Static/Dynamic Balance Facility; (12) the High Speed Centrifuge Facility; (13) the Optical Thin Film Deposition Facility; (14) the Gold Plating Facility; (15) the Paint Formulation and Application Laboratory; (16) the Propulsion Research Laboratory; (17) the Wallops Range Facility; (18) the Optical Instrument Assembly and Test Facility; (19) the Massively Parallel Processor Facility; (20) the X-Ray Diffraction and Scanning Auger Microscopy/Spectroscopy Laboratory; (21) the Parts Analysis Laboratory; (22) the Radiation Test Facility; (23) the Ainsworth Vacuum Balance Facility; (24) the Metallography Laboratory; (25) the Scanning Electron Microscope Laboratory; (26) the Organic Analysis Laboratory; (27) the Outgassing Test Facility; and (28) the Fatigue, Fracture Mechanics and Mechanical Testing Laboratory.

Source record↗

Predictive Modeling of Carbon Ablators

Efforts to build a Predictive Material Modeling (PMM) framework from the micro-scale to the macro-scale are presented in this abstract. To reduce the need for extensive testing, accelerate the design cycle process, and reduce uncertainty margins applied to final designs, NASA is developing simulation and modeling tools that enable characterization of material properties and response to high-enthalpy environments. The Porous Microstructure Analysis (PuMA) code has been developed for computing macroscale (volume averaged) properties of porous materials using microscale images from micro-computed tomography (micro-CT). Microscale modeling requires a realistic representation of a material microstructure; these are obtained either synthetically during the design of the material or through X-ray micro-CT. Volume averaged properties are then used to inform macroscale material response models, such as those implemented in the Porous-material Analysis Toolbox based on OpenFOAM (PATO) software, also actively developed by NASA. The computational model in PATO is a generic heat and mass transfer model for porous reactive materials containing several solid phases and a single gas phase. The detailed chemical interactions occurring between the solid phases and the gas phase are modeled at the pore scale assuming local thermal equilibrium. These tools were developed to efficiently interface with other pre-existing codes such as SPARTA (direct simulation Monte Carlo), DPLR (hypersonic CFD), NEQAIR (radiative transport) and DAKOTA (uncertainty quantification and optimization). Detailed flight data (Mars Science Laboratory [MSL] Entry Descent and Landing Instrument [MEDLI]) is critical for validating these computational tools for NASA applications. Examples of modeling ablative material response using these codes will be presented including 3D simulations of the full-scale heatshield of the MSL capsule. The simulations demonstrate the ability of the modern material response code, PATO, to handle the material response of geometrically complex and large domains, through the use of massively parallel computations.

Thermal Protection Systems↗

Predictive Modeling of Carbon Ablators Using Micro and Macro-Scale Modeling

Efforts to build a Predictive Material Modeling (PMM) framework from the micro-scale to the macro-scale are presented in this abstract. To reduce the need for extensive testing, accelerate the design cycle process, and reduce uncertainty margins applied to final designs, NASA is developing simulation and modeling tools that enable characterization of material properties and response to high-enthalpy environments. The Porous Microstructure Analysis (PuMA) code has been developed for computing macroscale (volume averaged) properties of porous materials using microscale images from micro-computed tomography (micro-CT). Microscale modeling requires a realistic representation of a material microstructure; these are obtained either synthetically during the design of the material or through X-ray micro-CT. Volume averaged properties are then used to inform macroscale material response models, such as those implemented in the Porous-material Analysis Toolbox based on OpenFOAM (PATO) software, also actively developed by NASA. The computational model in PATO is a generic heat and mass transfer model for porous reactive materials containing several solid phases and a single gas phase. The detailed chemical interactions occurring between the solid phases and the gas phase are modeled at the pore scale assuming local thermal equilibrium. These tools were developed to efficiently interface with other pre-existing codes such as SPARTA (direct simulation Monte Carlo), DPLR (hypersonic CFD), NEQAIR (radiative transport) and DAKOTA (uncertainty quantification and optimization). Detailed flight data (Mars Science Laboratory [MSL] Entry Descent and Landing Instrument [MEDLI]) is critical for validating these computational tools for NASA applications. Examples of modeling ablative material response using these codes will be presented including 3D simulations of the full-scale heatshield of the MSL capsule. The simulations demonstrate the ability of the modern material response code, PATO, to handle the material response of geometrically complex and large domains, through the use of massively parallel computations.

Thermal Protection Systems↗

Lightforce Photon-Pressure Collision Avoidance: Efficiency Analysis in the Current Debris Environment and Long-Term Simulation Perspective

This work provides an efficiency analysis of the LightForce space debris collision avoidance scheme in the current debris environment and describes a simulation approach to assess its impact on the long-term evolution of the space debris environment. LightForce aims to provide just-in-time collision avoidance by utilizing photon pressure from ground-based industrial lasers. These ground stations impart minimal accelerations to increase the miss distance for a predicted conjunction between two objects. In the first part of this paper we will present research that investigates the short-term effect of a few systems consisting of 20-kilowatt-class lasers directed by 1.5-meter-diameter telescopes using adaptive optics. The results found such a network of ground stations to mitigate more than 85 percent of conjunctions and could lower the expected number of collisions in Low Earth Orbit (LEO) by an order of magnitude. While these are impressive numbers that indicate LightForce's utility in the short-term, the remaining 15 percent of possible collisions contain (among others) conjunctions between two massive objects that would add large amount of debris if they collide. Still, conjunctions between massive objects and smaller objects can be mitigated. Hence, we choose to expand the capabilities of the simulation software to investigate the overall effect of a network of LightForce stations on the long-term debris evolution. In the second part of this paper, we will present the planned simulation approach for that effort. For the efficiency analysis of collision avoidance in the current debris environment, we utilize a simulation approach that uses the entire Two Line Element (TLE) catalog in LEO for a given day as initial input. These objects are propagated for one year and an all-on-all conjunction analysis is performed. For conjunctions that fall below a range threshold, we calculate the probability of collision and record those values. To assess efficiency, we compare a baseline (without collision avoidance) conjunction analysis with an analysis where LightForce is active. Using that approach, we take into account that collision avoidance maneuvers could have effects on third objects. Performing all-on-all conjunction analyses for extended period of time requires significant computer resources; hence we implemented this simulation utilizing a highly parallel approach on the NASA Pleiades supercomputer.

space debris mitigation↗

Evolution of Flexible Multibody Dynamics for Simulation Applications Supporting Human Spaceflight

During the course of transition from the Space Shuttle and International Space Station programs to the Orion and Journey to Mars exploration programs, a generic flexible multibody dynamics formulation and associated software implementation has evolved to meet an ever changing set of requirements at the NASA Johnson Space Center (JSC). Challenging problems related to large transitional topologies and robotic free-flyer vehicle capture/ release, contact dynamics, and exploration missions concept evaluation through simulation (e.g., asteroid surface operations) have driven this continued development. Coupled with this need is the requirement to oftentimes support human spaceflight operations in real-time. Moreover, it has been desirable to allow even more rapid prototyping of on-orbit manipulator and spacecraft systems, to support less complex infrastructure software for massively integrated simulations, to yield further computational efficiencies, and to take advantage of recent advances and availability of multi-core computing platforms. Since engineering analysis, procedures development, and crew familiarity/training for human spaceflight is fundamental to JSC's charter, there is also a strong desire to share and reuse models in both the non-realtime and real-time domains, with the goal of retaining as much multibody dynamics fidelity as possible. Three specific enhancements are reviewed here: (1) linked list organization to address large transitional topologies, (2) body level model order reduction, and (3) parallel formulation/implementation. This paper provides a detailed overview of these primary updates to JSC's flexible multibody dynamics algorithms as well as a comparison of numerical results to previous formulations and associated software.

Multibody dynamics↗

Resistive Switching of Spinel Li 4 Ti 5 O 12 Lithium-Ion Battery Material for Neuromorphic Computing

The rapid rise of AI has exposed significant limitations in conventional Von Neumann computing architecture, particularly in regard to speed and energy efficiency. To address these challenges, researchers are exploring a brain-inspired neuromorphic architecture that mimics biological neural networks, enabling massive parallel processing with reduced power consumption for complex AI computational demands. Recent interest has focused on utilizing battery electrodes and solid electrolyte materials for their resistive switching properties in developing a neuromorphic architecture. These properties are precisely tuned through local- and bulk-level chemical composition modifications via voltage bias stimuli. In this study, we demonstrate fabricating a three-terminal lithium-ion electrochemical transistor based on lithium titanium oxide (Li 4 Ti 5 O 12 ), a popular lithium-ion battery anode material. We deposited and characterized LTO thin films using RF sputtering, demonstrating a 6 orders of magnitude increase in electronic conductivity upon lithiation, with conductivity plateauing after 20% lithiation. Density functional theory calculations revealed transformation from the insulating to conducting state, supported by experimental characterization through X-Ray Photoelectron Spectroscopy (XPS) and Direct Current (DC) polarization analyses. The fabricated transistor consisted of LTO as the channel layer, gold as source/drain terminals, lithium phosphorus oxynitride (LiPON) as the lithium-ion conductor, and copper as the gate terminal. The device exhibited clear hysteresis in transfer characteristics due to lithium insertion/extraction processes. Long-term potentiation (LTP) and long-term depression (LTD) measurements showed an asymmetric ratio of 1.425 and maximum/minimum conductance ratio of 7.83. When implemented in a deep neural network (DNN) for MNIST handwritten digit recognition, the device achieved 92.03% accuracy over 20 training epochs. Detailed transport mechanism analysis revealed the crucial role of oxygen vacancies and interface effects in device operation. Our preliminary findings establish LTO-based lithium-ion electrochemical transistors as promising candidates for energy-efficient neuromorphic computing applications, offering potential solutions to traditional Von Neumann architecture limitations.

25 ENERGY STORAGE↗

Introduction to: Atlantic Meridional Overturning Circulation(AMOC)

A striking conclusion of the Intergovernmental Panel on Climate Change 2007 report is the crucial role that the Atlantic Meridional Overturning Circulation (AMOC) may play in anthropogenic climate change. However, these IPCC coupled climate simulations show a broad range of uncertainty in the magnitude and timing of AMOC transport change ranging from none to nearly complete collapse within the 21st century. The potential consequences of large changes in the characteristics of AMOC have motivated the creation in the United States of an interagency program and implementation plan to develop monitoring and prediction capabilities for the AMOC This program parallels the development of substantial monitoring efforts by European, South American and African countries -- notably the UK Rapid and Rapid-Watch programs. The papers contained in this volume are derived from presentations at the First U.S. Atlantic Meridional Overturning Circulation (AMOC) Meeting held 4 - 6 May, 2009 to review the US implementation plan and its coordination with other monitoring activities. The Atlantic Meridional Overturning Circulation consists of multiple components illustrated in an attached figure. Water enters the South Atlantic at upper and intermediate depths through both western and eastern routes (where eddy transport is especially important) and is transported northward across the equator, where it recirculates within the northern subtropical and subpolar gyres. The northern end is defined by the sinking regions of the Nordic Seas and the Labrador Sea where the waters that eventually form the upper and lower branches of North Atlantic Deep Water are conditioned. High surface salinities, the result of high net evaporation in the tropics and subtropics (including the Mediterranean Sea), and presence of regions of the Arctic Ocean that remain ice-free even in winter allow for the rapid cooling and thus densification of surface water. This dense surface water becomes the source of deep water formation in the sinking regions. In addition to transporting mass, the AMOC transports roughly half of the total amount of heat carried northward through the northern subtropics (down the temperature-gradient) by the ocean. In contrast in the Southern Hemisphere AMOC transports heat up-gradient from the cool Circumpolar Current to the warm tropics. Paleoevidence suggests that AMOC heat transport in the two hemispheres has varied over time in ways intimately tied to millennial changes in the Earth's climate. In one example, the abrupt Younger Dryas spell of cold weather over the North Atlantic, which began 13,000 years ago, has generally been linked to a millennial shutdown of the AMOC as a result of massive freshwater discharge from the North American continent. The current AMOC monitoring array consists of a series of instrumented transects located across key passages (see Cunningham et al., 2010 for a recent review). In the Arctic and sub-Arctic, transects cross Fram Strait, Denmark Strait and the Faroe Channel (connecting Greenland, Iceland, and the United Kingdom), as well as the entrance to the Labrador Sea. Further south and extending outwards from the east coast of North America there are a series of monitoring arrays including arrays of the Canadian Atlantic Zone Monitoring Program, deployments of the Rapid Western Atlantic Variability Experiment (WAVE), Line W at 39 N, as well as the Rapid-MOC moored array. The latter spans the entire Atlantic basin along 26.5 N. At tropical latitudes we have the Meridional Overturning Variability Experiment (MOVE) array at 16 N, while in the Southern Hemisphere a corresponding basin-spanning transect is being established at the latitude of Cape of Good Hope, complemented by arrays at Drake Passage.

Hakkinen, Sirpa↗