Search NASA⌕ Search

SEARCH · Search NASA

Results for “distributed parallel computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 595 records · Page 33

Software For Drawing Design Details Concurrently

Software system containing five computer-aided-design programs enables more than one designer to work on same part or assembly at same time. Reduces time necessary to produce design by implementing concept of parallel or concurrent detailing, in which all detail drawings documenting three-dimensional model of part or assembly produced simultaneously, rather than sequentially. Keeps various detail drawings consistent with each other and with overall design by distributing changes in each detail to all other affected details.

Crosby, Dewey C., III↗

Analysis of coseismic surface displacement gradients using radar interferometry: New insights into the Landers earthquake

The map of the coseismic displacement field generated by interferometric processing of synthetic aperture radar (SAR) images taken before and after the June 28, 1992, Landers earthquake sequence brings new insights into the nature of deformation caused by these earthquakes. We use the interferometric map generated by Massonnet et al. (1993) to analyze the surface displacement field in the vicinity of the fault trace. Complexities in the fringe pattern near the fault reflect short-wavelength variations of the surface rupture and slip distribution, and attest to large displacement gradients. Along two sections of the fault, characteristic fringe patterns can be recognized, contrasting in density and direction with patterns observed away from the rupture. In order to understand the observed fringe patterns, we compute synthetic interferograms in three simple cases: (1) rigid-body rotations about a vertical axis, (2) about a horizontal axis (tilt), and (3) distributed, simple shear. The orientation and spatial separation of interferometric fringes predicted by these models help constrain near-field deformation and rupture parameters. Where the Kickapoo fault connects with the Homestead Valley fault, the interferogram shows a clear pattern of parallel N20 deg W fringes separated by about 160 m. This pattern and vertical offsets measured along the Kickapoo fault suggest that the block between this fault and the Johnson Valley fault may have been tilted, down to the west. A 5-km block lifted by 1 m on one side would be tilted by an angle of 0.01 deg (190 microrad), producing fringes separated by about 160 m, parallel to the tilt axis. Such a tilt, parallel to a N20 deg W direction, would account for the gradual, northward increase of the vertical slip component observed along the Kickapoo fault. This tilt may also explain the 1 m of reverse slip observed along the 'slip gap' section of the Homestead Valley break. Between the southern end of the Johnson Valley fault and the Eureka Peak fault, where no surface rupture has been mapped, the dense pattern of fringes implies distributed shear, probably resulting from fault slip at depth. The density and direction of the fringes in the gap are consistent with a right-lateral slip of 1.2-3.8 m on a blind fault locked above the depth of 1.5-2 km. Such observations of small wavelength features in the SAR interferogram bring new insights into the near-field displacement gradient and thus on response of the uppermost crust to seismic rupture.

Peltzer, Gilles↗

Validation of a Pressure-Based Combustion Simulation Tool Using a Single Element Injector Test Problem

The traditional design and analysis practice for advanced propulsion systems, particularly chemical rocket engines, relies heavily on expensive full-scale prototype development and testing. Over the past decade, use of high-fidelity analysis and design tools such as CFD early in the product development cycle has been identified as one way to alleviate testing costs and to develop these devices better, faster and cheaper. Increased emphasis is being placed on developing and applying CFD models to simulate the flow field environments and performance of advanced propulsion systems. This necessitates the development of next generation computational tools which can be used effectively and reliably in a design environment by non-CFD specialists. A computational tool, called Loci-STREAM is being developed for this purpose. It is a pressure-based, Reynolds-averaged Navier-Stokes (RANS) solver for generalized unstructured grids, which is designed to handle all-speed flows (incompressible to hypersonic) and is particularly suitable for solving multi-species flow in fixed-frame combustion devices. Loci-STREAM integrates proven numerical methods for generalized grids and state-of-the-art physical models in a novel rule-based programming framework called Loci which allows: (a) seamless integration of multidisciplinary physics in a unified manner, and (b) automatic handling of massively parallel computing. The objective of the ongoing work is to develop a robust simulation capability for combustion problems in rocket engines. As an initial step towards validating this capability, a model problem is investigated in the present study which involves a gaseous oxygen/gaseous hydrogen (GO2/GH2) shear coaxial single element injector, for which experimental data are available. The sensitivity of the computed solutions to grid density, grid distribution, different turbulence models, and different near-wall treatments is investigated. A refined grid, which is clustered in the vicinity of the solid walls as well as the flame, is used to obtain a steady state solution which may be considered as the best solution attainable with the steady-state RANS methodology. From a design point of view, quick turnaround times are desirable; with this in mind, coarser grids are also employed and the resulting solutions are evaluated with respect to the fine grid solution.

Thakur, Siddarth↗

Practical Implementation of GPU-based Computing at the Grid Edge for Resilience Scenarios

This paper presents a practical implementation of GPU-accelerated computing at the grid edge to enhance power system resilience through next-generation smart meters. Advanced Metering Infrastructure (AMI) systems rely predominantly on centralized processing architectures, which limit real-time response capabilities during grid disturbances. This work proposes the integration of GPU-enabled computational platforms directly within smart meter to enable local execution support for power system analytics, fault detection algorithms, and optimization routines. The proposed framework uses the Julia programming language to leverage highperformance parallel computing capabilities while maintaining code portability and development efficiency. We use two experimental scenarios to benchmark the computational feasibility of this approach: sparse linear system solutions representative of power flow analyses, and multi-stage production cost simulations incorporating unit commitment and economic dispatch operations. Results demonstrate that computationally intensive power system algorithms, such as those supporting resilience scenario calculations, can be effectively executed at the distribution edge using commercially available embedded GPU hardware. Keywords—GPU acceleration, edge computing, smart meters, grid resilience, AMI, resilience.

De Souza, Reubun [School of Electrical Engineering↗

MAX - An advanced parallel computer for space applications

MAX is a fault-tolerant multicomputer hardware and software architecture designed to meet the needs of NASA spacecraft systems. It consists of conventional computing modules (computers) connected via a dual network topology. One network is used to transfer data among the computers and between computers and I/O devices. This network's topology is arbitrary. The second network operates as a broadcast medium for operating system synchronization messages and supports the operating system's Byzantine resilience. A fully distributed operating system supports multitasking in an asynchronous event and data driven environment. A large grain dataflow paradigm is used to coordinate the multitasking and provide easy control of concurrency. It is the basis of the system's fault tolerance and allows both static and dynamical location of tasks. Redundant execution of tasks with software voting of results may be specified for critical tasks. The dataflow paradigm also supports simplified software design, test and maintenance. A unique feature is a method for reliably patching code in an executing dataflow application.

Lewis, Blair F.↗

Implementing Atmospheric Infrared Sounder (AIRS) and Cross-Track Infrared Sounder (CrIS) Cloud-Clearing Algorithm into the NASA GEOS: Focus on the 2017 Atlantic Tropical Cyclone Season

Numerical Weather Prediction (NWP) centers assimilate cloud-free infrared (IR) radiances because the assimilation of all-sky IR radiances is not yet operationally achievable. The cloud-clearing procedure offers a simpler, but effective strategy that produces cloud-affected radiances suitable for assimilation in partially cloudy regions. Several studies conducted by this team have demonstrated that IR Cloud-Cleared Radiances (CCRs), if thinned more aggressively than clear-sky radiances, can improve analysis and forecasts, particularly in meteorologically active areas. However, CCRs are not used by operational centers due partly to the thought that the process of cloud-clearing may affect latency and introduce difficult-to-control external dependencies. This study presents the results of implementing an Atmospheric Infrared Sounder (AIRS) and Cross-Track Infrared Sounder (CrIS) cloud-clearing procedure into the NASA Goddard Earth Observing System (GEOS) to demonstrate the portability of the procedure. The AIRS and CrIS cloud-clearing algorithms have been deprived of external dependencies, made customizable to any specific model, and the computational efficiency has been improved via parallelization. The revised AIRS and CrIS cloud-clearing algorithms allow a customized choice of channel selection, the use of a user-specified model's fields as first guess, and can perform in real time. Data assimilation experiments with the hybrid 4DEnVar GEOS system were successfully performed for the 2017 tropical cyclones (TC) season with a focus on three major hurricanes (Harvey, Irma, and Maria). This study shows that assimilation of locally-generated CCRs have a positive impact on both global skill and TC representation, compared to the assimilation of AIRS and CrIS clear-sky radiances, and a comparable or slightly improved impact compared to assimilation of CCRs produced by external sources, such as NASA's Distributed Active Archive Centers and NOAA’s Comprehensive Large Array-data Stewardship System. The customization and computational efficiency of the revised procedure would enable its usability in a real-time forecast context.

Niama Boukachaba↗

Array distribution in data-parallel programs

We consider distribution at compile time of the array data in a distributed-memory implementation of a data-parallel program written in a language like Fortran 90. We allow dynamic redistribution of data and define a heuristic algorithmic framework that chooses distribution parameters to minimize an estimate of program completion time. We represent the program as an alignment-distribution graph. We propose a divide-and-conquer algorithm for distribution that initially assigns a common distribution to each node of the graph and successively refines this assignment, taking computation, realignment, and redistribution costs into account. We explain how to estimate the effect of distribution on computation cost and how to choose a candidate set of distributions. We present the results of an implementation of our algorithms on several test problems.

Chatterjee, Siddhartha↗

11.5 micron emission from smokes computed using finite cloud geometry

This study presents the infrared radiative transport properties of smoke produced by brush fires under conditions when the smoke is confined to a finite horizontal area and the usual plane parallel model for radiative transfer in an absorbing and scattering medium may not be valid. The transport model is a three-dimensional version of the two-stream approximation applied to finite cuboidal clouds of water, carbon and silicates with known optical properties; assumptions are made regarding the particle size distribution of the polydispersion.

Harshvardhan, MR.↗

Computing architecture for telerobots in earth orbit

Based on generic operational and computational requirements associated with the control of telerobots in earth orbit, a multibus-based distributed but integrated computing architecture is proposed. An experimental system of that kind under development at the Jet Propulsion Laboratory (JPL) is briefly described. It uses Intel Multibus I at both control station and remote robot (telerobot) computing nodes. An essential element within each multibus is a Unified (or Universal) Computer Control Subsystem (UCCS) for telerobot and control station motor components. The two multibus-based computing nodes can be linked by parallel or high speed serial links for real-time data transmission and for closing the real-time bilateral (force-reflecting) control loop between telerobot and control station. The experimental system is briefly commented, followed by a brief discussion of future development plans and possibilities.

Bejczy, A. K.↗

Upwind-biased, point-implicit relaxation strategies for hypersonic flowfield simulations on supercomputers

An upwind-biased, point-implicit relaxation algorithm for obtaining the numerical solution to the governing equations for three-dimensional, viscous, hypersonic flows in chemical and thermal nonequilibrium is described. The algorithm is derived using a finite-volume formulation in which the inviscid components of flux across cell walls are described with Roe's averaging and Harten's entropy fix with second-order corrections based on Yee's Symmetric Total Variation Diminishing scheme. The relaxation strategy is well suited for computers employing either vector or parallel architectures, and the relation between computer architecture and algorithm is emphasized. It is also well suited to the numerical solution of the governing equations on unstructured grids. Because of the point-implicit relaxation strategy, the algorithm remains stable at large Courant numbers without the necessity of solving large. block tri-diagonal systems. A single relaxation step depends only on information from nearest neighbors. Predictions for pressure distributions, surface heating, and aerodynamic coefficients compare well with experimental data for Mach 10 flow over a blunt body. Predictions for the hypersonic flow of air in chemical and thermal nonequilibrium (velocity = 8917 m/s, altitude = 78 km.) over the Aeroassist Flight Experiment (AFE) configuration obtained on a multi-domain grid are discussed.

Gnoffo, Peter A.↗

Mechanisms of Active Aerodynamic Load Reduction on a Rotorcraft Fuselage With Rotor Effects

The reduction of the aerodynamic load that acts on a generic rotorcraft fuselage by the application of active flow control was investigated in a wind tunnel test conducted on an approximately 1/3-scale powered rotorcraft model simulating forward flight. The aerodynamic mechanisms that make these reductions, in both the drag and the download, possible were examined in detail through the use of the measured surface pressure distribution on the fuselage, velocity field measurements made in the wake directly behind the ramp of the fuselage and computational simulations. The fuselage tested was the ROBIN-mod7, which was equipped with a series of eight slots located on the ramp section through which flow control excitation was introduced. These slots were arranged in a U-shaped pattern located slightly downstream of the baseline separation line and parallel to it. The flow control excitation took the form of either synthetic jets, also known as zero-net-mass-flux blowing, and steady blowing. The same set of slots were used for both types of excitation. The differences between the two excitation types and between flow control excitation from different combinations of slots were examined. The flow control is shown to alter the size of the wake and its trajectory relative to the ramp and the tailboom and it is these changes to the wake that result in a reduction in the aerodynamic load.

Schaeffler, Norman W.↗

Particle Interaction Physics Model Formulation for Plume-Surface Interaction Erosion and Cratering

The Predictive Simulation Capability development team of the STMD Game Changing Development sponsored PSI project is implementing computational simulation capability for the efficient and accurate prediction of Plume-Surface Interaction induced surface erosion and cratering in Martian and Lunar environments. The status of the Focus Area 3 of the PSI project in the generation and efficient application of accurate soil particle composition modeling in the Gas-Granular Flow Solver (GGFS) computational framework is presented. The process of constitutive closure model database generation using DEM particle interaction modeling for capturing the effects of irregular particle shape and poly-disperse mixture distribution effects is outlined. This capability has now been ported to NASA supercomputer assets and NASA engineers successfully demonstrated technology and skillset transfer in model generation for spherical and irregularly shaped, mono-disperse and bi-disperse mixture compositions. Assessment of the computational efficiency and practicality of the academic serially executed DEM tools on NASA supercomputers identified the need to migrate to a DEM framework capable of performing parallel simulations in a simultaneous process orchestrated in an automated setup, execution, database extraction, and dataset delivery ready for application simulations. The LIGGGHTS DEM toolset has been selected as the most suitable tool to migrate the DEM simulations. Once the soil model generation process is implemented, models capturing the shape and poly-dispersity effects will be generated to perform much refined validation simulations against the experiments performed under the PSI project. The application readiness of the soil models currently operational in GGFS was presented for the example of a full scale, 3-D simulation of the plume induced erosion and crater formation of the Apollo LM at an elevation of 5m above ground in a low pressure, near vacuum background.

Peter A Liever↗

Particle Interaction Physics Model Formulation for Plume-Surface Interaction Erosion and Cratering

The Predictive Simulation Capability development team of the STMD Game Changing Development sponsored PSI project is implementing computational simulation capability for the efficient and accurate prediction of Plume-Surface Interaction induced surface erosion and cratering in Martian and Lunar environments. The status of the Focus Area 3 of the PSI project in the generation and efficient application of accurate soil particle composition modeling in the Gas-Granular Flow Solver (GGFS) computational framework is presented. The process of constitutive closure model database generation using DEM particle interaction modeling for capturing the effects of irregular particle shape and poly-disperse mixture distribution effects is outlined. This capability has now been ported to NASA supercomputer assets and NASA engineers successfully demonstrated technology and skillset transfer in model generation for spherical and irregularly shaped, mono-disperse and bi-disperse mixture compositions. Assessment of the computational efficiency and practicality of the academic serially executed DEM tools on NASA supercomputers identified the need to migrate to a DEM framework capable of performing parallel simulations in a simultaneous process orchestrated in an automated setup, execution, database extraction, and dataset delivery ready for application simulations. The LIGGGHTS DEM toolset has been selected as the most suitable tool to migrate the DEM simulations. Once the soil model generation process is implemented, models capturing the shape and poly-dispersity effects will be generated to perform much refined validation simulations against the experiments performed under the PSI project. The application readiness of the soil models currently operational in GGFS was presented for the example of a full scale, 3-D simulation of the plume induced erosion and crater formation of the Apollo LM at an elevation of 5m above ground in a low pressure, near vacuum background.

Peter A Liever↗

Effects of Ordering Strategies and Programming Paradigms on Sparse Matrix Computations

The Conjugate Gradient (CG) algorithm is perhaps the best-known iterative technique to solve sparse linear systems that are symmetric and positive definite. For systems that are ill-conditioned, it is often necessary to use a preconditioning technique. In this paper, we investigate the effects of various ordering and partitioning strategies on the performance of parallel CG and ILU(O) preconditioned CG (PCG) using different programming paradigms and architectures. Results show that for this class of applications: ordering significantly improves overall performance on both distributed and distributed shared-memory systems, that cache reuse may be more important than reducing communication, that it is possible to achieve message-passing performance using shared-memory constructs through careful data ordering and distribution, and that a hybrid MPI+OpenMP paradigm increases programming complexity with little performance gains. A implementation of CG on the Cray MTA does not require special ordering or partitioning to obtain high efficiency and scalability, giving it a distinct advantage for adaptive applications; however, it shows limited scalability for PCG due to a lack of thread level parallelism.

Oliker, Leonid↗

Accurate pressure gradient calculations in hydrostatic atmospheric models

A method for the accurate calculation of the horizontal pressure gradient acceleration in hydrostatic atmospheric models is presented which is especially useful in situations where the isothermal surfaces are not parallel to the vertical coordinate surfaces. The present method is shown to be exact if the potential temperature lapse rate is constant between the vertical pressure integration limits. The technique is applied to both the integration of the hydrostatic equation and the computation of the slope correction term in the horizontal pressure gradient. A fixed vertical grid and a dynamic grid defined by the significant levels in the vertical temperature distribution are employed.

Carroll, John J.↗

Using parallel banded linear system solvers in generalized eigenvalue problems

Subspace iteration is a reliable and cost effective method for solving positive definite banded symmetric generalized eigenproblems, especially in the case of large scale problems. This paper discusses an algorithm that makes use of two parallel banded solvers in subspace iteration. A shift is introduced to decompose the banded linear systems into relatively independent subsystems and to accelerate the iterations. With this shift, an eigenproblem is mapped efficiently into the memories of a multiprocessor and a high speedup is obtained for parallel implementations. An optimal shift is a shift that balances total computation and communication costs. Under certain conditions, we show how to estimate an optimal shift analytically using the decay rate for the inverse of a banded matrix, and how to improve this estimate. Computational results on iPSC/2 and iPSC/860 multiprocessors are presented.

DISTRIBUTED MEMORY MULTIPROCES↗

Load Balancing Unstructured Adaptive Grids for CFD Problems

Mesh adaption is a powerful tool for efficient unstructured-grid computations but causes load imbalance among processors on a parallel machine. A dynamic load balancing method is presented that balances the workload across all processors with a global view. After each parallel tetrahedral mesh adaption, the method first determines if the new mesh is sufficiently unbalanced to warrant a repartitioning. If so, the adapted mesh is repartitioned, with new partitions assigned to processors so that the redistribution cost is minimized. The new partitions are accepted only if the remapping cost is compensated by the improved load balance. Results indicate that this strategy is effective for large-scale scientific computations on distributed-memory multiprocessors.

Biswas, Rupak↗

Diffuse ions produced by electromagnetic ion beam instabilities

The evolution of the electromagnetic ion beam instability driven by the reflected ion component backstreaming away from the earth's bow shock into the foreshock region is studied by means of computer simulation. The linear and quasi-linear stages of the instability are found to be in good agreement with known results for the resonant mode propagating parallel to the beam along the magnetic field and with theory developed in this paper for the nonresonant mode, which propagates antiparallel to the beam direction. The quasi-linear stage, which produces large amplitude delta B approximately B, sinusoidal transverse waves and 'intermediate' ion distributions, is terminated by a nonlinear phase in which strongly nonlinear, compressive waves and 'diffuse' ion distributions are produced. Additional processes by which the diffuse ions are accelerated to observed high energies are not addressed. The results are discussed in terms of the ion distributions and hydromagnetic waves observed in the foreshock of the earth's bow shock and of interplanetary shocks.

Winske, D.↗