Search NASA⌕ Search

SEARCH · Search NASA

Results for “parallel processing (computers)”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 847 records · Page 47

The Design and Evaluation of "CAPTools"--A Computer Aided Parallelization Toolkit

Writing applications for high performance computers is a challenging task. Although writing code by hand still offers the best performance, it is extremely costly and often not very portable. The Computer Aided Parallelization Tools (CAPTools) are a toolkit designed to help automate the mapping of sequential FORTRAN scientific applications onto multiprocessors. CAPTools consists of the following major components: an inter-procedural dependence analysis module that incorporates user knowledge; a 'self-propagating' data partitioning module driven via user guidance; an execution control mask generation and optimization module for the user to fine tune parallel processing of individual partitions; a program transformation/restructuring facility for source code clean up and optimization; a set of browsers through which the user interacts with CAPTools at each stage of the parallelization process; and a code generator supporting multiple programming paradigms on various multiprocessors. Besides describing the rationale behind the architecture of CAPTools, the parallelization process is illustrated via case studies involving structured and unstructured meshes. The programming process and the performance of the generated parallel programs are compared against other programming alternatives based on the NAS Parallel Benchmarks, ARC3D and other scientific applications. Based on these results, a discussion on the feasibility of constructing architectural independent parallel applications is presented.

Yan, Jerry↗

On k-ary n-cubes: Theory and applications

Many parallel processing networks can be viewed as graphs called k-ary n-cubes, whose special cases include rings, hypercubes and toruses. In this paper, combinatorial properties of k-ary n-cubes are explored. In particular, the problem of characterizing the subgraph of a given number of nodes with the maximum edge count is studied. These theoretical results are then used to compute a lower bounding function in branch-and-bound partitioning algorithms and to establish the optimality of some irregular partitions.

Mao, Weizhen↗

Use of networked workstations for parallel nonlinear structural dynamic simulations of rotating bladed-disk assemblies

The principal objective of this research is to investigate, develop and demonstrate coarse-grained, parallel-processing strategies for nonlinear dynamic simulations for rotating bladed-disk assemblies. The parallel -processing strategies addressed include numerical algorithms for parallel nonlinear solutions and techniques to effect load balancing among processors. The parallel environment employed is a distributed-memory, coarse-grained one consisting of networked workstations. A parallel explicit time integration method has been implemented for transient nonlinear solutions of rotationg bladed-disk assemblies. Automatic domain partitioning techniques have been investigated for load balancing among processors. Advanced computing environments, data structures and interactive computer graphics all contribute to an integrated parallel finite element analysis system to facilitate more efficient and powerful dynamic simulations.

Hsieh, Shang-Hsien↗

Modeling of Failure for Analysis of Triaxial Braided Carbon Fiber Composites

In the development of advanced aircraft-engine fan cases and containment systems, composite materials are beginning to be used due to their low weight and high strength. The design of these structures must include the capability of withstanding impact loads from a released fan blade. Relatively complex triaxially braided fiber architectures have been found to yield the best performance for the fan cases. To properly work with and design these structures, robust analytical tools are required that can be used in the design process. A new analytical approach models triaxially braided carbon fiber composite materials within the environment of a transient dynamic finite-element code, specifically the commercially available transient dynamic finite-element code LS-DYNA. The geometry of the braided composites is approximated by a series of parallel laminated composites. The composite is modeled by using shell finite elements. The material property data are computed by examining test data from static tests on braided composites, where optical strain measurement techniques are used to examine the local strain variations within the material. These local strain data from the braided composite tests are used along with a judicious application of composite micromechanics- based methods to compute the stiffness properties of an equivalent unidirectional laminated composite required for the shell elements. The local strain data from the braided composite tests are also applied to back out strength and failure properties of the equivalent unidirectional composite. The properties utilized are geared towards the application of a continuum damage mechanics-based composite constitutive model available within LS-DYNA. The developed model can be applied to conduct impact simulations of structures composed of triaxially braided composites. The advantage of this technology is that it facilitates the analysis of the deformation and damage response of a triaxially braided polymer matrix composite within the environment of a transient dynamic finite-element code such as LS-DYNA in a manner which accounts for the local physical mechanisms but is still computationally efficient. This methodology is tightly coupled to experimental tests on the braided composite, which ensures that the material properties have physical significance. Aerospace or automotive companies interested in using triaxially braided composites in their structures, particularly for impact or crash applications, would find the technology useful. By the development of improved design tools, the amount of very expensive impact testing that will need to be performed can be significantly reduced.

Goldberg, Robert K.↗

Archive Management of NASA Earth Observation Data to Support Cloud Analysis

NASA collects, processes and distributes petabytes of Earth Observation (EO) data from satellites, aircraft, in situ instruments and model output, with an order of magnitude increase expected by 2024. Cloud-based web object storage (WOS) of these data can simplify the execution of such an increase. More importantly, it can also facilitate user analysis of those volumes by making the data available to the massively parallel computing power in the cloud. However, storing EO data in cloud WOS has a ripple effect throughout the NASA archive system with unexpected challenges and opportunities. One challenge is modifying data servicing software (such as Web Coverage Service servers) to access and subset data that are no longer on a directly accessible file system, but rather in cloud WOS. Opportunities include refactoring of the archive software to a cloud-native architecture; virtualizing data products by computing on demand; and reorganizing data to be more analysis-friendly.

Lynnes, Christopher↗

Archive Management of NASA Earth Observation Data to Support Cloud Analysis

NASA collects, processes and distributes petabytes of Earth Observation (EO) data from satellites, aircraft, in situ instruments and model output, with an order of magnitude increase expected by 2024. Cloud-based web object storage (WOS) of these data can simplify the execution of such an increase. More importantly, it can also facilitate user analysis of those volumes by making the data available to the massively parallel computing power in the cloud. However, storing EO data in cloud WOS has a ripple effect throughout the NASA archive system with unexpected challenges and opportunities. One challenge is modifying data servicing software (such as Web Coverage Service servers) to access and subset data that are no longer on a directly accessible file system, but rather in cloud WOS. Opportunities include refactoring of the archive software to a cloud-native architecture; virtualizing data products by computing on demand; and reorganizing data to be more analysis-friendly. Reviewed by Mark McInerney ESDIS Deputy Project Manager.

Lynnes, Christopher↗

Parallelization and visual analysis of multidimensional fields: Application to ozone production, destruction, and transport in three dimensions

Atmospheric modeling is a grand challenge problem for several reasons, including its inordinate computational requirements and its generation of large amounts of data concurrent with its use of very large data sets derived from measurement instruments like satellites. In addition, atmospheric models are typically run several times, on new data sets or to reprocess existing data sets, to investigate or reinvestigate specific chemical or physical processes occurring in the earth's atmosphere, to understand model fidelity with respect to observational data, or simply to experiment with specific model parameters or components.

Schwan, Karsten↗

Computational Study of Oxidative Etch Pitting in FiberForm and the Effect on Its Material Properties

Erosion of carbon due to oxidation does not occur uniformly but through the formation of localized etch pits because of active surface sites. These active sites are formed due to the presence of atomic defects on the carbon surface, and have much higher reactivity compared to average non-defective sites. Thus, these active sites are first to react during ablation, resulting in their removal. This causes all the neighboring atoms to be defective and increase their reactivity, thus leading to the localized carbon removal around these “active” sites. In this manner, these highly reactive defective sites serve as nucleation sites for the formation and growth of etch pits with potentially detrimental effects on the structural integrity. In order to understand the influence of these etch pits on the material properties of carbon fiber microstructures, we have developed a new capability within direct simulation Monte Carlo (DSMC) to capture the etch pit formation process. This capability is developed within the DSMC code SPARTA (Stochastic PArallel Rarefied-gas Time-accurate Analyzer) and can model the material removal in presence of active sites leading to the formation of etch pits. The focus of the current work will be to study the effect of etch pits on the material properties of FiberForm, a commonly used base material within many thermal protection system materials (TPS). The microstructure of virgin FiberForm obtained directly from X-ray microtomography experiments is used within SPARTA to obtain the ablated geometries with etch pits. These pitted microstructures are then imported within the Porous Microstructure Analysis (PuMA) software and various material properties such as elasticity, thermal conductivity, and permeability are computed. The variation of these properties as a result of the complex evolution of the surface topology due to etch pit formation is studied and analyzed. Furthermore, the effect of pitting is compared to the case of shrinking fibers, which has been the standard for modelling ablation of carbon structures; and significant differences are observed. Thus, such a physically realistic modeling of material removal through the formation of etch pits will be helpful in predicting the degradation of carbon-based TPS more accurately during oxidation; as well as other mechanisms such as spallation, which involves the removal of chunks of material into the flow due to etch pit growth. This will ultimately improve our understanding of the failure modes in these materials due to ablation.

Carbon Ablators↗

Computational Study of Oxidative Etch Pitting in FiberForm and the Effect on Its Material Properties

Erosion of carbon due to oxidation does not occur uniformly but through the formation of localized etch pits because of active surface sites. These active sites are formed due to the presence of atomic defects on the carbon surface, and have much higher reactivity compared to average non-defective sites. Thus, these active sites are first to react during ablation, resulting in their removal. This causes all the neighboring atoms to be defective and increase their reactivity, thus leading to the localized carbon removal around these “active” sites. In this manner, these highly reactive defective sites serve as nucleation sites for the formation and growth of etch pits with potentially detrimental effects on the structural integrity. In order to understand the influence of these etch pits on the material properties of carbon fiber microstructures, we have developed a new capability within direct simulation Monte Carlo (DSMC) to capture the etch pit formation process. This capability is developed within the DSMC code SPARTA (Stochastic PArallel Rarefied-gas Time-accurate Analyzer) and can model the material removal in presence of active sites leading to the formation of etch pits. The focus of the current work will be to study the effect of etch pits on the material properties of FiberForm, a commonly used base material within many thermal protection system materials (TPS). The microstructure of virgin FiberForm obtained directly from X-ray microtomography experiments is used within SPARTA to obtain the ablated geometries with etch pits. These pitted microstructures are then imported within the Porous Microstructure Analysis (PuMA) software and various material properties such as elasticity, thermal conductivity, and permeability are computed. The variation of these properties as a result of the complex evolution of the surface topology due to etch pit formation is studied and analyzed. Furthermore, the effect of pitting is compared to the case of shrinking fibers, which has been the standard for modelling ablation of carbon structures; and significant differences are observed. Thus, such a physically realistic modeling of material removal through the formation of etch pits will be helpful in predicting the degradation of carbon-based TPS more accurately during oxidation; as well as other mechanisms such as spallation, which involves the removal of chunks of material into the flow due to etch pit growth. This will ultimately improve our understanding of the failure modes in these materials due to ablation.

Carbon Ablators↗

Implementing Connected Component Labeling as a User Defined Operator for SciDB

We have implemented a flexible User Defined Operator (UDO) for labeling connected components of a binary mask expressed as an array in SciDB, a parallel distributed database management system based on the array data model. This UDO is able to process very large multidimensional arrays by exploiting SciDB's memory management mechanism that efficiently manipulates arrays whose memory requirements far exceed available physical memory. The UDO takes as primary inputs a binary mask array and a binary stencil array that specifies the connectivity of a given cell to its neighbors. The UDO returns an array of the same shape as the input mask array with each foreground cell containing the label of the component it belongs to. By default, dimensions are treated as non-periodic, but the UDO also accepts optional input parameters to specify periodicity in any of the array dimensions. The UDO requires four stages to completely label connected components. In the first stage, labels are computed for each subarray or chunk of the mask array in parallel across SciDB instances using the weighted quick union (WQU) with half-path compression algorithm. In the second stage, labels around chunk boundaries from the first stage are stored in a temporary SciDB array that is then replicated across all SciDB instances. Equivalences are resolved by again applying the WQU algorithm to these boundary labels. In the third stage, relabeling is done for each chunk using the resolved equivalences. In the fourth stage, the resolved labels, which so far are "flattened" coordinates of the original binary mask array, are renamed with sequential integers for legibility. The UDO is demonstrated on a 3-D mask of O(1011) elements, with O(108) foreground cells and O(106) connected components. The operator completes in 19 minutes using 84 SciDB instances.

UDO↗

Consequences of using nonlinear particle trajectories to compute spatial diffusion coefficients

The propagation of charged particles through interstellar and interplanetary space has often been described as a random process in which the particles are scattered by ambient electromagnetic turbulence. In general, this changes both the magnitude and direction of the particles' momentum. Some situations for which scattering in direction (pitch angle) is of primary interest were studied. A perturbed orbit, resonant scattering theory for pitch-angle diffusion in magnetostatic turbulence was slightly generalized and then utilized to compute the diffusion coefficient for spatial propagation parallel to the mean magnetic field, Kappa. All divergences inherent in the quasilinear formalism when the power spectrum of the fluctuation field falls off as K to the minus Q power (Q less than 2) were removed. Various methods of computing Kappa were compared and limits on the validity of the theory discussed. For Q less than 1 or 2, the various methods give roughly comparable values of Kappa, but use of perturbed orbits systematically results in a somewhat smaller Kappa than can be obtained from quasilinear theory.

Goldstein, M. L.↗

Magnetic configuration of the Venus magnetosheath

A data set consisting of nearly six Venus years of Pioneer Venus Orbiter magnetometer data is analyzed in order to describe the configuration of the magnetosheath. A coordinate system is developed which isolates the effects of the solar wind flow and the interplanetary magnetic field (IMF). The data indicate that the magnetosheath field magnitude is responsive to solar wind dynamic pressure and that the compression of the upstream field is controlled by magnetosonic Mach number. The draping of the field is similar to that predicted by gasdynamic modeling; specific draping features include distinct dependence on the radial and the transverse components of the IMF, and a tendency of the field to encircle the planet at low altitude. Certain features of the processes of mass loading by magnetosheath fluctuations and by the motional electric field are examined. Magnetic fluctuations can dominate the magnetosheath field for periods of low interplanetary cone angle and in general for the magnetosheath regions near the quasi-parallel part of the bow shock. The magnetosheath magnetic field magnitude displays a hemispherical asymmetry; the sense of this asymmetry is controlled by the cross-flow component of the IMF in a manner which indicates that ion pickup by the convective electric field occurs preferentially in one hemisphere. The results of this survey indicate that mass loading is a significant process which should be incorporated into computational models but that a fluid treatment of the planetary ion component may not be appropriate.

Phillips, J. L.↗

SHARP: A multi-mission artificial intelligence system for spacecraft telemetry monitoring and diagnosis

The Spacecraft Health Automated Reasoning Prototype (SHARP) is a system designed to demonstrate automated health and status analysis for multi-mission spacecraft and ground data systems operations. Telecommunications link analysis of the Voyager 2 spacecraft is the initial focus for the SHARP system demonstration which will occur during Voyager's encounter with the planet Neptune in August, 1989, in parallel with real time Voyager operations. The SHARP system combines conventional computer science methodologies with artificial intelligence techniques to produce an effective method for detecting and analyzing potential spacecraft and ground systems problems. The system performs real time analysis of spacecraft and other related telemetry, and is also capable of examining data in historical context. A brief introduction is given to the spacecraft and ground systems monitoring process at the Jet Propulsion Laboratory. The current method of operation for monitoring the Voyager Telecommunications subsystem is described, and the difficulties associated with the existing technology are highlighted. The approach taken in the SHARP system to overcome the current limitations is also described, as well as both the conventional and artificial intelligence solutions developed in SHARP.

Lawson, Denise L.↗

Automated target recognition and tracking using an optical pattern recognition neural network

The on-going development of an automatic target recognition and tracking system at the Jet Propulsion Laboratory is presented. This system is an optical pattern recognition neural network (OPRNN) that is an integration of an innovative optical parallel processor and a feature extraction based neural net training algorithm. The parallel optical processor provides high speed and vast parallelism as well as full shift invariance. The neural network algorithm enables simultaneous discrimination of multiple noisy targets in spite of their scales, rotations, perspectives, and various deformations. This fully developed OPRNN system can be effectively utilized for the automated spacecraft recognition and tracking that will lead to success in the Automated Rendezvous and Capture (AR&C) of the unmanned Cargo Transfer Vehicle (CTV). One of the most powerful optical parallel processors for automatic target recognition is the multichannel correlator. With the inherent advantages of parallel processing capability and shift invariance, multiple objects can be simultaneously recognized and tracked using this multichannel correlator. This target tracking capability can be greatly enhanced by utilizing a powerful feature extraction based neural network training algorithm such as the neocognitron. The OPRNN, currently under investigation at JPL, is constructed with an optical multichannel correlator where holographic filters have been prepared using the neocognitron training algorithm. The computation speed of the neocognitron-type OPRNN is up to 10(exp 14) analog connections/sec that enabling the OPRNN to outperform its state-of-the-art electronics counterpart by at least two orders of magnitude.

Chao, Tien-Hsin↗

Implementation and analysis of a Navier-Stokes algorithm on parallel computers

The results of the implementation of a Navier-Stokes algorithm on three parallel/vector computers are presented. The object of this research is to determine how well, or poorly, a single numerical algorithm would map onto three different architectures. The algorithm is a compact difference scheme for the solution of the incompressible, two-dimensional, time-dependent Navier-Stokes equations. The computers were chosen so as to encompass a variety of architectures. They are the following: the MPP, an SIMD machine with 16K bit serial processors; Flex/32, an MIMD machine with 20 processors; and Cray/2. The implementation of the algorithm is discussed in relation to these architectures and measures of the performance on each machine are given. The basic comparison is among SIMD instruction parallelism on the MPP, MIMD process parallelism on the Flex/32, and vectorization of a serial code on the Cray/2. Simple performance models are used to describe the performance. These models highlight the bottlenecks and limiting factors for this algorithm on these architectures. Finally, conclusions are presented.

Fatoohi, Raad A.↗

Supercomputing systems - A projection to 2000

Advances in computer architecture, computer science, computational methods, and constituent technologies are expected to lead to significant advances in the performance of scientific supercomputing system capabilities over the next decade. By the year 2000, single 1-in-sq dies are projected to incorporate four processors, each of which would be operating faster than 750 million instructions per second (MIPS) for a total on-chip processing performance in excess of 2000 MIPS. Scalable parallel processors can be expected to contain thousands of such multiple processor chips. In general, semiconductor performance advances appear to change about one order of magnitude every five years. Rotating magnetic memory and communications technology are not advancing as rapidly, with the result that the allocation of functions within the system configurations fo future supercomputer systems will require important changes. Availability of massively parallel heterogeneous processing capabilities should be a catalyst leading to new approaches for applications.

Lundstrom, S. F.↗

Object-Oriented Implementation of the NAS Parallel Benchmarks using Charm++

This report describes experiences with implementing the NAS Computational Fluid Dynamics benchmarks using a parallel object-oriented language, Charm++. Our main objective in implementing the NAS CFD kernel benchmarks was to develop a code that could be used to easily experiment with different domain decomposition strategies and dynamic load balancing. We also wished to leverage the object-orientation provided by the Charm++ parallel object-oriented language, to develop reusable abstractions that would simplify the process of developing parallel applications. We first describe the Charm++ parallel programming model and the parallel object array abstraction, then go into detail about each of the Scalar Pentadiagonal (SP) and Lower/Upper Triangular (LU) benchmarks, along with performance results. Finally we conclude with an evaluation of the methodology used.

Krishnan, Sanjeev↗

Recent Advancements in the PATO Material Response Code

Introduction: Predicting the complicated multiphysics phenomena during atmospheric entry requires high-fidelity modeling tools to refine estimates of mission risks during entry. To this end, new capabilities are being added to the Porous-material Analysis Toolbox based on OpenFOAM (PATO). PATO is an open-source software for Computational Material Response (CMR) of reactive porous materials submitted to high-temperature environments. The objective of this work is to highlight current efforts to add to and improve upon the modeling capabilities of PATO. These include efforts to loosely couple PATO with other discipline specialized codes including hypersonic Computational Fluid Dynamics (CFD), to assess the interaction effects between pyrolysis gas blowing and the boundary layer, and Computational Solid Mechanics (CSM), to address modeling of mechanical erosion. Other refinements include surface phenomena modeling capabilities to address the effects of silicone-based coatings applied to the TPS during flight preparation, and a unified multiphase solver for a mixed porous-material and plain-fluid domain. Coupling CMR with CFD (CMR/CFD): A loose coupling between PATO and the Data Parallel Line Relaxation (DPLR) CFD code has been achieved by making use of a blowing boundary condition at the heatshield surface available in DPLR. Starting with heat flux estimates with no pyrolysis gas blowing at the surface, blowing gases are computed by the CMR and passed to the CFD such that aerothermal properties of the environment can be recomputed for a new CMR computation. This leads to an iterative process which is supplemented with an estimate of the radiative heat flux using the Nonequilibrium air radiation (NEQAIR) program. The entire iterative process is illustrated in Figure 1. This coupling strategy has been utilized in computing the MSL material response. The goal is to compare the coupled CMR/CFD results with material response results obtained using traditional blowing corrections. Coupling CMS with CMR: A mechanical erosion model is currently being implemented in PATO to account for the additional mass removal induced by high shear conditions. The modeling process at each timestep consists of updating the mechanical properties as a function of temperature and computing the stress tensor and displacement fields of the material. Then, a failure criteria model determines the regions in which the stress exceeds the ultimate strength values resulting in mesh movement to account for mass removal. This model allows the material response simulation to compute the recession due to both oxidation and shear-induced erosion. The model is demonstrated by computing material response of sphere-cone arc jet samples. Surface Modeling Capabilities: NuSil, a silicone-based coating, was sprayed onto the MSL and Mars 2020 heatshields to mitigate shedding of phenolic dust. To better understand the effects of the NuSil coating on the material response, a novel model has been implemented in PATO. In this model, the equilibrium of the charred NuSil surface is modeled as pure silica, and a constant offset, inspired by the classical spallation model, is added to the the char blowing rate and wall enthalpy to reproduce HyMETS experimental results. The model has also been used to estimate the 3D material response of the MSL heatshield. Unified Solver: In addition to the iterative loose coupling approach mentioned above, a multiphase unified solver is being developed to couple the environment (plain-fluid phase) and the porous-material phase. The solver is based on the volume averaged conservation of mass, momentum, and energy for the macroscale with closure models which include microscale effects through effective physicochemical properties. The unified solver has been used to compute flow through a porous plug and solve the Beavers and Joseph problem. Since the strong coupling between phases is inherent to this solver, modeling assumptions present in other coupling methods of material response are mitigated. This strategy also makes it feasible to capture the competition between surface and volume ablation in the same computational domain, which is usually not possible with other coupling approaches.

Material Response↗