Search NASA⌕ Search

SEARCH · Search NASA

Results for “parallel processing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,081 records · Page 60

Engineered Carbon Nanotube Materials for High-Q Nanomechanical Resonators

This document represents a presentation offered by the Jet Propulsion Laboratory, with assistance from researchers from Brown University and Northrop Grumman. The presentation took place in Seoul, Korea in July 2003 and attempted to demonstrate the fabrication approach regarding the development of high quality factor (high-Q) mechanical oscillators (in the forms of a tunable nanotube resonator and a nanotube array radio frequency [RF] filter) aimed at signal processing and based on carbon nanotubes. The presentation also addressed parallel efforts to develop both in-plane single nanotube resonators as well as vertical array power devices.

chemical vapor depositions↗

Flight Results from the HST SM4 Relative Navigation Sensor System

On May 11, 2009, Space Shuttle Atlantis roared off of Launch Pad 39A enroute to the Hubble Space Telescope (HST) to undertake its final servicing of HST, Servicing Mission 4. Onboard Atlantis was a small payload called the Relative Navigation Sensor experiment, which included three cameras of varying focal ranges, avionics to record images and estimate, in real time, the relative position and attitude (aka "pose") of the telescope during rendezvous and deploy. The avionics package, known as SpaceCube and developed at the Goddard Space Flight Center, performed image processing using field programmable gate arrays to accelerate this process, and in addition executed two different pose algorithms in parallel, the Goddard Natural Feature Image Recognition and the ULTOR Passive Pose and Position Engine (P3E) algorithms

Naasz, Bo↗

Recent Developments in the Code RITRACKS (Relativistic Ion Tracks)

The code RITRACKS (Relativistic Ion Tracks) was developed to simulate detailed stochastic radiation track structures of ions of different types and energies. Many new capabilities were added to the code during the recent years. Several options were added to specify the times at which the tracks appear in the irradiated volume, allowing the simulation of dose-rate effects. The code has been used to simulate energy deposition in several targets: spherical, ellipsoidal and cylindrical. More recently, density changes as well as a spherical shell were implemented for spherical targets, in order to simulate energy deposition in walled tissue equivalent proportional counters. RITRACKS is used as a part of the new program BDSTracks (Biological Damage by Stochastic Tracks) to simulate several types of chromosome aberrations in various irradiation conditions. The simulation of damage to various DNA structures (linear and chromatin fiber) by direct and indirect effects has been improved and is ongoing. Many improvements were also made to the graphic user interface (GUI), including the addition of several labels allowing changes of units. A new GUI has been added to display the electron ejection vectors. The parallel calculation capabilities, notably the pre- and post-simulation processing on Windows and Linux machines have been reviewed to make them more portable between different systems. The calculation part is currently maintained in an Atlassian Stash® repository for code tracking and possibly future collaboration.

Plante, Ianik↗

Recent Improvements to the LAURA and HARA Codes

This paper describes recent improvements to the LAURA and HARA codes. LAURA is a CFD code for aerothermodynamics, and HARA evaluates the shock-layer radiation that provides the radiative source term for the flowfield energy equations and radiative heating to a surface. The next release of LAURA and HARA includes a variety of new capabilities. These new capabilities include an automated uncertainty quantification workflow for radiative heat transfer, options for specifying surface roughness and turbulent transition location in the algebraic turbulence models, and improved grid and solution interpolation techniques. Additionally, the computational efficiency of both LAURA and HARA have been improved. Optimization of the MPI communication routines in LAURA are shown to improve the parallel efficiency of the primary flow when running with multiple processes per block, and recent optimization of HARA leverage graphics processing unit (GPU) acceleration in the radiation calculations. Using GPU acceleration of HARA is shown to decrease the cost of the radiation line-of-sight calculation by approximately one order of magnitude for a 10.5 km/s Earth entry simulation.

LAURA HARA CFD 5.6↗

Collaborative Pose Estimation of An Unknown Target Using Multiple Spacecraft

A reliable method for pose estimation of an unknown and uncooperative space target using monocular vision remains an open problem. Vision-based pose determination can be challenging in case of unfavorable illumination, time-varying conditions due to rotational motion and relative orbit, and scale ambiguity resolution. To address these challenges, we propose a novel collaborative pose determination algorithm called Multi- Spacecraft Simultaneous Estimation of Pose and Shape algorithm or M-SEPS.Within M-SEPS, a team of chaser spacecraft, each equipped with a monocular camera, exchange information over a local network to jointly estimate the relative kinematic state of the target and its sparse shape landmarks. In this approach, each spacecraft processes its own images and observes particular target landmarks in parallel and in a distributed fashion. Then, the local network is exploited by the spacecraft to share their consensus proposals and aggregate them to achieve the joint estimate. We validate our algorithm using simulations of relative orbits and observations, captured by each chaser spacecraft. To the best of the authors’ knowledge, this is the first cooperative, vision-based algorithm for estimating the pose and shape of a space object for an arbitrary number of spacecraft.

Chung, Soon-Jo↗

Computational Investigation of Oxidative Etch Pitting in FiberForm and Its Impact on Material Properties

Oxidation-driven carbon erosion does not occur uniformly but rather through the development of localized etch pits at active surface sites. These active sites form due to atomic defects on the carbon surface, making them significantly more reactive than the surrounding, non-defective areas. As a result, these sites are the first to react during ablation, leading to their removal. This process creates new defects in neighboring atoms, increasing their reactivity and causing localized carbon removal around these active sites. In this way, the highly reactive defective areas serve as nucleation points for the formation and growth of etch pits, which can have adverse effects on structural integrity of FiberForm. To better understand how these etch pits impact the material properties of carbon fiber microstructures, we have developed a new capability within the direct simulation Monte Carlo (DSMC) framework to capture the etch pit formation process. This capability, integrated into the DSMC code SPARTA (Stochastic Parallel Rarefied-gas Time-accurate Analyzer), models material removal in the presence of active sites, leading to the formation of etch pits. The current work focuses on studying the effects of these etch pits on the material properties of FiberForm, a widely used base material in thermal protection systems (TPS). The microstructure of virgin FiberForm, obtained via X-ray microtomography, is imported into SPARTA to generate the ablated geometries with etch pits. These modified microstructures are then analyzed using the Porous Microstructure Analysis (PuMA) software to compute various material properties, including elasticity, thermal conductivity, and permeability. We investigate the variation of these properties due to the complex surface topology changes caused by etch pit formation. Additionally, we compare the effects of pitting with the conventional model of shrinking fibers, traditionally used to simulate the ablation of carbon structures. Significant differences emerge between the two approaches. Consequently, this physically realistic model of material removal through etch pit formation offers improved accuracy in predicting the degradation of carbon-based TPS during oxidation. It also provides insights into other mechanisms, such as spallation, where chunks of material are removed into the flow due to etch pit growth. Ultimately, this model enhances our understanding of failure modes in these materials during ablation.

PuMA↗

Detection of edges using local geometry

Researchers described a new representation, the local geometry, for early visual processing which is motivated by results from biological vision. This representation is richer than is often used in image processing. It extracts more of the local structure available at each pixel in the image by using receptive fields that can be continuously rotated and that go to third order spatial variation. Early visual processing algorithms such as edge detectors and ridge detectors can be written in terms of various local geometries and are computationally tractable. For example, Canny's edge detector has been implemented in terms of a local geometry of order two, and a ridge detector in terms of a local geometry of order three. The edge detector in local geometry was applied to synthetic and real images and it was shown using simple interpolation schemes that sufficient information is available to locate edges with sub-pixel accuracy (to a resolution increase of at least a factor of five). This is reasonable even for noisy images because the local geometry fits a smooth surface - the Taylor series - to the discrete image data. Only local processing was used in the implementation so it can readily be implemented on parallel mesh machines such as the MPP. Researchers expect that other early visual algorithms, such as region growing, inflection point detection, and segmentation can also be implemented in terms of the local geometry and will provide sufficiently rich and robust representations for subsequent visual processing.

Gualtieri, J. A.↗

MLP: A Parallel Programming Alternative to MPI for New Shared Memory Parallel Systems

Recent developments at the NASA AMES Research Center's NAS Division have demonstrated that the new generation of NUMA based Symmetric Multi-Processing systems (SMPs), such as the Silicon Graphics Origin 2000, can successfully execute legacy vector oriented CFD production codes at sustained rates far exceeding processing rates possible on dedicated 16 CPU Cray C90 systems. This high level of performance is achieved via shared memory based Multi-Level Parallelism (MLP). This programming approach, developed at NAS and outlined below, is distinct from the message passing paradigm of MPI. It offers parallelism at both the fine and coarse grained level, with communication latencies that are approximately 50-100 times lower than typical MPI implementations on the same platform. Such latency reductions offer the promise of performance scaling to very large CPU counts. The method draws on, but is also distinct from, the newly defined OpenMP specification, which uses compiler directives to support a limited subset of multi-level parallel operations. The NAS MLP method is general, and applicable to a large class of NASA CFD codes.

Taft, James R.↗

Spatial distributions of magnetic field fluctuations in the dayside magnetosheath

In a study that tests the hypothesis that magnetosheath magnetic fields are disturbed on plasma streamlines which are connected to the quasi-parallel bow shock, magnetometer observations from the ISEE 2 and IMP 8 spacecraft are used to investigate the dayside spatial distributions of fluctuating magnetosheath fields for different interplanetary field orientations. The results suggest that although other sources such as Kelvin-Helmholtz instabilities and flux transfer processes probably contribute to the fluctuations in the magnetosheath field, the quasi-parallel shock source is an important contributor in the dayside region.

Luhmann, J. G.↗

Resource Management for Distributed Parallel Systems

Multiprocessor systems should exist in the the larger context of distributed systems, allowing multiprocessor resources to be shared by those that need them. Unfortunately, typical multiprocessor resource management techniques do not scale to large networks. The Prospero Resource Manager (PRM) is a scalable resource allocation system that supports the allocation of processing resources in large networks and multiprocessor systems. To manage resources in such distributed parallel systems, PRM employs three types of managers: system managers, job managers, and node managers. There exist multiple independent instances of each type of manager, reducing bottlenecks. The complexity of each manager is further reduced because each is designed to utilize information at an appropriate level of abstraction.

Neuman, B. Clifford↗

Parallel triangularization of substructured finite element problems

Much of the computational effort of the finite element process involves the solution of a system of linear equations. The coefficient matrix of this system, known as the global stiffness matrix, is symmetric, positive definite, and generally sparse. An important technique for reducing the time required to solve this system is substructuring or matrix partitioning. Substructuring is based on the idea of dividing a structure into pieces, each of which can then be analyzed relatively indepenently. As a result of this division, each point in the finite element discretization is either interior to a substructure or on a boundary between substructures. Contributions to the global stiffness matrix from connections between boundary points from the K(bb) matrix are reported. The triangularization of a general K(bb) matrix on a parallel machine is specifically discussed.

Leuze, M. R.↗

Performance and Application of Parallel OVERFLOW Codes on Distributed and Shared Memory Platforms

The presentation discusses recent studies on the performance of the two parallel versions of the aerodynamics CFD code, OVERFLOW_MPI and _MLP. Developed at NASA Ames, the serial version, OVERFLOW, is a multidimensional Navier-Stokes flow solver based on overset (Chimera) grid technology. The code has recently been parallelized in two ways. One is based on the explicit message-passing interface (MPI) across processors and uses the _MPI communication package. This approach is primarily suited for distributed memory systems and workstation clusters. The second, termed the multi-level parallel (MLP) method, is simple and uses shared memory for all communications. The _MLP code is suitable on distributed-shared memory systems. For both methods, the message passing takes place across the processors or processes at the advancement of each time step. This procedure is, in effect, the Chimera boundary conditions update, which is done in an explicit "Jacobi" style. In contrast, the update in the serial code is done in more of the "Gauss-Sidel" fashion. The programming efforts for the _MPI code is more complicated than for the _MLP code; the former requires modification of the outer and some inner shells of the serial code, whereas the latter focuses only on the outer shell of the code. The _MPI version offers a great deal of flexibility in distributing grid zones across a specified number of processors in order to achieve load balancing. The approach is capable of partitioning zones across multiple processors or sending each zone and/or cluster of several zones into a single processor. The message passing across the processors consists of Chimera boundary and/or an overlap of "halo" boundary points for each partitioned zone. The MLP version is a new coarse-grain parallel concept at the zonal and intra-zonal levels. A grouping strategy is used to distribute zones into several groups forming sub-processes which will run in parallel. The total volume of grid points in each group are approximately balanced. A proper number of threads are initially allocated to each group, and in subsequent iterations during the run-time, the number of threads are adjusted to achieve load balancing across the processes. Each process exploits the multitasking directives already established in Overflow.

Djomehri, M. Jahed↗

Applications of the massively parallel machine, the MasPar MP-1, to Earth sciences

The computational workload of upcoming NASA science missions, especially the ground data processing for the Earth Observing System, is projected to be quite large (in the 50 to 100 gigaFLOPS range) and corespondingly very expensive to perform using conventional supercomputer systems. High performance, general purpose massively parallel computer systems such as the MasPar MP-1 are being investigated by NASA as a more cost effective alternative. Massively parallel systems are targeted for accelerated development and maturation by NASA's upcoming five-year High Performance Computing and Communications Program. A summary of the broad range of applications currently running on the MP-1 at NASA/Goddard are presented in this paper along with descriptions of the parallel algorithmic techniques employed in five applications that have bearing on Earth sciences.

Fischer, James R.↗

X-Ray Imagery as the Record of All Data of Interest in Hypervelocity Impact Fragment Studies

Laboratory study of hypervelocity spacecraft fragmentation has traditionally involved the collection and analysis of fragments that were caught in deceleration material surrounding the impact. This process has typically involved the disintegration of the catchment material either through chemical dissolution, or through physical excavation to recover the fragments. Due to the scale of the three impact tests—the Satellite Orbital Debris Characterization Impact Test (SOCIT), the DebriSat satellite impact test, and the DebrisLV launch vehicle impact test—the latter two using more than 12 cubic meters of polyurethane foam to capture the fragments, hese projects have used x-ray imagery to precisely locate and thus, to more efficiently extract fragments in the soft-catch material. Three years into the DebriSat fragment extraction process, a side study was initiated to explore what additional information could be discerned from the x-rays, with significant results. This study was instrumental to a rapid replacement and retooling as the project was forced to replace the x-ray system around which the extraction process had been based. The revised process continues to map the debris for extraction. The project has, in parallel, systematically addressed the limits/tolerances of what x-rays can reveal about size, shape, density, mass, velocity, energy, and deformation/damage of the fragment during the deceleration in the catchment material while replicating the original extraction mapping function. All of these features have been optimized or have sufficient understanding to characterize the basic factors that will define a complete data set extracted solely from x-ray imagery. It is an ideal time to develop such a process, with extracted fragments providing “ground truth” against image-only data, and abundant available imagery of the same fragments under both the prior and replacement x-ray technologies, which have several fundamentally different characteristics. This paper addresses the types and quality of hypervelocity fragmentation data that can be and has been extracted from x-rays. It further addresses the question of whether and under what circumstances future hypervelocity experiments can use x-ray methods to largely—or to completely—avoid the extraction process in recording all appropriate results. Lastly, this paper addresses lessons learned and how future efforts can be further optimized.

John B. Bacon↗

Direct kinematics solution architectures for industrial robot manipulators: Bit-serial versus parallel

A Very Large Scale Integration (VLSI) architecture for robot direct kinematic computation suitable for industrial robot manipulators was investigated. The Denavit-Hartenberg transformations are reviewed to exploit a proper processing element, namely an augmented CORDIC. Specifically, two distinct implementations are elaborated on, such as the bit-serial and parallel. Performance of each scheme is analyzed with respect to the time to compute one location of the end-effector of a 6-links manipulator, and the number of transistors required.

Lee, J.↗

Technology and future ground processing systems

Land-observing satellites with multiple thematic mappers will produce data at rates of 100 to 300 Mbps. When coupled with a high daily scene production rate, these rates will require new approaches to ground processing. Consideration is given here to future downlink rates and data volumes, and requirements peculiar to the future user community are discussed. The advanced technologies required to attain an operational system in the years 1985-1990 are considered, together with advances foreseen in communications, mass storage, bulk memories, and data processing. Using advanced devices, a centralized data processing system capable of handling the 100 Mbps data rate is described. New approaches, among them a parallel pipelined calibration front-end, real-time browse image production, a high bandwidth optical disk archive, regional image broadcast and massively parallel product production, are considered. A distributed system capable of handling the 300 Mbps data rate is then described. Designs for a hub system and a regional processing center are presented.

Wood, B. J.↗

NiAl alloys for structural uses

Alloys based on the intermetallic compound NiAl are of technological interest as high temperature structural alloys. These alloys possess a relatively low density, high melting temperature, good thermal conductivity, and (usually) good oxidation resistance. However, NiAl and NiAl-base alloys suffer from poor fracture resistance at low temperatures as well as inadequate creep strength at elevated temperatures. This research program explored macroalloying additions to NiAl-base alloys in order to identify possible alloying and processing routes which promote both low temperature fracture toughness and high temperature strength. Initial results from the study examined the additions of Fe, Co, and Hf on the microstructure, deformation, and fracture resistance of NiAl-based alloys. Of significance were the observations that the presence of the gamma-prime phase, based on Ni3Al, could enhance the fracture resistance if the gamma-prime were present as a continuous grain boundary film or 'necklace'; and the Ni-35Al-20Fe alloy was ductile in ribbon form despite a microstructure consisting solely of the B2 beta phase based on NiAl. The ductility inherent in the Ni-35Al-20Fe alloy was explored further in subsequent studies. Those results confirm the presence of ductility in the Ni-35Al-20Fe alloy after rapid cooling from 750 - 1000 C. However exposure at 550 C caused embrittlement; this was associated with an age-hardening reaction caused by the formation of Fe-rich precipitates. In contrast, to the Ni-35Al-20Fe alloy, exploratory research indicated that compositions in the range of Ni-35Al-12Fe retain the ordered B2 structure of NiAl, are ductile, and do not age-harden or embrittle after thermal exposure. Thus, our recent efforts have focused on the behavior of the Ni-35Al-12Fe alloy. A second parallel effort initiated in this program was to use an alternate processing technique, mechanical alloying, to improve the properties of NiAl-alloys. Mechanical alloying in the conventional sense requires ductile powder particles which, through a cold welding and fracture process, can be dispersion strengthened by submicron-sized oxide particles. Using both the Ni-35Al-Fe alloys to contain approx. 1 v/o Y2O3. Preliminary results indicate that mechanically alloyed and extruded NiAl-Fe + Y2O3 alloys when heat treated to a grain-coarsened condition, exhibit improved creep resistance at 1000 C when compared to NiAl; oxidation resistance comparable to NiAl; and fracture toughness values a factor of three better than NiAl. As a result of the research initiated on this NASA program, a subsequent project with support from Inco Alloys International is underway.

Koss, D. A.↗

Algorithms and programming tools for image processing on the MPP

Topics addressed include: data mapping and rotational algorithms for the Massively Parallel Processor (MPP); Parallel Pascal language; documentation for the Parallel Pascal Development system; and a description of the Parallel Pascal language used on the MPP.

Reeves, A. P.↗