Search NASA⌕ Search

SEARCH · Search NASA

Results for “parallel programming”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,009 records · Page 56

Fast, Real-Time, Animated Displays

Displays for Advanced Concepts Simulator (ACS) generated on Adage Raster Display System 3000 (RDS 3000). Improved programming techniques developed, and revisions to language implementation made. Both types of changes took better advantage of high-speed characteristics of RDS 3000 hardware. Increases in speed resulted from: utilization of parallel-processing capabilities of AGG4, and use of AGG4 to take advantage of certain high-speed characteristics of display memory not previously used. Result was fourfold increase in animation-update rate to 16 frames per second.

Kahlbaum, William M.↗

Inhomogeneities of stratocumulus liquid water

There is a growing body of observational evidence on inhomogeneous cloud structure, most recently from the extensive measurements of the FIRE field program. Knowledge of cloud structure is important because it strongly influences the cloud radiative properties, one of the major factors in determining the global energy balance. Current atmospheric circulation models use plane-parallel radiation, so that the liquid water in each gridbox is assumed to be uniform, which gives an unrealistically large albedo. In reality cloud liquid water occupies only a subset of each gridbox, greatly reducing the mean albedo. If future climate models are to treat the hydrological cycle in a manner consistent with energy balance, a better treatment of cloud liquid is needed. FIRE concentrated upon two cloud types of special interest: cirrus and marine stratocumulus. Cirrus tend to be high and optically thin, thus reducing the effective radiative temperature without increasing the albedo significantly, leading to an enhanced greenhouse heating. In contrast, marine stratocumulus are low and optically thick, thus producing a large increase in reflected radiation with a small change in emitted radiation, giving a net cooling which could potentially mitigate the expected greenhouse warming. The FIRE measurements in California stratocumulus during June and July of 1987 show variations in cloud liquid water on all scales. Such variations are associated with inhomogeneous entrainment, in which entrained dry air, rather than mixing uniformly with cloudy air, remains intact in blobs of all sizes, which decay only slowly by invasion of cloudy air. Two important stratocumulus observations are described, followed by a simple fractal model which reproduces these properties, and finally, the model radiative properties are discussed.

Cahalan, Robert F.↗

Discrete approximations to optimal trajectories using direct transcription and nonlinear programming

A recently developed method for solving optimal trajectory problems uses a piecewise-polynomial representation of the state and control variables, enforces the equations of motion via a collocation procedure, and thus approximates the original calculus-of-variations problem with a nonlinear-programming problem, which is solved numerically. This paper identifies this method as a direct transcription method and proceeds to investigate the relationship between the original optimal-control problem and the nonlinear-programming problem. The discretized adjoint equation of the collocation method is found to have deficient accuracy, and an alternate scheme which discretizes the equations of motion using an explicit Runge-Kutta parallel-shooting approach is developed. Both methods are applied to finite-thrust spacecraft trajectory problems, including a low-thrust escape spiral, a three-burn rendezvous, and a low-thrust transfer to the moon.

Enright, Paul J.↗

Visualization and Tracking of Parallel CFD Simulations

We describe a system for interactive visualization and tracking of a 3-D unsteady computational fluid dynamics (CFD) simulation on a parallel computer. CM/AVS, a distributed, parallel implementation of a visualization environment (AVS) runs on the CM-5 parallel supercomputer. A CFD solver is run as a CM/AVS module on the CM-5. Data communication between the solver, other parallel visualization modules, and a graphics workstation, which is running AVS, are handled by CM/AVS. Partitioning of the visualization task, between CM-5 and the workstation, can be done interactively in the visual programming environment provided by AVS. Flow solver parameters can also be altered by programmable interactive widgets. This system partially removes the requirement of storing large solution files at frequent time steps, a characteristic of the traditional 'simulate (yields) store (yields) visualize' post-processing approach.

Vaziri, Arsi↗

Integrated System Test of an Airbreathing Rocket (ISTAR)

Rocket Based Combined Cycle (RBCC) propulsion system development and ground test is being conducted as part of the NASA Marshall Space Flight Center Integrated System Test of an Airbreathing Rocket (ISTAR) program. Rocketdyne, Aerojet and Pratt & Whitney have teamed as the Rocket Based Combined Cycle Consortium (RBC3) to work the propulsion system development. Each company offered unique RBCC propulsion concepts as candidates for the ISTAR propulsion system. A team of engine contractor, vehicle contractor and NASA representatives reviewed the concepts proposed by each company, reviewed the available data and selected the Aerojet RBCC propulsion system concept as the team propulsion system baseline for the ISTAR program. The ISTAR program is currently in a "Jumpstart" phase for development of the engine system leading to ground test of a thermally and power balanced RBCC propulsion system at Stennis Space Center in 2005. A parallel flight test demonstration of this propulsion system is anticipated to lead to first flight in the 2007 timeframe.

Faulkner, Robert F.↗

Multidisciplinary Optimization Methods for Aircraft Preliminary Design

This paper describes a research program aimed at improved methods for multidisciplinary design and optimization of large-scale aeronautical systems. The research involves new approaches to system decomposition, interdisciplinary communication, and methods of exploiting coarse-grained parallelism for analysis and optimization. A new architecture, that involves a tight coupling between optimization and analysis, is intended to improve efficiency while simplifying the structure of multidisciplinary, computation-intensive design problems involving many analysis disciplines and perhaps hundreds of design variables. Work in two areas is described here: system decomposition using compatibility constraints to simplify the analysis structure and take advantage of coarse-grained parallelism; and collaborative optimization, a decomposition of the optimization process to permit parallel design and to simplify interdisciplinary communication requirements.

Kroo, Ilan↗

Highly parallel structured adaptive mesh refinement using parallel language-based approaches

Adaptive mesh refinement (AMR) calculations carried out on structured meshes play an exceedingly important role in several areas of science and engineering. A strategy for using Fortran 90 in an object-oriented fashion is presented. This permits AMR applications to be expressed in terms of familiar abstractions that are natural to the process of solving AMR hierarchies. The OpenMP features that are useful for parallel processing of AMR hierarchies in a load balanced fashion on multiprocessors is described.

computational↗

A Parallel Genetic Algorithm for Automated Electronic Circuit Design

Parallelized versions of genetic algorithms (GAs) are popular primarily for three reasons: the GA is an inherently parallel algorithm, typical GA applications are very compute intensive, and powerful computing platforms, especially Beowulf-style computing clusters, are becoming more affordable and easier to implement. In addition, the low communication bandwidth required allows the use of inexpensive networking hardware such as standard office ethernet. In this paper we describe a parallel GA and its use in automated high-level circuit design. Genetic algorithms are a type of trial-and-error search technique that are guided by principles of Darwinian evolution. Just as the genetic material of two living organisms can intermix to produce offspring that are better adapted to their environment, GAs expose genetic material, frequently strings of 1s and Os, to the forces of artificial evolution: selection, mutation, recombination, etc. GAs start with a pool of randomly-generated candidate solutions which are then tested and scored with respect to their utility. Solutions are then bred by probabilistically selecting high quality parents and recombining their genetic representations to produce offspring solutions. Offspring are typically subjected to a small amount of random mutation. After a pool of offspring is produced, this process iterates until a satisfactory solution is found or an iteration limit is reached. Genetic algorithms have been applied to a wide variety of problems in many fields, including chemistry, biology, and many engineering disciplines. There are many styles of parallelism used in implementing parallel GAs. One such method is called the master-slave or processor farm approach. In this technique, slave nodes are used solely to compute fitness evaluations (the most time consuming part). The master processor collects fitness scores from the nodes and performs the genetic operators (selection, reproduction, variation, etc.). Because of dependency issues in the GA, it is possible to have idle processors. However, as long as the load at each processing node is similar, the processors are kept busy nearly all of the time. In applying GAs to circuit design, a suitable genetic representation 'is that of a circuit-construction program. We discuss one such circuit-construction programming language and show how evolution can generate useful analog circuit designs. This language has the desirable property that virtually all sets of combinations of primitives result in valid circuit graphs. Our system allows circuit size (number of devices), circuit topology, and device values to be evolved. Using a parallel genetic algorithm and circuit simulation software, we present experimental results as applied to three analog filter and two amplifier design tasks. For example, a figure shows an 85 dB amplifier design evolved by our system, and another figure shows the performance of that circuit (gain and frequency response). In all tasks, our system is able to generate circuits that achieve the target specifications.

Long, Jason D.↗

Diagnostics of wear in aeronautical systems

The use of appropriate diagnostic tools for aircraft oil wetted components is reviewed, noting that it can reduce direct operating costs through reduced unscheduled maintenance, particularly in helicopter engine and transmission systems where bearing failures are a significant cost factor. Engine and transmission wear modes are described, and diagnostic methods for oil and wet particle analysis, the spectrometric oil analysis program, chip detectors, ferrography, in-line oil monitor and radioactive isotope tagging are discussed, noting that they are effective over a limited range of particle sizes but compliment each other if used in parallel. Fine filtration can potentially increase time between overhauls, but reduces the effectiveness of conventional oil monitoring techniques so that alternative diagnostic techniques must be used. It is concluded that the development of a diagnostic system should be parallel and integral with the development of a mechanical system.

Wedeven, L. D.↗

Flow rate/pressure drop data gathered from testing a sample of the Space Shuttle Strain Isolation Pad (SIP): Effects of ambient pressure combined with tension and compression conditions

Tests were conducted on a sample of strain isolation pad (SIP) typical of that used in the shuttle orbiter thermal protection system to determine the characteristics of SIP internal flow. Data obtained were pressure drop as a function of flow rate for a range of ambient pressures representing various points along the Shuttle trajectory and for stretched and compressed conditions of the SIP. Flow was in the direction of the weave parallel to most of the fibers. The data are plotted in several standard engineering formats in order to be of maximum utility to the user. In addition to providing support to the Space Shuttle Program, these data are a source of experimental information on flow through fiberous (rather than the more usual sand bed type) porous media.

Springfield, R. D.↗

Determination of hot-spot susceptibility of multistring photovoltaic modules in a central-station application

Part of the effort of the Jet Propulsion Laboratory (JPL) Flat-Plate Solar Array Project (FSA) includes a program to improve module and array reliability. A collaborative activity with industry dealing with the problem of hot-spot heating due to the shadowing of photovoltaic cells in modules and arrays containing several paralleled cell strings is described. The use of multiparallel strings in large central-station arrays introduces the likelihood of unequal current sharing and increased heating levels. Test results that relate power dissipated, current imbalance, cross-strapping frequency, and shadow configuration to hot-spot heating levels are presented. Recommendations for circuit design configurations appropriate to central-station applications that reduce the risk of hot-spot problems are offered. Guidelines are provided for developing hot-spot tests for arrays when current imbalance is a threat.

Gonzalez, C. C.↗

A class Hierarchical, object-oriented approach to virtual memory management

The Choices family of operating systems exploits class hierarchies and object-oriented programming to facilitate the construction of customized operating systems for shared memory and networked multiprocessors. The software is being used in the Tapestry laboratory to study the performance of algorithms, mechanisms, and policies for parallel systems. Described here are the architectural design and class hierarchy of the Choices virtual memory management system. The software and hardware mechanisms and policies of a virtual memory system implement a memory hierarchy that exploits the trade-off between response times and storage capacities. In Choices, the notion of a memory hierarchy is captured by abstract classes. Concrete subclasses of those abstractions implement a virtual address space, segmentation, paging, physical memory management, secondary storage, and remote (that is, networked) storage. Captured in the notion of a memory hierarchy are classes that represent memory objects. These classes provide a storage mechanism that contains encapsulated data and have methods to read or write the memory object. Each of these classes provides specializations to represent the memory hierarchy.

Russo, Vincent F.↗

Massively parallel computing for the simulation of unsteady flows in turbomachinery

This paper deals with evaluating the capabilities of the massively parallel Connection Machine CM2 in predicting unsteady flows in turbomachines. The implementation on the CM2 of an implicit, time-accurate, zonal algorithm for the Navier-Stokes equations in two dimensions is described. Programming issues and modifications made to the original sequential algorithm to improve performance on the CM2 are briefly discussed. Performance is compared to a functionally equivalent code for the Cray YMP.

Madavan, Nateri K.↗

Cycle life status of SAFT VOS nickel-cadmium cells

The SAFT prismatic VOS Ni-Cd cells have been flown in geosynchronous orbit since 1977 and in low earth orbit since 1983. Parallel cycling tests are performed by several space agencies in order to determine the cycle life for a wide range of temperature and depth of discharge (DOD). In low Earth orbit (LEO), the ELAN program is conducted on 24 Ah cells by CNES and ESA at the European Battery Test Center at temperatures ranging from 0 to 27 C and DOD from 10 to 40 percent. Data are presented up to 37,000 cycles. One pack (X-80) has achieved 49,000 cycles at 10 C and 23 percent DOD. The geosynchronous orbit simulation of a high DOD test is conducted by ESA on 3 batteries at 10 C and 70, 90, and 100 percent DOD. Thirty-one eclipse seasons are completed, and no signs of degradation have been found. The Air Force test at CRANE on 24 Ah and 40 Ah cells at 20 C and 80 percent DOD has achieved 19 shadow periods. Life expectancy is discussed. The VOS cell technology could be used for the following: (1) in geosynchronous conditions--15 yrs at 10-15 C and 80 percent DOD; and (2) in low earth orbit--10 yrs at 5-15 C and 25-30 percent DOD.

Goualard, Jacques↗

Multidisciplinary propulsion simulation using NPSS

The current status of the Numerical Propulsion System Simulation (NPSS) program, a cooperative effort of NASA, industry, and universities to reduce the cost and time of advanced technology propulsion system development, is reviewed. The technologies required for this program include (1) interdisciplinary analysis to couple the relevant disciplines, such as aerodynamics, structures, heat transfer, combustion, acoustics, controls, and materials; (2) integrated systems analysis; (3) a high-performance computing platform, including massively parallel processing; and (4) a simulation environment providing a user-friendly interface. Several research efforts to develop these technologies are discussed.

Claus, Russell W.↗

Working toward a three-dimensional fatigue closure model for surface cracks

The first reliable elastic fracture mechanics solutions for a surface crack in a plate were obtained by Newman and Raju. The authors, both from the Mechanics of Materials Branch at NASA-Langley, used a highly detailed finite element solution requiring substantial computational resources. Computers have since become more powerful and available; however, many important related problems remain computationally expensive. The problem of three-dimensional fatigue crack growth taking into account plasticity-induced crack closure is one such problem. It is the goal of this research to provide an efficient method to account for three-dimensional crack closure in fatigue. Newman developed a two-dimensional plasticity-induced crack closure model for center cracked specimens. This model requires iterations to determine both the contact solution at each growth step and the extent of the plastic zone at the crack tip. A three-dimensional version of this model would require obtaining these nonlinear variables all along the crack front. This model must be efficient so that repeated calculations can be performed during crack growth simulations. The highly versatile line spring model (LSM) with contact, fatigue, and plasticity will form the basis of the closure model. There are several required additions to past work to address the three-dimensional crack closure problem. Initially, these additions will include (1) an improved LSM to more accurately obtain the crack opening displacement, stress intensity factors, and elastic T-stress near the ends of the surface crack; (2) a method to determine the extent of the plastic zone all along the crack front; (3) a method to determine the contact zone given a perfectly plastic layer of material on the crack surfaces; (4) a method to determine the magnitude of the compressive contact stress; and (5) a way to implement the degree of constraint along the curved crack front. During the summer ASEE program an enhanced LSM was developed. A method similar to that of 'strip synthesis' first introduced by Fujimoto was used. Briefly, the crack opening displacements of 'slices' of the surface crack in a direction parallel to the plate surface are considered in addition to the standard LSM approach that makes use of springs obtained from slices perpendicular to the plate surface. This enhancement is necessary so that an accurate three-dimensional representation of quantities such as contact zone size, plastic zone size, stress intensity factors, T-stress, and crack opening displacement can be determined. By combining results of previous investigations with the LSM, the problem of three-dimensional crack closure will be addressed. In addition to a closure model, the enhanced LSM can be used for many other problems including interacting surface cracks and fatigue crack growth of a through crack with a curved crack front.

Joseph, Paul F.↗

Automatic Data Distribution for CFD Applications on Structured Grids

Data distribution is an important step in implementation of any parallel algorithm. The data distribution determines data traffic, utilization of the interconnection network and affects the overall code efficiency. In recent years a number data distribution methods have been developed and used in real programs for improving data traffic. We use some of the methods for translating data dependence and affinity relations into data distribution directives. We describe an automatic data alignment and placement tool (ADAPT) which implements these methods and show it results for some CFD codes (NPB and ARC3D). Algorithms for program analysis and derivation of data distribution implemented in ADAPT are efficient three pass algorithms. Most algorithms have linear complexity with the exception of some graph algorithms having complexity O(n(sup 4)) in the worst case.

Frumkin, Michael↗

Automatic Data Distribution for CFD Applications on Structured Grids

Data distribution is an important step in implementation of any parallel algorithm. The data distribution determines data traffic, utilization of the interconnection network and affects the overall code efficiency. In recent years a number data distribution methods have been developed and used in real programs for improving data traffic. We use some of the methods for translating data dependence and affinity relations into data distribution directives. We describe an automatic data alignment and placement tool (ADAFT) which implements these methods and show it results for some CFD codes (NPB and ARC3D). Algorithms for program analysis and derivation of data distribution implemented in ADAFT are efficient three pass algorithms. Most algorithms have linear complexity with the exception of some graph algorithms having complexity O(n(sup 4)) in the worst case.

Frumkin, Michael↗