Search NASA⌕ Search

SEARCH · Search NASA

Results for “parallel programming”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 883 records · Page 49

An Airborne Onboard Parallel Processing Testbed

This presentation provides information on the progress the Intelligent Payload Module (IPM) development effort. In addition, a vision is presented on integration of the IPM architecture with the GeoSocial Application Program Interface (API) architecture to enable efficient distribution of satellite data products.

Onboard processing↗

Energy efficient engine sector combustor rig test program

Under the NASA-sponsored Energy Efficient Engine program, Pratt & Whitney Aircraft has successfully completed a comprehensive combustor rig test using a 90-degree sector of an advanced two-stage combustor with a segmented liner. Initial testing utilized a combustor with a conventional louvered liner and demonstrated that the Energy Efficient Engine two-stage combustor configuration is a viable system for controlling exhaust emissions, with the capability to meet all aerothermal performance goals. Goals for both carbon monoxide and unburned hydrocarbons were surpassed and the goal for oxides of nitrogen was closely approached. In another series of tests, an advanced segmented liner configuration with a unique counter-parallel FINWALL cooling system was evaluated at engine sea level takeoff pressure and temperature levels. These tests verified the structural integrity of this liner design. Overall, the results from the program have provided a high level of confidence to proceed with the scheduled Combustor Component Rig Test Program.

Dubiel, D. J.↗

Detection of faults and software reliability analysis

Multi-version or N-version programming is proposed as a method of providing fault tolerance in software. The approach requires the separate, independent preparation of multiple versions of a piece of software for some application. These versions are executed in parallel in the application environment; each receives identical inputs and each produces its version of the required outputs. The outputs are collected by a voter and, in principle, they should all be the same. In practice there may be some disagreement. If this occurs, the results of the majority are taken to be the correct output, and that is the output used by the system. A total of 27 programs were produced. Each of these programs was then subjected to one million randomly-generated test cases. The experiment yielded a number of programs containing faults that are useful for general studies of software reliability as well as studies of N-version programming. Fault tolerance through data diversity and analytic models of comparison testing are discussed.

Knight, John C.↗

Reducing False Positives in Runtime Analysis of Deadlocks

This paper presents an improvement of a standard algorithm for detecting dead-lock potentials in multi-threaded programs, in that it reduces the number of false positives. The standard algorithm works as follows. The multi-threaded program under observation is executed, while lock and unlock events are observed. A graph of locks is built, with edges between locks symbolizing locking orders. Any cycle in the graph signifies a potential for a deadlock. The typical standard example is the group of dining philosophers sharing forks. The algorithm is interesting because it can catch deadlock potentials even though no deadlocks occur in the examined trace, and at the same time it scales very well in contrast t o more formal approaches to deadlock detection. The algorithm, however, can yield false positives (as well as false negatives). The extension of the algorithm described in this paper reduces the amount of false positives for three particular cases: when a gate lock protects a cycle, when a single thread introduces a cycle, and when the code segments in different threads that cause the cycle can actually not execute in parallel. The paper formalizes a theory for dynamic deadlock detection and compares it to model checking and static analysis techniques. It furthermore describes an implementation for analyzing Java programs and its application to two case studies: a planetary rover and a space craft altitude control system.

Bensalem, Saddek↗

Spur, helical, and spiral bevel transmission life modeling

A computer program, TLIFE, which estimates the life, dynamic capacity, and reliability of aircraft transmissions, is presented. The program enables comparisons of transmission service life at the design stage for optimization. A variety of transmissions may be analyzed including: spur, helical, and spiral bevel reductions as well as series combinations of these reductions. The basic spur and helical reductions include: single mesh, compound, and parallel path plus revert star and planetary gear trains. A variety of straddle and overhung bearing configurations on the gear shafts are possible as is the use of a ring gear for the output. The spiral bevel reductions include single and dual input drives with arbitrary shaft angles. The program is written in FORTRAN 77 and has been executed both in the personal computer DOS environment and on UNIX workstations. The analysis may be performed in either the SI metric or the English inch system of units. The reliability and life analysis is based on the two-parameter Weibull distribution lives of the component gears and bearings. The program output file describes the overall transmission and each constituent transmission, its components, and their locations, capacities, and loads. Primary output is the dynamic capacity and 90-percent reliability and mean lives of the unit transmissions and the overall system which can be used to estimate service overhaul frequency requirements. Two examples are presented to illustrate the information available for single element and series transmissions.

Savage, Michael↗

CARA Devolution ESMO Pilot Program

Information regarding the current effort ongoing with CARA to determine the steps needed to have a mission move away from being supported by CARA, otherwise referred to as Devolving/Devolution. Information regarding the working group established to undertake this effort, documents agreed that needed to be written and plans for a parallel operations period. The communication between ESMO and CARA during the Parallel Operations period will be as the following: – A daily email will be sent from ESMO to CARA describing each HIE that is being actively planned and worked, and the action that ESMO is planning on taking – A weekly email will be sent from ESMO to CARA listing the progress toward achieving the parallel operations success criteria – The CARA team will remain on distribution for ESMO HIE communications and CAM invitations associated with DAMs/RMMs during the parallel operations period – At a minimum, 2 members of the CARA team will be in attendance at all ESMO planning and review discussions – A monthly email will be sent from ESMO to CARA with the HIE metrics • During parallel operations, ESMO will contact NASA Headquarters to report any redDevolution Working Group laid out and completed all steps needed to begin parallel operations • Parallel Ops for ESMO began on 11/26 (CARA in shadow mode only) • Any lessons learned from parallel ops will be folded back into documents/templates • Hope to achieve all success criteria < 6 months • Decision for permanent devolution will be based on completion of parallel ops and other future documentation creation (CARA Standard and Handbook, etc) and reviews (ORR) event which is less than 2 days from TCA, including the proposed remediation, on an as-needed daily basis flow using SpaceTrack to send and receive data

ESMO pilot devolution program↗

Progress in developing ultrathin solar cell blanket technology

A program was conducted to develop technologies for welding interconnects to three types of 50-micron-thick, 2 by 2-cm solar cells. Parallel-gap resistance welding was used for interconnect attachment. Weld schedules were independently developed for each of the three cell types and were coincidentally identical. Six 48-cell modules were assembled with 50-micron (nominal) thick cells, frosted fused-silica covers, silver-plated Invar interconnectors, and four different substrate designs. Three modules (one for each cell type) have single-layer Kapton (50-micron-thick) substrates. The other three modules each have a different substrate (Kapton-Kevlar-Kapton, Kapton-graphite-Kapton, and Kapton-graphite-aluminum honeycomb-graphite). All six modules were subjected to 4112 thermal cycles from -175 to 65 C (corresponding to over 40 years of simulated geosynchronous orbit thermal cycling) and experienced only negligible electrical degradation (1.1 percent average of six 48-cell modules).

Patterson, R. E.↗

Optimization of V-groove radiator configuration

In the design of spacecraft radiators intended to provide cryogenic cooling, it is important to minimize the space occupied by radiator shielding while satisfying the performance requirements. This study develops and tests the first step toward optimizing radiator shield configurations and examines the design of the highly effective V-groove radiator concept. This investigation, which makes use of a special purpose Monte Carlo/ray tracing computer program, directly identifies the two shield configuration which minimizes radiator footprint for a given temperature drop. Furthermore, multiple shield configurations may be analyzed through a combination of analysis and program statistics. The results presented demonstrate the sensitivity of intershield temperature drop to shield angle and/or offset and provide a comparison of angled and parallel shield configurations. Good agreement with experimental data is shown for a multiple shield configuration.

Schember, Helene R.↗

Internal fluid mechanics research on supercomputers for aerospace propulsion systems

The Internal Fluid Mechanics Division of the NASA Lewis Research Center is combining the key elements of computational fluid dynamics, aerothermodynamic experiments, and advanced computational technology to bring internal computational fluid mechanics (ICFM) to a state of practical application for aerospace propulsion systems. The strategies used to achieve this goal are to: (1) pursue an understanding of flow physics, surface heat transfer, and combustion via analysis and fundamental experiments, (2) incorporate improved understanding of these phenomena into verified 3-D CFD codes, and (3) utilize state-of-the-art computational technology to enhance experimental and CFD research. Presented is an overview of the ICFM program in high-speed propulsion, including work in inlets, turbomachinery, and chemical reacting flows. Ongoing efforts to integrate new computer technologies, such as parallel computing and artificial intelligence, into high-speed aeropropulsion research are described.

Miller, Brent A.↗

Numerical solutions of the compressible 3-D boundary-layer equations for aerospace configurations with emphasis on LFC

The application of stability theory in Laminar Flow Control (LFC) research requires that density and velocity profiles be specified throughout the viscous flow field of interest. These profile values must be as numerically accurate as possible and free of any numerically induced oscillations. Guidelines for the present research project are presented: develop an efficient and accurate procedure for solving the 3-D boundary layer equation for aerospace configurations; develop an interface program to couple selected 3-D inviscid programs that span the subsonic to hypersonic Mach number range; and document and release software to the LFC community. The interface program was found to be a dependable approach for developing a user friendly procedure for generating the boundary-layer grid and transforming an inviscid solution from a relatively coarse grid to a sufficiently fine boundary-layer grid. The boundary-layer program was shown to be fourth-order accurate in the direction normal to the wall boundary and second-order accurate in planes parallel to the boundary. The fourth-order accuracy allows accurate calculations with as few as one-fifth the number of grid points required for conventional second-order schemes.

Harris, Julius E.↗

Model atmosphere analysis of selected luminous B stars

The general scientific goal of this program has been to determine whether the atmospheric structure of the B-type stars can be represented by the current generation of plane parallel, line-blanketed, LTE stellar atmosphere models sufficiently well to allow accurate effective temperatures and surface gravities to be deduced. The B stars cover a wide range of temperature and luminosity. For the hottest such stars (with T approximately 30,000 K) the applicability of the models may be compromised by departures from LTE in the stellar atmospheres ('non-LTE effects'). At the highest luminosities (the B 'super giants'), the models may be invalidated by departures from plane parallel geometry. Thus we seek to identify the temperature and luminosity range within which these effects are unimportant and where the models may be relied upon.

Fitzpatrick, Edward L.↗

A Comparative Propulsion System Analysis for the High-Speed Civil Transport

Six of the candidate propulsion systems for the High-Speed Civil Transport are the turbojet, turbine bypass engine, mixed flow turbofan, variable cycle engine, Flade engine, and the inverting flow valve engine. A comparison of these propulsion systems by NASA's Glenn Research Center, paralleling studies within the aircraft industry, is presented. This report describes the Glenn Aeropropulsion Analysis Office's contribution to the High-Speed Research Program's 1993 and 1994 propulsion system selections. A parametric investigation of each propulsion cycle's primary design variables is analytically performed. Performance, weight, and geometric data are calculated for each engine. The resulting engines are then evaluated on two airframer-derived supersonic commercial aircraft for a 5000 nautical mile, Mach 2.4 cruise design mission. The effects of takeoff noise, cruise emissions, and cycle design rules are examined.

Berton, Jeffrey J.↗

The UDF05 Follow-up of the HUDF: I. The Faint-End Slope of the Lyman-Break Galaxy Population at zeta approx. 5

We present the UDF05 project, a HST Large Program of deep ACS (F606W, F775W, F850LP, and NICMOS (Fll0W, Fl60W) imaging of three fields, two of which coincide with the NICP1-4 NICMOS parallel observations of the Hubble Ultra Deep Field (HUDF). In this first paper we use the ACS data for the NICP12 field, as well as the original HUDF ACS data, to measure the UV Luminosity Function (LF) of z approximately 5 Lyman Break Galaxies (LBGs) down to very faint levels. Specifically, based on a V - i, i - z selection criterion, we identify a sample of 101 and 133 candidate z approximately 5 galaxies down to z(sub 850) = 28.5 and 29.25 magnitudes in the NICP12 field and in the HUDF, respectively. Using an extensive set of Monte Carlo simulations we derive corrections for observational biases and selection effects, and construct the rest-frame 1400 Angstroms LBG LF over the range M(sub 1400) = [-22.2, -17.1], i.e. down to approximately 0.04 L(sub *) at z = 5. We show that: (i) Different assumptions for the SED distribution of the LBG population, dust properties and intergalactic absorption result in a 25% variation in the number density of LBGs at z = 5 (ii) Under consistent assumptions for dust properties and intergalactic absorption, the HUDF is about 30% under-dense in z = 5 LBGs relative to the NICP12 field, a variation which is well explained by cosmic variance; (iii) The faint-end slope of the LF is independent of the specific assumptions for the input physical parameters, and has a value of alpha approximately -1.6, similar to the faint-end slope of the LF that has been measured for LBGs at z = 3 and z = 6. Our study therefore supports no variation in the faint-end of the LBG LF over the whole redshift range z = 3 to z = 6. The comparison with theoretical predictions suggests that (a,) the majority of the stars in the z = 5 LBG population are produced with a Top-Heavy IMF in merger-driven starbursts, and that (b) possibly, either the fraction of stellar mass produced in starburst, or the fraction of high mass stars in the bursts is increased towards the bright end of the LF.

Oesch, P. A.↗

Partial Overhaul and Initial Parallel Optimization of KINETICS, a Coupled Dynamics and Chemistry Atmosphere Model

KINETICS is a coupled dynamics and chemistry atmosphere model that is data intensive and computationally demanding. The potential performance gain from using a supercomputer motivates the adaptation from a serial version to a parallelized one. Although the initial parallelization had been done, bottlenecks caused by an abundance of communication calls between processors led to an unfavorable drop in performance. Before starting on the parallel optimization process, a partial overhaul was required because a large emphasis was placed on streamlining the code for user convenience and revising the program to accommodate the new supercomputers at Caltech and JPL. After the first round of optimizations, the partial runtime was reduced by a factor of 23; however, performance gains are dependent on the size of the data, the number of processors requested, and the computer used.

KINETICS↗

Performance Enhancement of a Computational Persistent Homology Package

In recent years, persistent homology has become an attractive method for data analysis. It captures topological features, such as connected components, holes, voids, etc., from a point cloud by finding out when these features appear and disappear in the filtration sequence. In this project, we focus on improving the performance of Eirene, a fancy computational persistent homology package. Eirene is a 5000-line opensource software implemented by using the dynamic programming language Julia. We use the Julia profiling tools to identify the performance bottlenecks and develop different methods to manage the bottlenecks, including the parallelization of some time-consuming functions on the multicore/manycore hardware. The empirical results show that the performance can be greatly improved.

Profiling↗

Assessment of Cislunar Staging Orbits to Support the Artemis III Lunar Surface Mission

Since NASA’s selection of an L2 9:2 lunar synodic resonant Near Rectilinear Halo Orbit (NRHO) as the baseline for the Gateway Program, the agency has worked to mature its understanding of this orbit and its use for the Artemis III, IV, and V missions. In parallel with these efforts, NASA has investigated alternative staging orbits to perform the Artemis III lunar surface landing mission and compared those options to the baseline NRHO. This paper evaluates a number of alternative orbits on their feasibility and favorability and compares them to the agency baseline NRHO.

Artemis↗

Substructure analysis using NICE/SPAR and applications of force to linear and nonlinear structures

Parallel computing studies are presented for a variety of structural analysis problems. Included are the substructure planar analysis of rectangular panels with and without a hole, the static analysis of space mast, using NICE/SPAR and FORCE, and substructure analysis of plane rigid-jointed frames using FORCE. The computations are carried out on the Flex/32 MultiComputer using one to eighteen processors. The NICE/SPAR runstream samples are documented for the panel problem. For the substructure analysis of plane frames, a computer program is developed to demonstrate the effectiveness of a substructuring technique when FORCE is enforced. Ongoing research activities for an elasto-plastic stability analysis problem using FORCE, and stability analysis of the focus problem using NICE/SPAR are briefly summarized. Speedup curves for the panel, the mast, and the frame problems provide a basic understanding of the effectiveness of parallel computing procedures utilized or developed, within the domain of the parameters considered. Although the speedup curves obtained exhibit various levels of computational efficiency, they clearly demonstrate the excellent promise which parallel computing holds for the structural analysis problem. Source code is given for the elasto-plastic stability problem and the FORCE program.

Razzaq, Zia↗

Computer-Aided Parallelizer and Optimizer

The Computer-Aided Parallelizer and Optimizer (CAPO) automates the insertion of compiler directives (see figure) to facilitate parallel processing on Shared Memory Parallel (SMP) machines. While CAPO currently is integrated seamlessly into CAPTools (developed at the University of Greenwich, now marketed as ParaWise), CAPO was independently developed at Ames Research Center as one of the components for the Legacy Code Modernization (LCM) project. The current version takes serial FORTRAN programs, performs interprocedural data dependence analysis, and generates OpenMP directives. Due to the widely supported OpenMP standard, the generated OpenMP codes have the potential to run on a wide range of SMP machines. CAPO relies on accurate interprocedural data dependence information currently provided by CAPTools. Compiler directives are generated through identification of parallel loops in the outermost level, construction of parallel regions around parallel loops and optimization of parallel regions, and insertion of directives with automatic identification of private, reduction, induction, and shared variables. Attempts also have been made to identify potential pipeline parallelism (implemented with point-to-point synchronization). Although directives are generated automatically, user interaction with the tool is still important for producing good parallel codes. A comprehensive graphical user interface is included for users to interact with the parallelization process.

Jin, Haoqiang↗