Search NASA⌕ Search

SEARCH · Search NASA

Results for “parallel processing (computers)”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 811 records · Page 45

Cascade model of gamma-ray bursts

If, in a neutron star magnetosphere, an electron is accelerated to an energy of 10 to the 11th or 12th power eV by an electric field parallel to the magnetic field, motion of the electron along the curved field line leads to a cascade of gamma rays and electron-positron pairs. This process is believed to occur in radio pulsars and gamma ray burst sources. Results are presented from numerical simulations of the radiation and photon annihilation pair production processes, using a computer code previously developed for the study of radio pulsars. A range of values of initial energy of a primary electron was considered along with initial injection position, and magnetic dipole moment of the neutron star. The resulting spectra was found to exhibit complex forms that are typically power law over a substantial range of photon energy, and typically include a dip in the spectrum near the electron gyro-frequency at the injection point. The results of a number of models are compared with data for the 5 Mar., 1979 gamma ray burst. A good fit was found to the gamma ray part of the spectrum, including the equivalent width of the annihilation line.

Sturrock, P. A.↗

Solving unstructured grid problems on massively parallel computers

A highly parallel graph mapping technique that enables one to efficiently solve unstructured grid problems on massively parallel computers is presented. Many implicit and explicit methods for solving discretized partial differential equations require each point in the discretization to exchange data with its neighboring points every time step or iteration. The cost of this communication can negate the high performance promised by massively parallel computing. To eliminate this bottleneck, the graph of the irregular problem is mapped into the graph representing the interconnection topology of the computer such that the sum of the distances that the messages travel is minimized. It is shown that using the heuristic mapping algorithm significantly reduces the communication time compared to a naive assignment of processes to processors.

Hammond, Steven W.↗

Surrogates for numerical simulations; optimization of eddy-promoter heat exchangers

Although the advent of fast and inexpensive parallel computers has rendered numerous previously intractable calculations feasible, many numerical simulations remain too resource-intensive to be directly inserted in engineering optimization efforts. An attractive alternative to direct insertion considers models for computational systems: the expensive simulation is evoked only to construct and validate a simplified, input-output model; this simplified input-output model then serves as a simulation surrogate in subsequent engineering optimization studies. A simple 'Bayesian-validated' statistical framework for the construction, validation, and purposive application of static computer simulation surrogates is presented. As an example, dissipation-transport optimization of laminar-flow eddy-promoter heat exchangers are considered: parallel spectral element Navier-Stokes calculations serve to construct and validate surrogates for the flowrate and Nusselt number; these surrogates then represent the originating Navier-Stokes equations in the ensuing design process.

Patera, Anthony T.↗

NASA Approach to HPCCP Support Software and Tools

The NASA HPCC Program, together with other agencies participating in the Federal HPCC Program, intends to advance technologies to enable the execution of grand challenge applications at sustained rates up to TeraFLOPS. During 1995-6 NASA undertook two major systems software efforts to improve the state of high performance support software and tools. The first of these activities was a replanning of support software and tools activities internal to the Agency. In replanning the software activities emphasis was placed on Meeting the needs of Grand Challenge Uses Few projects Near term useful results. The revised NASA plan calls for support software and tools activities in four areas: Application Creation Process Support Application Usage/Operations Support Advanced Support Software and Tools Concepts Metrics Based Monitoring and Management The second major activity undertaken was participation in a multiagency Task Force resulting from the Second Pasadena Workshop on System Software and Tools. The task force developed the Guidelines for Writing System Software and Tools Requirements for Parallel and Clustered Computers.

Blaylock, Bruce↗

Automated Generation of Message-Passing Programs: An Evaluation of CAPTools using NAS Benchmarks

Scientists at NASA Ames Research Center have been developing computational aeroscience applications on highly parallel architectures over the past ten years. During the same time period, a steady transition of hardware and system software also occurred, forcing us to expand great efforts into migrating and receding our applications. As applications and machine architectures continue to become increasingly complex, the cost and time required for this process will become prohibitive. Various attempts to exploit software tools to assist and automate the parallelization process have not produced favorable results. In this paper, we evaluate an interactive parallelization tool, CAPTools, for parallelizing serial versions of the NAB Parallel Benchmarks. Finally, we compare the performance of the resulting CAPTools generated code to the hand-coded benchmarks on the Origin 2000 and IBM SP2. Based on these results, a discussion on the feasibility of automated parallelization of aerospace applications is presented along with suggestions for future work.

Hribar, Michelle R.↗

Computational structural mechanics and fluid dynamics: Advances and trends; Proceedings of the Symposium, Washington, DC, Oct. 17-19, 1988

Recent advances in computational structural and fluid dynamics are discussed in reviews and reports. Topics addressed include fluid-structure interaction and aeroelasticity, CFD techniques for reacting flows, micromechanics, stability and eigenproblems, probabilistic methods and chaotic dynamics, and perturbation and spectral methods. Consideration is given to finite-element, finite-volume, and boundary-element methods; adaptive methods; parallel processing machines and applications; and visualization, mesh generation, and AI interfaces.

Noor, Ahmed K.↗

DSN Beowulf Cluster-Based VLBI Correlator

The NASA Deep Space Network (DSN) requires a broadband VLBI (very long baseline interferometry) correlator to process data routinely taken as part of the VLBI source Catalogue Maintenance and Enhancement task (CAT M&E) and the Time and Earth Motion Precision Observations task (TEMPO). The data provided by these measurements are a crucial ingredient in the formation of precision deep-space navigation models. In addition, a VLBI correlator is needed to provide support for other VLBI related activities for both internal and external customers. The JPL VLBI Correlator (JVC) was designed, developed, and delivered to the DSN as a successor to the legacy Block II Correlator. The JVC is a full-capability VLBI correlator that uses software processes running on multiple computers to cross-correlate two-antenna broadband noise data. Components of this new system (see Figure 1) consist of Linux PCs integrated into a Beowulf Cluster, an existing Mark5 data storage system, a RAID array, an existing software correlator package (SoftC) originally developed for Delta DOR Navigation processing, and various custom- developed software processes and scripts. Parallel processing on the JVC is achieved by assigning slave nodes of the Beowulf cluster to process separate scans in parallel until all scans have been processed. Due to the single stream sequential playback of the Mark5 data, some ramp-up time is required before all nodes can have access to required scan data. Core functions of each processing step are accomplished using optimized C programs. The coordination and execution of these programs across the cluster is accomplished using Pearl scripts, PostgreSQL commands, and a handful of miscellaneous system utilities. Mark5 data modules are loaded on Mark5 Data systems playback units, one per station. Data processing is started when the operator scans the Mark5 systems and runs a script that reads various configuration files and then creates an experiment-dependent status database used to delegate parallel tasks between nodes and storage areas (see Figure 2). This script forks into three processes: extract, translate, and correlate. Each of these processes iterates on available scan data and updates the status database as the work for each scan is completed. The extract process coordinates and monitors the transfer of data from each of the Mark5s to the Beowulf RAID storage systems. The translate process monitors and executes the data conversion processes on available scan files, and writes the translated files to the slave nodes. The correlate process monitors the execution of SoftC correlation processes on the slave nodes for scans that have completed translation. A comparison of the JVC and the legacy Block II correlator outputs reveals they are well within a formal error, and that the data are comparable with respect to their use in flight navigation. The processing speed of the JVC is improved over the Block II correlator by a factor of 4, largely due to the elimination of the reel-to-reel tape drives used in the Block II correlator.

Rogstad, Stephen P.↗

A numerical study of the vertical transport of momentum in a tropical rainband

The vertical transport of horizontal momentum in a convective tropical rainband is studied using a two-dimensional cloud ensemble model. Twelve simulations are made under the same large-scale conditions. The vertical transports of v momentum (parallel to the rainband) are essentially the same in all of the simulations, even though the structure of the clouds is different in each of the runs. The magnitude of the v-momentum transport by clouds is fairly large. It takes only half of a day to smooth out the tropical low-level easterly jet parallel to the rainband if no other processes are operating. The vertical transports of u momentum (perpendicular to the rainband) are quite different in all of the simulations. This difference can be explained by the dissimilarities in the distributions of horizontal momentum associated with various cloud configurations. The simulated vertical transports of horizontal momentum are compared with those computed with the Schneider and Lindzen scheme. The results suggest that their scheme is basically correct and usable if some improvements are made.

Soong, S.-T.↗

Design and Implementation of a Parallel Multivariate Ensemble Kalman Filter for the Poseidon Ocean General Circulation Model

A multivariate ensemble Kalman filter (MvEnKF) implemented on a massively parallel computer architecture has been implemented for the Poseidon ocean circulation model and tested with a Pacific Basin model configuration. There are about two million prognostic state-vector variables. Parallelism for the data assimilation step is achieved by regionalization of the background-error covariances that are calculated from the phase-space distribution of the ensemble. Each processing element (PE) collects elements of a matrix measurement functional from nearby PEs. To avoid the introduction of spurious long-range covariances associated with finite ensemble sizes, the background-error covariances are given compact support by means of a Hadamard (element by element) product with a three-dimensional canonical correlation function. The methodology and the MvEnKF configuration are discussed. It is shown that the regionalization of the background covariances; has a negligible impact on the quality of the analyses. The parallel algorithm is very efficient for large numbers of observations but does not scale well beyond 100 PEs at the current model resolution. On a platform with distributed memory, memory rather than speed is the limiting factor.

Keppenne, Christian L.↗

PPM Receiver Implemented in Software

A computer program has been written as a tool for developing optical pulse-position- modulation (PPM) receivers in which photodetector outputs are fed to analog-to-digital converters (ADCs) and all subsequent signal processing is performed digitally. The program can be used, for example, to simulate an all-digital version of the PPM receiver described in Parallel Processing of Broad-Band PPM Signals (NPO-40711), which appears elsewhere in this issue of NASA Tech Briefs. The program can also be translated into a design for digital PPM receiver hardware. The most notable innovation embodied in the software and the underlying PPM-reception concept is a digital processing subsystem that performs synchronization of PPM time slots, even though the digital processing is, itself, asynchronous in the sense that no attempt is made to synchronize it with the incoming optical signal a priori and there is no feedback to analog signal processing subsystems or ADCs. Functions performed by the software receiver include time-slot synchronization, symbol synchronization, coding preprocessing, and diagnostic functions. The program is written in the MATLAB and Simulink software system. The software receiver is highly parameterized and, hence, programmable: for example, slot- and symbol-synchronization filters have programmable bandwidths.

Gray, Andrew↗

Comparison of DAC and MONACO DSMC Codes with Flat Plate Simulation

Various implementations of the direct simulation Monte Carlo (DSMC) method exist in academia, government and industry. By comparing implementations, deficiencies and merits of each can be discovered. This document reports comparisons between DSMC Analysis Code (DAC) and MONACO. DAC is NASA's standard DSMC production code and MONACO is a research DSMC code developed in academia. These codes have various differences; in particular, they employ distinct computational grid definitions. In this study, DAC and MONACO are compared by having each simulate a blunted flat plate wind tunnel test, using an identical volume mesh. Simulation expense and DSMC metrics are compared. In addition, flow results are compared with available laboratory data. Overall, this study revealed that both codes, excluding grid adaptation, performed similarly. For parallel processing, DAC was generally more efficient. As expected, code accuracy was mainly dependent on physical models employed.

Padilla, Jose F.↗

The force on the flex: Global parallelism and portability

A parallel programming methodology, called the force, supports the construction of programs to be executed in parallel by an unspecified, but potentially large, number of processes. The methodology was originally developed on a pipelined, shared memory multiprocessor, the Denelcor HEP, and embodies the primitive operations of the force in a set of macros which expand into multiprocessor Fortran code. A small set of primitives is sufficient to write large parallel programs, and the system has been used to produce 10,000 line programs in computational fluid dynamics. The level of complexity of the force primitives is intermediate. It is high enough to mask detailed architectural differences between multiprocessors but low enough to give the user control over performance. The system is being ported to a medium scale multiprocessor, the Flex/32, which is a 20 processor system with a mixture of shared and local memory. Memory organization and the type of processor synchronization supported by the hardware on the two machines lead to some differences in efficient implementations of the force primitives, but the user interface remains the same. An initial implementation was done by retargeting the macros to Flexible Computer Corporation's ConCurrent C language. Subsequently, the macros were caused to directly produce the system calls which form the basis for ConCurrent C. The implementation of the Fortran based system is in step with Flexible Computer Corporations's implementation of a Fortran system in the parallel environment.

Jordan, H. F.↗

Techniques for increasing the update rate of real-time dynamic computer graphic displays

This paper describes several techniques which may be used to increase the animation update rate of real-time computer raster graphic displays. The techniques were developed on the ADAGE RDS 3000 graphic system in support of the Advanced Concepts Simulator at the NASA Langley Research Center. The first technique involves pre-processing of the next animation frame while the previous one is being erased from the screen memory. The second technique involves the use of a parallel processor, the AGG4, for high speed character generation. The description of the AGG4 includes the Barrel Shifter which is a part of the hardware and is the key to the high speed character rendition. The final result of this total effort was a four fold increase in the update rate of an existing primary flight display from 4 to 16 frames per second.

Kahlbaum, W. M., Jr.↗

Utility of Emulation and Simulation Computer Modeling of Space Station Environmental Control and Life Support Systems

Over the years, computer modeling has been used extensively in many disciplines to solve engineering problems. A set of computer program tools is proposed to assist the engineer in the various phases of the Space Station program from technology selection through flight operations. The development and application of emulation and simulation transient performance modeling tools for life support systems are examined. The results of the development and the demonstration of the utility of three computer models are presented. The first model is a detailed computer model (emulation) of a solid amine water desorbed (SAWD) CO2 removal subsystem combined with much less detailed models (simulations) of a cabin, crew, and heat exchangers. This model was used in parallel with the hardware design and test of this CO2 removal subsystem. The second model is a simulation of an air revitalization system combined with a wastewater processing system to demonstrate the capabilities to study subsystem integration. The third model is that of a Space Station total air revitalization system. The station configuration consists of a habitat module, a lab module, two crews, and four connecting nodes.

Yanosy, James L.↗

The Meteorological Measurement System on the NASA ER-2 aircraft

A Meteorological Measurement System (MMS) was designed and installed on one of the NASA high-altitude ER-2 aircraft (NASA 706). The MMS provides in situ measurements of free-stream pressure (+ or - 0.3 mb), temperature (+ or - 0.3 C), and wind vector (+ or - 1 m/s). It incorporates a high-resolution inertial navigation system specially configured for scientific applications, a radome differential pressure system for measurements of the airflow angles, and a compact, computer-controlled data acquisition system to sample, process and store 45 variables on tape and on disk. The MMS hardware and software development is described, and resolution and accuracy of the instrumentation discussed. Custom software facilitates preflight system checkout, inflight data acquisition, and fast postflight data download. It accommodates various modes of MMS data: analog and digital, serial and parallel, and synchronous and asynchronous. Flight results are presented to demonstrate the capability of the system.

Scott, Stan G.↗

Oceanic Observations

For many years, merchant ships and the naval fleets of various countries have been the major source of data over and in the open ocean. Oceanographic research experiments and process studies in the field have also contributed to the climatological data bases for the global ocean, but, for the most part, these have been limited in duration and extent. However, over the last 10 years under the auspices of the World Climate Research Program and the International Geosphere Biosphere Program the role of the oceans in global and climate change has taken on increased significance. This has created a need for a considerably improved understanding of the seasonal, interannual, decadal and longer time-scale variability of the physical and biogeochemical attributes of the global ocean. As a result, over the past 10 years several major international field programs have been implemented and have had a tremendous impact on the number of in situ observations obtained for the global ocean. The Tropical Ocean Global Atmosphere (TOGA) program, the World Ocean Circulation Experiment (WOCE), and the Joint Global Ocean Flux Study (JGOFS) were designed with observational, modelling, and process study components aimed at analyzing different aspects of the ocean's role in the coupled climate system. In parallel with the field programs, continuous space-based observations of sea surface temperature, sea surface topography, and sea surface winds spanning nearly a decade or longer have become a reality. During this same time period, numerical ocean models and computational power have advanced to the point where the oceanographic observations, both in situ and remotely sensed, can be assimilated into numerical ocean models in order to provide a four-dimensional (x-y-z-t) depiction of the evolving state of the global ocean.

Busalacchi, Antonio J.↗

Flame-Generated Vorticity Production in Premixed Flame-Vortex Interactions

In this study, we use detailed time-dependent, multi-dimensional numerical simulations to investigate the relative importance of the processes leading to FGV in flame-vortex interactions in normal gravity and microgravity and to determine if the production of vorticity in flames in gravity is the same as that in zero gravity except for the contribution of the gravity term. The numerical simulations will be performed using the computational model developed at NRL, FLAME3D. FLAME3D is a parallel, multi-dimensional (either two- or three-dimensional) flame model based on FLIC2D, which has been used extensively to study the structure and stability of premixed hydrogen and methane flames.

Patnaik, G.↗

Computational Investigation of Oxidative Etch Pitting in FiberForm and Its Impact on Material Properties

Oxidation-driven carbon erosion does not occur uniformly but rather through the development of localized etch pits at active surface sites. These active sites form due to atomic defects on the carbon surface, making them significantly more reactive than the surrounding, non-defective areas. As a result, these sites are the first to react during ablation, leading to their removal. This process creates new defects in neighboring atoms, increasing their reactivity and causing localized carbon removal around these active sites. In this way, the highly reactive defective areas serve as nucleation points for the formation and growth of etch pits, which can have adverse effects on structural integrity of FiberForm. To better understand how these etch pits impact the material properties of carbon fiber microstructures, we have developed a new capability within the direct simulation Monte Carlo (DSMC) framework to capture the etch pit formation process. This capability, integrated into the DSMC code SPARTA (Stochastic Parallel Rarefied-gas Time-accurate Analyzer), models material removal in the presence of active sites, leading to the formation of etch pits. The current work focuses on studying the effects of these etch pits on the material properties of FiberForm, a widely used base material in thermal protection systems (TPS). The microstructure of virgin FiberForm, obtained via X-ray microtomography, is imported into SPARTA to generate the ablated geometries with etch pits. These modified microstructures are then analyzed using the Porous Microstructure Analysis (PuMA) software to compute various material properties, including elasticity, thermal conductivity, and permeability. We investigate the variation of these properties due to the complex surface topology changes caused by etch pit formation. Additionally, we compare the effects of pitting with the conventional model of shrinking fibers, traditionally used to simulate the ablation of carbon structures. Significant differences emerge between the two approaches. Consequently, this physically realistic model of material removal through etch pit formation offers improved accuracy in predicting the degradation of carbon-based TPS during oxidation. It also provides insights into other mechanisms, such as spallation, where chunks of material are removed into the flow due to etch pit growth. Ultimately, this model enhances our understanding of failure modes in these materials during ablation.

PuMA↗