Search NASA⌕ Search

SEARCH · Search NASA

Results for “Supercomputers”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 685 records · Page 38

A parallel finite-difference method for computational aerodynamics

A finite-difference scheme for solving complex three-dimensional aerodynamic flow on parallel-processing supercomputers is presented. The method consists of a basic flow solver with multigrid convergence acceleration, embedded grid refinements, and a zonal equation scheme. Multitasking and vectorization have been incorporated into the algorithm. Results obtained include multiprocessed flow simulations from the Cray X-MP and Cray-2. Speedups as high as 3.3 for the two-dimensional case and 3.5 for segments of the three-dimensional case have been achieved on the Cray-2. The entire solver attained a factor of 2.7 improvement over its unitasked version on the Cray-2. The performance of the parallel algorithm on each machine is analyzed.

Swisshelm, Julie M.↗

Reorientation of rotating fluid in microgravity environment with and without gravity jitters

In a spacecraft design, the requirements of settled propellant are different for tank pressurization, engine restart, venting, or propellant transfer. The requirement to settle or to position liquid fuel over the outlet end of the spacecraft propellant tank prior main engine restart poses a microgravity fluid behavior problem. In this paper, the dynamical behavior of liquid propellant, fluid reorientation, and propellant resettling have been carried out through the execution of supercomputer CRAY X-MP to simulate the fluid management in a microgravity environment. Results show that the resettlement of fluid can be accomplished more efficiently for fluid in rotating tank than in nonrotating tank, and also better performance for gravity jitters imposed on fluid settlement than without gravity jitters based on the amount of time needed to carry out resettlement period of time between the initiation and termination of geysering.

Hung, R. J.↗

Machine characterization based on an abstract high-level language machine

Measurements are presented for a large number of machines ranging from small workstations to supercomputers. The authors combine these measurements into groups of parameters which relate to specific aspects of the machine implementation, and use these groups to provide overall machine characterizations. The authors also define the concept of pershapes, which represent the level of performance of a machine for different types of computation. A metric based on pershapes is introduced that provides a quantitative way of measuring how similar two machines are in terms of their performance distributions. The metric is related to the extent to which pairs of machines have varying relative performance levels depending on which benchmark is used.

Saavedra-Barrera, Rafael H.↗

An interactive adaptive remeshing algorithm for the two-dimensional Euler equations

An interactive adaptive remeshing algorithm utilizing a frontal grid generator and a variety of time integration schemes for the two-dimensional Euler equations on unstructured meshes is presented. Several device dependent interactive graphics interfaces have been developed along with a device independent DI-3000 interface which can be employed on any computer that has the supporting software including the Cray-2 supercomputers Voyager and Navier. The time integration methods available include: an explicit four stage Runge-Kutta and a fully implicit LU decomposition. A cell-centered finite volume upwind scheme utilizing Roe's approximate Riemann solver is developed. To obtain higher order accurate results a monotone linear reconstruction procedure proposed by Barth is utilized. Results for flow over a transonic circular arc and flow through a supersonic nozzle are examined.

Slack, David C.↗

Characteristic-based algorithms for flows in thermo-chemical nonequilibrium

A generalized finite-rate chemistry algorithm with Steger-Warming, Van Leer, and Roe characteristic-based flux splittings is presented in three-dimensional generalized coordinates for the Navier-Stokes equations. Attention is placed on convergence to steady-state solutions with fully coupled chemistry. Time integration schemes including explicit m-stage Runge-Kutta, implicit approximate-factorization, relaxation and LU decomposition are investigated and compared in terms of residual reduction per unit of CPU time. Practical issues such as code vectorization and memory usage on modern supercomputers are discussed.

Walters, Robert W.↗

Chemical calculations on Cray computers

The influence of recent developments in supercomputing on computational chemistry is discussed with particular reference to Cray computers and their pipelined vector/limited parallel architectures. After reviewing Cray hardware and software the performance of different elementary program structures are examined, and effective methods for improving program performance are outlined. The computational strategies appropriate for obtaining optimum performance in applications to quantum chemistry and dynamics are discussed. Finally, some discussion is given of new developments and future hardware and software improvements.

Taylor, Peter R.↗

A vectorized Lanczos eigensolver for high-performance computers

The computational strategies used to implement a Lanczos-based-method eigensolver on the latest generation of supercomputers are described. Several examples of structural vibration and buckling problems are presented that show the effects of using optimization techniques to increase the vectorization of the computational steps. The data storage and access schemes and the tools and strategies that best exploit the computer resources are presented. The method is implemented on the Convex C220, the Cray 2, and the Cray Y-MP computers. Results show that very good computation rates are achieved for the most computationally intensive steps of the Lanczos algorithm and that the Lanczos algorithm is many times faster than other methods extensively used in the past.

Bostic, Susan W.↗

A parallel-vector algorithm for rapid structural analysis on high-performance computers

A fast, accurate Choleski method for the solution of symmetric systems of linear equations is presented. This direct method is based on a variable-band storage scheme and takes advantage of column heights to reduce the number of operations in the Choleski factorization. The method employs parallel computation in the outermost DO-loop and vector computation via the 'loop unrolling' technique in the innermost DO-loop. The method avoids computations with zeros outside the column heights, and as an option, zeros inside the band. The close relationship between Choleski and Gauss elimination methods is examined. The minor changes required to convert the Choleski code to a Gauss code to solve non-positive-definite symmetric systems of equations are identified. The results for two large-scale structural analyses performed on supercomputers, demonstrate the accuracy and speed of the method.

Storaasli, Olaf O.↗

Aeroelastic problems in turbomachines

A review of the field of turbomachinery aeroelasticity is presented. Developments over the past decade are emphasized, and an assessment of possible future directions of research is offered. The paper reviews the areas of unsteady cascade flows, structural modeling, and flutter prediction methods. Representative results for unsteady flow calculations and flutter boundary predictions in subsonic, transonic, and supersonic flows are discussed, including recent calculations based on the methods of computational fluid mechanics. Results from current attempts to correlate experimental data with theoretical predictions are discussed briefly. It is recommended that future research include investigations of novel approaches to flutter calculations that can take full advantage of parallel processing supercomputers. The feasibility of using mistuning and aeroelastic tailoring as passive flutter suppression techniques should also be pursued.

Bendiksen, Oddvar O.↗

Precise orbit computation for the Geosat Exact Repeat Mission

Results are reported from an extensive investigation of orbit-determination strategies for the Geosat Exact Repeat Mission (ERM). The goal is to establish optimum geodetic parameters and procedures for the computation of the most accurate Geosat orbits possible and to apply these procedures for routine computation during the ERM for the following purposes: (1) to enhance the value of the Geosat oceanographic investigations by providing the user community with improved ephemerides, (2) to develop orbit determination techniques for the upcoming altimetric mission Topex/Poseidon, and (3) to assess the radial orbit accuracy obtainable with recently developed gravity models. To this end, ephemerides for the entire first year of the ERM have been computed using the GEODYN II orbit program on the Cyber 205 supercomputer system at the NASA Goddard.

Haines, Bruce J.↗

Maximum likelihood estimation for distributed parameter models of flexible spacecraft

A distributed-parameter model of the NASA Solar Array Flight Experiment spacecraft structure is constructed on the basis of measurement data and analyzed to generate a priori estimates of modal frequencies and mode shapes. A Newton-Raphson maximum-likelihood algorithm is applied to determine the unknown parameters, using a truncated model for the estimation and the full model for the computation of the higher modes. Numerical results are presented in a series of graphs and briefly discussed, and the significant improvement in computation speed obtained by parallel implementation of the method on a supercomputer is noted.

Taylor, L. W., Jr.↗

FFTs in external or hierarchical memory

A description is given of advanced techniques for computing an ordered FFT on a computer with external or hierarchical memory. These algorithms (1) require as few as two passes through the external data set, (2) use strictly unit stride, long vector transfers between main memory and external storage, (3) require only a modest amount of scratch space in main memory, and (4) are well suited for vector and parallel computation. Performance figures are included for implementations of some of these algorithms on Cray supercomputers. Of interest is the fact that a main memory version outperforms the current Cray library FFT routines on the Cray-2, the Cray X-MP, and the Cray Y-MP systems. Using all eight processors on the Cray Y-MP, this main memory routine runs at nearly 2 Gflops.

Bailey, David H.↗

Merlin - Massively parallel heterogeneous computing

Hardware and software for Merlin, a new kind of massively parallel computing system, are described. Eight computers are linked as a 300-MIPS prototype to develop system software for a larger Merlin network with 16 to 64 nodes, totaling 600 to 3000 MIPS. These working prototypes help refine a mapped reflective memory technique that offers a new, very general way of linking many types of computer to form supercomputers. Processors share data selectively and rapidly on a word-by-word basis. Fast firmware virtual circuits are reconfigured to match topological needs of individual application programs. Merlin's low-latency memory-sharing interfaces solve many problems in the design of high-performance computing systems. The Merlin prototypes are intended to run parallel programs for scientific applications and to determine hardware and software needs for a future Teraflops Merlin network.

Wittie, Larry↗

A collision-selection rule for a particle simulation method suited to vector computers

A theory is developed for a selection rule governing collisions in a particle simulation of rarefied gas-dynamic flows. The selection rule leads to an algorithmic form highly compatible with fine grain parallel decomposition, allowing for efficient utilization of supercomputers having vector or massively parallel single instruction multiple data architectures. A comparison of shock-wave profiles obtained using both the selection rule and Bird's direct simulation Monte Carlo (DSMC) method show excellent agreement. The equation on which the selection rule is based is shown to be directly related to the time-counter procedure in the DSMC method. The results of several example simulations of representative rarefied flows are presented, for which the number of particles used ranged from 10 to the 6th to 10 to the 7th demonstrating the greatly improved computational efficiency of the method.

Baganoff, D.↗

Adaptive domain decomposition for Monte Carlo simulations on parallel processors

A method is described for performing direct simulation Monte Carlo (DSMC) calculations on parallel processors using adaptive domain decomposition to distribute the computational work load. The method has been implemented on a commercially available hypercube and benchmark results are presented which show the performance of the method relative to current supercomputers. The problems studied were simulations of equilibrium conditions in a closed, stationary box, a two-dimensional vortex flow, and the hypersonic, rarefield flow in a two-dimensional channel. For these problems, the parallel DSMC method ran 5 to 13 times faster than on a single processor of a Cray-2. The adaptive decomposition method worked well in uniformly distributing the computational work over an arbitrary number of processors and reduced the average computational time by over a factor of two in certain cases.

Wilmoth, Richard G.↗

Numerical Aerodynamic Simulation Program

Report describes developments that occurred at National Aerodynamics Simulation (NAS) facility at Ames Research Center during years 1987 and 1988. Begins with description of NAS processing network, which is large network of computers. Followed by summary of early achievements of program and advanced features of network, including installation of supercomputers, installation of standard operating system and communication software on all processors, accessibility to users at remote facilities across United States, and emphasis on development of graphics workstations to display results of numerical simulation.

Bailey, F. R.↗

Numerical simulation of rotorcraft

The objective of the research is to develop and validate accurate, user-oriented viscous CFD codes (with inviscid options) for three-dimensional, unsteady aerodynamic flows about arbitrary rotorcraft configurations. Unsteady, three-dimensional Euler and Navier-Stokes codes are developed, adapted, and extended to rotor-body combinations. Flow solvers are coupled with zonal grid topologies, including rotating and nonrotating blocks. Special grid clustering and wave-fitting techniques were developed to capture low-level radiating acoustic waves. Significant progress was made in computing the propagation of acoustic waves due to the interaction of a concentrated vortex and a helicopter airfoil. The need for higher-order schemes was firmly established in relatively inexpensive two-dimensional calculations. In three dimensions, the number of grid points required to capture the low-level acoustic waves becomes very large, so that large supercomputer memory becomes essential. Good agreement was obtained between the numerical results obtained with a thin-layer Navier-Stokes code and experimental data from a model rotor. In addition, several nonrotating configurations that are sometimes proposed to simulate rotor blade tips in conventional wind tunnels were examined, and the complex flow around the radical tip shape of the world's fastest helicopter is under investigation. These studies demonstrate the flexibility and power of CFD to gain physical insight, study novel ideas, and examine various possibilities that might be difficult or impossible to set up in physical experiments. As a prelude to studies of rotor-body aerodynamic interactions, a preliminary grid topology and moving-interface strategy were developed. A new Euler/Navier-Stokes code using these techniques computes the vortical wake directly, rather than modeling it, as in most previous rotorcraft studies. Several hover cases were run for conventional and advanced-geometry blades. Numerical schemes using multi-zones and/or adaptive grids appear to be necessary to simulate the complex vortical flows in rotor wakes.

Mccroskey, William J.↗

Simulation of turbomachinery flows

Significant advancements have been made in the last five years in the ability to model turbomachinery flows of engineering interest. This advancement can be directly attributed to the second generation of supercomputers like the Cray XMP and Cray 2 and advanced instrumentation techniques. Early on, the National Aeronautics and Space Administration Lewis Research Center recognized the potential gains in turbomachinery performance and life that could be achieved by taking advantage of this technology and instituted a comprehensive research program in turbomachinery flow modeling. This activity combined the areas of fluid flow analysis, computational fluid dynamics, and experimental fluid mechanics. As a result of this activity, Lewis has become an internationally recognized leader in turbomachinery flow modeling. Many of the research activities conducted under this program are utilized by industry. The presentation gives an overview of this program and provides sample illustration of simulation performed to date.

Adamczyk, John J.↗