Search NASASearch

SEARCH · Search NASA

Results for “High-Performance Computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

Integration of tools for the Design and Assessment of High-Performance, Highly Reliable Computing Systems (DAHPHRS), phase 1

Systems for Space Defense Initiative (SDI) space applications typically require both high performance and very high reliability. These requirements present the systems engineer evaluating such systems with the extremely difficult problem of conducting performance and reliability trade-offs over large design spaces. A controlled development process supported by appropriate automated tools must be used to assure that the system will meet design objectives. This report describes an investigation of methods, tools, and techniques necessary to support performance and reliability modeling for SDI systems development. Models of the JPL Hypercubes, the Encore Multimax, and the C.S. Draper Lab Fault-Tolerant Parallel Processor (FTPP) parallel-computing architectures using candidate SDI weapons-to-target assignment algorithms as workloads were built and analyzed as a means of identifying the necessary system models, how the models interact, and what experiments and analyses should be performed. As a result of this effort, weaknesses in the existing methods and tools were revealed and capabilities that will be required for both individual tools and an integrated toolset were identified.

Scheper, C.

A parallel-vector equation solver for unsymmetric matrices on supercomputers

A parallel-vector unsymmetric equation solver is presented. The solver exploits both vector and parallel capabilities provided by modern, high-performance supercomputers. A special storage scheme and loop-unrolling technique are used to optimize the vector performance. A parallel FORTRAN language is used to develop the solver on the CRAY 2 and CRAY Y-MP multiple processing computer environment. Three numerical examples are presented which demonstrate the efficiency and accuracy of this equation solver. The first two examples demonstrate the improved performance, and the third example utilizes the proposed solver to solve a highly nonlinear, unsymmetric finite element formulation for panel flutter.

Qin, J.

NASA Aeronautics: Research and Technology Program Highlights

This report contains numerous color illustrations to describe the NASA programs in aeronautics. The basic ideas involved are explained in brief paragraphs. The seven chapters deal with Subsonic aircraft, High-speed transport, High-performance military aircraft, Hypersonic/Transatmospheric vehicles, Critical disciplines, National facilities and Organizations & installations. Some individual aircraft discussed are : the SR-71 aircraft, aerospace planes, the high-speed civil transport (HSCT), the X-29 forward-swept wing research aircraft, and the X-31 aircraft. Critical disciplines discussed are numerical aerodynamic simulation, computational fluid dynamics, computational structural dynamics and new experimental testing techniques.

Source record

A high-fidelity, six-degree-of-freedom batch simulation environment for tactical guidance research and evaluation

A batch air combat simulation environment, the tactical maneuvering simulator (TMS), is presented. The TMS is a tool for developing and evaluating tactical maneuvering logics, but it can also be used to evaluate the tactical implications of perturbations to aircraft performance or supporting systems. The TMS can simulate air combat between any number of engagement participants, with practical limits imposed by computer memory and processing power. Aircraft are modeled using equations of motion, control laws, aerodynamics, and propulsive characteristics equivalent to those used in high-fidelity piloted simulations. Data bases representative of a modern high-performance aircraft with and without thrust-vectoring capability are included. To simplify the task of developing and implementing maneuvering logics in the TMS, an outer-loop control system, the tactical autopilot (TA), is implemented in the aircraft simulation model. The TA converts guidance commands by computerized maneuvering logics from desired angle of attack and wind-axis bank-angle inputs to the inner loop control augmentation system of the aircraft. The capabilities and operation of the TMS and the TA are described.

Goodrich, Kenneth H.

Applying and validating the RANS-3D flow-solver for evaluating a subsonic serpentine diffuser geometry

Subsonic inlet ducts for advanced, high-performance aircraft are evolving towards complex three-dimensional shapes for reasons of overall integration and weight. These factors lead to diffuser geometries that may sacrifice inlet performance, unless careful attention to design details and boundary layer management techniques are employed. The ability of viscous computational fluid dynamic (CFD) analysis of such geometries to aid the aircraft configurator in this complex design problem is herein examined. The RANS-3D Reynolds-Averaged Navier-Stokes solver is applied to model the complex flowfield occurring in a representative diffuser geometry and the solutions are compared to experimental results from a static test of the inlet duct. The computational results are shown to compare very favorably with experimental results over a range of mass flow rates, including those involving large amounts of separation in the diffuser. In addition, a novel grid topology is presented, and two turbulence models are evaluated in this study as part of the RANS-3D code.

Fletcher, Michael J.

High performance pipelined multiplier with fast carry-save adder

A high-performance pipelined multiplier is described. Its high performance results from the fast carry-save adder basic cell which has a simple structure and is suitable for the Gate Forest semi-custom environment. The carry-save adder computes the sum and carry within two gate delay. Results show that the proposed adder can operate at 200 MHz for a 2-micron CMOS process; better performance is expected in a Gate Forest realization.

Wu, Angus

Turbulence modeling of free shear layers for high-performance aircraft

The High Performance Aircraft (HPA) Grand Challenge of the High Performance Computing and Communications (HPCC) program involves the computation of the flow over a high performance aircraft. A variety of free shear layers, including mixing layers over cavities, impinging jets, blown flaps, and exhaust plumes, may be encountered in such flowfields. Since these free shear layers are usually turbulent, appropriate turbulence models must be utilized in computations in order to accurately simulate these flow features. The HPCC program is relying heavily on parallel computers. A Navier-Stokes solver (POVERFLOW) utilizing the Baldwin-Lomax algebraic turbulence model was developed and tested on a 128-node Intel iPSC/860. Algebraic turbulence models run very fast, and give good results for many flowfields. For complex flowfields such as those mentioned above, however, they are often inadequate. It was therefore deemed that a two-equation turbulence model will be required for the HPA computations. The k-epsilon two-equation turbulence model was implemented on the Intel iPSC/860. Both the Chien low-Reynolds-number model and a generalized wall-function formulation were included.

Sondak, Douglas L.

Brief summary of the evolution of high-temperature creep-fatigue life prediction models for crack initiation

The evolution of high-temperature, creep-fatigue, life-prediction methods used for cyclic crack initiation is traced from inception in the late 1940's. The methods reviewed are material models as opposed to structural life prediction models. Material life models are used by both structural durability analysts and by material scientists. The latter use micromechanistic models as guidance to improve a material's crack initiation resistance. Nearly one hundred approaches and their variations have been proposed to date. This proliferation poses a problem in deciding which method is most appropriate for a given application. Approaches were identified as being combinations of thirteen different classifications. This review is intended to aid both developers and users of high-temperature fatigue life prediction methods by providing a background from which choices can be made. The need for high-temperature, fatigue-life prediction methods followed immediately on the heels of the development of large, costly, high-technology industrial and aerospace equipment immediately following the second world war. Major advances were made in the design and manufacture of high-temperature, high-pressure boilers and steam turbines, nuclear reactors, high-temperature forming dies, high-performance poppet valves, aeronautical gas turbine engines, reusable rocket engines, etc. These advances could no longer be accomplished simply by trial and error using the 'build-em and bust-em' approach. Development lead times were too great and costs too prohibitive to retain such an approach. Analytic assessments of anticipated performance, cost, and durability were introduced to cut costs and shorten lead times. The analytic tools were quite primitive at first and out of necessity evolved in parallel with hardware development. After forty years more descriptive, more accurate, and more efficient analytic tools are being developed. These include thermal-structural finite element and boundary element analyses, advanced constitutive stress-strain-temperature-time relations, and creep-fatigue-environmental models for crack initiation and propagation. The high-temperature durability methods that have evolved for calculating high-temperature fatigue crack initiation lives of structural engineering materials are addressed. Only a few of the methods were refined to the point of being directly useable in design. Recently, two of the methods were transcribed into computer software for use with personal computers.

Halford, Gary R.

Installed F/A-18 inlet flow calculations at 30 degrees angle-of-attack: A comparative study

NASA Lewis is currently engaged in a research effort as a team member of the High Alpha Technology Program (HATP) within NASA. This program utilizes a specially equipped F/A-18, the High Alpha Research Vehicle (HARV), in an ambitious effort to improve the maneuverability of high-performance military aircraft at low subsonic speed, high angle of attack conditions. The overall objective of the Lewis effort is to develop inlet technology that will ensure efficient airflow delivery to the engine during these maneuvers. One part of the Lewis approach utilizes computational fluid dynamics codes to predict the installed performance of inlets for these highly maneuverable aircraft. Full Navier-Stokes (FNS) calculations on the installed F/A-18 inlet at 30 degrees angle of attack, 0 degrees yaw, and a freestream Mach number of 0.2 have been obtained in this study using an algebraic turbulence model with two grids (original and revised). Results obtained with the original grid were used to determine where further grid refinements and additional geometry were needed. In order to account properly for the external effects, the forebody, leading edge extension (LEX), ramp, and wing were included with inlet geometry. In the original grid, the diverter, LEX slot, and leading edge flap were not included due to insufficient geometry definition, but were included in a revised grid. In addition, a thin-layer Navier-Stokes (TLNS) code is used with the revised grid and the numerical results are compared to those obtained with the FNS code. The TLNS code was used to evaluate the effects on the solution using a code with more recent CFD developments such as upwinding with TVD schemes versus central differencing with artificial dissipation. The calculations are compared to a limited amount of available experimental data. The predicted forebody/fuselage surface static pressures compared well with data of all solutions. The predicted trajectory of the vortex generated under the LEX was different for each solution. These discrepancies are attributed to differences in the grid resolution and turbulence modeling. All solutions predict that this vortex is ingested by the inlet. The predicted inlet total pressure recoveries are lower than data and the distortions are higher than data. The results obtained with the revised grid were significantly improved from the original grid results. The original grid results indicated the ingested vortex migrated to the engine face and caused additional distortions to those already present due to secondary flow development. The revised grid results indicate that the ingested vortex is dissipated along the inlet duct inboard wall. The TLNS results indicate the flow at the engine face was much more distorted than the FNS results and is attributed to the pole boundary condition introducing numerical distortions into the flow field.

Smith, C. Frederic

Characterizing parallel file-access patterns on a large-scale multiprocessor

Rapid increases in the computational speeds of multiprocessors have not been matched by corresponding performance enhancements in the I/O subsystem. To satisfy the large and growing I/O requirements of some parallel scientific applications, we need parallel file systems that can provide high-bandwidth and high-volume data transfer between the I/O subsystem and thousands of processors. Design of such high-performance parallel file systems depends on a thorough grasp of the expected workload. So far there have been no comprehensive usage studies of multiprocessor file systems. Our CHARISMA project intends to fill this void. The first results from our study involve an iPSC/860 at NASA Ames. This paper presents results from a different platform, the CM-5 at the National Center for Supercomputing Applications. The CHARISMA studies are unique because we collect information about every individual read and write request and about the entire mix of applications running on the machines. The results of our trace analysis lead to recommendations for parallel file system design. First the file system should support efficient concurrent access to many files, and I/O requests from many jobs under varying load conditions. Second, it must efficiently manage large files kept open for long periods. Third, it should expect to see small requests predominantly sequential access patterns, application-wide synchronous access, no concurrent file-sharing between jobs appreciable byte and block sharing between processes within jobs, and strong interprocess locality. Finally, the trace data suggest that node-level write caches and collective I/O request interfaces may be useful in certain environments.

Purakayastha, Apratim

VLSI neuroprocessors

Electronic and optoelectronic hardware implementations of highly parallel computing architectures address several ill-defined and/or computation-intensive problems not easily solved by conventional computing techniques. The concurrent processing architectures developed are derived from a variety of advanced computing paradigms including neural network models, fuzzy logic, and cellular automata. Hardware implementation technologies range from state-of-the-art digital/analog custom-VLSI to advanced optoelectronic devices such as computer-generated holograms and e-beam fabricated Dammann gratings. JPL's concurrent processing devices group has developed a broad technology base in hardware implementable parallel algorithms, low-power and high-speed VLSI designs and building block VLSI chips, leading to application-specific high-performance embeddable processors. Application areas include high throughput map-data classification using feedforward neural networks, terrain based tactical movement planner using cellular automata, resource optimization (weapon-target assignment) using a multidimensional feedback network with lateral inhibition, and classification of rocks using an inner-product scheme on thematic mapper data. In addition to addressing specific functional needs of DOD and NASA, the JPL-developed concurrent processing device technology is also being customized for a variety of commercial applications (in collaboration with industrial partners), and is being transferred to U.S. industries. This viewgraph p resentation focuses on two application-specific processors which solve the computation intensive tasks of resource allocation (weapon-target assignment) and terrain based tactical movement planning using two extremely different topologies. Resource allocation is implemented as an asynchronous analog competitive assignment architecture inspired by the Hopfield network. Hardware realization leads to a two to four order of magnitude speed-up over conventional techniques and enables multiple assignments, (many to many), not achievable with standard statistical approaches. Tactical movement planning (finding the best path from A to B) is accomplished with a digital two-dimensional concurrent processor array. By exploiting the natural parallel decomposition of the problem in silicon, a four order of magnitude speed-up over optimized software approaches has been demonstrated.

Kemeny, Sabrina E.

Particle simulation on heterogeneous distributed supercomputers

We describe the implementation and performance of a three dimensional particle simulation distributed between a Thinking Machines CM-2 and a Cray Y-MP. These are connected by a combination of two high-speed networks: a high-performance parallel interface (HIPPI) and an optical network (UltraNet). This is the first application to use this configuration at NASA Ames Research Center. We describe our experience implementing and using the application and report the results of several timing measurements. We show that the distribution of applications across disparate supercomputing platforms is feasible and has reasonable performance. In addition, several practical aspects of the computing environment are discussed.

Becker, Jeffrey C.

Program For Analyzing Designs Of Liquid-Propellant Rockets

Rocket Combustor Interactive Design Computer Methodology (ROCCID) computer program provides standardized methodology, using state-of-art codes and procedures, for analysis of combustion performance and stability of liquid-propellant rocket engine. Provides combustion analyst with software tool to analyze existing combustor design (point-analysis option), or design high-performance, stable combustor, given set of input design requirements (point-design option). Written in ANSI FORTRAN 77 and VAX FORTRAN.

Klem, Mark D.

VLSI Processor For Vector Quantization

Pixel intensities in each kernel compared simultaneously with all code vectors. Prototype high-performance, low-power, very-large-scale integrated (VLSI) circuit designed to perform compression of image data by vector-quantization method. Contains relatively simple analog computational cells operating on direct or buffered outputs of photodetectors grouped into blocks in imaging array, yielding vector-quantization code word for each such block in sequence. Scheme exploits parallel-processing nature of vector-quantization architecture, with consequent increase in speed.

Tawel, Raoul

File-access characteristics of parallel scientific workloads

Phenomenal improvements in the computational performance of multiprocessors have not been matched by comparable gains in I/O system performance. This imbalance has resulted in I/O becoming a significant bottleneck for many scientific applications. One key to overcoming this bottleneck is improving the performance of parallel file systems. The design of a high-performance parallel file system requires a comprehensive understanding of the expected workload. Unfortunately, until recently, no general workload studies of parallel file systems have been conducted. The goal of the CHARISMA project was to remedy this problem by characterizing the behavior of several production workloads, on different machines, at the level of individual reads and writes. The first set of results from the CHARISMA project describe the workloads observed on an Intel iPSC/860 and a Thinking Machines CM-5. This paper is intended to compare and contrast these two workloads for an understanding of their essential similarities and differences, isolating common trends and platform-dependent variances. Using this comparison, we are able to gain more insight into the general principles that should guide parallel file-system design.

Nieuwejaar, Nils

Numerical Stability and Control Analysis Towards Falling-Leaf Prediction Capabilities of Splitflow for Two Generic High-Performance Aircraft Models

Aerodynamic analysis are performed using the Lockheed-Martin Tactical Aircraft Systems (LMTAS) Splitflow computational fluid dynamics code to investigate the computational prediction capabilities for vortex-dominated flow fields of two different tailless aircraft models at large angles of attack and sideslip. These computations are performed with the goal of providing useful stability and control data to designers of high performance aircraft. Appropriate metrics for accuracy, time, and ease of use are determined in consultations with both the LMTAS Advanced Design and Stability and Control groups. Results are obtained and compared to wind-tunnel data for all six components of forces and moments. Moment data is combined to form a "falling leaf" stability analysis. Finally, a handful of viscous simulations were also performed to further investigate nonlinearities and possible viscous effects in the differences between the accumulated inviscid computational and experimental data.

Charlton, Eric F.

Efficient GO2/GH2 Injector Design: A NASA, Industry and University Cooperative Effort

Developing new propulsion components in the face of shrinking budgets presents a significant challenge. The technical, schedule and funding issues common to any design/development program are complicated by the ramifications of the continuing decrease in funding for the aerospace industry. As a result, new working arrangements are evolving in the rocket industry. This paper documents a successful NASA, industry, and university cooperative effort to design efficient high performance GO2/GH2 rocket injector elements in the current budget environment. The NASA Reusable Launch Vehicle (RLV) Program initially consisted of three vehicle/engine concepts targeted at achieving single stage to orbit. One of the Rocketdyne propulsion concepts, the RS 2100 engine, used a full-flow staged-combustion cycle. Therefore, the RS 2100 main injector would combust GO2/GH 2 propellants. Early in the design phase, but after budget levels and contractual arrangements had been set the limitations of the current gas/gas injector database were identified. Most of the relevant information was at least twenty years old. Designing high performance injectors to meet the RS 2100 requirements would require the database to be updated and significantly enhanced. However, there was no funding available to address the need for more data. NASA proposed a teaming arrangement to acquire the updated information without additional funds from the RLV Program. A determination of the types and amounts of data needed was made along with test facilities with capabilities to meet the data requirements, budget constraints, and schedule. After several iterations a program was finalized and a team established to satisfy the program goals. The Gas/Gas Injector Technology (GGIT) Program had the overall goal of increasing the ability of the rocket engine community to design efficient high-performance, durable gas/gas injectors relevant to RLV requirements. First, the program would provide Rocketdyne with data on preliminary gas/gas injector designs which would enable discrimination among candidate injector designs. Secondly, the program would enhance the national gas/gas database by obtaining high-quality data that increases the understanding of gas/gas injector physics and is suitable for computational fluid dynamics (CFD) code validation. The third program objective was to validate CFD codes for future gas/gas injector design in the RLV program.

Tucker, P. K.

High Performance Computing at NASA

The speaker will give an overview of high performance computing in the U.S. in general and within NASA in particular, including a description of the recently signed NASA-IBM cooperative agreement. The latest performance figures of various parallel systems on the NAS Parallel Benchmarks will be presented. The speaker was one of the authors of the NAS (National Aerospace Standards) Parallel Benchmarks, which are now widely cited in the industry as a measure of sustained performance on realistic high-end scientific applications. It will be shown that significant progress has been made by the highly parallel supercomputer industry during the past year or so, with several new systems, based on high-performance RISC processors, that now deliver superior performance per dollar compared to conventional supercomputers. Various pitfalls in reporting performance will be discussed. The speaker will then conclude by assessing the general state of the high performance computing field.

Bailey, David H.