Search NASA⌕ Search

SEARCH · Search NASA

Results for “parallel programming”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,081 records · Page 60

Factors which Limit the Value of Additional Redundancy in Human Rated Launch Vehicle Systems

The National Aeronautics and Space Administration (NASA) has embarked on an ambitious program to return humans to the moon and beyond. As NASA moves forward in the development and design of new launch vehicles for future space exploration, it must fully consider the implications that rule-based requirements of redundancy or fault tolerance have on system reliability/risk. These considerations include common cause failure, increased system complexity, combined serial and parallel configurations, and the impact of design features implemented to control premature activation. These factors and others must be considered in trade studies to support design decisions that balance safety, reliability, performance and system complexity to achieve a relatively simple, operable system that provides the safest and most reliable system within the specified performance requirements. This paper describes conditions under which additional functional redundancy can impede improved system reliability. Examples from current NASA programs including the Ares I Upper Stage will be shown.

Anderson, Joel M.↗

Path Planning: Differential Dynamic Programming and Model Predictive Path Integral Control on VTOL Aircraft

This paper explores two optimal control approaches, widely used in robotics, to establish their viability as real-time trajectory planners for vehicle configurations envisioned for the emerging aviation sector of Urban Air Mobility (UAM). Differential Dynamic Programming (DDP) enables planning over highly nonlinear dynamics using second-order approximations along a nominal trajectory, and displays quadratic convergence to a local solution. Model Predictive Path Integral (MPPI) is a stochastic sampling-based algorithm that can optimize for general cost criteria, including potentially highly nonlinear formulations, and supports parallel computation through the use of modern GPU hardware. In this work, DDP and MPPI were implemented using model predictive control (MPC), and the results indicate they are able to successfully transition the aircraft over different flight envelopes and generate trajectories unique to UAM vehicles.

Differential Dynamic Programming↗

Application of Portable Parallelization Strategies for GPUs on track reconstruction kernels

Utilizing the computational power of GPUs is one of the key ingredients to meet the computing challenges presented to the next generation of High-Energy Physics (HEP) experiments. Unlike CPUs, developing software for GPUs often involves using architecturespecific programming languages promoted by the GPU vendors and hence limits the platform that the code can run on. Various portability solutions have been developed to achieve portable, performant software across different GPU vendors. Given the rapid evolution of these portability solutions, an early adoption of them in simple HEP testbed applications will help us understand the strengths and weaknesses of respective approaches.We apply several portability solutions, including Alpaka, Kokkos, SYCL and std::execution::par, on kernels for track propagation extracted from the mkFit project. We report on the development experience of the same application with different portability solutions, as well as their performance on GPUs, measured as the throughput of the kernels, from different manufacturers such as NVIDIA, AMD and Intel.

Kwok, Martin [Fermilab] (ORCID:0000000286936146)↗

The energy distributions of B supergiants in the Large Magellanic Cloud

It is shown that line-blanketed, LTE, plane-parallel model atmosphere calculations provide excellent fits to the ultraviolet-through-visual energy distributions of B supergiants in the Large Magellanic Cloud. The models were computed using Kurucz's (1979) ATLAS atmosphere program, but with lower gravities than were contained in Kurucz's published model grid. The ultraviolet continua of low gravity stars are found to be sensitive to changes in temperature and gravity. Measurements of Teff and log g for ten LMC B supergiants from model atmosphere fits to the energy distributions yield estimates of their radii, luminosities, and masses. Model atmosphere fits suggest that the late B supergiants have significantly lower masses than the earlier B types of the same luminosity, contrary to stellar evolution theory which predicts that B supergiants are in a post-core hydrogen burning phase and should evolve very quickly and at essentially constant mass.

Fitzpatrick, Edward L.↗

Acoustically excited heated jets. 1: Internal excitation

The effects of relatively strong upstream acoustic excitation on the mixing of heated jets with the surrounding air are investigated. To determine the extent of the available information on experiments and theories dealing with acoustically excited heated jets, an extensive literature survey was carried out. The experimental program consisted of flow visualization and flowfield velocity and temperature measurements for a broad range of jet operating and flow excitation conditions. A 50.8-mm-diam nozzle was used for this purpose. Parallel to the experimental study, an existing theoretical model of excited jets was refined to include the region downstream of the jet potential core. Excellent agreement was found between theory and experiment in moderately heated jets. However, the theory has not yet been confirmed for highly heated jets. It was found that the sensitivity of heated jets to upstream acoustic excitation varies strongly with the jet operating conditions and that the threshold excitation level increases with increasing jet temperature. Furthermore, the preferential Strouhal number is found not to change significantly with a change of the jet operating conditions. Finally, the effects of the nozzle exit boundary layer thickness appear to be similar for both heated and unheated jets at low Mach numbers.

Lepicovsky, J.↗

Fast, Real-Time, Animated Displays

Displays for Advanced Concepts Simulator (ACS) generated on Adage Raster Display System 3000 (RDS 3000). Improved programming techniques developed, and revisions to language implementation made. Both types of changes took better advantage of high-speed characteristics of RDS 3000 hardware. Increases in speed resulted from: utilization of parallel-processing capabilities of AGG4, and use of AGG4 to take advantage of certain high-speed characteristics of display memory not previously used. Result was fourfold increase in animation-update rate to 16 frames per second.

Kahlbaum, William M.↗

Inhomogeneities of stratocumulus liquid water

There is a growing body of observational evidence on inhomogeneous cloud structure, most recently from the extensive measurements of the FIRE field program. Knowledge of cloud structure is important because it strongly influences the cloud radiative properties, one of the major factors in determining the global energy balance. Current atmospheric circulation models use plane-parallel radiation, so that the liquid water in each gridbox is assumed to be uniform, which gives an unrealistically large albedo. In reality cloud liquid water occupies only a subset of each gridbox, greatly reducing the mean albedo. If future climate models are to treat the hydrological cycle in a manner consistent with energy balance, a better treatment of cloud liquid is needed. FIRE concentrated upon two cloud types of special interest: cirrus and marine stratocumulus. Cirrus tend to be high and optically thin, thus reducing the effective radiative temperature without increasing the albedo significantly, leading to an enhanced greenhouse heating. In contrast, marine stratocumulus are low and optically thick, thus producing a large increase in reflected radiation with a small change in emitted radiation, giving a net cooling which could potentially mitigate the expected greenhouse warming. The FIRE measurements in California stratocumulus during June and July of 1987 show variations in cloud liquid water on all scales. Such variations are associated with inhomogeneous entrainment, in which entrained dry air, rather than mixing uniformly with cloudy air, remains intact in blobs of all sizes, which decay only slowly by invasion of cloudy air. Two important stratocumulus observations are described, followed by a simple fractal model which reproduces these properties, and finally, the model radiative properties are discussed.

Cahalan, Robert F.↗

Discrete approximations to optimal trajectories using direct transcription and nonlinear programming

A recently developed method for solving optimal trajectory problems uses a piecewise-polynomial representation of the state and control variables, enforces the equations of motion via a collocation procedure, and thus approximates the original calculus-of-variations problem with a nonlinear-programming problem, which is solved numerically. This paper identifies this method as a direct transcription method and proceeds to investigate the relationship between the original optimal-control problem and the nonlinear-programming problem. The discretized adjoint equation of the collocation method is found to have deficient accuracy, and an alternate scheme which discretizes the equations of motion using an explicit Runge-Kutta parallel-shooting approach is developed. Both methods are applied to finite-thrust spacecraft trajectory problems, including a low-thrust escape spiral, a three-burn rendezvous, and a low-thrust transfer to the moon.

Enright, Paul J.↗

Visualization and Tracking of Parallel CFD Simulations

We describe a system for interactive visualization and tracking of a 3-D unsteady computational fluid dynamics (CFD) simulation on a parallel computer. CM/AVS, a distributed, parallel implementation of a visualization environment (AVS) runs on the CM-5 parallel supercomputer. A CFD solver is run as a CM/AVS module on the CM-5. Data communication between the solver, other parallel visualization modules, and a graphics workstation, which is running AVS, are handled by CM/AVS. Partitioning of the visualization task, between CM-5 and the workstation, can be done interactively in the visual programming environment provided by AVS. Flow solver parameters can also be altered by programmable interactive widgets. This system partially removes the requirement of storing large solution files at frequent time steps, a characteristic of the traditional 'simulate (yields) store (yields) visualize' post-processing approach.

Vaziri, Arsi↗

Integrated System Test of an Airbreathing Rocket (ISTAR)

Rocket Based Combined Cycle (RBCC) propulsion system development and ground test is being conducted as part of the NASA Marshall Space Flight Center Integrated System Test of an Airbreathing Rocket (ISTAR) program. Rocketdyne, Aerojet and Pratt & Whitney have teamed as the Rocket Based Combined Cycle Consortium (RBC3) to work the propulsion system development. Each company offered unique RBCC propulsion concepts as candidates for the ISTAR propulsion system. A team of engine contractor, vehicle contractor and NASA representatives reviewed the concepts proposed by each company, reviewed the available data and selected the Aerojet RBCC propulsion system concept as the team propulsion system baseline for the ISTAR program. The ISTAR program is currently in a "Jumpstart" phase for development of the engine system leading to ground test of a thermally and power balanced RBCC propulsion system at Stennis Space Center in 2005. A parallel flight test demonstration of this propulsion system is anticipated to lead to first flight in the 2007 timeframe.

Faulkner, Robert F.↗

Multidisciplinary Optimization Methods for Aircraft Preliminary Design

This paper describes a research program aimed at improved methods for multidisciplinary design and optimization of large-scale aeronautical systems. The research involves new approaches to system decomposition, interdisciplinary communication, and methods of exploiting coarse-grained parallelism for analysis and optimization. A new architecture, that involves a tight coupling between optimization and analysis, is intended to improve efficiency while simplifying the structure of multidisciplinary, computation-intensive design problems involving many analysis disciplines and perhaps hundreds of design variables. Work in two areas is described here: system decomposition using compatibility constraints to simplify the analysis structure and take advantage of coarse-grained parallelism; and collaborative optimization, a decomposition of the optimization process to permit parallel design and to simplify interdisciplinary communication requirements.

Kroo, Ilan↗

Highly parallel structured adaptive mesh refinement using parallel language-based approaches

Adaptive mesh refinement (AMR) calculations carried out on structured meshes play an exceedingly important role in several areas of science and engineering. A strategy for using Fortran 90 in an object-oriented fashion is presented. This permits AMR applications to be expressed in terms of familiar abstractions that are natural to the process of solving AMR hierarchies. The OpenMP features that are useful for parallel processing of AMR hierarchies in a load balanced fashion on multiprocessors is described.

computational↗

A Parallel Genetic Algorithm for Automated Electronic Circuit Design

Parallelized versions of genetic algorithms (GAs) are popular primarily for three reasons: the GA is an inherently parallel algorithm, typical GA applications are very compute intensive, and powerful computing platforms, especially Beowulf-style computing clusters, are becoming more affordable and easier to implement. In addition, the low communication bandwidth required allows the use of inexpensive networking hardware such as standard office ethernet. In this paper we describe a parallel GA and its use in automated high-level circuit design. Genetic algorithms are a type of trial-and-error search technique that are guided by principles of Darwinian evolution. Just as the genetic material of two living organisms can intermix to produce offspring that are better adapted to their environment, GAs expose genetic material, frequently strings of 1s and Os, to the forces of artificial evolution: selection, mutation, recombination, etc. GAs start with a pool of randomly-generated candidate solutions which are then tested and scored with respect to their utility. Solutions are then bred by probabilistically selecting high quality parents and recombining their genetic representations to produce offspring solutions. Offspring are typically subjected to a small amount of random mutation. After a pool of offspring is produced, this process iterates until a satisfactory solution is found or an iteration limit is reached. Genetic algorithms have been applied to a wide variety of problems in many fields, including chemistry, biology, and many engineering disciplines. There are many styles of parallelism used in implementing parallel GAs. One such method is called the master-slave or processor farm approach. In this technique, slave nodes are used solely to compute fitness evaluations (the most time consuming part). The master processor collects fitness scores from the nodes and performs the genetic operators (selection, reproduction, variation, etc.). Because of dependency issues in the GA, it is possible to have idle processors. However, as long as the load at each processing node is similar, the processors are kept busy nearly all of the time. In applying GAs to circuit design, a suitable genetic representation 'is that of a circuit-construction program. We discuss one such circuit-construction programming language and show how evolution can generate useful analog circuit designs. This language has the desirable property that virtually all sets of combinations of primitives result in valid circuit graphs. Our system allows circuit size (number of devices), circuit topology, and device values to be evolved. Using a parallel genetic algorithm and circuit simulation software, we present experimental results as applied to three analog filter and two amplifier design tasks. For example, a figure shows an 85 dB amplifier design evolved by our system, and another figure shows the performance of that circuit (gain and frequency response). In all tasks, our system is able to generate circuits that achieve the target specifications.

Long, Jason D.↗

Design and Preliminary Experimental Results on a Uniform-Field Test Fixture for Power Flow Experiments

Electrodes that produce a uniform electric field across their surfaces are desired for a variety of high-voltage systems including spark-gap switches and transversely excited atmospheric lasers. Here, the implementation of uniform-field electrodes to a parallel-plate load is described for the study of vacuum power flow under pulsed-power applications. We begin with an overview of general uniform-field geometries, and provide an open-source program that generates these geometries as described by Ernst. Following that, a modification to an existing power-flow experimental platform, that adopts a uniform-field geometry, is introduced and its utility is described. Lastly, preliminary experimental results are presented and future steps are outlined.

Capacitors↗

PaRSEC: Scalability, flexibility, and hybrid architecture support for task-based applications in ECP

This paper highlights the most significant enhancements made to PaRSEC, a scalable task-based runtime system designed for hybrid machines, during the Exascale Computing Project (ECP). The enhancements focus on expanding the capabilities of PaRSEC to address the evolving landscape of parallel computing. Notable achievements include the integration of support for three major types of accelerators (NVIDIA, AMD, and Intel GPUs), the refinement and increased flexibility of the communication subsystem, and the introduction of new programming interfaces tailored for irregular applications. Additionally, the project resulted in the development of powerful debugging and performance analysis tools aimed at assisting users in understanding and optimizing their applications. We present a comprehensive demonstration of these advancements through a series of benchmarks and applications within ECP and beyond, thereby showcasing the enhanced capabilities of PaRSEC across the diverse architectures within the ECP, providing valuable insights into the runtime system’s adaptability and performance across varied computing environments.

Bouteiller, Aurelien↗

Climate Model Output Rewriter

The Climate Model Output Rewriter (CMOR) software was first developed by LLNL’s PCMDI program in early 2000s and was formally released with v1.0 (July 2006), v2.0 (January 2011), and v3.1(June 2016). CMOR is used to produce Climate and Forecast Convention (http://cfconventions.org/) CF-compliant netCDF files, in the standard format required to satisfy the World Climate Research Program (WCRP) Coupled Model Intercomparison Project (CMIP). The software has been used across multiple phases of the Earth System Modeling (ESM) project CMIP (CMIP3, CMIP5, CMIP6, and planned use in CMIP7) along with numerous parallel projects focused on preparation observations for use in model evaluation (obs4MIPs) and forcing datasets (input4MIPs) to guide ESM simulations to meet strict experimental protocols. More information can be obtained from the CMOR website and code repositories: https://cmor.llnl.gov/; https://github.com/pcmdi/cmor; https://github.com/PCMDI/cmor3_documentation The ESM variable definitions used as input for CMOR can also be viewed in code repositories: https://github.com/PCMDI/cmip3-cmor-tables/; https://github.com/PCMDI/cmip5-cmor-tables/; https://github.com/PCMDI/cmip6-cmor-tables/

Mauzey, ChristopherF↗

Diagnostics of wear in aeronautical systems

The use of appropriate diagnostic tools for aircraft oil wetted components is reviewed, noting that it can reduce direct operating costs through reduced unscheduled maintenance, particularly in helicopter engine and transmission systems where bearing failures are a significant cost factor. Engine and transmission wear modes are described, and diagnostic methods for oil and wet particle analysis, the spectrometric oil analysis program, chip detectors, ferrography, in-line oil monitor and radioactive isotope tagging are discussed, noting that they are effective over a limited range of particle sizes but compliment each other if used in parallel. Fine filtration can potentially increase time between overhauls, but reduces the effectiveness of conventional oil monitoring techniques so that alternative diagnostic techniques must be used. It is concluded that the development of a diagnostic system should be parallel and integral with the development of a mechanical system.

Wedeven, L. D.↗

Flow rate/pressure drop data gathered from testing a sample of the Space Shuttle Strain Isolation Pad (SIP): Effects of ambient pressure combined with tension and compression conditions

Tests were conducted on a sample of strain isolation pad (SIP) typical of that used in the shuttle orbiter thermal protection system to determine the characteristics of SIP internal flow. Data obtained were pressure drop as a function of flow rate for a range of ambient pressures representing various points along the Shuttle trajectory and for stretched and compressed conditions of the SIP. Flow was in the direction of the weave parallel to most of the fibers. The data are plotted in several standard engineering formats in order to be of maximum utility to the user. In addition to providing support to the Space Shuttle Program, these data are a source of experimental information on flow through fiberous (rather than the more usual sand bed type) porous media.

Springfield, R. D.↗