Search NASA⌕ Search

SEARCH · Search NASA

Results for “parallel computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,135 records · Page 63

Pressure Loss Predictions of the Reactor Simulator Subsystem at NASA Glenn Research Center

Testing of the Fission Power System (FPS) Technology Demonstration Unit (TDU) is being conducted at NASA Glenn Research Center. The TDU consists of three subsystems: the reactor simulator (RxSim), the Stirling Power Conversion Unit (PCU), and the heat exchanger manifold (HXM). An annular linear induction pump (ALIP) is used to drive the working fluid. A preliminary version of the TDU system (which excludes the PCU for now) is referred to as the "RxSim subsystem" and was used to conduct flow tests in Vacuum Facility 6 (VF 6). In parallel, a computational model of the RxSim subsystem was created based on the computer-aided-design (CAD) model and was used to predict loop pressure losses over a range of mass flows. This was done to assess the ability of the pump to meet the design intent mass flow demand. Measured data indicates that the pump can produce 2.333 kg/sec of flow, which is enough to supply the RxSim subsystem with a nominal flow of 1.75 kg/sec. Computational predictions indicated that the pump could provide 2.157 kg/sec (using the Spalart-Allmaras (S‒A) turbulence model) and 2.223 kg/sec (using the k- turbulence model). The computational error of the predictions for the available mass flow is ‒0.176 kg/sec (with the S-A turbulence model) and -0.110 kg/sec (with the k- turbulence model) when compared to measured data.

Mass Flow↗

Transition Analysis for the CRM-NLF Wind Tunnel Configuration

This paper reports the results of a comprehensive linear stability analysis of the boundary layer flow over the common research model with natural laminar flow (CRM-NLF) aircraft configuration. The flow conditions match selected test conditions from a recent wind tunnel experiment in the National Transonic Facility at the NASA Langley Research Center. Previous work has shown that the measured onset of laminar-turbulent transition during the experiments can be correlated with the linear amplification of Tollmien-Schlichting (TS) and stationary crossflow (CF) instabilities in the swept wing boundary layer. However, a significant scatter ( N ∈ (4,9)) was observed in the values of the logarithmic amplification factor along the measured transition front. This previous analysis was based on an approximate basic state (based on a boundary layer code with conical flow approximation) and parallel stability computations without surface curvature effects. Here, we examine the effects of these various approximations with the goal of quantifying the resulting changes in the N-factor correlations. Specifically, both linear stability theory (LST) and the parabolized stability equations (PSE)are used in conjunction with an accurate definition of the laminar boundary layer flow as computed with a Navier-Stokes solver with a Reynolds-Averaged-Navier-Stokes (RANS) based turbulence model within the turbulent parts of the flow. Furthermore, the effects of instability wave propagation within a fully three-dimensional boundary layer are also evaluated by integrating the disturbance growth rates along suitably chosen, curvilinear (i.e., nonplanar) propagation trajectories. The results of this analysis are also used in an accompanying paper by Venkatachari et al. to develop improved, physics based transition predictions for the same CRM-NLF configuration.

Boundary layer transition↗

Transition Analysis for the CRM-NLF Wind Tunnel Configuration

This paper presents the results of an ongoing study into the linear stability characteristics of the boundary layer flow over the common research model with natural laminar flow (CRMNLF) aircraft configuration. The flow conditions match selected test conditions from a recent wind tunnel experiment in the National Transonic Facility at the NASA Langley Research Center. Previous work involving parallel stability computations of a boundary layer flow based on the conical flow approximation has shown that the measured onset of laminar-turbulent transition during the experiments can be correlated with the linear amplification of Tollmien- Schlichting (TS) and stationary crossflow (CF) instabilities in the swept wing boundary layer. Here, we examine the effects of the simplifying approximations in both basic state computation and the stability analysis, with the goal of quantifying the resulting changes in the N-factor correlations. Specifically, the basic states are computed by using full Navier-Stokes equations and the stability analysis is performed by using a nonorthogonal coordinate system that allows a clear distinction between planar TS and CF instabilities. Furthermore, the effects of curvature and nonparallel mean flow have been included in the stability computations based on the parabolized stability equations (PSE). The fully turbulent Reynolds-Averaged-Navier-Stokes (RANS) mean flow solutions show good agreement with the measured wall pressure distribution. Viscous-inviscid interactive effects are observed to be important because the shock fronts along the suction surface are influenced by the imposed transition front. The stability results confirm the previous findings related to TS amplification within the inboard region of the wing and the dominance of stationary CF modes in the outboard region. However, given the close proximity of the measured transition front and the dual shock system within the outer part of the wing, the onset of transition may well be shock limited within the outboard region. In general, the transition criterion based on the dual N-factor method with N TS = N CF = 6 is reasonably successful at correlating with the measured transition fronts at Re MAC = 15 million and AoA = 1.5, 2 degrees; however, the low values of the correlating N-factors at Re MAC = 17.5 million support the hypothesis that the measured transition at the higher Reynolds number may have been strongly influenced by the merging of turbulent wedges that originate from surface imperfections near the leading edge.

Boundary layer transition↗

Transition Analysis for the CRM-NLF Wind Tunnel Configuration

This paper presents the results of an ongoing study into the linear stability characteristics of the boundary layer flow over the common research model with natural laminar flow (CRMNLF) aircraft configuration. The flow conditions match selected test conditions from a recent wind tunnel experiment in the National Transonic Facility at the NASA Langley Research Center. Previous work involving parallel stability computations of a boundary layer flow based on the conical flow approximation has shown that the measured onset of laminar-turbulent transition during the experiments can be correlated with the linear amplification of Tollmien- Schlichting (TS) and stationary crossflow (CF) instabilities in the swept wing boundary layer. Here, we examine the effects of the simplifying approximations in both basic state computation and the stability analysis, with the goal of quantifying the resulting changes in the N-factor correlations. Specifically, the basic states are computed by using full Navier-Stokes equations and the stability analysis is performed by using a nonorthogonal coordinate system that allows a clear distinction between planar TS and CF instabilities. Furthermore, the effects of curvature and nonparallel mean flow have been included in the stability computations based on the parabolized stability equations (PSE). The fully turbulent Reynolds-Averaged-Navier-Stokes (RANS) mean flow solutions show good agreement with the measured wall pressure distribution. Viscous-inviscid interactive effects are observed to be important because the shock fronts along the suction surface are influenced by the imposed transition front. The stability results confirm the previous findings related to TS amplification within the inboard region of the wing and the dominance of stationary CF modes in the outboard region. However, given the close proximity of the measured transition front and the dual shock system within the outer part of the wing, the onset of transition may well be shock limited within the outboard region. In general, the transition criterion based on the dual N-factor method with N TS = N CF = 6 is reasonably successful at correlating with the measured transition fronts at Re MAC = 15 million and AoA = 1.5, 2 degrees; however, the low values of the correlating N-factors at Re MAC = 17.5 million support the hypothesis that the measured transition at the higher Reynolds number may have been strongly influenced by the merging of turbulent wedges that originate from surface imperfections near the leading edge.

Boundary layer transition↗

A strategy for reducing turnaround time in design optimization using a distributed computer system

There is a need to explore methods for reducing lengthly computer turnaround or clock time associated with engineering design problems. Different strategies can be employed to reduce this turnaround time. One strategy is to run validated analysis software on a network of existing smaller computers so that portions of the computation can be done in parallel. This paper focuses on the implementation of this method using two types of problems. The first type is a traditional structural design optimization problem, which is characterized by a simple data flow and a complicated analysis. The second type of problem uses an existing computer program designed to study multilevel optimization techniques. This problem is characterized by complicated data flow and a simple analysis. The paper shows that distributed computing can be a viable means for reducing computational turnaround time for engineering design problems that lend themselves to decomposition. Parallel computing can be accomplished with a minimal cost in terms of hardware and software.

Young, Katherine C.↗

Gust Acoustics Computation with a Space-Time CE/SE Parallel 3D Solver

The benchmark Problem 2 in Category 3 of the Third Computational Aero-Acoustics (CAA) Workshop is solved using the space-time conservation element and solution element (CE/SE) method. This problem concerns the unsteady response of an isolated finite-span swept flat-plate airfoil bounded by two parallel walls to an incident gust. The acoustic field generated by the interaction of the gust with the flat-plate airfoil is computed by solving the 3D (three-dimensional) Euler equations in the time domain using a parallel version of a 3D CE/SE solver. The effect of the gust orientation on the far-field directivity is studied. Numerical solutions are presented and compared with analytical solutions, showing a reasonable agreement.

Wang, X. Y.↗

Scheduling Tasks In Parallel Processing

Algorithms sought to minimize time and cost of computation. Report describes research on scheduling of computations tasks in system of multiple identical data processors operating in parallel. Computational intractability requires use of suboptimal heuristic algorithms. First algorithm called "list heuristic", variation of classical list scheduling. Second algorithm called "cluster heuristic" applied to tightly coupled tasks and consists of four phases. Third algorithm called "exchange heuristic", iterative-improvement algorithm beginning with initial feasible assignment of tasks to processors and periods of time. Fourth algorithm is iterative one for optimal assignment of tasks and based on concept called "simulated annealing" because of mathematical resemblance to aspects of physical annealing processes.

Price, Camille C.↗

Parallelized modelling and solution scheme for hierarchically scaled simulations

This two-part paper presents the results of a benchmarked analytical-numerical investigation into the operational characteristics of a unified parallel processing strategy for implicit fluid mechanics formulations. This hierarchical poly tree (HPT) strategy is based on multilevel substructural decomposition. The Tree morphology is chosen to minimize memory, communications and computational effort. The methodology is general enough to apply to existing finite difference (FD), finite element (FEM), finite volume (FV) or spectral element (SE) based computer programs without an extensive rewrite of code. In addition to finding large reductions in memory, communications, and computational effort associated with a parallel computing environment, substantial reductions are generated in the sequential mode of application. Such improvements grow with increasing problem size. Along with a theoretical development of general 2-D and 3-D HPT, several techniques for expanding the problem size that the current generation of computers are capable of solving, are presented and discussed. Among these techniques are several interpolative reduction methods. It was found that by combining several of these techniques that a relatively small interpolative reduction resulted in substantial performance gains. Several other unique features/benefits are discussed in this paper. Along with Part 1's theoretical development, Part 2 presents a numerical approach to the HPT along with four prototype CFD applications. These demonstrate the potential of the HPT strategy.

Padovan, Joe↗

A Novel Method for Characterization of Superconductors: Physical Measurements and Modeling of Thin Films

A method for characterization of granular superconducting thin films has been developed which encompasses both the morphological state of the sample and its fabrication process parameters. The broad scope of this technique is due to the synergism between experimental measurements and their interpretation using numerical simulation. Two novel technologies form the substance of this system: the magnetically modulated resistance method for characterizing superconductors; and a powerful new computer peripheral, the Parallel Information Processor card, which provides enhanced computing capability for PC computers. This enhancement allows PC computers to operate at speeds approaching that of supercomputers. This makes atomic scale simulations possible on low cost machines. The present development of this system involves the integration of these two technologies using mesoscale simulations of thin film growth. A future stage of development will incorporate atomic scale modeling.

Kim, B. F.↗

Parallel Monotonic Basin Hopping for Low Thrust Trajectory Optimization

Monotonic Basin Hopping has been shown to be an effective method of solving low thrust trajectory optimization problems. This paper outlines an extension to the common serial implementation by parallelizing it over any number of available compute cores. The Parallel Monotonic Basin Hopping algorithm described herein is shown to be an effective way to more quickly locate feasible solutions, and improve locally optimal solutions in an automated way without requiring a feasible initial guess. The increased speed achieved through parallelization enables the algorithm to be applied to more complex problems that would otherwise be impractical for a serial implementation. Low thrust cislunar transfers and a hybrid Mars example case demonstrate the effectiveness of the algorithm. Finally, a preliminary scaling study quantifies the expected decrease in solve time compared to a serial implementation.,

McCarty, Steven L.↗

Parallel Monotonic Basin Hopping for Low Thrust Trajectory Optimization

Monotonic Basin Hopping has been shown to be an effective method of solving low thrust trajectory optimization problems. This paper outlines an extension to the common serial implementation by parallelizing it over any number of available compute cores. The Parallel Monotonic Basin Hopping algorithm described herein is shown to be an effective way to more quickly locate feasible solutions, and improve locally optimal solutions in an automated way without requiring a feasible initial guess. The increased speed achieved through parallelization enables the algorithm to be applied to more complex problems that would otherwise be impractical for a serial implementation. Low thrust cislunar transfers and a hybrid Mars example case demonstrate the effectiveness of the algorithm. Finally, a preliminary scaling study quantifies the expected decrease in solve time compared to a serial implementation.

McCarty, Steven L.↗

Parallel language constructs for tensor product computations on loosely coupled architectures

Distributed memory architectures offer high levels of performance and flexibility, but have proven awkard to program. Current languages for nonshared memory architectures provide a relatively low level programming environment, and are poorly suited to modular programming, and to the construction of libraries. A set of language primitives designed to allow the specification of parallel numerical algorithms at a higher level is described. Tensor product array computations are focused on along with a simple but important class of numerical algorithms. The problem of programming 1-D kernal routines is focused on first, such as parallel tridiagonal solvers, and then how such parallel kernels can be combined to form parallel tensor product algorithms is examined.

Mehrotra, Piyush↗

An Open-Source Parallel EMT Simulation Framework

As the integration level of inverter-based resources (IBRs) increases, ensuring the reliable operation of the bulk power systems requires the use of electromagnetic transient (EMT) simulation tools to identify and mitigate system-wide stability risks. Conducting EMT studies for large-scale, IBR-rich grids, however, is challenging due to the inherent computational bottleneck caused by the underlying high-fidelity models and required small time steps. This paper introduces ParaEMT: an open-source, generic EMT simulation framework designed to accelerate simulations by leveraging advanced parallel computational technologies, such as high-performance computers. This paper presents a comprehensive exposition of ParaEMT, covering its modeling library, simulation strategy, framework structure, operational procedures, and auxiliary features, alongside its extensible parallel computational architecture. Notably, ParaEMT is a publicly accessible and modularized framework written in Python, thereby facilitating future development and the integration of new models and algorithms. The accuracy and efficiency of ParaEMT are demonstrated by rigorous validations via multiple case studies.

electromagnetic transient simulation↗

Dynamic Load Balancing For Grid Partitioning on a SP-2 Multiprocessor: A Framework

Computational requirements of full scale computational fluid dynamics change as computation progresses on a parallel machine. The change in computational intensity causes workload imbalance of processors, which in turn requires a large amount of data movement at runtime. If parallel CFD is to be successful on a parallel or massively parallel machine, balancing of the runtime load is indispensable. Here a framework is presented for dynamic load balancing for CFD applications, called Jove. One processor is designated as a decision maker Jove while others are assigned to computational fluid dynamics. Processors running CFD send flags to Jove in a predetermined number of iterations to initiate load balancing. Jove starts working on load balancing while other processors continue working with the current data and load distribution. Jove goes through several steps to decide if the new data should be taken, including preliminary evaluate, partition, processor reassignment, cost evaluation, and decision. Jove running on a single IBM SP2 node has been completely implemented. Preliminary experimental results show that the Jove approach to dynamic load balancing can be effective for full scale grid partitioning on the target machine IBM SP2.

Sohn, Andrew↗

Dynamic Load Balancing for Grid Partitioning on a SP-2 Multiprocessor: A Framework

Computational requirements of full scale computational fluid dynamics change as computation progresses on a parallel machine. The change in computational intensity causes workload imbalance of processors, which in turn requires a large amount of data movement at runtime. If parallel CFD is to be successful on a parallel or massively parallel machine, balancing of the runtime load is indispensable. Here a framework is presented for dynamic load balancing for CFD applications, called Jove. One processor is designated as a decision maker Jove while others are assigned to computational fluid dynamics. Processors running CFD send flags to Jove in a predetermined number of iterations to initiate load balancing. Jove starts working on load balancing while other processors continue working with the current data and load distribution. Jove goes through several steps to decide if the new data should be taken, including preliminary evaluate, partition, processor reassignment, cost evaluation, and decision. Jove running on a single EBM SP2 node has been completely implemented. Preliminary experimental results show that the Jove approach to dynamic load balancing can be effective for full scale grid partitioning on the target machine IBM SP2.

Sohn, Andrew↗