Search NASA⌕ Search

SEARCH · Search NASA

Results for “parallel hybrid”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

Asynchronous-many-task systems: Challenges and opportunities - Scaling an AMR astrophysics code on exascale machines using Kokkos and HPX

Dynamic and adaptive mesh refinement is pivotal in high-resolution, multi-physics, multi-model simulations, necessitating precise physics resolution in localized areas across expansive domains. Today’s supercomputers’ extreme heterogeneity presents a significant challenge for dynamically adaptive codes, highlighting the importance of achieving performance portability at scale. Our research focuses on astrophysical simulations, particularly stellar mergers, to elucidate early universe dynamics. Here, we present Octo-Tiger, leveraging Kokkos, HPX, and SIMD for portable performance at scale in complex, massively parallel adaptive multi-physics simulations. Octo-Tiger supports diverse processors, accelerators, and network backends. Experiments demonstrate exceptional scalability across several heterogeneous supercomputers including Perlmutter, Frontier, and Fugaku, encompassing major GPU architectures and x86, ARM, and RISC-V CPUs. Parallel efficiency of 47.59% (110,080 cores and 6880 hybrid A100 GPUs) on a full-system run on Perlmutter (26% HPCG peak performance) and 51.37% (using 32,768 cores and 2048 MI250X) on Frontier are achieved.

97 MATHEMATICS AND COMPUTING↗

Hot flow anomaly formation by magnetic deflection

Hot flow anomalies (HFAs) are localized plasma structures observed in the solar wind and magnetosheath near the earth's quasi-parallel bow shock. This paper presents one-dimensional hybrid computer simulations illustrating a formation mechanism for HFAs in which the single hot ion population results from a spatial separation of two counterstreaming ion beams. The higher-density cooler regions are dominated by the background (solar wind) ions, and the lower-density hotter internal regions are dominated by the beam ions. The spatial separation of the beam and background is caused by the deflection of the ions in large-amplitude magnetic fields which are generated by ion/ion streaming instabilities.

Onsager, T. G.↗

Parallel computing for probabilistic fatigue analysis

This paper presents the results of Phase I research to investigate the most effective parallel processing software strategies and hardware configurations for probabilistic structural analysis. We investigate the efficiency of both shared and distributed-memory architectures via a probabilistic fatigue life analysis problem. We also present a parallel programming approach, the virtual shared-memory paradigm, that is applicable across both types of hardware. Using this approach, problems can be solved on a variety of parallel configurations, including networks of single or multiprocessor workstations. We conclude that it is possible to effectively parallelize probabilistic fatigue analysis codes; however, special strategies will be needed to achieve large-scale parallelism to keep large number of processors busy and to treat problems with the large memory requirements encountered in practice. We also conclude that distributed-memory architecture is preferable to shared-memory for achieving large scale parallelism; however, in the future, the currently emerging hybrid-memory architectures will likely be optimal.

Sues, Robert H.↗

Nonlinear Landau damping and Alfven wave dissipation

Nonlinear Landau damping has been often suggested to be the cause of the dissipation of Alfven waves in the solar wind as well as the mechanism for ion heating and selective preacceleration in solar flares. We discuss the viability of these processes in light of our theoretical and numerical results. We present one-dimensional hybrid plasma simulations of the nonlinear Landau damping of parallel Alfven waves. In this scenario, two Alfven waves nonresonantly combine to create second-order magnetic field pressure gradients, which then drive density fluctuations, which in turn drive a second-order longitudinal electric field. Under certain conditions, this electric field strongly interacts with the ambient ions via the Landau resonance which leads to a rapid dissipation of the Alfven wave energy. While there is a net flux of energy from the waves to the ions, one of the Alfven waves will grow if both have the same polarization. We compare damping and growth rates from plasma simulations with those predicted by Lee and Volk (1973), and also discuss the evolution of the ambient ion distribution. We then consider this nonlinear interaction in the presence of a spectrum of Alfven waves, and discuss the spectrum's influence on the growth or damping of a single wave. We also discuss the implications for wave dissipation and ion heating in the solar wind.

Vinas, Adolfo F.↗

Aerodynamic Shape Optimization Using Hybridized Differential Evolution

An aerodynamic shape optimization method that uses an evolutionary algorithm known at Differential Evolution (DE) in conjunction with various hybridization strategies is described. DE is a simple and robust evolutionary strategy that has been proven effective in determining the global optimum for several difficult optimization problems. Various hybridization strategies for DE are explored, including the use of neural networks as well as traditional local search methods. A Navier-Stokes solver is used to evaluate the various intermediate designs and provide inputs to the hybrid DE optimizer. The method is implemented on distributed parallel computers so that new designs can be obtained within reasonable turnaround times. Results are presented for the inverse design of a turbine airfoil from a modern jet engine. (The final paper will include at least one other aerodynamic design application). The capability of the method to search large design spaces and obtain the optimal airfoils in an automatic fashion is demonstrated.

Madavan, Nateri K.↗

Parallel Anisotropic Tetrahedral Adaptation

An adaptive method that robustly produces high aspect ratio tetrahedra to a general 3D metric specification without introducing hybrid semi-structured regions is presented. The elemental operators and higher-level logic is described with their respective domain-decomposed parallelizations. An anisotropic tetrahedral grid adaptation scheme is demonstrated for 1000-1 stretching for a simple cube geometry. This form of adaptation is applicable to more complex domain boundaries via a cut-cell approach as demonstrated by a parallel 3D supersonic simulation of a complex fighter aircraft. To avoid the assumptions and approximations required to form a metric to specify adaptation, an approach is introduced that directly evaluates interpolation error. The grid is adapted to reduce and equidistribute this interpolation error calculation without the use of an intervening anisotropic metric. Direct interpolation error adaptation is illustrated for 1D and 3D domains.

Park, Michael A.↗

T RI M E ++: Multi-threaded triangular meshing in two dimensions

We present T RI M E ++, a multi-threaded software library designed for generating two-dimensional meshes for intricate geometric shapes using the Delaunay triangulation. Multi-threaded parallel computing is implemented throughout the meshing procedure, making it suitable for fast generation of large-scale meshes. Three iterative meshing algorithms are implemented: the DistMesh algorithm, the centroidal Voronoi diagram meshing, and a hybrid of the two. We compare the performance of the three meshing methods in T RI M E ++, and show that the hybrid method retains the advantages of the other two. The software library achieves significant parallel speedup when generating large-scale meshes containing between 10 4 to 10 7 points. T RI M E ++ can handle complicated geometries and generates adaptive meshes of high quality.

97 MATHEMATICS AND COMPUTING↗

V/STOL tilt rotor aircraft study mathematical model for a real time simulation of a tilt rotor aircraft (Boeing Vertol Model 222), volume 8

This report documents the development of a real time mathematical model of a tilt rotor aircraft. This mathematical model is to be used in conjunction with the NASA Flight Simulator for Advanced Aircraft (FSAA) at Ames Research Center for evaluation of aircraft performance and handling qualities. In addition to developing the mathematical model, a parallel programming effort was conducted utilizing Boeing-Vertol's Hybrid Simulation Laboratory for the purpose of developing and evaluating model simplification. The mathematical model is an eleven degree of freedom total force model. This model includes the basic six degree of freedom rigid body .outer loop equations written about the instantaneous center of gravity with the inertial and aerodynamic terms included. The rotor is treated as a point source of forces and moments with appropriate response time lags and actuator dynamics. The wing has one vertical bending and one wing torsion degree of freedom. These structural degrees of freedom are treated on a "quasistatic" basis; i.e., the natural frequencies of vibration of the structure are much higher than the· frequencies of the rigid body motion, and the coupling is in the aerodynamic terms. Each nacelle has an independent pitch degree of freedom about the wing pivot. The aerodynamics of the wing, tail, rotors, landing gear and fuselage are included. Wing and tail mutual interference effects and turbine engine performance and dynamic responses' are represented.

H Rosenstein↗

Ion heating in the cusp

Data from satellite observations and theoretical simulations of ion heating in the magnetospheric cusp region are compiled in tables, graphs, and diagrams and discussed. Consideration is given to the mixing of ionospheric and magnetosheath plasmas, the instability of downward-flowing ring distributions of H(+) and He(2+) to lower-hybrid waves, and oxygen and hydrogen heating at finite k(parallel). A range of unstable propagation angles of + or - 20 deg about the perpendicular is estimated for M(H)/M(e) = 50, including superthermal and background electron dynamics.

Hudson, M. K.↗

Acquisition Of Spread-Spectrum Code

Effects of Doppler shift and data modulation taken into account. Two advanced schemes for acquisition of direct-sequence spread-spectrum codes proposed. M1-Lag correlator in each strip of spread-spectrum-code detector operates at different offset code-chip time. Each offset represents assumed (tentative) Doppler shift. Schemes have highly parallel architecture implemented with currently available technology. Possible to use hybrid parallel/serial architecture in which acquisition time varies in inverse proportion to number of correlators and fast-Fourier-transform processors.

Cheng, Unjeng↗

Steepening of parallel propagating hydromagnetic waves into magnetic pulsations - A simulation study

The steepening mechanism of parallel propagating low-frequency MHD-like waves observed upstream of the earth's quasi-parallel bow shock has been investigated by means of electromagnetic hybrid simulations. It is shown that an ion beam through the resonant electromagnetic ion/ion instability excites large-amplitude waves, which consequently pitch angle scatter, decelerate, and eventually magnetically trap beam ions in regions where the wave amplitudes are largest. As a result, the beam ions become bunched in both space and gyrophase. As these higher-density, nongyrotropic beam segments are formed, the hydromagnetic waves rapidly steepen, resulting in magnetic pulsations, with properties generally in agreement with observations. This steepening process operates on the scale of the linear growth time of the resonant ion/ion instability. Many of the pulsations generated by this mechanism are left-hand polarized in the spacecraft frame.

Akimoto, K.↗

Structure of medium Mach number quasi-parallel shocks - Upstream and downstream waves

The transition from steady low-Mach-number to unsteady high-Mach-number quasi-parallel shocks was investigated by performing large-scale 1D hybrid code simulations at increasing Mach numbers. It was found that only at very low Mach number shocks the steepening is limited by upstream phase-standing whistlers, as predicted by the classical theory (Tidman and Northrop, 1968). In the intermediate region of Mach numbers between 1.5 and 3.5, a very diverse behavior is observed. Backstreaming ions generate fast magnetosonic waves which dominate the upstream, with wavelengths longer than phase-standing whistlers. At increasing Mach numbers, the phase and group velocities of the dominant waves are reduced until they point back toward the shock; when there is sufficient energy flux in these waves, they lead to unsteady shock behavior and eventually to shock reformation.

Krauss-Varban, D.↗

Experimental studies of the properties of 'simulated' upstream turbulence using a statistical multipoint method

In this report we present a different approach to the multipoint measurement of magnetic fields and plasma. This is called the multi-spacecraft ensemble technique (MET), essentially free of process restrictions, such as linearity and stationarity. We comprehensively discuss the other conditions and limitations intrinsic to this statistical method. We also show the results of the application of the ensemble method to the synthetic data obtained from a hybrid simulation in the region upstream of a quasi-parallel shock. The important implications of the above approach for the CLUSTER mission are discussed.

Orlowski, D. S.↗

Mechanisms for the Dissipation of Alfven Waves in Near-Earth Space Plasma

Alfven waves are a major mechanism for the transport of electromagnetic energy from the distant part of the magnetosphere to the near-Earth space. This is especially true for the auroral and polar regions of the Earth. However, the mechanisms for their dissipation have remained illusive. One of the mechanisms is the formation of double layers when the current associated with Alfven waves in the inertial regime interact with density cavities, which either are generated nonlinearly by the waves themselves or are a part of the ambient plasma turbulence. Depending on the strength of the cavities, weak and strong double layers could form. Such double layers are transient; their lifetimes depend on that of the cavities. Thus they impulsively accelerate ions and electrons. Another mechanism is the resonant absorption of broadband Alfven- wave noise by the ions at the ion cyclotron frequencies. But this resonant absorption may not be possible for the very low frequency waves, and it may be more suited for electromagnetic ion cyclotron waves. A third mechanism is the excitation of secondary waves by the drifts of electrons and ions in the Alfven wave fields. It is found that under suitable conditions, the relative drifts between different ion species and/or between electrons and ions are large enough to drive lower hybrid waves, which could cause transverse accelerations of ions and parallel accelerations of electrons. This mechanism is being further studied by means of kinetic simulations using 2.5- and 3-D particle-in-cell codes. The ongoing modeling efforts on space weather require quantitative estimates of energy inputs of various kinds, including the electromagnetic energy. Our studies described here contribute to the methods of determining the estimates of the input from ubiquitous Alfven waves.

Singh, Nagendra↗

Thermal Vibrational Convection in a Two-phase Stratified Liquid

The response of a two-phase stratified liquid system subject to a vibration parallel to an imposed temperature gradient is analyzed using a hybrid thermal lattice Boltzmann method (HTLB). The vibrations considered correspond to sinusoidal translations of a rigid cavity at a fixed frequency. The layers are thermally and mechanically coupled. Interaction between gravity-induced and vibration-induced thermal convection is studied. The ability of applied vibration to enhance the flow, heat transfer and interface distortion is investigated. For the range of conditions investigated, the results reveal that the effect of vibrational Rayleigh number and vibrational frequency on a two-phase stratified fluid system is much different than that for a single-phase fluid system. Comparisons of the response of a two-phase stratified fluid system with a single-phase fluid system are discussed.

Chang, Qingming↗

Sparse Linear Algebra Toolkit for Computational Aerodynamics

Finding solutions to sparse linear systems of equations is an essential step in Computational Engineering applications of interest to NASA. Linear systems of equations are composed and solved in almost every computational engineering application. The characteristics of linear systems vary greatly from one application to another. Accordingly, there are a wide variety of methods for the solution of linear systems of equations. The operations and methods prepared by the authors are focused on linear systems of interest to NASA, primarily those associated with Computational Fluid Dynamics (CFD), Aeroelasticity, and Aeroacoustics. The Sparse Linear Algebra Toolkit (SLAT) is a coordinated collection of software featuring operations, methods, and data structures that are useful when solving sparse linear systems of equations on modern computer architectures. The implemented operations and methods are designed and tuned for parallelism in shared memory, in distributed memory, and across the hybrid combination of distributed-shared memory. The toolkit includes novel methods and implementations for modern architectures and facilitates development of new approaches for meeting NASA’s evolving computational engineering challenges using evolving computer architectures that are not available in vendor libraries. In this paper, significant features and interfaces within SLAT are presented and verified for simulations performed with NASA’s CFD solver, FUN3D. The runtime and scaling performance of the Generalized Minimum Residual (GMRES) method implemented in SLAT is analyzed for the linear subproblems within the solution of turbulent Navier-Stokes equations employed in the simulation of high-lift configurations. Prior to this work, the SPARSKIT GMRES implementation was the only Krylov subspace method available within FUN3D. A strong scaling study shows the SLAT GMRES implementation facilitates accurate Reynolds-averaged Navier-Stokes CFD solutions between 15% and 56% faster than the SPARSKIT GMRES implementation.

Stephen L Wood↗

Parallel Domain Decomposition Formulation and Software for Large-Scale Sparse Symmetrical/Unsymmetrical Aeroacoustic Applications

The overall objectives of this research work are to formulate and validate efficient parallel algorithms, and to efficiently design/implement computer software for solving large-scale acoustic problems, arised from the unified frameworks of the finite element procedures. The adopted parallel Finite Element (FE) Domain Decomposition (DD) procedures should fully take advantages of multiple processing capabilities offered by most modern high performance computing platforms for efficient parallel computation. To achieve this objective. the formulation needs to integrate efficient sparse (and dense) assembly techniques, hybrid (or mixed) direct and iterative equation solvers, proper pre-conditioned strategies, unrolling strategies, and effective processors' communicating schemes. Finally, the numerical performance of the developed parallel finite element procedures will be evaluated by solving series of structural, and acoustic (symmetrical and un-symmetrical) problems (in different computing platforms). Comparisons with existing "commercialized" and/or "public domain" software are also included, whenever possible.

Nguyen, D. T.↗

Accelerating Climate and Weather Simulations through Hybrid Computing

Unconventional multi- and many-core processors (e.g. IBM (R) Cell B.E.(TM) and NVIDIA (R) GPU) have emerged as effective accelerators in trial climate and weather simulations. Yet these climate and weather models typically run on parallel computers with conventional processors (e.g. Intel, AMD, and IBM) using Message Passing Interface. To address challenges involved in efficiently and easily connecting accelerators to parallel computers, we investigated using IBM's Dynamic Application Virtualization (TM) (IBM DAV) software in a prototype hybrid computing system with representative climate and weather model components. The hybrid system comprises two Intel blades and two IBM QS22 Cell B.E. blades, connected with both InfiniBand(R) (IB) and 1-Gigabit Ethernet. The system significantly accelerates a solar radiation model component by offloading compute-intensive calculations to the Cell blades. Systematic tests show that IBM DAV can seamlessly offload compute-intensive calculations from Intel blades to Cell B.E. blades in a scalable, load-balanced manner. However, noticeable communication overhead was observed, mainly due to IP over the IB protocol. Full utilization of IB Sockets Direct Protocol and the lower latency production version of IBM DAV will reduce this overhead.

hybrid computing↗