Search NASA⌕ Search

SEARCH · Search NASA

Results for “Partitioned scheme”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Global Load Balancing with Parallel Mesh Adaption on Distributed-Memory Systems

Dynamic mesh adaptation on unstructured grids is a powerful tool for efficiently computing unsteady problems to resolve solution features of interest. Unfortunately, this causes load inbalances among processors on a parallel machine. This paper described the parallel implementation of a tetrahedral mesh adaption scheme and a new global load balancing method. A heuristic remapping algorithm is presented that assigns partitions to processors such that the redistribution coast is minimized. Results indicate that the parallel performance of the mesh adaption code depends on the nature of the adaption region and show a 35.5X speedup on 64 processors of an SP2 when 35 percent of the mesh is randomly adapted. For large scale scientific computations, our load balancing strategy gives an almost sixfold reduction in solver execution times over non-balanced loads. Furthermore, our heuristic remappier yields processor assignments that are less than 3 percent of the optimal solutions, but requires only 1 percent of the computational time.

Biswas, Rupak↗

MATRIX-VBS Condensing Organic Aerosols in an Aerosol Microphysics Model

The condensation of organic aerosols is represented in a newly developed box-model scheme, where its effect on the growth and composition of particles are examined. We implemented the volatility-basis set (VBS) framework into the aerosol mixing state resolving microphysical scheme Multiconfiguration Aerosol TRacker of mIXing state (MATRIX). This new scheme is unique and advances the representation of organic aerosols in models in that, contrary to the traditional treatment of organic aerosols as non-volatile in most climate models and in the original version of MATRIX, this new scheme treats them as semi-volatile. Such treatment is important because low-volatility organics contribute significantly to the growth of particles. The new scheme includes several classes of semi-volatile organic compounds from the VBS framework that can partition among aerosol populations in MATRIX, thus representing the growth of particles via condensation of low volatility organic vapors. Results from test cases representing Mexico City and a Finish forrest condistions show good representation of the time evolutions of concentration for VBS species in the gas phase and in the condensed particulate phase. Emitted semi-volatile primary organic aerosols evaporate almost completely in the high volatile range, and they condense more efficiently in the low volatility range.

Aerosols↗

An efficient massively parallel Euler solver for unstructured grids

A data parallel mesh-vertex upwind finite-volume scheme for solving the Euler equations on triangular unstructured meshes is described. A novel vertex-based partitioning of the problem is introduced which minimizes the computation and communication costs associated with distributing the computation to the processors of a massively parallel computer. Finally, the performance of this unstructured computation on 8K processors of the Connection Machine CM-2 is compared with one processor of a Cray-YMP. The experiments show that 8K processors of the CM-2 achieve approximately 70 percent of the performance of one processor of the Cray-YMP on the unstructured mesh computations described here.

Hammond, Steven W.↗

A Large-Grain Mapping Approach for Multiprocessor Systems Through Data Flow Model Ph.D. Thesis

A large-grain level mapping method is presented of numerical oriented applications onto multiprocessor systems. The method is based on the large-grain data flow representation of the input application and it assumes a general interconnection topology of the multiprocessor system. The large-grain data flow model was used because such representation best exhibits inherited parallelism in many important applications, e.g., CFD models based on partial differential equations can be presented in large-grain data flow format, very effectively. A generalized interconnection topology of the multiprocessor architecture is considered, including such architectural issues as interprocessor communication cost, with the aim to identify the 'best matching' between the application and the multiprocessor structure. The objective is to minimize the total execution time of the input algorithm running on the target system. The mapping strategy consists of the following: (1) large-grain data flow graph generation from the input application using compilation techniques; (2) data flow graph partitioning into basic computation blocks; and (3) physical mapping onto the target multiprocessor using a priority allocation scheme for the computation blocks.

Kim, Hwa-Soo↗

A Fully Implicit Time Accurate Method for Hypersonic Combustion: Application to Shock-induced Combustion Instability

A new fully implicit, time accurate algorithm suitable for chemically reacting, viscous flows in the transonic-to-hypersonic regime is described. The method is based on a class of Total Variation Diminishing (TVD) schemes and uses successive Gauss-Siedel relaxation sweeps. The inversion of large matrices is avoided by partitioning the system into reacting and nonreacting parts, but still maintaining a fully coupled interaction. As a result, the matrices that have to be inverted are of the same size as those obtained with the commonly used point implicit methods. In this paper we illustrate the applicability of the new algorithm to hypervelocity unsteady combustion applications. We present a series of numerical simulations of the periodic combustion instabilities observed in ballistic-range experiments of blunt projectiles flying at subdetonative speeds through hydrogen-air mixtures. The computed frequencies of oscillation are in excellent agreement with experimental data.

Yungster, Shaye↗

Simulations of Turbine Cooling Flows Using a Multiblock-Multigrid Scheme

Results from numerical simulations of air flow and heat transfer in a 'branched duct' geometry are presented. The geometry contains features, including pins and a partition, as are found in coolant passages of turbine blades. The simulations were performed using a multi-block structured grid system and a finite volume discretization of the governing equations (the compressible Navier-Stokes equations). The effects of turbulence on the mean flow and heat transfer were modeled using the Baldwin-Lomax turbulence model. The computed results are compared to experimental data. It was found that the extent of some regions of high heat transfer was somewhat under predicted. It is conjectured that the underlying reason is the local nature of the turbulence model which cannot account for upstream influence on the turbulence field. In general, however, the comparison with the experimental data is favorable.

Steinthorsson, Erlendur↗

Performance studies of the multigrid algorithms implemented on hypercube multiprocessor systems

In this paper, we analyze and compare the performance on a hypercube multiprocessor of some of the major multigrid techniques used in practice. The model problem considered here is that of solving the 2-D incompressible Navier-Stokes equations representing the flow between two parallel plates. Results obtained by implementing the different multigrid schemes on an iPSC are presented. Effects on the overall performance of various parameters of the algorithms, of the partitioning strategies employed, and of some of the characteristics of the underlying architecture are discussed.

Naik, Vijay K.↗

Propulsion system performance resulting from an Integrated Flight/Propulsion Control design

Propulsion system specific results are presented from the application of the Integrated Methodology for Propulsion and Airframe Control (IMPAC) design approach to Integrated Flight/Propulsion Control design for a STOVL aircraft in transition flight. The IMPAC method is briefly discussed and the propulsion system specifications for the integrated control design are examined. The structure of a linear engine controller that results from partitioning a linear centralized controller is discussed. The details of a nonlinear propulsion control system are presented, including a scheme to protect the engine operational limits: the fan surge margin and the acceleration/deceleration schedule which limits the fuel flow. Also, a simple but effective multivariable integrator windup protection scheme is investigated. Nonlinear closed-loop simulation results are presented for two typical pilot commands for transition flight: acceleration while maintaining flight path angle and a change in flight path angle while maintaining airspeed. The simulation nonlinearities include the airframe/engine coupling, the actuator and sensor dynamics and limits, the protection scheme for the engine operational limits, and the integrator windup protection. Satisfactory performance of the total airframe plus engine system for transition flight, as defined by the specifications, is maintained during the limit operation of the closed-loop engine subsystem.

Mattern, Duane↗

Propulsion system performance resulting from an integrated flight/propulsion control design

Propulsion-system-specific results are presented from the application of the integrated methodology for propulsion and airframe control (IMPAC) design approach to integrated flight/propulsion control design for a 'short takeoff and vertical landing' (STOVL) aircraft in transition flight. The IMPAC method is briefly discussed and the propulsion system specifications for the integrated control design are examined. The structure of a linear engine controller that results from partitioning a linear centralized controller is discussed. The details of a nonlinear propulsion control system are presented, including a scheme to protect the engine operational limits: the fan surge margin and the acceleration/deceleration schedule that limits the fuel flow. Also, a simple but effective multivariable integrator windup protection scheme is examined. Nonlinear closed-loop simulation results are presented for two typical pilot commands for transition flight: acceleration while maintaining flightpath angle and a change in flightpath angle while maintaining airspeed. The simulation nonlinearities include the airframe/engine coupling, the actuator and sensor dynamics and limits, the protection scheme for the engine operational limits, and the integrator windup protection. Satisfactory performance of the total airframe plus engine system for transition flight, as defined by the specifications, was maintained during the limit operation of the closed-loop engine subsystem.

Mattern, Duane↗

The Use of Indirect Estimates of Soil Moisture to Initialize Coupled Models and its Impact on Short-Term and Seasonal Simulations

It is well known that soil moisture is a characteristic of the land surface that strongly affects the partitioning of outgoing radiation into sensible and latent heat which significantly impacts both weather and climate. Detailed land surface schemes are now being coupled to mesoscale atmospheric models in order to represent the effect of soil moisture upon atmospheric simulations. However, there is little direct soil moisture data available to initialize these models on regional to continental scales. As a result, a Soil Hydrology Model (SHM) is currently being used to generate an indirect estimate of the soil moisture conditions over the continental United States at a grid resolution of 36 Km on a daily basis since 8 May 1995. The SHM is forced by analyses of atmospheric observations including precipitation and contains detailed information on slope soil and landcover characteristics.The purpose of this paper is to evaluate the utility of initializing a detailed coupled model with the soil moisture data produced by SHM.

Lapenta, William M.↗

Computational analysis of methods for reduction of induced drag

The purpose of this effort was to perform a computational flow analysis of a design concept centered around induced drag reduction and tip-vortex energy recovery. The flow model solves the unsteady three-dimensional Euler equations, discretized as a finite-volume method, utilizing a high-resolution approximate Riemann solver for cell interface flux definitions. The numerical scheme is an approximately-factored block LU implicit Newton iterative-refinement method. Multiblock domain decomposition is used to partition the field into an ordered arrangement of blocks. Three configurations are analyzed: a baseline fuselage-wing, a fuselage-wing-nacelle, and a fuselage-wing-nacelle-propfan. Aerodynamic force coefficients, propfan performance coefficients, and flowfield maps are used to qualitatively access design efficacy. Where appropriate, comparisons are made with available experimental data.

Janus, J. M.↗

Assimilation of Gridded Terrestrial Water Storage Observations from GRACE into a Land Surface Model

Observations of terrestrial water storage (TWS) from the Gravity Recovery and Climate Experiment (GRACE) satellite mission have a coarse resolution in time (monthly) and space (roughly 150,000 km(sup 2) at midlatitudes) and vertically integrate all water storage components over land, including soil moisture and groundwater. Data assimilation can be used to horizontally downscale and vertically partition GRACE-TWS observations. This work proposes a variant of existing ensemble-based GRACE-TWS data assimilation schemes. The new algorithm differs in how the analysis increments are computed and applied. Existing schemes correlate the uncertainty in the modeled monthly TWS estimates with errors in the soil moisture profile state variables at a single instant in the month and then apply the increment either at the end of the month or gradually throughout the month. The proposed new scheme first computes increments for each day of the month and then applies the average of those increments at the beginning of the month. The new scheme therefore better reflects submonthly variations in TWS errors. The new and existing schemes are investigated here using gridded GRACE-TWS observations. The assimilation results are validated at the monthly time scale, using in situ measurements of groundwater depth and soil moisture across the U.S. The new assimilation scheme yields improved (although not in a statistically significant sense) skill metrics for groundwater compared to the open-loop (no assimilation) simulations and compared to the existing assimilation schemes. A smaller impact is seen for surface and root-zone soil moisture, which have a shorter memory and receive smaller increments from TWS assimilation than groundwater. These results motivate future efforts to combine GRACE-TWS observations with observations that are more sensitive to surface soil moisture, such as L-band brightness temperature observations from Soil Moisture Ocean Salinity (SMOS) or Soil Moisture Active Passive (SMAP). Finally, we demonstrate that the scaling parameters that are applied to the GRACE observations prior to assimilation should be consistent with the land surface model that is used within the assimilation system.

GRACE↗

A Partitioned - Task Parallel Implementation of the NASA Multiscale Analysis Tool for High Performance Computing

The NASA Multiscale Analysis Tool (NASMAT) is a platform for multiscale modeling of composites which can perform analysis of materials with any arbitrary number of length scales. The platform supports modularity, scalability, and interoperability using recursive procedures and data structures. A Macro solver driven parallelization scheme often limits the capability of NASMAT to scale as it has access to limited memory and number of cores (often one core/thread) and often forces to implement macro solver specific changes to the platform. In this work, a partitioned task-parallel approach is adopted, where the parallelization strategy adopted for NASMAT is independent of the macro solver and the computational resources are managed independently. The programming architecture takes into account the hierarchy of multiple scales (task-dependence) and the heterogeneous nature (dynamic load balancing) of computation through implementation of a hierarchy-informed task parallel model. The partitioned nature of the framework further extends the “plug and play” capability of NASMAT. preCICE, an open-source library for coupling multiphysics solver in a partitioned manner, is adopted to integrate NASMAT with an external macro solver by implementing a NASMAT adapter for preCICE. Speedup and scalability of the framework is studied for micromechanical models of varying size.

task-parallel↗

Automated Classification of Thermal Infrared Spectra Using Self-organizing Maps

Existing and planned space missions to a variety of planetary and satellite surfaces produce an ever increasing volume of spectral data. Understanding the scientific informational content in this large data volume is a daunting task. Fortunately various statistical approaches are available to assess such data sets. Here we discuss an automated classification scheme based on Kohonen Self-organizing maps (SOM) we have developed. The SUM process produces an output layer were spectra having similar properties lie in close proximity to each other. One major effort is partitioning this output layer into appropriate regions. This is prefonned by defining dosed regions based upon the strength of the boundaries between adjacent cells in the SOM output layer. We use the Davies-Bouldin index as a measure of the inter-class similarities and intra-class dissimilarities that determines the optimum partition of the output layer, and hence number of SOM clusters. This allows us to identify the natural number of clusters formed from the spectral data. Mineral spectral libraries prepared at Arizona State University (ASU) and John Hopkins University (JHU) are used to test and evaluate the classification scheme. We label the library sample spectra in a hierarchical scheme with class, subclass, and mineral group names. We use a portion of the spectra to train the SOM, i.e. produce the output layer, while the remaining spectra are used to test the SOM. The test spectra are presented to the SOM output layer and assigned membership to the appropriate cluster. We then evaluate these assignments to assess the scientific meaning and accuracy of the derived SOM classes as they relate to the labels. We demonstrate that unsupervised classification by SOMs can be a useful component in autonomous systems designed to identify mineral species from reflectance and emissivity spectra in the therrnal IR.

Roush, Ted L.↗

RANS-MP: A Portable Parallel Navier-Stokes Solver

RANS-MP, a new implementation of a single-grid Navier-Stokes solver using the diagonalized Beam-Warming approximate-factorization scheme, is presented. This first release of the completely rewritten solver employs the following optimizations: (1) Bi-directional multi-partition method for the ADI solver part; this improves granularity and load balance; (2) Improved cache usage through elimination of non-unit-stride array access (possible in part due to multi-partitioning); (3) Preprocessing of communicating boundary conditions to streamline logic during time stepping; (4) Truly parallel, high-performance I/O using the newly-developed MPI-IO library; (5) Elimination of large amounts of redundant operations through efficient use of workspace. Results of some realistic wing computations on the IBM SP2 computer will be presented. We will demonstrate that excellent absolute performance and scalability are obtained with RANS-MP, even for relatively small grid sizes. Besides high performance, an outstanding feature of RANS-MP is its true portability, due to the use of the portable message passing and I/O libraries MPI and MPI-IO.

VanderWijngaart, Rob F.↗

Steady and transient least square solvers for thermal problems

This paper develops a hierarchical least square solution algorithm for highly nonlinear heat transfer problems. The methodology's capability is such that both steady and transient implicit formulations can be handled. This includes problems arising from highly nonlinear heat transfer systems modeled by either finite-element or finite-difference schemes. The overall procedure developed enables localized updating, iteration, and convergence checking as well as constraint application. The localized updating can be performed at a variety of hierarchical levels, i.e., degree of freedom, substructural, material-nonlinear groups, and/or boundary groups. The choice of such partitions can be made via energy partitioning or nonlinearity levels as well as by user selection. Overall, this leads to extremely robust computational characteristics. To demonstrate the methodology, problems are drawn from nonlinear heat conduction. These are used to quantify the robust capabilities of the hierarchical least square scheme.

Padovan, Joe↗

Flux vector splitting and approximate Newton methods

In the present investigation, the basic approach is employed to view an iterative scheme as Newton's method or as a modified Newton's method. Attention is given to various modified Newton methods which can arise from differencing schemes for the Euler equations. Flux vector splitting is considered as the basic spatial differencing technique. This technique is based on the partition of a flux vector into groups which have certain properties. The Euler equations fluxes can be split into two groups, the first group having a flux Jacobian with all positive eigenvalues, and the second group having a flux Jacobian with all negative eigenvalues. Flux vector splitting based on a velocity-sound speed split is considered along with the use of numerical techniques to analyze nonlinear systems, and the steady Euler equations for quasi-one-dimensional flow in a nozzle. Results are given for steady flows with shocks.

Jespersen, D. C.↗