Search NASA⌕ Search

SEARCH · Search NASA

Results for “Massively Parallel Simulations”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

137 records · Page 8

Computational chaos in massively parallel neural networks

A fundamental issue which directly impacts the scalability of current theoretical neural network models to massively parallel embodiments, in both software as well as hardware, is the inherent and unavoidable concurrent asynchronicity of emerging fine-grained computational ensembles and the possible emergence of chaotic manifestations. Previous analyses attributed dynamical instability to the topology of the interconnection matrix, to parasitic components or to propagation delays. However, researchers have observed the existence of emergent computational chaos in a concurrently asynchronous framework, independent of the network topology. Researcher present a methodology enabling the effective asynchronous operation of large-scale neural networks. Necessary and sufficient conditions guaranteeing concurrent asynchronous convergence are established in terms of contracting operators. Lyapunov exponents are computed formally to characterize the underlying nonlinear dynamics. Simulation results are presented to illustrate network convergence to the correct results, even in the presence of large delays.

Barhen, Jacob↗

Improving NASA's Multiscale Modeling Framework for Tropical Cyclone Climate Study

One of the current challenges in tropical cyclone (TC) research is how to improve our understanding of TC interannual variability and the impact of climate change on TCs. Recent advances in global modeling, visualization, and supercomputing technologies at NASA show potential for such studies. In this article, the authors discuss recent scalability improvement to the multiscale modeling framework (MMF) that makes it feasible to perform long-term TC-resolving simulations. The MMF consists of the finite-volume general circulation model (fvGCM), supplemented by a copy of the Goddard cumulus ensemble model (GCE) at each of the fvGCM grid points, giving 13,104 GCE copies. The original fvGCM implementation has a 1D data decomposition; the revised MMF implementation retains the 1D decomposition for most of the code, but uses a 2D decomposition for the massive copies of GCEs. Because the vast majority of computation time in the MMF is spent computing the GCEs, this approach can achieve excellent speedup without incurring the cost of modifying the entire code. Intelligent process mapping allows differing numbers of processes to be assigned to each domain for load balancing. The revised parallel implementation shows highly promising scalability, obtaining a nearly 80-fold speedup by increasing the number of cores from 30 to 3,335.

tropical cyclone interannual variability↗

Toward Automatic Scalability Analysis of Message Passing Programs: A Case Study

Scalability analysis forms an important component of any performance debugging cycle, for massively parallel machines. However, tools that help in performing such analysis for parallel programs are non-existent. The primary reason for lack of such tools is the complexity involved in capturing program dynamics such as communication-computation overlap, communication latencies and memory hierarchy reference patterns. In this paper, we highlight some simple techniques that can be used to study scalability of explicit message-passing parallel programs that consider the above issues. We start from the high level source code and use a methodology for deducing communication characteristics and its impact on the total execution time of the program. The approach is validated with the help of a pipelined method for solving scalar tri-diagonal systems, using both simulations and symbolic cost models on the Intel hypercube.

Sarukkai, Sekhar R.↗

Digital Technologies at NASA for Science and Engineering

While scientific and engineering advancements used to rely primarily on theoretical studies and physical experiments, today digital technology enabled by petaflops-scale supercomputers is an equal, if not a greater, contributor to such achievements. In addition, computational modeling and simulation serves as a predictive tool that is not otherwise available. As a result, the use of high performance computing is integral to NASA's work in all mission areas such as space exploration, aeronautics, and scientific discovery. But traditional supercomputing alone is not sufficient for all of the space agency's needs. The success of many NASA missions depends on solving complex computing challenges, some of which are NP-hard (decision theory) if using classical solution methods. Quantum computing promises an unprecedented ability to solve such intractable problems by harnessing quantum mechanical effects such as tunneling, superposition, and entanglement. Another disruptive digital technology is neuromorphic computing that uses brain-inspired lessons to generate new architectures that are much more energy efficient, and capable of massive parallel processing and learning in-situ. Finally, with large amounts of observational and computational data sets, the opportunities of big data and data analytics can be leveraged to enable deep learning and knowledge discovery - it's all a massive digital transformation. This talk will be an overview how NASA utilizes digital technologies for its science and engineering efforts.

Biswas, Rupak↗

High-performance parallel analysis of coupled problems for aircraft propulsion

This research program deals with the application of high-performance computing methods to the numerical simulation of complete jet engines. The program was initiated in 1993 by applying two-dimensional parallel aeroelastic codes to the interior gas flow problem of a by-pass jet engine. The fluid mesh generation, domain decomposition and solution capabilities were successfully tested. Attention was then focused on methodology for the partitioned analysis of the interaction of the gas flow with a flexible structure and with the fluid mesh motion driven by these structural displacements. The latter is treated by an ALE technique that models the fluid mesh motion as that of a fictitious mechanical network laid along the edges of near-field fluid elements. New partitioned analysis procedures to treat this coupled 3-component problem were developed in 1994. These procedures involved delayed corrections and subcycling, and have been successfully tested on several massively parallel computers. For the global steady-state axisymmetric analysis of a complete engine we have decided to use the NASA-sponsored ENG10 program, which uses a regular FV-multiblock-grid discretization in conjunction with circumferential averaging to include effects of blade forces, loss, combustor heat addition, blockage, bleeds and convective mixing. A load-balancing preprocessor for parallel versions of ENG10 has been developed. It is planned to use the steady-state global solution provided by ENG10 as input to a localized three-dimensional FSI analysis for engine regions where aeroelastic effects may be important.

Felippa, C. A.↗

Analytical Assessment of Simultaneous Parallel Approach Feasibility from Total System Error

In a simultaneous paired approach to closely-spaced parallel runways, a pair of aircraft flies in close proximity on parallel approach paths. The aircraft pair must maintain a longitudinal separation within a range that avoids wake encounters and, if one of the aircraft blunders, avoids collision. Wake avoidance defines the rear gate of the longitudinal separation. The lead aircraft generates a wake vortex that, with the aid of crosswinds, can travel laterally onto the path of the trail aircraft. As runway separation decreases, the wake has less distance to traverse to reach the path of the trail aircraft. The total system error of each aircraft further reduces this distance. The total system error is often modeled as a probability distribution function. Therefore, Monte-Carlo simulations are a favored tool for assessing a "safe" rear-gate. However, safety for paired approaches typically requires that a catastrophic wake encounter be a rare one-in-a-billion event during normal operation. Using a Monte-Carlo simulation to assert this event rarity with confidence requires a massive number of runs. Such large runs do not lend themselves to rapid turn-around during the early stages of investigation when the goal is to eliminate the infeasible regions of the solution space and to perform trades among the independent variables in the operational concept. One can employ statistical analysis using simplified models more efficiently to narrow the solution space and identify promising trades for more in-depth investigation using Monte-Carlo simulations. These simple, analytical models not only have to address the uncertainty of the total system error but also the uncertainty in navigation sources used to alert an abort of the procedure. This paper presents a method for integrating total system error, procedure abort rates, avionics failures, and surveillance errors into a statistical analysis that identifies the likely feasible runway separations for simultaneous paired approaches.

Madden, Michael M.↗

Cosmic Pathways for Compact Groups in the Milli-Millennium Simulation

We detected 10 compact galaxy groups (CGs) at z=0 in the semianalytic galaxy catalog of Guo et al. for themilli-Millennium Cosmological Simulation (sCGs in mGuo2010a). We aimed to identify potential canonicalpathways for compact group evolution and thus illuminate the history of observed nearby CGs. By constructingmerger trees for z=0 sCG galaxies, we studied the cosmological evolution of key properties and compared themwith z=0 Hickson CGs (HCGs). We found that, once sCG galaxies come within 1 (0.5) Mpc of their mostmassive galaxy, they remain within that distance until z=0, suggesting sCG "birth redshifts." At z=0 stellarmasses of sCG most massive galaxies are within 1010M*/Me1011. In several cases, especially in the twofour- and five-member systems, the amount of cold gas mass anticorrelates with stellar mass, which in turncorrelates with hot gas mass. We define the angular difference between group members' 3D velocity vectors,vel, and note that many of the groups are long-lived because their small values of vel indicate a significantparallel component. For triplets in particular, vel values range between 20° and 40° so that galaxies are comingtogether along roughly parallel paths, and pairwise separations do not show large pronounced changes after closeencounters. The best agreement between sCG and HCG physical properties is for M* galaxy values, but HCGvalues are higher overall, including for star formation rates (SFRs). Unlike HCGs, due to a tail at low SFR and M*and a lack of M*1011Me galaxies, only a few sCG galaxies are on the star-forming main sequence

Tzanavaris, Panayiotis↗

Lightforce Photon-Pressure Collision Avoidance: Efficiency Analysis in the Current Debris Environment and Long-Term Simulation Perspective

This work provides an efficiency analysis of the LightForce space debris collision avoidance scheme in the current debris environment and describes a simulation approach to assess its impact on the long-term evolution of the space debris environment. LightForce aims to provide just-in-time collision avoidance by utilizing photon pressure from ground-based industrial lasers. These ground stations impart minimal accelerations to increase the miss distance for a predicted conjunction between two objects. In the first part of this paper we will present research that investigates the short-term effect of a few systems consisting of 20-kilowatt-class lasers directed by 1.5-meter-diameter telescopes using adaptive optics. The results found such a network of ground stations to mitigate more than 85 percent of conjunctions and could lower the expected number of collisions in Low Earth Orbit (LEO) by an order of magnitude. While these are impressive numbers that indicate LightForce's utility in the short-term, the remaining 15 percent of possible collisions contain (among others) conjunctions between two massive objects that would add large amount of debris if they collide. Still, conjunctions between massive objects and smaller objects can be mitigated. Hence, we choose to expand the capabilities of the simulation software to investigate the overall effect of a network of LightForce stations on the long-term debris evolution. In the second part of this paper, we will present the planned simulation approach for that effort. For the efficiency analysis of collision avoidance in the current debris environment, we utilize a simulation approach that uses the entire Two Line Element (TLE) catalog in LEO for a given day as initial input. These objects are propagated for one year and an all-on-all conjunction analysis is performed. For conjunctions that fall below a range threshold, we calculate the probability of collision and record those values. To assess efficiency, we compare a baseline (without collision avoidance) conjunction analysis with an analysis where LightForce is active. Using that approach, we take into account that collision avoidance maneuvers could have effects on third objects. Performing all-on-all conjunction analyses for extended period of time requires significant computer resources; hence we implemented this simulation utilizing a highly parallel approach on the NASA Pleiades supercomputer.

space debris mitigation↗

Fibonacci Grids

Recent years have seen a resurgence of interest in a variety of non-standard computational grids for global numerical prediction. The motivation has been to reduce problems associated with the converging meridians and the polar singularities of conventional regular latitude-longitude grids. A further impetus has come from the adoption of massively parallel computers, for which it is necessary to distribute work equitably across the processors; this is more practicable for some non-standard grids. Desirable attributes of a grid for high-order spatial finite differencing are: (i) geometrical regularity; (ii) a homogeneous and approximately isotropic spatial resolution; (iii) a low proportion of the grid points where the numerical procedures require special customization (such as near coordinate singularities or grid edges). One family of grid arrangements which, to our knowledge, has never before been applied to numerical weather prediction, but which appears to offer several technical advantages, are what we shall refer to as "Fibonacci grids". They can be thought of as mathematically ideal generalizations of the patterns occurring naturally in the spiral arrangements of seeds and fruit found in sunflower heads and pineapples (to give two of the many botanical examples). These grids possess virtually uniform and highly isotropic resolution, with an equal area for each grid point. There are only two compact singular regions on a sphere that require customized numerics. We demonstrate the practicality of these grids in shallow water simulations, and discuss the prospects for efficiently using these frameworks in three-dimensional semi-implicit and semi-Lagrangian weather prediction or climate models.

Swinbank, Richard↗

Geophysics of Small Planetary Bodies

As a SETI Institute PI from 1996-1998, Erik Asphaug studied impact and tidal physics and other geophysical processes associated with small (low-gravity) planetary bodies. This work included: a numerical impact simulation linking basaltic achondrite meteorites to asteroid 4 Vesta (Asphaug 1997), which laid the groundwork for an ongoing study of Martian meteorite ejection; cratering and catastrophic evolution of small bodies (with implications for their internal structure; Asphaug et al. 1996); genesis of grooved and degraded terrains in response to impact; maturation of regolith (Asphaug et al. 1997a); and the variation of crater outcome with impact angle, speed, and target structure. Research of impacts into porous, layered and prefractured targets (Asphaug et al. 1997b, 1998a) showed how shape, rheology and structure dramatically affects sizes and velocities of ejecta, and the survivability and impact-modification of comets and asteroids (Asphaug et al. 1998a). As an affiliate of the Galileo SSI Team, the PI studied problems related to cratering, tectonics, and regolith evolution, including an estimate of the impactor flux around Jupiter and the effect of impact on local and regional tectonics (Asphaug et al. 1998b). Other research included tidal breakup modeling (Asphaug and Benz 1996; Schenk et al. 1996), which is leading to a general understanding of the role of tides in planetesimal evolution. As a Guest Computational Investigator for NASA's BPCC/ESS supercomputer testbed, helped graft SPH3D onto an existing tree code tuned for the massively parallel Cray T3E (Olson and Asphaug, in preparation), obtaining a factor xIO00 speedup in code execution time (on 512 cpus). Runs which once took months are now completed in hours.

Asphaug, Erik I.↗

Introduction to: Atlantic Meridional Overturning Circulation(AMOC)

A striking conclusion of the Intergovernmental Panel on Climate Change 2007 report is the crucial role that the Atlantic Meridional Overturning Circulation (AMOC) may play in anthropogenic climate change. However, these IPCC coupled climate simulations show a broad range of uncertainty in the magnitude and timing of AMOC transport change ranging from none to nearly complete collapse within the 21st century. The potential consequences of large changes in the characteristics of AMOC have motivated the creation in the United States of an interagency program and implementation plan to develop monitoring and prediction capabilities for the AMOC This program parallels the development of substantial monitoring efforts by European, South American and African countries -- notably the UK Rapid and Rapid-Watch programs. The papers contained in this volume are derived from presentations at the First U.S. Atlantic Meridional Overturning Circulation (AMOC) Meeting held 4 - 6 May, 2009 to review the US implementation plan and its coordination with other monitoring activities. The Atlantic Meridional Overturning Circulation consists of multiple components illustrated in an attached figure. Water enters the South Atlantic at upper and intermediate depths through both western and eastern routes (where eddy transport is especially important) and is transported northward across the equator, where it recirculates within the northern subtropical and subpolar gyres. The northern end is defined by the sinking regions of the Nordic Seas and the Labrador Sea where the waters that eventually form the upper and lower branches of North Atlantic Deep Water are conditioned. High surface salinities, the result of high net evaporation in the tropics and subtropics (including the Mediterranean Sea), and presence of regions of the Arctic Ocean that remain ice-free even in winter allow for the rapid cooling and thus densification of surface water. This dense surface water becomes the source of deep water formation in the sinking regions. In addition to transporting mass, the AMOC transports roughly half of the total amount of heat carried northward through the northern subtropics (down the temperature-gradient) by the ocean. In contrast in the Southern Hemisphere AMOC transports heat up-gradient from the cool Circumpolar Current to the warm tropics. Paleoevidence suggests that AMOC heat transport in the two hemispheres has varied over time in ways intimately tied to millennial changes in the Earth's climate. In one example, the abrupt Younger Dryas spell of cold weather over the North Atlantic, which began 13,000 years ago, has generally been linked to a millennial shutdown of the AMOC as a result of massive freshwater discharge from the North American continent. The current AMOC monitoring array consists of a series of instrumented transects located across key passages (see Cunningham et al., 2010 for a recent review). In the Arctic and sub-Arctic, transects cross Fram Strait, Denmark Strait and the Faroe Channel (connecting Greenland, Iceland, and the United Kingdom), as well as the entrance to the Labrador Sea. Further south and extending outwards from the east coast of North America there are a series of monitoring arrays including arrays of the Canadian Atlantic Zone Monitoring Program, deployments of the Rapid Western Atlantic Variability Experiment (WAVE), Line W at 39 N, as well as the Rapid-MOC moored array. The latter spans the entire Atlantic basin along 26.5 N. At tropical latitudes we have the Meridional Overturning Variability Experiment (MOVE) array at 16 N, while in the Southern Hemisphere a corresponding basin-spanning transect is being established at the latitude of Cape of Good Hope, complemented by arrays at Drake Passage.

Hakkinen, Sirpa↗