Search NASA⌕ Search

SEARCH · Search NASA

Results for “distributed computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 1,135 records · Page 63

Sparse distributed memory overview

The Sparse Distributed Memory (SDM) project is investigating the theory and applications of massively parallel computing architecture, called sparse distributed memory, that will support the storage and retrieval of sensory and motor patterns characteristic of autonomous systems. The immediate objectives of the project are centered in studies of the memory itself and in the use of the memory to solve problems in speech, vision, and robotics. Investigation of methods for encoding sensory data is an important part of the research. Examples of NASA missions that may benefit from this work are Space Station, planetary rovers, and solar exploration. Sparse distributed memory offers promising technology for systems that must learn through experience and be capable of adapting to new circumstances, and for operating any large complex system requiring automatic monitoring and control. Sparse distributed memory is a massively parallel architecture motivated by efforts to understand how the human brain works. Sparse distributed memory is an associative memory, able to retrieve information from cues that only partially match patterns stored in the memory. It is able to store long temporal sequences derived from the behavior of a complex system, such as progressive records of the system's sensory data and correlated records of the system's motor controls.

Raugh, Mike↗

Computation of load performance and other parameters of extra high speed modified Lundell alternators from 3D-FE magnetic field solutions

The combined magnetic vector potential - magnetic scalar potential method of computation of 3D magnetic fields by finite elements, introduced in a companion paper, in combination with state modeling in the abc-frame of reference, are used for global 3D magnetic field analysis and machine performance computation under rated load and overload condition in an example 14.3 kVA modified Lundell alternator. The results vividly demonstrate the 3D nature of the magnetic field in such machines, and show how this model can be used as an excellent tool for computation of flux density distributions, armature current and voltage waveform profiles and harmonic contents, as well as computation of torque profiles and ripples. Use of the model in gaining insight into locations of regions in the magnetic circuit with heavy degrees of saturation is demonstrated. Experimental results which correlate well with the simulations of the load case are given.

Wang, R.↗

Job Superscheduler Architecture and Performance in Computational Grid Environments

Computational grids hold great promise in utilizing geographically separated heterogeneous resources to solve large-scale complex scientific problems. However, a number of major technical hurdles, including distributed resource management and effective job scheduling, stand in the way of realizing these gains. In this paper, we propose a novel grid superscheduler architecture and three distributed job migration algorithms. We also model the critical interaction between the superscheduler and autonomous local schedulers. Extensive performance comparisons with ideal, central, and local schemes using real workloads from leading computational centers are conducted in a simulation environment. Additionally, synthetic workloads are used to perform a detailed sensitivity analysis of our superscheduler. Several key metrics demonstrate that substantial performance gains can be achieved via smart superscheduling in distributed computational grids.

Shan, Hongzhang↗

Preliminary analysis of the span-distributed-load concept for cargo aircraft design

A simplified computer analysis of the span-distributed-load airplane (in which payload is placed within the wing structure) has shown that the span-distributed-load concept has high potential for application to future air cargo transport design. Significant increases in payload fraction over current wide-bodied freighters are shown for gross weights in excess of 0.5 Gg (1,000,000 lb). A cruise-matching calculation shows that the trend toward higher aspect ratio improves overall efficiency; that is, less thrust and fuel are required. The optimal aspect ratio probably is not determined by structural limitations. Terminal-area constraints and increasing design-payload density, however, tend to limit aspect ratio.

Whitehead, A. H., Jr.↗

An evaluation of the method for determining the Whitham F-function using distributions of downwash and sidewash angles

The method of computing the Whitham F function using distributions of downwash and sidewash angles was evaluated with two different models. F functions which were calculated for a half angle cone cylinder at M infinites = 2.01, using theoretically and experimentally derived flow angles, show that the method is sensitive to small inaccuracies in the measured flow angles. An oblique wing transport model was tested at 0 deg angle of attack at M infinitely = 2.01. In this test, two different probes were used at two different distances from the model. The pressure signature derived from the F function was extrapolated and compared to the pressure signature measured at the distance of 0.87 body lengths with the static pressure probe. The agreement between the two pressure signatures was poor due to the many inaccuracies involved in using a probe designed to measure flow angularity.

Mendoza, J. P.↗

Electrophoresis experiment. Experiment MA-014

A continuous free-flow electrophoresis study was conducted during the Apollo-Soyuz Test Project Mission to investigate and evaluate the increase in sample flow rate and sample resolution achievable in space. The electrophoresis equipment was designed for the separation of four mixtures of biological cells with variable sample flow rates, buffer flow rates, and electric field gradients. Separation quality was assessed by measuring the light from a quartz lamp through the electrophoresis channel and onto a photodiode system. The data evaluation indicates that all monitored systems operated correctly during the experiment. The optical system produced a light that was too bright to discern true cell distributions, but final analysis of scientific data by computer processing shows the expected distribution of separated cells.

Hannig, K. H.↗

Computer programs for smoothing and scaling airfoil coordinates

Detailed descriptions are given of the theoretical methods and associated computer codes of a program to smooth and a program to scale arbitrary airfoil coordinates. The smoothing program utilizes both least-squares polynomial and least-squares cubic spline techniques to smooth interatively the second derivatives of the y-axis airfoil coordinates with respect to a transformed x-axis system which unwraps the airfoil and stretches the nose and trailing-edge regions. The corresponding smooth airfoil coordinates are then determined by solving a tridiagonal matrix of simultaneous cubic-spline equations relating the y-axis coordinates and their corresponding second derivatives. A technique for computing the camber and thickness distribution of the smoothed airfoil is also discussed. The scaling program can then be used to scale the thickness distribution generated by the smoothing program to a specific maximum thickness which is then combined with the camber distribution to obtain the final scaled airfoil contour. Computer listings of the smoothing and scaling programs are included.

Morgan, H. L., Jr.↗

Hypercluster Parallel Processor

Hypercluster computer system includes multiple digital processors, operation of which coordinated through specialized software. Configurable according to various parallel-computing architectures of shared-memory or distributed-memory class, including scalar computer, vector computer, reduced-instruction-set computer, and complex-instruction-set computer. Designed as flexible, relatively inexpensive system that provides single programming and operating environment within which one can investigate effects of various parallel-computing architectures and combinations on performance in solution of complicated problems like those of three-dimensional flows in turbomachines. Hypercluster software and architectural concepts are in public domain.

Blech, Richard A.↗

Pressure Distribution on Joukowski Wings and Graphic Construction of Joukowski Wings

In the first article, in connection with a lecture on the hydrodynamic basis of flight and the potential flow about a Joukowski wing, the pressure distribution on several wings is computed and plotted. The diagrams of the pressure distributions are presented accompanied with a qualitative discussion of the pressure distribution. In the second article, the the cross-sectional outline (or profile) a Joukowski wing are plotted.

PRESSURE DISTRIBUTION - AIRFOILS - JOUKOWSKI↗

An MPI-IO Interface to HPSS

This paper describes an implementation of the proposed MPI-IO standard for parallel I/O. Our system uses third-party transfer to move data over an external network between the processors where it is used and the I/O devices where it resides. Data travels directly from source to destination, without the need for shuffling it among processors or funneling it through a central node. Our distributed server model lets multiple compute nodes share the burden of coordinating data transfers. The system is built on the High Performance Storage System (HPSS), and a prototype version runs on a Meiko CS-2 parallel computer.

Parallel Processing↗

Wing-Body Aeroelasticity Using Finite-Difference Fluid/Finite-Element Structural Equations on Parallel Computers

In recent years significant advances have been made for parallel computers in both hardware and software. Now parallel computers have become viable tools in computational mechanics. Many application codes developed on conventional computers have been modified to benefit from parallel computers. Significant speedups in some areas have been achieved by parallel computations. For single-discipline use of both fluid dynamics and structural dynamics, computations have been made on wing-body configurations using parallel computers. However, only a limited amount of work has been completed in combining these two disciplines for multidisciplinary applications. The prime reason is the increased level of complication associated with a multidisciplinary approach. In this work, procedures to compute aeroelasticity on parallel computers using direct coupling of fluid and structural equations will be investigated for wing-body configurations. The parallel computer selected for computations is an Intel iPSC/860 computer which is a distributed-memory, multiple-instruction, multiple data (MIMD) computer with 128 processors. In this study, the computational efficiency issues of parallel integration of both fluid and structural equations will be investigated in detail. The fluid and structural domains will be modeled using finite-difference and finite-element approaches, respectively. Results from the parallel computer will be compared with those from the conventional computers using a single processor. This study will provide an efficient computational tool for the aeroelastic analysis of wing-body structures on MIMD type parallel computers.

Byun, Chansup↗

Computer programs for the interpretation of low resolution mass spectra: Program for calculation of molecular isotopic distribution and program for assignment of molecular formulas

Two FORTRAN computer programs for the interpretation of low resolution mass spectra were prepared and tested. One is for the calculation of the molecular isotopic distribution of any species from stored elemental distributions. The program requires only the input of the molecular formula and was designed for compatability with any computer system. The other program is for the determination of all possible combinations of atoms (and radicals) which may form an ion having a particular integer mass. It also uses a simplified input scheme and was designed for compatability with any system.

Miller, R. A.↗

A Review of Edge Computing Technology and Its Applications in Power Systems

Recent advancements in network-connected devices have led to a rapid increase in the deployment of smart devices and enhanced grid connectivity, resulting in a surge in data generation and expanded deployment to the edge of systems. Classic cloud computing infrastructures are increasingly challenged by the demands for large bandwidth, low latency, fast response speed, and strong security. Therefore, edge computing has emerged as a critical technology to address these challenges, gaining widespread adoption across various sectors. This paper introduces the advent and capabilities of edge computing, reviews its state-of-the-art architectural advancements, and explores its communication techniques. A comprehensive analysis of edge computing technologies is also presented. Furthermore, this paper highlights the transformative role of edge computing in various areas, particularly emphasizing its role in power systems. It summarizes edge computing applications in power systems that are oriented from the architectures, such as power system monitoring, smart meter management, data collection and analysis, resource management, etc. Additionally, the paper discusses the future opportunities of edge computing in enhancing power system applications.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Probability of lensing magnification by cosmologically distributed galaxies

We present the analytical formulae for computing the magnification probability caused by cosmologically distributed galaxies. The galaxies are assumed to be singular, truncated-isothermal spheres without both evolution and clustering in redshift. We find that, for a fixed total mass, extended galaxies produce a broader shape in the magnification probability distribution and hence are less efficient as gravitational lenses than compact galaxies. The high-magnification tail caused by large galaxies is well approximated by an A exp -3 form, while the tail by small galaxies is slightly shallower. The mean magnification as a function of redshift is, however, found to be independent of the size of the lensing galaxies. In terms of the flux conservation, our formulae for the isothermal galaxy model predict a mean magnification to within a few percent with the Dyer-Roeder model of a clumpy universe.

Pei, Yichuan C.↗

Plasma Dispersion Function for the Kappa Distribution

The plasma dispersion function is computed for a homogeneous isotropic plasma in which the particle velocities are distributed according to a Kappa distribution. An ordinary differential equation is derived for the plasma dispersion function and it is shown that the solution can be written in terms of Gauss' hypergeometric function. Using the extensive theory of the hypergeometric function, various mathematical properties of the plasma dispersion function are derived including symmetry relations, series expansions, integral representations, and closed form expressions for integer and half-integer values of K.

Podesta, John J.↗

The chemical composition of the dust-free Martian atmosphere - Preliminary results of a two-dimensional model

This paper describes a two-dimensional model of the Martian atmosphere, in which chemical, radiative and dynamical processes are treated interactively. The model is developed for a carbon dioxide-hydrogen-oxygen-nitrogen atmosphere and provides estimates of concentrations for 19 chemical species. The dynamical equations are expressed in the transformed Eulerian coordinates. The wave driving and eddy mixing coefficients resulting from gravity and Rossby wave absorption are computed consistently with the evolving distribution of the mean zonal wind. The net diabatic heating/cooling rate is derived from a detailed radiative scheme including the contributions of CO2, O3, H2O and O2, and is computed consistently with the calculated distribution of temperature and trace species quantities. The computed temperature field as well as the meridional and seasonal variations of ozone column abundance are in good agreement with the distributions observed by Mariner 9 and Viking spacecrafts and the results obtained by previous studies. The present version of the model does not include the effects of dust, clouds and polar hood and only the chemistry in a dust-free atmosphere is considered.

Moreau, D.↗

Analysis of spanwise temperature distribution in three types of air-cooled turbine blade

Methods for computing spanwise blade-temperature distributions are derived for air-cooled hollow blades, air-cooled hollow blades with inserts, and air-cooled blades containing internal cooling fins. Individual and combined effects on spanwise blade-temperature distributions of cooling-air and radial heat conduction are determined. In general, the effects of radiation and radial heat conduction were found to be small and the omission of these variations permitted the construction of nondimensional charts for use in determining spanwise temperature distribution through air-cooled turbine blades. An approximate method for determining the allowable stress-limited blade-temperature distribution is included, with brief accounts of a method for determining the maximum allowable effective gas temperatures and the cooling-air requirements. Numerical examples that illustrate the use of the various temperature-distribution equations and of the nondimensional charts are also included.

Livingood, John N B↗

Distributed-Memory Sparse Deep Neural Network Inference Using Global Arrays

Partitioned Global Address Space (PGAS) models exhibit tremendous promise in developing efficient and productive distributed-memory parallel applications. They have been used extensively in scientific computations due to conveniently offering a ``shared-memory''-like model and convenient interfaces that separate communication with synchronization. Traditionally, PGAS communication models have been applied to dense/contiguously distributed data, but most modern applications depict varied levels of sparsity. Existing PGAS models require certain adaptations to support distributed sparse computations, since associated computations often require matrix arithmetic, in addition to data movement. The Global Arrays toolkit from Pacific Northwest National Laboratory (PNNL) is one of the earliest PGAS models to combine one-sided data communication and distributed matrix operations and is still used in the popular NWChem quantum chemistry suite. Recently, we have expanded the Global Arrays toolkit to support common sparse operations, like sparse matrix-dense matrix multiplies (SpMM), sparse matrix-sparse matrix multiplication (SpGEMM) and Sampled Dense-Dense Matrix Multiplication (SDDMM). As it turns out, these operations are the bedrock of sparse Deep Learning (DL); sparse deep neural networks and Graph Neural Networks (GNNs) have gained increasing attention recently in achieving speedups on training and inference with reduced memory footprints. Unlike scientific applications in High Performance Computing (HPC), modern (distributed-memory capable) DL toolkits often rely on non-standardized and closed-source vendor software optimizations, creating challenges in software-hardware co-design at scale. Our goal is to support a variety of distributed-memory sparse matrix operations and helper functions in the newly created Sparse Global Arrays (SGA), such that it is possible to build portable and productive Machine Learning scenarios for algorithm/software and hardware codesign purposes. Contemporary data-parallel schemes for training/inference are undergoing a major overhaul since model replication limits scalability and causes resource inefficiencies. As such, we have adopted tensor parallelism in decomposing the model and inputs, to mitigate memory issues. Current implementation is built on top of MPI and uses CPUs to maximize the portability across the platforms.

Distributed computing, machine learning↗