Search NASA⌕ Search

SEARCH · Search NASA

Results for “distributed computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 559 records · Page 31

How Distributed Energy Resources Can Support Resilience in Utility Distribution Networks

The goal of this webinar is to engage with electric utilities in the Midwest, particularly small public utilities, to understand the industry's needs for science tools to plan for winter resilience in the future, designing tools that will benefit electric power resilience in all communities. Michigan Tech leads this project with partners from multiple academic, government, and industry groups and asked NLR to present on DERs and laboratory tools and resources.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Development of a Dynamically Configurable, Object-Oriented Framework for Distributed, Multi-modal Computational Aerospace Systems Simulation: Second Year Progress Report

Mesh generation has long been recognized as a bottleneck in the CFD process. While much research on automating the volume mesh generation process have been relatively successful,these methods rely on appropriate initial surface triangulation to work properly. Surface discretization has been one of the least automated steps in computational simulation due to its dependence on implicitly defined CAD surfaces and curves. Differences in CAD peometry engines manifest themselves in discrepancies in their interpretation of the same entities. This lack of "good" geometry causes significant problems for mesh generators, requiring users to "repair" the CAD geometry before mesh generation. The problem is exacerbated when CAD geometry is translated to other forms (e.g., IGES )which do not include important topological and construction information in addition to entity geometry. One technique to avoid these problems is to access the CAD geometry directly from the mesh generating software, rather than through files. By accessing the geometry model (not a discretized version) in its native environment, t h s a proach avoids translation to a format which can deplete the model of topological information. Our approach to enable models developed in the Denali software environment to directly access CAD geometry and functions is through an Application Programming Interface (API) known as CAPRI. CAPRI provides a layer of indirection through which CAD-specific data may be accessed by an application program using CAD-system neutral C and FORTRAN language function calls. CAPRI supports a general set of CAD operations such as truth testing, geometry construction and entity queries.

Afjeh, Abdollah A.↗

Investigating Global Ion and Neutral Atom Populations with IBEX and Voyager

The main objective of this project was to investigate pickup ion (PUI) production in the solar wind and heliosheath (the region between the termination shock and the heliopause) and compute the distributed energetic neutral atom fluxes throughout the helioshpere. The simulations were constrained by comparing the model output against observations from Ulysses, New Horizons, Voyager 1 and 2, and IBEX space probes. As evidenced by the number of peer reviewed journal publications resulting from the project (13 plus three submitted) and their citation rate (156 citations over three years), the project has made a lasting contribution to the field. The outcome is a significant improvement of our understanding of the pickup ion production and distribution in the distant heliosphere. The team has accomplished the entire set of tasks A-H set forth in the proposal. Namely, the transport modeling framework has been augmented with two populations of pickup ions (PUIs), the boundary conditions for the plasma and interstellar neutral hydrogen were verified against Ulysses and New Horizons PUI and an optimal set of velocity diffusion parameters established. The multi-component fluxes of PUIs were computed and isotropic velocity distributions generated for each cell in the computer simulation that covered the heliosphere from 1.5 AU to the heliopause. The distributions were carefully compared with in situ measurements at 3 AU (Ulysses), 12 AU (New Horizons), and 80-90 AU (Voyager 1 and 2) as well as those inferred from ENA fluxes measured by Cassini and IBEX (Wu et al., 2016). Some examples of modeldata comparison are shown in Figure 1. We have used coupled MHD-plasma and kinetic-neutral code to investigate the likely range of plasma and magnetic field parameters in the local interstellar medium (LISM), based on the assumption that the shape of the IBEX ribbon could be used to determine the orientation of the interstellar magnetic field. While the magnetic field is believed to be oriented toward the center of the ribbon, constraining its strength requires comparing the model-predicted angular diameter and circularity of the ribbon with the observations. The study, published in Heerikhuisen et al. (2014), found that the most likely range for the LISM magnetic field strength is between 0.2 and 0.3 nT, which is less than previously thought. Figure 2 shows the IBEX data (left) and compares it to the simulation with a 0.2 nT interstellar magnetic field (center) and a 0.4 nT (right).

Florinski, Vladimir↗

A Comparison of PETSC Library and HPF Implementations of an Archetypal PDE Computation

Two paradigms for distributed-memory parallel computation that free the application programmer from the details of message passing are compared for an archetypal structured scientific computation a nonlinear, structured-grid partial differential equation boundary value problem using the same algorithm on the same hardware. Both paradigms, parallel libraries represented by Argonne's PETSC, and parallel languages represented by the Portland Group's HPF, are found to be easy to use for this problem class, and both are reasonably effective in exploiting concurrency after a short learning curve. The level of involvement required by the application programmer under either paradigm includes specification of the data partitioning (corresponding to a geometrically simple decomposition of the domain of the PDE). Programming in SPAM style for the PETSC library requires writing the routines that discretize the PDE and its Jacobian, managing subdomain-to-processor mappings (affine global- to-local index mappings), and interfacing to library solver routines. Programming for HPF requires a complete sequential implementation of the same algorithm, introducing concurrency through subdomain blocking (an effort similar to the index mapping), and modest experimentation with rewriting loops to elucidate to the compiler the latent concurrency. Correctness and scalability are cross-validated on up to 32 nodes of an IBM SP2.

Hayder, M. Ehtesham↗

Computational Component Build-Up for the X-57 Distributed Electric Propulsion Aircraft

A computational study of the wing for the distributed electric propulsion X-57 Maxwell airplane configuration at cruise and takeoff/landing conditions was completed. Three unstructured-mesh, Navier-Stokes computational fluid dynamics methods, FUN3D, USM3D and Kestrel, were used to predict the performance buildup of components to the full X-57 configuration. The goal of the X-57 wing and distributed electric propulsion system design was to meet or exceed the required lift coefficient of 3.95 for a stall speed of 58 knots. The X-57 Maxwell airplane was designed with a small, high aspect ratio cruise wing that was designed for a high cruise lift coefficient of 0.75 at a cruise speed of 150 knots and altitude of 8,000 ft, with an angle of attack of approximately 0deg. The computational data indicates that the X-57 full aircraft drag would meet the cruise drag goal with a 25 count drag margin. The cruise configuration maximum lift coefficient is 2.07 and without including the stabilator is 1.86 at an angle of attack of 14 deg, predicted with the USM3D flow solver using the Spalart-Allmaras turbulence model. The maximum lift coefficient for the high-lift wing (with the 30deg flap deflection) without the stabilator contribution is 2.60 at an angle of attack of 13 deg. For high-lift blowing conditions with 13.7 hp/prop, the maximum lift coefficient excluding the stabilator is 4.426 at (alpha) = 13 deg. Therefore, the lift augmentation from the high-lift propellers is 1.7 and the total lift augmentation from the high-lift system (30 deg flap deflection and the high-lift blowing) is 2.38. The drag for the high-lift wing with 30 deg flap deflection is much higher than the cruise wing configuration, but the high-lift system is used only during a small portion of the flight envelope. The pitching moment is relatively constant for both blown and unblown conditions when the stabilator is excluded. Modeling the full geometry has indicated some adverse effects from the fuselage on the wing and stabilator. At high angles of attack, the solutions with the USM3D flow solver using the Spalart-Allmaras turbulence model indicates large flow separation on the wing upper surface between the two high-lift nacelles near the fuselage, and also a reduction in sectional lift on the stabilator in the first 50 percent of the stabilator semispan. However, the large flow separation near the fuselage is mostly eliminated in the solutions predicted with two codes, USM3D and Kestrel, using Hybrid Reynolds-averaged Navier Stokes/Large Eddy Simulation turbulence models.

Deere, Karen A.↗

Unorthodox parallelization for Bayesian quantum state estimation

Quantum state tomography (QST) allows for the reconstruction of quantum states through measurements and some inference technique under the assumption of repeated state preparations. Bayesian inference provides a promising platform to achieve both efficient QST and accurate uncertainty quantification, yet is generally plagued by the computational limitations associated with long Markov chains. In this work, we present a novel Bayesian QST approach that leverages modern distributed parallel computer architectures to efficiently sample a D-dimensional Hilbert space. Using a parallelized preconditioned Crank–Nicholson Metropolis–Hastings algorithm, we demonstrate our approach on simulated data and experimental results from IBM Quantum systems up to four qubits, showing significant speedups through parallelization. Although highly unorthodox in pooling independent Markov chains, our method proves remarkably practical, with validation ex post facto via diagnostics like the intrachain autocorrelation time. We conclude by discussing scalability to higher-dimensional systems, offering a path toward efficient and accurate Bayesian characterization of large quantum systems.

Bayesian inference↗

An analysis for high speed propeller-nacelle aerodynamic performance prediction. Volume 1: Theory and application

A computer program, the Propeller Nacelle Aerodynamic Performance Prediction Analysis (PANPER), was developed for the prediction and analysis of the performance and airflow of propeller-nacelle configurations operating over a forward speed range inclusive of high speed flight typical of recent propfan designs. A propeller lifting line, wake program was combined with a compressible, viscous center body interaction program, originally developed for diffusers, to compute the propeller-nacelle flow field, blade loading distribution, propeller performance, and the nacelle forebody pressure and viscous drag distributions. The computer analysis is applicable to single and coaxial counterrotating propellers. The blade geometries can include spanwise variations in sweep, droop, taper, thickness, and airfoil section type. In the coaxial mode of operation the analysis can treat both equal and unequal blade number and rotational speeds on the propeller disks. The nacelle portion of the analysis can treat both free air and tunnel wall configurations including wall bleed. The analysis was applied to many different sets of flight conditions using selected aerodynamic modeling options. The influence of different propeller nacelle-tunnel wall configurations was studied. Comparisons with available test data for both single and coaxial propeller configurations are presented along with a discussion of the results.

Egolf, T. Alan↗

Systematic Uncertainties from Gribov Copies in Lattice Calculation of Parton Distributions in the Coulomb Gauge

Recently, a new method has been proposed to compute parton distributions using boosted correlators fixed in the Coulomb gauge (CG) within the framework of large-momentum effective theory. This approach, which does not involve Wilson lines, could greatly improve the efficiency and precision of lattice quantum chromodynamics calculations. However, concerns remain regarding whether systematic uncertainties from Gribov copies, which correspond to ambiguities in lattice gauge-fixing, are adequately controlled. This work assesses the effects of Gribov copies on Coulomb-gauge-fixed quark correlators. We utilize different strategies for Coulomb-gauge fixing, selecting two different groups of Gribov copies based on lattice gauge configurations. We examine the differences in the resulting spatial quark correlators in both vacuum and pion states. Our findings indicate that the statistical errors of the matrix elements from both Gribov copies, regardless of the correlation range, decrease proportionally to the square root of the number of gauge configurations. The difference between the strategies does not show statistical significance compared to the gauge noise, demonstrating that the effect of the Gribov copies can be neglected in practical lattice calculations of quark parton distributions.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

The sensitivity of fracture location distribution in brittle materials

The Weibull weak-link theory allows for the computation of distribution functions for both the fracture location and the applied far-field stress. Several authors have suggested using the fracture location information from tests to infer Weibull parameters, and others have used the predictive capabilities of the theory to calculate average fracture locations for brittle bodies. By a simple set of example calculations, it is shown that the fracture location distribution function is distinctly more sensitive to perturbations in the stress state than the fracture stress distribution function is. In general, the average fracture location is more subject to stress perturbations than the average fracture stress. The results indicate that care must be exercised in applying fracture location theory.

Wetherhold, Robert C.↗

Cosmic Reionization on Computers: Statistical Properties of the Distributions of Mean Opacities

Quasar absorption lines provide a unique window to the relationship between galaxies and the intergalactic medium during the Epoch of Reionization. In particular, high redshift quasars enable measurements of the neutral hydrogen content of the universe. However, the limited sample size of observed quasar spectra, particularly at the highest redshifts, hampers our ability to fully characterize the intergalactic medium during this epoch from observations alone. In this work, we characterize the distributions of mean opacities of the intergalactic medium in simulations from the Cosmic Reionization on Computers (CROC) project. We find that the distribution of mean opacities along sightlines follows a non-trivial distribution that cannot be easily approximated by a known distribution. When comparing the cumulative distribution function of mean opacities measurements in subsamples of sample sizes similar to observational measurements from the literature, we find consistency between CROC and observations at redshifts $z\lesssim 5.7$. However, at higher redshifts ($z\gtrsim5.7$), the cumulative distribution function of mean opacities from CROC is notably narrower than those from observed quasar sightlines implying that observations probe a systematically more opaque intergalactic medium at higher redshifts than the intergalactic medium in CROC boxes at these same redshifts. This is consistent with previous analyses that indicate that the universe is reionized too early in CROC simulations.

79 ASTRONOMY AND ASTROPHYSICS↗

Investigation at Mach Numbers of 0.60 to 3.50 of Blended Wing-Body Combinations with Cambered and Twisted Wings with Diamond, Delta and Arrow Plan Forms

This investigation is a continuation of the experimental and theoretical evaluation of blended wing-body combinations. The basic diamond, delta, and arrow plan forms which had an aspect ratio of 2 with leading-edge sweeps of 45.00 deg., 59.04 deg., and 70.82 deg. and trailing edge of -45.00 deg., -18.43 deg., and 41.19 deg., respectively, are used herein as standards for evaluating the effects of camber and warp. The wing thickness distributions were computed by varying the section shape along with the body radii (blending process) to match the prescribed area distribution and wing plan form. The wing camber and warp were computed to try to obtain nearly elliptical spanwise and chordwise load distributions for each plan form and thus to obtain low drag due to lift for a range of Mach numbers for which the velocities normal to the wing leading edge are subsonic. Elliptical chordwise load distributions were not possible for the plan forms and design conditions selected, so these distributions were somewhat different for each plan form. The models were tested with transition fixed at Mach numbers from 0.60 to 3.50 and at Reynolds numbers, based on the mean aerodynamic chord of the wing, of roughly 4,000,000 to 9,000,000. At speeds where the velocities normal to the wing leading edges were supersonic, an increase in the experimental wave-drag coefficients due to camber and twist was evident, but this penalty decreased with increased sweep. Thus the minimum wave-drag coefficients for the cambered arrow model were almost identical with the zero-lift wave- drag coefficients for the uncambered arrow model at all test Mach numbers.

Holdaway, George H.↗

Parallel Simulation of Unsteady Turbulent Flames

Time-accurate simulation of turbulent flames in high Reynolds number flows is a challenging task since both fluid dynamics and combustion must be modeled accurately. To numerically simulate this phenomenon, very large computer resources (both time and memory) are required. Although current vector supercomputers are capable of providing adequate resources for simulations of this nature, the high cost and their limited availability, makes practical use of such machines less than satisfactory. At the same time, the explicit time integration algorithms used in unsteady flow simulations often possess a very high degree of parallelism, making them very amenable to efficient implementation on large-scale parallel computers. Under these circumstances, distributed memory parallel computers offer an excellent near-term solution for greatly increased computational speed and memory, at a cost that may render the unsteady simulations of the type discussed above more feasible and affordable.This paper discusses the study of unsteady turbulent flames using a simulation algorithm that is capable of retaining high parallel efficiency on distributed memory parallel architectures. Numerical studies are carried out using large-eddy simulation (LES). In LES, the scales larger than the grid are computed using a time- and space-accurate scheme, while the unresolved small scales are modeled using eddy viscosity based subgrid models. This is acceptable for the moment/energy closure since the small scales primarily provide a dissipative mechanism for the energy transferred from the large scales. However, for combustion to occur, the species must first undergo mixing at the small scales and then come into molecular contact. Therefore, global models cannot be used. Recently, a new model for turbulent combustion was developed, in which the combustion is modeled, within the subgrid (small-scales) using a methodology that simulates the mixing and the molecular transport and the chemical kinetics within each LES grid cell. Finite-rate kinetics can be included without any closure and this approach actually provides a means to predict the turbulent rates and the turbulent flame speed. The subgrid combustion model requires resolution of the local time scales associated with small-scale mixing, molecular diffusion and chemical kinetics and, therefore, within each grid cell, a significant amount of computations must be carried out before the large-scale (LES resolved) effects are incorporated. Therefore, this approach is uniquely suited for parallel processing and has been implemented on various systems such as: Intel Paragon, IBM SP-2, Cray T3D and SGI Power Challenge (PC) using the system independent Message Passing Interface (MPI) compiler. In this paper, timing data on these machines is reported along with some characteristic results.

Menon, Suresh↗

Demonstration of Cost-Effective, High-Performance Computing at Performance and Reliability Levels Equivalent to a 1994 Vector Supercomputer

The Affordable High Performance Computing (AHPC) project demonstrated that high-performance computing based on a distributed network of computer workstations is a cost-effective alternative to vector supercomputers for running CPU and memory intensive design and analysis tools. The AHPC project created an integrated system called a Network Supercomputer. By connecting computer work-stations through a network and utilizing the workstations when they are idle, the resulting distributed-workstation environment has the same performance and reliability levels as the Cray C90 vector Supercomputer at less than 25 percent of the C90 cost. In fact, the cost comparison between a Cray C90 Supercomputer and Sun workstations showed that the number of distributed networked workstations equivalent to a C90 costs approximately 8 percent of the C90.

Babrauckas, Theresa↗

Aerodynamic Shape Optimization of Supersonic Aircraft Configurations via an Adjoint Formulation on Parallel Computers

This work describes the application of a control theory-based aerodynamic shape optimization method to the problem of supersonic aircraft design. The design process is greatly accelerated through the use of both control theory and a parallel implementation on distributed memory computers. Control theory is employed to derive the adjoint differential equations whose solution allows for the evaluation of design gradient information at a fraction of the computational cost required by previous design methods. The resulting problem is then implemented on parallel distributed memory architectures using a domain decomposition approach, an optimized communication schedule, and the MPI (Message Passing Interface) Standard for portability and efficiency. The final result achieves very rapid aerodynamic design based on higher order computational fluid dynamics methods (CFD). In our earlier studies, the serial implementation of this design method was shown to be effective for the optimization of airfoils, wings, wing-bodies, and complex aircraft configurations using both the potential equation and the Euler equations. In our most recent paper, the Euler method was extended to treat complete aircraft configurations via a new multiblock implementation. Furthermore, during the same conference, we also presented preliminary results demonstrating that this basic methodology could be ported to distributed memory parallel computing architectures. In this paper, our concern will be to demonstrate that the combined power of these new technologies can be used routinely in an industrial design environment by applying it to the case study of the design of typical supersonic transport configurations. A particular difficulty of this test case is posed by the propulsion/airframe integration.

Reuther, James↗

Velocity-space synthesis of ISEE-1 measurements of the three dimensional electron distribution function

A computer package which produces contour plots of the three dimensional electron distribution function measured by an electron spectrometer aboard ISEE-1 is described. Examples of the contour plots and an explanation of how to use the program, including the necessary computer code for running the program on the GSFC 360/91 computer is presented. The method by which the discrete measurements of the distribution function, given by points on the four dimensional surface are synthesized into a smooth surface in a three dimensional space which can be contoured is described. The velocity components are parallel and perpendicular to the magnetic field, respectively, in the proper frame of the electrons.

Fitzenreiter, R. J.↗

Enhancing Distribution System Resilience: A First-Order Meta-RL Algorithm for Critical Load Restoration

The increasing frequency of extreme events and the integration of distributed energy resources (DERs) into modern grids have elevated the need for resilient and efficient critical load restoration strategies in distribution systems. However, the stochastic nature of renewable DERs, limited energy resource availability and the intricate nonlinearities inherent in complex grid control problem make the problem challenging. Although reinforcement learning (RL) and warm-start RL methods have shown promising results, their performance often falls short in rapidly adapting to new, unseen situations and typically requires exhaustive problem-specific tuning. To address these gaps, we propose a First-Order Meta-based RL (FOM-RL) algorithm within an online framework for adaptive and robust critical load restoration. By harnessing local DERs as the enabling technology, FOM-RL allows the RL agent to swiftly adapt to new unseen scenarios by leveraging previously acquired knowledge of different tasks. Experimental results provide evidence that proposed algorithm learns more efficiently and showcases generalization capabilities across diverse set of operational scenarios. Moreover, a rigorous theoretical analysis yields a tight sublinear regret bound, sensitive to temporal variability, with a task-averaged optimality gap bounded by O(VM+D*/(Tsquare root(M))). These results suggest that optimality improves with task similarity and an increased number of tasks M, reaffirming the efficacy and scalability of the proposed approach in addressing the complexities of critical load restoration in distribution systems.

complexity theory↗