Search NASA⌕ Search

SEARCH · Search NASA

Results for “cluster computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 613 records · Page 34

Impact of amorphous pockets on displacement damage evolution in silicon

Silicon has long been known to exhibit amorphization in response to heavy particle bombardment. For doses below the total amorphization threshold, partial amorphization is observed in the form of scattered amorphous pockets. While extensive research has gone into modeling the formation and evolution of amorphous pockets in response to irradiation, no studies yet investigate their impact on the evolution of other damage such as interstitial supersaturation and clustering. In this study, we survey the impact of amorphous pockets on defect evolution in silicon when treated as static sinks. MD is first used to show that amorphous pockets provide energetically favorable sites for point defects relative to the crystalline bulk, supporting the hypothesis that they act as sinks. A 0-D cluster dynamics model is then constructed, taking an interstitial clustering model from the literature and including amorphous pockets as a sink species. We conduct our survey for temperatures between 30 and 400 °C and sink strengths between 1 to 6 x 10 10 cm −2 . Both implantation- and radiation-induced damage states are investigated using interstitial and vacancy concentrations as initial condition variables. We find that, due to the differing migration rates of the interstitial and the vacancy, amorphous pockets have a non-monotonic impact on the final damage state depending on the effective sink strength of the amorphous pockets, resulting in increased damage formation in regimes of intermediate amorphization. In conclusion, this result emphasizes the important role of amorphous pockets in governing the evolution of damage in partially amorphized crystalline materials.

36 MATERIALS SCIENCE↗

A static quantum embedding scheme based on coupled cluster theory

Here, we develop a static quantum embedding scheme that utilizes different levels of approximations to coupled cluster (CC) theory for an active fragment region and its environment. To reduce the computational cost, we solve the local fragment problem using a high-level CC method and address the environment problem with a lower-level Møller–Plesset (MP) perturbative method. This embedding approach inherits many conceptual developments from the hybrid second-order Møller–Plesset (MP2) and CC works by Nooijen [J. Chem. Phys. 111, 10815 (1999)] and Bochevarov and Sherrill [J. Chem. Phys. 122, 234110 (2005)]. We go beyond those works here by primarily targeting a specific localized fragment of a molecule and also introducing an alternative mechanism to relax the environment within this framework. We will call this approach MP-CC. We demonstrate the effectiveness of MP-CC on several potential energy curves and a set of thermochemical reaction energies, using CC with singles and doubles as the fragment solver, and MP2-like treatments of the environment. The results are substantially improved by the inclusion of orbital relaxation in the environment. Using localized bonds as the active fragment, we also report results for N=N bond breaking in azomethane and for the central C–C bond torsion in butadiene. We find that when the fragment Hilbert space size remains fixed (e.g., when determined by an intrinsic atomic orbital approach), the method achieves comparable accuracy with both a small and a large basis set. Additionally, our results indicate that increasing the fragment Hilbert space size systematically enhances the accuracy of observables, approaching the precision of the full CC solver.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

SIAM Conference on Parallel Processing for Scientific Computing, 4th, Chicago, IL, Dec. 11-13, 1989, Proceedings

Attention is given to such topics as an evaluation of block algorithm variants in LAPACK and presents a large-grain parallel sparse system solver, a multiprocessor method for the solution of the generalized Eigenvalue problem on an interval, and a parallel QR algorithm for iterative subspace methods on the CM2. A discussion of numerical methods includes the topics of asynchronous numerical solutions of PDEs on parallel computers, parallel homotopy curve tracking on a hypercube, and solving Navier-Stokes equations on the Cedar Multi-Cluster system. A section on differential equations includes a discussion of a six-color procedure for the parallel solution of elliptic systems using the finite quadtree structure, data parallel algorithms for the finite element method, and domain decomposition methods in aerodynamics. Topics dealing with massively parallel computing include hypercube vs. 2-dimensional meshes and massively parallel computation of conservation laws. Performance and tools are also discussed.

Dongarra, Jack↗

Enabling Innovative Analysis on Heterogeneous Clusters through HTCdaskgateway

High energy particle (HEP) physics research is going through fundamental changes as we move to collect larger amounts of data from the Large Hadron Collider (LHC). Analysis facilities and distributed computing, through HTCs, have come together to create the next pythonic generation of analysis by utilizing HTCdaskgateway, a Dask gateway extension, allowing users to spawn workers compatible with both their analysis and heterogeneous clusters in line with authentication requirements. This is enabling physicists to engage with scientific python in ways they had not before because of domain specific C++ tools. An example of HTCdaskgateway’s use is Fermilab’s Elastic Analysis Facility.

Chavez, Elise [U. Wisconsin, Madison (main)]↗

Electra: A Modular-Based Expansion of NASA's Supercomputing Capability

NASA has increasingly relied on high-performance computing (HPC) re- sources for computational modeling, simulation, and data analysis to meet the science and engineering goals of its missions in space exploration, aeronautics, and Earth and space science. The NASA Advanced Supercomputing (NAS) Division at Ames Research Center in Silicon Valley, Calif., hosts NASA’s premier supercomputing resources, integral to achieving and enhancing the success of the agency’s missions. NAS provides a balanced environment, funded under the High-End Computing Capability (HECC) project, comprised of world-class supercomputers, including its flagship distributed-memory cluster, Pleiades; high-speed networking; and massive data storage facilities, along with multi-disciplinary support teams for user support, code porting and optimization, and large-scale data analysis and scientific visualization. However, as scientists have increased the fidelity of their simulations and engineers are conducting larger parameter-space studies, the requirements for supercomputing resources have been growing by leaps and bounds. With the facility housing the HECC systems reaching its power and cooling capacity, NAS undertook a prototype project to investigate an alternative approach for housing supercomputers. Modular supercomputing, or container-based computing, is an innovative concept for expanding NASA’s HPC capabilities. With modular supercomputing, additional containers—similar to portable storage pods—can be connected together as needed to accommodate the agency’s ever-increasing demand for computing resources. In addition, taking advantage of the local weather permits the use of cooling technologies that would additionally save energy and reduce annual water usage. The first stage of NASA’s Modular Supercomputing Facility (MSF) prototype, which resulted in a 1,000 square-foot module on a concrete pad with room for 16 compute racks, was completed in Fall 2016 and an SGI (now HPE) computer system, named Electra, was deployed there in early 2017. Cooling is performed via an evaporative system built into the module, and preliminary experience shows a Power Usage Effectiveness (PUE) measurement of 1.03. Electra achieved over a petaflop on the LINPACK benchmark, sufficient to rank number 96 on the November 2016 TOP500 list [14]. The system consists of 1,152 InfiniBand-connected Intel Xeon Broadwell-based nodes. Its users access their files on a facility-wide file system shared by all HECC compute assets via Mellanox MetroX InfiniBand extenders, which connect the Electra fabric to Lustre routers in the primary facility over fiber-optic links about 900 feet long. The MSF prototype has exceeded expectations and is serving as a blueprint for future expansions. In the remainder of this chapter, we detail how modular data center technology can be used to expand an existing compute resource. We begin by describing NASA’s requirements for supercomputing and how resources were provided prior to the integration of the Electra module-based system.

Biswas, Rupak↗

A robust multilevel simultaneous eigenvalue solver

Multilevel (ML) algorithms for eigenvalue problems are often faced with several types of difficulties such as: the mixing of approximated eigenvectors by the solution process, the approximation of incomplete clusters of eigenvectors, the poor representation of solution on coarse levels, and the existence of close or equal eigenvalues. Algorithms that do not treat appropriately these difficulties usually fail, or their performance degrades when facing them. These issues motivated the development of a robust adaptive ML algorithm which treats these difficulties, for the calculation of a few eigenvectors and their corresponding eigenvalues. The main techniques used in the new algorithm include: the adaptive completion and separation of the relevant clusters on different levels, the simultaneous treatment of solutions within each cluster, and the robustness tests which monitor the algorithm's efficiency and convergence. The eigenvectors' separation efficiency is based on a new ML projection technique generalizing the Rayleigh Ritz projection, combined with a technique, the backrotations. These separation techniques, when combined with an FMG formulation, in many cases lead to algorithms of O(qN) complexity, for q eigenvectors of size N on the finest level. Previously developed ML algorithms are less focused on the mentioned difficulties. Moreover, algorithms which employ fine level separation techniques are of O(q(sub 2)N) complexity and usually do not overcome all these difficulties. Computational examples are presented where Schrodinger type eigenvalue problems in 2-D and 3-D, having equal and closely clustered eigenvalues, are solved with the efficiency of the Poisson multigrid solver. A second order approximation is obtained in O(qN) work, where the total computational work is equivalent to only a few fine level relaxations per eigenvector.

Costiner, Sorin↗

Condition-Based Maintenance of a Circulating Water System of a Canadian Nuclear Power Plant using Machine Learning and Statistical Tools

Canada Deuterium Uranium pressurized-heavy-water reactors (PHWR) are a type of nuclear power plant that generate clean and reliable energy. The scope of this work is to automate data analysis methodologies to inform a condition-based maintenance strategy of a circulating water system (CWS) of a PHWR. The multiunit CWS provides a continuous supply of water to cool steam condensers, even during transient scenarios, thereby improving the thermal efficiency. This work aims to develop a machine learning (ML) based approach to detect anomalies in heterogeneous data of a CWS in a PHWR to help inform a predictive maintenance strategy. The heterogeneous data include textual and numeric time series data for a PHWR. Natural-language-processing (NLP)-based models are used to analyze textual data contained in work orders and operator logs and an event-timeseries correlation detection method is applied to assist anomalies diagnoses for CWS. An ML model Robust Linear Model (RLM) is also used to remove the seasonal variations in the system variable distributions based on distributions of environmental variables. A machine learning model, Density-Based Spatial Clustering of Applications with Noise (DBSCAN), trained on both original data and data without any seasonal variations will then be used to detect if an anomaly exists. Thus, by moving to an automated methodology to detect, classify, and forecast anomalies, the maintenance strategy would be based on component condition instead of a time-based schedule.

97 - MATHEMATICS AND COMPUTING↗

Vapor-Phase Heteroatom Incorporation into Semiconductive Molecular-Scale Magic-Size Clusters

Magic-size metal chalcogenide clusters of molecular size exhibit well-defined structure and unique properties that might be further expanded with the incorporation or substitution of a second metal. Here, we report the postmodification of magic-size clusters synthesized in polymer thin films via exposure to volatile metal organic precursors commonly utilized for atomic layer deposition. Exposure of In 6 S 6 (CH 3 ) 6 clusters to dimethylcadmium results in exposure-dependent incorporation of Cd 2+ , which extends the optical absorbance of the clusters into the visible spectrum. The mechanism for Cd 2+ incorporation is consistent with Cd 2+ replacement of In 3+ that includes methyl ligand removal to maintain charge neutrality. Even for clusters embedded in a polymer matrix, ligand loss leads to sintering and transformation into larger nanoscale aggregates with zinc blende-type structure. The extent of Cd incorporation can be modulated by varying the process temperature and volatile metal organic exposure as well as the choice of volatile metal organic precursor. A computational thermodynamic analysis of heteroatom incorporation for several metals and chemistries reveals that both the stability of the substituted cluster and the favorability of reaction byproducts jointly determine the favorability of cation incorporation.

atomic layer deposition↗

Accretion flows in elliptical galaxies

A steady-state infall model of gas in elliptical galaxies is developed to investigate the properties and structure of the X-ray-emitting gas observed in these systems. Models have been computed for galaxies with an external pressure (as might be important for ellipticals in clusters), and for varying supernova heating rates. All the models exhibit cooling flows, with mass accretion rates of 0.1 - 0.5 solar mass/yr. A correlation between the radio luminosity and the X-ray luminosity of elliptical galaxies is examined which, in the context of the infall models, may suggest that the radio emission arises from nuclear sources that are powered by the gas accretion flow. These radio sources may also be confined effectively by the X-ray emitting gas.

Vedder, Peter W.↗

On the origin of extinction in the Coma cluster of galaxies

Visual extinction of distant clusters seen through the Coma cluster seem to suggest that dust may be present in the hot x ray emitting intracluster gas. However, the Infrared Astronomy Satellite (IRAS) failed to detect any infrared emission from the cluster at the level expected from the extinction measurements. Researchers carried out a detailed analysis of the properties of intracluster dust in the context of a model which includes continuous injection of dust by the cluster galaxies, grain destruction by sputtering, and transient grain heating by the hot plasma. Computed infrared fluxes are in agreement with the upper limit obtained from the IRAS. The calculations, and the constraint implied by the IRAS observations, suggest that the intracluster dust must be significantly depleted compared to interstellar abundances. Researchers discuss possible explanations for the discrepancy between the observed visual extinction and the IRAS upper limit.

Rephaeli, Y.↗

The three-dimensional Multi-Block Advanced Grid Generation System (3DMAGGS)

As the size and complexity of three dimensional volume grids increases, there is a growing need for fast and efficient 3D volumetric elliptic grid solvers. Present day solvers are limited by computational speed and do not have all the capabilities such as interior volume grid clustering control, viscous grid clustering at the wall of a configuration, truncation error limiters, and convergence optimization residing in one code. A new volume grid generator, 3DMAGGS (Three-Dimensional Multi-Block Advanced Grid Generation System), which is based on the 3DGRAPE code, has evolved to meet these needs. This is a manual for the usage of 3DMAGGS and contains five sections, including the motivations and usage, a GRIDGEN interface, a grid quality analysis tool, a sample case for verifying correct operation of the code, and a comparison to both 3DGRAPE and GRIDGEN3D. Since it was derived from 3DGRAPE, this technical memorandum should be used in conjunction with the 3DGRAPE manual (NASA TM-102224).

Alter, Stephen J.↗

Development of a Three-Dimensional PSE Code for Compressible Flows: Stability of Three-Dimensional Compressible Boundary Layers

A program is developed to investigate the linear stability of three-dimensional compressible boundary layer flows over bodies of revolutions. The problem is formulated as a two dimensional (2D) eigenvalue problem incorporating the meanflow variations in the normal and azimuthal directions. Normal mode solutions are sought in the whole plane rather than in a line normal to the wall as is done in the classical one dimensional (1D) stability theory. The stability characteristics of a supersonic boundary layer over a sharp cone with 50 half-angle at 2 degrees angle of attack is investigated. The 1D eigenvalue computations showed that the most amplified disturbances occur around x(sub 2) = 90 degrees and the azimuthal mode number for the most amplified disturbances range between m = -30 to -40. The frequencies of the most amplified waves are smaller in the middle region where the crossflow dominates the instability than the most amplified frequencies near the windward and leeward planes. The 2D eigenvalue computations showed that due to the variations in the azimuthal direction, the eigenmodes are clustered into isolated confined regions. For some eigenvalues, the eigenfunctions are clustered in two regions. Due to the nonparallel effect in the azimuthal direction, the eigenmodes are clustered into isolated confined regions. For some eigenvalues, the eigenfunctions are clustered in two regions. Due to the nonparallel effect in the azimuthal direction, the most amplified disturbances are shifted to 120 degrees compared to 90 degrees for the parallel theory. It is also observed that the nonparallel amplification rates are smaller than that is obtained from the parallel theory.

Balakumar, P.↗

Research on Spectroscopy, Opacity, and Atmospheres

I propose to continue providing observers with basic data for interpreting spectra from stars, novas, supernovas, clusters, and galaxies. These data will include allowed forbidden line lists both laboratory and computed, for the first five to ten ions of all atoms and for all relevant diatomic molecules. I will eventually expend to all ions of the first thirty elements to treat far UV end X-ray spectra, and for envelope opacities. I also include triatomic molecules providing by other researchers. I have made CDs with Partridge and Schwanke's water data for work on M stars.The luna data also serve as input to my model atmosphere and synthesis programs that generated energy distributions, photometry, limb darkening, and spectra that can be used for planning observations and for fitting observed spectra. The spectrum synthesis programs produce detailed plots with the line identified. Grids of stellar spectra can be used for radial velocity-, rotation-, or abundance templates and for population synthesis. I am fitting spectra of bright stars to test the data and to produce atlases to guide observer. For each star the whole spectrum is computed from the UV to the far IR. The line data, opacities, models, spectra, and programs are freely distributed on CDs and on my web site and represent a unique resource for many NASA programs.

Oliversen, Ronald↗

Benchmark Comparison of Dual- and Quad-Core Processor Linux Clusters with Two Global Climate Modeling Workloads

This viewgraph presentation details the science and systems environments that NASA High End computing program serves. Included is a discussion of the workload that is involved in the processing for the Global Climate Modeling. The Goddard Earth Observing System Model, Version 5 (GEOS-5) is a system of models integrated using the Earth System Modeling Framework (ESMF). The GEOS-5 system was used for the Benchmark tests, and the results of the tests are shown and discussed. Tests were also run for the Cubed Sphere system, results for these test are also shown.

McGalliard, James↗

Monte Carlo Explicitly Correlated Second-Order Many-Body Green’s Function Calculations of Semiconductor Band Gaps

A systematically converging series of ab initio, post-density-functional, size-consistent, electron-correlated approximations is desired for predictive computing of felectronic band structures of insulating, semiconducting, and metallic solids. A series that meets all of these desiderata (except the applicability to metals) is ab initio many-body Green's function theory based on Gaussian-type-orbital (GTO) basis sets. Here, its leading-order approximation, the second-order Green's function (GF2) method in the diagonal and frequency-independent approximations with the aug-cc-pVDZ basis set, is applied to the fundamental band gaps of three semiconductors (diamond, silicon, and silicon carbide in the zincblende structure) using cluster models. Corrections are made to the basis-set-incompleteness errors by the explicit-correlation (F12) ansatz (GF2-F12) for the valence band edges. The crystals are modeled as surface-passivated clusters of increasing sizes, whose wave functions are expanded by up to 2709 GTO basis functions. Immense computational costs of these calculations are overcome by the highly scalable stochastic algorithm of the Monte Carlo GF2-F12 method, whose operation cost per state increases only as a cubic power of system size, which has a tiny memory footprint and easily achieves near-perfect parallel efficiency on thousands of CPUs or on hundreds of GPUs. The correlated, F12-corrected highest-occupied and lowest-unoccupied molecular-orbital energy (HOMO-LUMO) gap is 5.78 ± 0.07 eV for C 87 H 76 as compared with the experimental value of the fundamental (indirect) band gap of bulk diamond at 5.48 eV. The correlated, F12-corrected HOMO-LUMO gaps for Si 75 H 76 and Si 32 C 43 H 76 are 2.56 ± 0.15 eV and 3.50 ± 0.12 eV, respectively, which are expected to decrease further with increasing cluster sizes. As a result, the experimental fundamental (indirect) band gaps of bulk silicon and silicon carbide are 1.17 eV and 2.42 eV, respectively.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Stochastic tensor contraction for quantum chemistry

Many computational methods in ab initio quantum chemistry are formulated in terms of high-order tensor contractions, whose cost determines the size of system that can be studied. We introduce stochastic tensor contraction to perform such operations with greatly reduced cost, and present its application to the gold-standard quantum chemistry method, coupled cluster theory with up to perturbative triples. For total energy errors more stringent than chemical accuracy, we reduce the computational scaling to that of mean-field theory, while starting to approach the mean-field absolute cost, thereby challenging the existing cost-to-accuracy landscape. Benchmarks against state-of-the-art local correlation approximations further show that we achieve an order-of-magnitude improvement in both total computation time and error, with significantly reduced sensitivity to system dimensionality and electron delocalization. We conclude that stochastic tensor contraction is a powerful computational primitive to accelerate a wide range of quantum chemistry.

Chemical Physics (physics.chem-ph)↗

Optical and near infrared photometry of Butcher-Oemler clusters

Rich clusters of galaxies at moderate redshifts (z approx. .3) have a larger proportion of optically blue galaxies than their low redshift counterparts. Spectroscopic examination of the blue galaxies by various authors has shown that the blue galaxies are generally Seyferts, show evidence for recent star formation, or are foreground objects. Unfortunately, spectroscopy is too time consuming to be used on large samples. Thus, we have looked for a way to separate Seyferts, starbursts, ellipticals and nonmembers using photometry alone. Five moderate redshift clusters, Abell numbers 777, 963, 1758, 1961 and 2218, have been observed in the V, R and K bands. We model the spectral energy distributions of various kinds of galaxies found in clusters and derive observed colors. We have modeled the spectral energy distributions (SED) of several kinds of galaxies and compute their colors as a function of redshift. We expect to see ellipticals, spirals, starbursts, post-starburst and Seyfert galaxies. The SED of elliptical and Sbc galaxies was observed by Rieke and Rieke. The SEDs for the starburst galaxies was created by adding a reddened 10(exp 8) year old burst to a spiral galaxy SED. The post-starburst (E+A) galaxy SEDs are composed of a slightly reddened 10(exp 9) year old burst and elliptical galaxy SED. SEDs for the Seyferts were created by adding a v(exp -1.1) power law, and a hot dust thermal spectrum to the Sbc. From the SEDs the colors of galaxies at various redshifts with assorted filters were computed. Lilly & Gunn (1985) have optical and infrared photometry for a sample of galaxies in CL0024+1654 observed spectroscopically by Dressler, Gunn and Schneider (1985). We have used this data to choose the most appropriate SEDs for our starburst and post-starburst models. The most likely explanation for the optically blue colors in most cluster galaxies is star formation. Very few galaxies lie in the Seyfert locus. Abel 1758 has more Seyfert candidates than the other clusters, we observed. It seems possible to roughly sort types of galaxies in clusters by color alone. The cluster population seems to vary considerably between clusters, but our K selected sample has few Seyferts in any cluster.

Shier, Lisa M.↗

A GPU-based compressible combustion solver for applications exhibiting disparate space and time scales

High-speed chemically active flows pose significant computational challenges due to their disparate space and time scales, with stiff chemistry often dominating simulation time. While modern scientific computing programs achieve exascale performance by leveraging graphics processing units (GPUs), existing GPU-based compressible combustion solvers face critical limitations in memory management, load balancing, and handling the highly localized nature of chemical reactions. To this end, we present a high-performance compressible reacting flow solver built on the AMReX framework and optimized for multi-GPU settings. Here, our approach addresses three GPU performance bottlenecks: memory access patterns through column-major storage optimization, computational workload variability via a bulk-sparse integration strategy for chemical kinetics, and multi-GPU load distribution for adaptive mesh refinement applications. The solver adapts existing matrix-based chemical kinetics formulations to multi-grid contexts. Using representative combustion applications, including 2D and 3D detonations and a 3D jet-in-crossflow configuration, we demonstrate 1.4–5× performance improvements over initial implementations on an in-house cluster of NVIDIA H100 GPUs, and near-ideal weak scaling on the Frontier supercomputer (Oak Ridge Leadership Computing Facility) with up to 1024 AMD Instinct MI250X GPUs. Roofline analysis reveals substantial improvements in arithmetic intensity for both convection (∼ 10 ×) and chemistry (∼ 4 ×) routines, confirming efficient utilization of GPU memory bandwidth and computational resources.

42 ENGINEERING↗