Search NASA⌕ Search

SEARCH · Search NASA

Results for “cluster computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 433 records · Page 24

Power Profile Monitoring and Tracking Evolution of System-Wide HPC Workloads

The power & energy demands of HPC machines have grown significantly. Modern exascale HPC systems require tens of megawatts of combined power for computing resources and cooling facilities at full capacity. The current energy trend is not sustainable for future HPC systems, and there is a need to work toward the energy efficiency aspect of HPC performance. Energy awareness of the HPC applications at the job level is essential for running an efficient HPC system. This work aims to develop a pipeline to provide a production-level system-wide overview of the HPC workloads' power profile while handling evolving workloads exhibiting new power trends. We developed an open-set classification model for HPC jobs based on the properties of power profiles to continuously provide a system-wide holistic view of recently completed jobs. The pipeline helps continuously monitor the job-level power usage pattern of HPC and enables us to capture the new trends in applications' power behavior. We employed a comprehensive set of techniques to generate job-level data, custom-designed feature extraction methods to extract critical features from jobs' power profiles, clustering techniques powered by generative modeling, and open-set classification for identifying job profiles into known classes or an unknown set. With extensive evaluations, we demonstrate the effectiveness of each component in our pipeline. We provide an analysis of the resulting clusters that characterize the power profile landscape of the Summit supercomputer from more than 60K jobs executed in a year. The open-set classification classifies the known data sets into known classes with high accuracy and identifies unknown data noints with over 85% accuracy.

Karimi, Ahmad Maroof↗

Neutron-Friendly Li-Ion Battery Coin Cell for In Situ 3D Visualization of Li Plating

Advanced battery characterization using in situ/operando neutron imaging is critical for uncovering degradation modes such as lithium (Li) plating in Li-ion batteries (LIBs). However, conventional LIBs hinder operando neutron radiography (NR) and in situ neutron micro-computed tomography (N-μCT) for visualizing Li plating near the graphite-separator interface due to strong attenuation from hydrogen-rich components like PP–PE–PP separators, electrolyte, and Fe-based spacers. In this work, we designed and tested a neutron-friendly battery (NFB) optimized for in situ Li detection during extreme fast charging (XFC). Guided by neutron attenuation cross-sections and material transmission, the NFB enables clear visualization at the graphite–separator interface, which is typically opaque in standard LIBs. Electrochemical tests show the NFB exhibits voltage/current responses like standard cells for up to 50 XFC cycles. However, its lower reversibility and capacity are likely due to Cu-coated Al spacer degradation from delamination or corrosion. We propose titanium spacers as a more stable alternative, albeit requiring custom machining. Using this optimized cell, we achieved simultaneous neutron tomography of multiple cells, capturing in situ 3D images of dead Li accumulation, particularly near graphite edges. These heterogeneous deposits and disconnected Li clusters suggest localized current density hotspots during XFC.

25 ENERGY STORAGE↗

Resolving the Solvation Structure and Transport Properties of Aqueous Zinc Electrolytes from Salt-in-Water to Water-in-Salt Using Neural Network Potential

Zn Cl 2 solutions are promising electrolytes for aqueous zinc-ion batteries. Here, we report a joint computational and experimental study of the structural and dynamic properties of aqueous Zn Cl 2 electrolytes with concentrations ranging from salt-in-water to water-in-salt (WIS). By developing a neural network potential (NNP) model, we perform molecular dynamics (MD) simulations with accuracy but at much larger lengths and longer timescales. The NNP predicted structures are validated by the structure factors measured by X-ray total scattering experiments. The MD trajectories provide a comprehensive and quantitative picture of the Zn 2 + solvation shell structures. Additionally, we find that the O − H covalent bonds in water are strengthened with increasing salt concentration, thus expanding the electrochemical stability window of aqueous electrolytes. In terms of dynamic properties, the calculated and experimentally measured conductivities are in good agreement. Through the analysis of the calculated cation transference number, we propose a three-stage charge carrier transport mechanism with increasing concentration: independent ion transport, strongly correlated ion transport, and small positive charge carrier diffusion through negatively charged polymeric clusters. Our study provides fundamental atomic scale insights into the structure and transport properties of the Zn Cl 2 electrolyte that can aid the optimization and development of WIS electrolytes. Published by the American Physical Society 2025

25 ENERGY STORAGE↗

ArborX 2.0

ArborX library tackles a problem of efficiently finding geometric objects that are close in space. Variations of this problem, such as finding the nearest neighbors of a point, or finding all objects within a certain distance, are inherent components of applications in many fields. The data may be large so that solving the problem efficiently may require significant computational resources, such as multiple processors or accelerators such as general purpose GPUs. ArborX' main advantage in its ability to solve large problems efficiently utilizing a combination of distributed and on-node parallelism. ArborX can be run efficiently on a wide variety of hardware, including GPUs from different vendors, which distinguishes it from other available libraries which typically choose only few of these. The other advantage is that it supports both types of user problems: spatial problems (useful for intersections and finding objects within certain distance), and nearest neighbor problems. ArborX also supports flexible interface in its interaction with a user. Particularly, it allows a user to call user's own function on a positive match, a functionality not rarely available in other libraries. ArborX implements construction and traversal algorithms using efficient tree structures, such as bounding volume hierarchy (BVH). At its core, ArborX uses linear BVH for its low construction cost and sufficient quality. ArborX implements both spatial and nearest-neighbor traversal algorithms. ArborX also provides several clustering algorithms (minimum spanning tree, DBSCAN, HDBSCAN*), interpolation using minimum least squares and ray tracing. ArborX is written using C++, and is parallelized using the message passing interface (MPI) for the distributed communication, and the Kokkos library for on-node parallelism. This approach allows ArborX to be run on a wide variety of hardware, from common laptops and desktops to supercomputers while using the same codebase.

Prokopenko, Andrey [Oak Ridge National Laboratory ↗

Predicting U 3 O 8 powder processing conditions: An AI/ML approach analyzing deep learning embeddings of SEM micrographs

High-resolution SEM images of uranium-oxide powders encode micro- and nanoscale clues to their synthesis route and calcination temperature. We trained a ResNet-50 model on 11 commercial-scale U₃O₈ classes, ammonium diuranate (ADU) or uranyl peroxide (H₂O₂) precursors calcined at temperatures ranging from 400 to 750 °C and added a 256-D projection head before the classifier to analyze the learned representation. The best of eight seeds reached 92.4 % accuracy on reserved testing data, but our focus is the structure of the embedding space rather than the accuracy and labels. We quantify class relatedness in the original 256-D space using centroid similarity and distributional distances, and we use Uniform Manifold Approximation Projection (UMAP) for visualization. ‘Unknown’ images from different preparation methods, SEM operators, and from the literature localized near the expected classes under a nearest-centroid analysis without retraining, as well as clustered in similar UMAP space. In conclusion, this embedding-centered workflow complements black-box classification by providing quantitative, similarity-based comparisons of U₃O₈ morphologies and reduces storage space by up to 98 % for image data used in millisecond vector search comparisons.

36 MATERIALS SCIENCE↗

Spatially Accelerated Winding Numbers for Curved Geometry

The generalized winding number (GWN) is a scalar field that supports robust containment queries on curved geometry, including non-watertight, overlapping, and nested boundary representations. While queries can be easily parallelized over samples, direct evaluation on parametric curves and surfaces remains costly for large and complex models. Fast, state-of-the-art GWN approaches leverage a spatial index to approximate the GWN, typically coupled with a Taylor expansion which approximates the GWN contribution for far clusters of geometric primitives. However, such methods operate only on discrete inputs such as triangle meshes and point clouds, and would introduce containment errors near boundaries if applied to curved input. We extend support for fast GWN evaluation over arbitrary collections of NURBS curves in 2D and trimmed NURBS patches in 3D via a Bounding Volume Hierarchy that stores efficiently precomputed moment data in the hierarchy nodes. When querying the hierarchy, approximations for far clusters are used alongside direct evaluation for nearby NURBS primitives, achieving sub-linear complexity while preserving the geometric features in the vicinity of the query point. Central to our performance improvements is an adaptive subdivision strategy for NURBS primitives during a preprocessing phase, creating better spatial partitions while retaining the same accuracy for containment decisions as a direct evaluation. We demonstrate the performance and accuracy of our approach across a large collection of 2D and 3D datasets.

Computer science↗

Superstructure Optimization for Brine Valorization from Brackish Water Desalination

This poster presents preliminary results from a superstructure optimization framework developed to identify cost-optimal brine valorization configurations for brackish water desalination plants across diverse U.S. regional feed chemistries. The study uses brackish groundwater compositions from Arizona, California, Florida, New Mexico, and Texas. Using Pyomo Generalized Disjunctive Programming (GDP) within the WaterTAP modeling environment, the optimization framework simultaneously evaluates thousands of candidate treatment configurations, spanning nanofiltration, reverse osmosis, and chemical precipitation, to minimize the levelized cost of water (LCOW) while meeting water recovery targets and product recovery constraints. Results across eight representative feed clusters demonstrate water recovery rates of 57–87% and net LCOW values ranging from -$0.032/m³ (net revenue-positive) to $0.80/m. Notably, no single process configuration was optimal across all feed types, underscoring the necessity of feed-specific optimization. Products targeted include calcium carbonate (CaCO₃) at $0.01/kg and sodium chloride (NaCl) at $0.10/kg, both at 95% purity, with product revenues offsetting treatment costs in several scenarios. The work advances NAWI's process systems engineering capabilities for multi-configuration screening.

58 GEOSCIENCES↗

Comparative genomic analysis of thermophilic fungi reveals convergent evolutionary adaptations and gene losses

Thermophily is a trait scattered across the fungal tree of life, with its highest prevalence within three fungal families (Chaetomiaceae, Thermoascaceae, and Trichocomaceae), as well as some members of the phylum Mucoromycota. We examined 37 thermophilic and thermotolerant species and 42 mesophilic species for this study and identified thermophily as the ancestral state of all three prominent families of thermophilic fungi. Thermophilic fungal genomes were found to encode various thermostable enzymes, including carbohydrate-active enzymes such as endoxylanases, which are useful for many industrial applications. At the same time, the overall gene counts, especially in gene families responsible for microbial defense such as secondary metabolism, are reduced in thermophiles compared to mesophiles. We also found a reduction in the core genome size of thermophiles in both the Chaetomiaceae family and the Eurotiomycetes class. The Gene Ontology terms lost in thermophilic fungi include primary metabolism, transporters, UV response, and O-methyltransferases. Comparative genomics analysis also revealed higher GC content in the third base of codons (GC3) and a lower effective number of codons in fungal thermophiles than in both thermotolerant and mesophilic fungi. Furthermore, using the Support Vector Machine classifier, we identified several Pfam domains capable of discriminating between genomes of thermophiles and mesophiles with 94% accuracy. Using AlphaFold2 to predict protein structures of endoxylanases (GH10), we built a similarity network based on the structures. We found that the number of disulfide bonds appears important for protein structure, and the network clusters based on protein structures correlate with the optimal activity temperature. Thus, comparative genomics offers new insights into the biology, adaptation, and evolutionary history of thermophilic fungi while providing a parts list for bioengineering applications.

59 BASIC BIOLOGICAL SCIENCES↗

Reduced-basis method for few-body bound-state emulation

Recent advances in both theoretical and computational methods have enabled large-scale, precision calculations of the properties of atomic nuclei. With the growing complexity of modern nuclear theory, however, also comes the need for novel methods to perform systematic studies and quantify the uncertainties of models when confronted with experimental data. Here, this study presents an application of such an approach, the reduced basis method, to substantially lower computational costs by constructing a significantly smaller Hamiltonian subspace informed by previous solutions. Our method shows comparable efficiency and accuracy to other dimensionality reduction techniques on an artificial three-body bound system while providing a richer representation of physical information in its projection and training subspace. This methodological advancement can be applied in other contexts and has the potential to greatly improve our ability to systematically explore theoretical models and thus enhance our understanding of the fundamental properties of nuclear systems.

cluster models↗

Feature engineering descriptors, transforms, and machine learning for grain boundaries and variable-sized atom clusters

Abstract Obtaining microscopic structure-property relationships for grain boundaries is challenging due to their complex atomic structures. Recent efforts use machine learning to derive these relationships, but the way the atomic grain boundary structure is represented can have a significant impact on the predictions. Key steps for property prediction common to grain boundaries and other variable-sized atom clustered structures include: (1) describing the atomic structure as a feature matrix, (2) transforming the variable-sized feature matrix to a fixed length common to all structures, and (3) applying a machine learning algorithm to predict properties from the transformed matrices. We examine how these steps and different combinations of engineered features impact the accuracy of grain boundary energy predictions using a database of over 7000 grain boundaries. Additionally, we assess how different engineered features support interpretability, offering insights into the physics of the structure-property relationships.

36 MATERIALS SCIENCE↗

Global Archaeal Diversity Revealed Through Massive Data Integration: Uncovering Just Tip of Iceberg

The domain of Archaea has gathered significant interest for its ecological and biotechnological potential and its role in helping us to understand the evolutionary history of Eukaryotes. In comparison to the bacterial domain, the number of adequately described members in Archaea is relatively low, with less than 1000 species described. It is not clear whether this is solely due to the cultivation difficulty of its members or, indeed, the domain is characterized by evolutionary constraints that keep the number of species relatively low. Based on molecular evidence that bypasses the difficulties of formal cultivation and characterization, several novel clades have been proposed, enabling insights into their metabolism and physiology. Given the extent of global sampling and sequencing efforts, it is now possible and meaningful to question the magnitude of global archaeal diversity based on molecular evidence. To do so, we extracted all sequences classified as Archaea from 500 thousand amplicon samples available in public repositories. After processing through our highly conservative pipeline, we named this comprehensive resource the ‘Global Archaea Diversity’ (GAD), which encompassed nearly 3 million molecular species clusters at 97% similarity, and organized it into over 500 thousand genera and nearly 100 thousand families. Saline environments have contributed the most to the novel taxa of this previously unseen diversity. The majority of those 16S rRNA gene sequence fragments were verified by matches in metagenomic datasets from IMG/M. These findings reveal a vast and previously overlooked diversity within the Archaea, offering insights into their ecological roles and evolutionary importance while establishing a foundation for the future study and characterization of this intriguing domain of life.

59 BASIC BIOLOGICAL SCIENCES↗

An interregional optimization approach for time series aggregation in continent-scale electricity system models

Modeling electric power systems with high shares of weather-dependent resources requires tradeoffs between temporal, spatial, and operational resolution. Many studies perform time series aggregation using clustering algorithms to reduce the temporal dimension, but when modeling continent-scale electricity systems that are large enough to contain multiple independent weather systems, this approach requires large numbers of representative periods to minimize errors in regional wind and solar capacity factors. Here, a new optimization-based approach for representative period selection and weighting is introduced that minimizes regional errors in average renewable capacity factors and electricity demand. The method delivers higher regional fidelity with fewer representative periods than alternative clustering methods when applied to wind, solar, and demand profiles for the contiguous United States. When representative periods are selected from multiple weather years, the optimized method reproduces regional averages with lower error than a complete 365-day time series from any single weather year. The method identifies only representative (as opposed to outlying) periods but can be combined with an iterative "stress period" identification approach to guide efficient decision-making considering both average and high-risk weather conditions.

24 POWER TRANSMISSION AND DISTRIBUTION↗

A comprehensive numerical investigation on spray models for Direct-Injection Spark-Ignition engines

Gasoline direct-injection spark-ignition (DISI) engines generate a large portion of their unburned hydrocarbon (UHC) and soot emissions during the cold-start phase. A predictive computational fluid dynamics (CFD) modeling framework can be used to understand the physical processes that characterize fuel spray evolution and fuel-film formation at cold start conditions, which can help to reduce engine-out particulate emissions. This study systematically evaluated spray submodels and developed a set of simulation best practices for physical-numerical submodels with the goal of enabling accurate simulations of liquid spray behavior in a DISI engine. Three comprehensive experimental datasets containing free-spray projected liquid volume (PLV), liquid volume fraction (LVF), and near-field X-ray radiography data were used to validate the simulation results and evaluate the spray submodels. Systematic analysis delved into injected parcel distribution, droplet collision, spray breakup, and evaporation via a detailed assessment of the relevant spray submodels. Moreover, the effects of turbulence models and the initial turbulent flow properties on the liquid spray evolution were examined. Based on extensive calibration efforts, a set of simulation best practices for the free spray was developed and validated against the PLV/LVF data. Simulation results indicated that the uniform distribution for parcel initialization, coupled with appropriate droplet collision submodels, provides an improved spray morphology compared to the cluster distribution. The findings also underscored the importance of calibrating the Kelvin-Helmholtz Rayleigh-Taylor (KH-RT) breakup model constants and droplet heat transfer coefficient scaling factor to achieve favorable agreement regarding measured liquid penetration and spray widths. In conclusion, this study marks a substantial stride towards accurately predicting fuel film evolution and soot formation within DISI engine performance.

ECN Spray G↗

Unsteady Land-Sea Breeze Circulations in the Presence of a Synoptic Pressure Forcing

Unsteady land-sea breezes (LSBs) that result from time-varying surface temperature contrasts Δθ(t) are explored in the presence of a constant synoptic pressure forcing, M g , oriented from sea to land (α = 0°) or land to sea (α = 180°). Large eddy simulations reveal the development of four distinctive regimes, depending on the joint interaction between M g , α, and Δθ(t) in modulating the fine-scale dynamics. Time lags, computed as the shifts that maximize correlation coefficients of the velocity between the unsteady and the corresponding steady scenarios at Δθ = Δθ max , are found to be significant and to extend 2 hr longer for α = 0° compared to α = 180°. These diurnal dynamics result in nonequilibrium conditions that are significantly affected by the flow history, and that behave differently over the two patches for the different α’s. Turbulence is found to be out of equilibrium with the mean flow, and the mean itself is found to be out of equilibrium with the thermal forcing. The sea surface heat flux is consistently more sensitive than its land counterpart to the time-varying external forcing Δθ(t), and more so for synoptic forcing from land to sea (α = 180°). Hence, although the land reaches equilibrium faster, the sea patch is found to exert a stronger control on the turbulence-mean flow equilibrium response. Finally, the vertical velocity profile at the shore and shore-normal velocity transects at the first grid level are shown to encode the multiscale regimes of the LSBs evolution and can thus be used to identify these regimes using k-means clustering.

58 GEOSCIENCES↗

Cosmological constraints from a joint DESI DR1 Full-Shape and DR2 BAO

We present a cosmological analysis combining full-shape (FS) clustering measurements from the Dark Energy Spectroscopic Instrument (DESI) DR1 with baryon acoustic oscillation (BAO) measurements from DESI DR2. To achieve a robust combination that accounts for the correlation between the two data releases, we employ the ShapeFit compression method and estimate the joint covariance using EZmocks. This compressed approach inherently mitigates the prior volume effects that have previously dominated Bayesian constraints from DESI data with minimal external priors. Consequently, we obtain — for the first time within a Bayesian framework — reliable DESI-only constraints on extensions to ΛCDM using only a Big Bang Nucleosynthesis prior on the baryon density and a wide prior on the spectral index. In flat ΛCDM, we find Ω m = 0.3035 ± 0.0085, h = 0.6876 ± 0.0059, and σ 8 = 0.822 ± 0.034. For the w 0 w a CDM dynamical dark energy model, we measure w 0 = -0.49 ± 0.25 and w a = -1.52 ± 0.77, improving constraints by ∼ 30% relative to the analogous DR1 measurement and reducing the discrepancy with ΛCDM to 1.4σ when compared to BAO only analyses. We also report competitive limits on the sum of neutrino masses and spatial curvature. This work demonstrates that the ShapeFit compression provides a prior-robust and computationally efficient pathway to constrain beyond-ΛCDM physics with large-scale structure.

baryon acoustic oscillations↗

Automated pipeline processing X-ray diffraction data from dynamic compression experiments on the Extreme Conditions Beamline of PETRA III

Presented and discussed here is the implementation of a software solution that provides prompt X-ray diffraction data analysis during fast dynamic compression experiments conducted within the dynamic diamond anvil cell technique. It includes efficient data collection, streaming of data and metadata to a high-performance cluster (HPC), fast azimuthal data integration on the cluster, and tools for controlling the data processing steps and visualizing the data using the DIOPTAS software package. This data processing pipeline is invaluable for a great number of studies. The potential of the pipeline is illustrated with two examples of data collected on ammonia–water mixtures and multiphase mineral assemblies under high pressure. The pipeline is designed to be generic in nature and could be readily adapted to provide rapid feedback for many other X-ray diffraction techniques, e.g. large-volume press studies, in situ stress/strain studies, phase transformation studies, chemical reactions studied with high-resolution diffraction etc.

97 MATHEMATICS AND COMPUTING↗

Phase separation, edge currents, and Hall effect for active matter with Magnus dynamics

Here we examine run-and-tumble disks in two-dimensional systems where the particles also have a Magnus component to their dynamics. For increased activity, we find that the system forms a motility-induced phase-separated (MIPS) state with chiral edge flow around the clusters, where the direction of the current is correlated with the sign of the Magnus term. The stability of the MIPS state is non-monotonic as a function of increasing Magnus term amplitude, with the MIPS region first extending down to lower activities followed by a break up of MIPS at large Magnus amplitudes into a gel-like state. We examine the dynamics in the presence of quenched disorder and a uniform drive and find that the bulk flow exhibits a drive-dependent Hall angle. This is a result of the side jump effect produced by scattering from the pinning sites and is similar to the behavior found for skyrmions in chiral magnets with quenched disorder.

97 MATHEMATICS AND COMPUTING↗