Search NASASearch

SEARCH · Search NASA

Results for “Distributed Computing”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

Transverse momentum distributions at large- x

We investigate the collinear matching of transverse momentum dependent (TMD) distributions at large values of x, computing and resumming the leading large-x asymptotics for matching coefficients. The large-x resummation is done directly within TMD distributions, ensuring the process-independence of the result. The derived resummation formulas are valid for all TMD distributions (except the pretzelosity). Their application improves perturbative convergence, provides practical estimation for unknown higher-order contributions, and sets restrictions for the nonperturbative part of models. Using the known anomalous dimensions, resummation can reach N 3 LL, often exceeding the accuracy of known coefficient functions.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

Machine Learning-Driven Solvent Screening for Biobased 2,3-Butanediol Extraction

Biobased 2,3-butanediol (2,3-BDO) is a valuable biomass-derived chemical due to its versatility in being transformed into a wide variety of products. However, the separation and purification of 2,3-BDO from fermentation broth remain a significant challenge owing to its high boiling point and hydrophilic nature. Herein, we developed a machine learning (ML)-based screening workflow that uses molecular calculations as training data and requires only a small number of experimental measurements for validation to identify alternative solvent candidates for the liquid–liquid extraction (LLE) of 2,3-BDO from aqueous solution. In particular, 130 density functional theory (DFT) calculations with the implicit solvation method not only built a correlation between the computational partition coefficient and the experimental distribution coefficient of 2,3-BDO but also parameterized an Extra-Trees ML model to screen the distribution coefficient for a wider range of 6717 organic solvents. The experimental measurements of only 24 solvents were needed to validate the computational results. A list of 50 prioritized solvents was proposed for 2,3-BDO LLE, and seven additional experimental measurements were conducted to further verify our selected solvents. The impact of the extraction temperature and solvent-to-feed ratio was also investigated for selected solvents in experiments. Furthermore, this work suggested alternative solvents for 2,3-BDO LLE and proposed a versatile workflow that requires fewer experiments and can be applied to a broader range of LLE studies.

Extraction

Multidimensional Distributional Neural Network Output Demonstrated in Super‐Resolution of Surface Wind Speed

Accurate quantification of uncertainty in neural network predictions remains a central challenge for scientific applications involving high-dimensional, correlated data. While existing methods capture either aleatoric or epistemic uncertainty, few offer closed-form, multidimensional distributions that preserve spatial correlation while remaining computationally tractable. In this work, we present a framework for training neural networks with a multidimensional Gaussian loss, generating a closed-form predictive distribution over outputs informed by non-identically distributed training data. Our approach captures aleatoric uncertainty by iteratively estimating the means and covariance matrices, and is demonstrated on a super-resolution example out-of-training-sample. We leverage a Fourier representation of the covariance matrix to stabilize network training and preserve spatial correlation. We introduce a novel regularization strategy—referred to as information sharing—that interpolates between image-specific and global covariance estimates, enabling convergence of the super-resolution downscaling network trained on image-specific distributional loss functions. This framework allows for efficient sampling, explicit correlation modeling, and extensions to more complex distribution families all without disrupting prediction performance. We demonstrate the method on a surface wind speed downscaling task and discuss its broader applicability to uncertainty-aware prediction in scientific models.

17 WIND ENERGY

Modeling and Characterizing the electron backscatter in a cylindrical anode-based distributed X-ray source

Upcoming advancements in computed tomography architectures warrants the investigation of new X-ray source designs and the impacts that electron backscatter can have on these designs. One such design being investigated is a distributed, cylindrical anode-based X-ray source. For such a distributed X-ray source, we developed a modeling pipeline for simulating electron optics and transport to characterize the quality of the primary X-ray beam and the electron backscatter behavior. We report our results on the energy distributions of the bremsstrahlung spectra; electron backscatter ratio; and spatial, temporal, and energy distributions of backscattered electrons that return to the anode.

46 - INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AN

Conservative velocity mappings for discontinuous Galerkin kinetics

Continuum computational kinetic plasma models evolve the distribution function of a plasma species f s on a phase-space grid over time. In many problems of interest the distribution function has limited extent in velocity space; hence, using a uniform, highly refined mesh would be costly and slow. Nonuniform velocity grids can reduce the computational cost by placing more degrees of freedom where f s is appreciable and fewer where it is not. In this work we introduce a first-of-its kind discontinuous Galerkin approach to nonuniform velocity-space discretization using mapped velocity coordinates. This new method is presented in the context of a gyrokinetic model used to study magnetized plasmas. We create discretizations of collisionless and collisional terms using mappings in a way that exactly conserves particles and energy. Numerical tests of such properties are presented, and we show that this new discretization can reproduce earlier gyrokinetic simulations using grids with up to 6–60 times fewer cells and 22X-60X speed-ups depending on dimensionality, geometry and plasma parameters.

Discontinuous Galerkin

Developing and Distributing HEP Software Stacks with Spack

The Computational Science and AI Directorate at Fermilab is using Spack to support the development efforts of a large number of scientific programmers, in many independent projects and experiments. While independent, these projects share many dependencies. They are typically under continuous and fairly rapid development. They have to support deployment on diverse hardware. This is a different context than is typical for the management of HPC software, where Spack was born. To support our community, we have created a model that enables users to develop code with greater efficiency than is possible with Spack’s current development facilities. In this talk we will present: - a brief introduction to the science we support (particle physics) - how the code we work with is naturally organized into several layers of packages - how we are using Spack to manage those layers - how we leverage the layering to provide efficient support for developers, using our Spack extension “MPD”. - some suggestions for changes or additions to Spack to make such work easier.

Knoepfel, Kyle J. [Fermilab]

Investigating the ecological fallacy through sampling distributions constructed from finite populations

Correlation coefficients and linear regression values computed from group averages can differ from correlation coefficients and linear regression values computed using individual scores. This observation known as the ecological fallacy often assumes that all the individual scores are available from a population. In many situations, one must use a sample from the larger population. In such cases, the computed correlation coefficient and linear regression values will depend on the sample that is chosen and the underlying sampling distribution. The sampling distribution of correlation coefficients and linear regression values for group averages will be identical to the sampling distribution for individuals for normally distributed variables for random samples drawn from infinitely large continuous distributions. However, data that is acquired in practice is often acquired when sampling without replacement from a finite population. Our objective is to demonstrate through Monte Carlo simulations that the sampling distributions for correlation and linear regression will also be similar for individuals and group averages when sampling without replacement from normally distributed variables. These simulations suggest that when a random sample from a population is selected, the correlation coefficients and linear regression values computed from individual scores will not be more accurate in estimating the entire population values compared to samples when group averages are used as long as the sample size is the same.

97 MATHEMATICS AND COMPUTING

Virtual Power Plants: Pilots, Challenges, and Innovations Shaping Future Development

Virtual Power Plants (VPPs) aggregate distributed energy resources (DERs) to provide grid services traditionally delivered by centralized power plants. This article reviews the current state of VPP deployment, highlighting business models, compensation mechanisms, and global pilot projects. While VPPs offer benefits such as grid flexibility, resilience, and cost savings, challenges remain in communication infrastructure, regulatory frameworks, market access, and customer engagement. To address these, we propose a scalable, privacy-preserving hierarchical VPP architecture that coordinates with distribution utilities and preserves customer data. We also present the Integrated T&D Control Room of the Future as a key test bed for validating and accelerating VPP adoption. These innovations can help transition VPPs from pilot programs to integral components of a modern, reliable power grid.

24 POWER TRANSMISSION AND DISTRIBUTION

Miscanthus × giganteus increases soil maximum water holding capacity compared to maize

Soil ecosystem services, like the ability to store water, have been depleted after a century of conventional, annual cropping, and perennial crops offer a solution to this and other agricultural environmental issues. We assessed the impact of Miscanthus × giganteus (miscanthus), a perennial biomass crop, on soil water holding capacity and structure compared to continuous maize (Zea mays L.) at two sites in Iowa. After three growing seasons, we measured the following: (1) maximum water holding capacity (MWHC) with and without soil structure, and (2) total porosity and pore size distribution (PSD) via micro-computed tomography (microCT). Miscanthus increased MWHC by 14.7% across both sites relative to maize (p = 0.002), and we attributed this to structural changes due to the lack of a crop effect when measured on structureless soils. No significant changes were detected in soil organic matter, texture, total porosity, or PSD that could explain the increase in MWHC under miscanthus. Our findings suggest that the increases in MWHC are primarily due to structural changes rather than increases in soil organic matter or porosity (at least porosity detectable by microCT). This study highlights miscanthus' potential to enhance soil water storage and underscores the need for further investigation to clarify the mechanisms through which this biomass crop influences soil structural properties.

60 APPLIED LIFE SCIENCES

Ability of x‐ray computed tomography to resolve critical flaw size in laser‐based, paste stereolithography ceramic printing of alumina

Abstract Complex alumina parts were printed using vat photopolymerization (VPP), which is a stereolithography‐based additive manufacturing (AM) technique used to shape ceramic preforms, or green parts. The critical flaw size was determined using classical fracture mechanics techniques. The strength and fracture toughness were measured and compared to flaws detected in x‐ray computed tomography (XCT or CT) distributions as well as the fracture surfaces. The strength was lower compared traditionally made alumina, and that is due to layering effects, slurry defects, and printing defects. The critical flaw size from fracture mechanics was 206 µm. XCT has high enough resolution to detect the critical flaw size and much smaller features, where the average flaw size observed in CT scans was around 80–100 µm. The fracture surfaces indicate that flaws causing failure are larger than that of the critical flaw size (∼300 µm), but fracture surfaces do not show definitive features compared to traditionally made ceramics. Since XCT can observe flaws smaller than the critical flaw size, this method can be used as a screening technique.

36 MATERIALS SCIENCE

Parallel sorting algorithm classification: is manual instrumentation necessary?

Understanding parallel algorithms is crucial for accelerating scientific simulations on complex, distributed memory, high-performance computers. Modern algorithm classification approaches learn semantics directly from source code to differentiate between algorithms, however, accessing source code is not always possible. We can learn about parallel algorithms from observing their performance, as programs running the same algorithms and using the same hardware should exhibit similar performance characteristics. We present an approach to learn algorithm classes from parallel performance data directly in order to classify algorithms without access to the source code. We extend previous work to enable classifying parallel sorting algorithms using automatic instrumentation instead of requiring manual region annotations in the source code. In this work, we design and demonstrate a study for classification of parallel sorting algorithms using parallel performance data collected from automatic instrumentation, and evaluate the performance of our new methodology on classification. We leverage Caliper to collect the performance data, Thicket for our exploratory data analysis (EDA), and PyTorch and Scikit-learn to evaluate the effectiveness of random forests, support vector machines (SVMs), decision trees, neural networks, and logistic regressions on parallel performance data. Additionally, we study noise in parallel performance data, whether the removal of noise and pre-processing of the data is necessary to accurately classify parallel sorting algorithms, and determine the effectiveness of features created from performance data. In conclusion, we demonstrate classification accuracy for these five different models of up to 97.7% across four different parallel algorithm classes.

Algorithm Classification

Scalable Generation of High-fidelity Synthetic Population Ensembles

Used within social simulations, synthetic population ensembles enable uncertainty quantification (UQ) methods for obtaining more robust model inference and prediction. A synthetic population ensemble is a series of plausible virtual reconstructions of an area’s population at the granularity of people and residences, generated stochastically to preserve privacy of the source population survey’s respondents. In this paper, we demonstrate the production of large synthetic population ensembles for the U.S. via Oak Ridge National Laboratory’s UrbanPop framework to support modeling of high spatial resolution energy affordability metrics from nationwide social surveys in collaboration with the fusionACS project. The study involves two scenarios: creating ensembles for (1) 17 U.S. metropolitan areas in 2019 and (2) full U.S. Census Divisions in 2023, with each scenario consisting of 41 population instances (a base realization and 40 replicates). To accomplish this task at scale, we configured an integrated system within a research cloud, comprised of virtual containerizations, GPU-enhanced functionality, and orchestrated deployments of UrbanPop’s maturing Likeness Python ecosystem. Results demonstrate we maintained high-fidelity approximations of residential totals by areas of interest and the demographic characteristics of neighborhoods while reducing manual workflow burdens. Finally, we discuss plans to fine-tune and further develop our automated workflows for truly distributed job orchestration to increase computational efficiency, as well as provide an outlook for broadening applications of the ensembles.

Cluster computing

Self-consistent modeling of tokamak edge plasma transport with lithium sources

Magnetic confinement fusion devices require effective heat and particle exhaust solutions on the divertor plates to operate sustainably, especially under reactor-relevant conditions. Liquid lithium divertors have been proposed to address two major challenges: control of excessive heat flux to plasma-facing components through vapor shielding and minimization of core plasma contamination from impurities. The National Spherical Torus Experiment-Upgrade (NSTX-U) will explore lithium as a divertor material due to its potential to meet both objectives. We present a self-consistent coupling framework between the plasma boundary transport code UEDGE and the lithium wall transport code Wall–Li to evaluate the feasibility and operational limits of lithium-based divertors. The model aims to optimize lithium sourcing levels to prevent core plasma contamination via fuel dilution while ensuring divertor protection through vapor shielding. This integrated framework, applicable to any tokamak with lithium sources, dynamically adjusts lithium sourcing based on local plasma conditions and surface temperature. The coupled model is tested using NSTX-like geometry and plasma conditions to assess its performance and reliability. Wall–Li calculates lithium fluxes from plasma-facing components, incorporating physical sputtering, thermally enhanced sputtering, and evaporation driven by surface temperature and ion flux. These fluxes are reintroduced into UEDGE as neutral lithium atoms, enabling simulation of their transport and distribution within the plasma. UEDGE computes plasma and neutral transport, surface heat flux, and iteratively feeds this information back to Wall–Li. A small time step is employed to ensure numerical stability and convergence, enabling accurate simulations over typical tokamak discharge durations. This integrated modeling approach provides a robust tool for identifying operational regimes that balance effective lithium sourcing with minimal core plasma contamination, offering critical insights for optimizing lithium-based divertor systems in current and future fusion devices.

Magnetic confinement fusion

Cosmic structure strikes back: The elimination of vector-mediated nonstandard interaction models as a mechanism for sterile neutrino dark matter production

We revisit sterile neutrino production enabled by nonstandard interactions (NSIs) among active neutrinos mediated by new bosons. We focus on vector mediators, including neutrinophilic, gauged 𝐿 𝜇 −𝐿 𝜏 , and 𝐵−𝐿 realizations, that modify in-medium dispersion and scattering, thereby altering the active-sterile conversion history. Building on a novel production framework with NSI thermal potentials and collision integrals, we compute nonthermal phase-space distributions across sterile neutrino mixing and NSI parameters and map each point to an equivalent thermal warm dark matter particle mass 𝑚 th via linear theory transfer function fitting with the cosmological structure formation Boltzmann solver. This enables a direct reinterpretation of state-of-the-art structure formation limits from Milky Way satellites, strong lensing, and the Lyman-𝛼 forest. These limits, in conjunction with x-ray decay searches, as well as results from a wide variety of particle physics experiments allow for a more complete examination of these models. We find that these vector-mediated models are ruled out when the full combination of current constraints, listed above, are taken into account. NSI scalar-mediated models and models with low reheating temperatures remain viable.

cosmology

Dispatch Manager for NEML2 Constitutive Model Calculations Embedded in MOOSE

This report describes the extended capabilities of the NEML2 constitutive modeling library, including a flexible and efficient work dispatching system designed to leverage both CPU and GPU resources. This enhancement addresses one of the primary computational challenges in large-scale simulations: the ability to distribute and execute batches of material model evaluations across heterogeneous computing devices. The new dispatch system introduces a modular set of dispatcher and scheduler classes that coordinate the flow of data and execution between devices. The dispatcher is responsible for efficiently packaging work, managing device-specific memory operations, and synchronizing results. This modularity allows for extensibility, making it straightforward to integrate additional computing backends in the future. From an implementation standpoint, the dispatcher system interfaces seamlessly with NEML2's existing models. They handle device-aware tensor operations, optimize memory transfers, and support asynchronous execution when applicable. This design ensures that batches of material points can be evaluated concurrently, substantially improving throughput compared to previous single-device or serial implementations. These improvements not only enhance the raw performance of NEML2 but also improve its usability in multiscale and high-fidelity simulations, where the simultaneous evaluation of large material point batches is critical. Benchmarks included in the report demonstrate the system’s scalability, highlighting its effectiveness when leveraging modern GPU architectures.

22 GENERAL STUDIES OF NUCLEAR REACTORS

Optimization Studies of Radiation Shielding for the PIP-II Project at Fermilab

The PIP-II project at Fermilab, which includes an 800-MeV superconducting LINAC, demands rigorous radiation shielding optimization to meet safety requirements. We updated the MARS geometry model to reflect new magnet and collimator designs and introduced high-resolution detector planes to better capture radiation field distributions. To overcome the significant computational demands, we implemented a well-known branching technique that drastically reduced simulation runtimes while maintaining statistical integrity. This was achieved through particle splitting and the application of Russian Roulette techniques. Additionally, new graphical tools were created to streamline data visualization and MARS code usability.

Makovec, Alajos [Fermilab] (ORCID:0000000286157492

Conditional Pseudo-Reversible Normalizing Flow for Surrogate Modeling in Quantifying Uncertainty Propagation

We introduce a conditional pseudo-reversible normalizing flow (PR-NF) that directly learns conditional probability distributions from noisy physical models to efficiently quantify both forward and inverse uncertainty propagation. Traditional surrogate modeling approaches approximate only the deterministic component of physical models, requiring separate noise characterization and computationally expensive sampling methods for inverse problems. Here, in this work, we develop the conditional PR-NF model to directly learn and efficiently generate samples from the conditional probability density functions (PDFs). The training process utilizes dataset consisting of input-output pairs without requiring prior knowledge about the noise and the function. Once trained, our model efficiently generates samples from conditional PDFs for any input within the training domain. Moreover, the pseudo-reversibility feature allows for the use of fully connected neural network architectures, which simplifies the implementation and enables theoretical analysis. We provide a rigorous convergence analysis of the conditional PR-NF model, showing its ability to converge to the target conditional PDF using the Kullback−Leibler divergence. To demonstrate the effectiveness of our method, we apply it to several benchmark tests and a real-world geologic carbon storage problem.

97 MATHEMATICS AND COMPUTING

Thermo-hydraulic steam pipe models for district heating simulations: Simplifications to balance accuracy and simulation speed

Steam piping networks are essential for optimizing performance in industrial processes and district heating systems. However, dynamic models that balance thermo-hydraulic accuracy with computational efficiency remain limited. In response, this paper presents a new discretized steam pipe model based on the plug flow approach, capturing key thermo-hydraulic behaviors while simplifying steam phase change processes. Implemented in Modelica, the model accurately calculates temperature and pressure distributions along steam pipelines. To improve computational efficiency for district-scale simulations, five model simplifications are introduced: lumped thermo-hydraulic functions, empirical correlations, fluid state approximations, steady-state dynamics and inclusion of flow derivatives. These simplified models achieve 85%-98% accuracy in predicting pressure drop and condensation losses, including dynamic condensate behavior during pipe warm-up—a factor often overlooked in existing models. The models support diverse network configurations, scaling effectively to systems with multiple distribution pipes and connected building loads. Discrete models provide detailed insights but exhibit a cubic increase in simulation time as the network scales by N connected building O(N 2.42 ). In contrast, lumped models simulate 10–28 times faster than discrete, offering quadratic scaling of simulation time O(N 1.73 ). However, they still require 6 times more computation time than a lossless network, highlighting the inherent computational challenges of modeling compressible fluid flow. In conclusion, the steady-state lumped variant, with its near-linear scalability in computational time O(N 1.01 ), emerges as an efficient solution for preliminary design evaluations and extensive parametric studies.

15 GEOTHERMAL ENERGY