Search NASA⌕ Search

SEARCH · Search NASA

Results for “distributed solutions”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 235 records · Page 13

Enabling Efficient Sparse Computations using Linear Algebra Aware Compilers

This project developed the LAPIS compiler framework, built on the Multilevel Intermediate Representation (MLIR), to optimize sparse linear algebra operations and support performance portability across diverse architectures. The main innovation of LAPIS is the Kokkos dialect, which allows for lowering codes from a high productivity language to different architectures in an elegant way. The dialect also allows the conversion of lower-level MLIR code to C++ Kokkos code, facilitating the integration of scientific machine learning (SciML) models into applications. To extend LAPIS for distributed memory architectures, a new partition dialect was created to manage the distribution of sparse tensors and express communication patterns for sparse linear algebra operations. This dialect also supports the distributed execution of operators and includes algorithmic optimizations to minimize communication to improve performance. The project also demonstrates that MLIR can enable effective linear algebra-level optimizations, improving performance on different GPUs for both sparse and dense linear algebra kernels. Key applications of LAPIS include sparse linear algebra and graph kernels, TenSQL, a relational database management solution built on GraphBLAS, and the development of subgraph isomorphism and monomorphism kernels, showcasing performance portability. In summary, the LAPIS framework supports productivity, performance, portability, and distributed memory execution, while also enabling linear algebra-level optimizations that are challenging in traditional programming languages, with successful applications ranging from simple sparse linear algebra to complex graph kernels.

97 MATHEMATICS AND COMPUTING↗

Axisymmetric hybrid Vlasov equilibria with applications to tokamak plasmas

We derive axisymmetric equilibrium equations in the context of the hybrid Vlasov model with kinetic ions and massless fluid electrons, assuming isothermal electrons and deformed Maxwellian distribution functions for the kinetic ions. The equilibrium system comprises a Grad–Shafranov partial differential equation and an integral equation. These equations can be utilized to calculate the equilibrium magnetic field and ion distribution function, respectively, for given particle density or given ion and electron toroidal current density profiles. The resulting solutions describe states characterized by toroidal plasma rotation and toroidal electric current density. Additionally, due to the presence of fluid electrons, these equilibria also exhibit a poloidal current density component. This is in contrast to the fully kinetic Vlasov model, where axisymmetric Jeans equilibria can only accommodate toroidal currents and flows, given the absence of a third integral of the microscopic motion.

Physics↗

New Results on Communication- and Memory-Aware Load Balancing Model and Algorithms

While load balancing in distributed-memory computing has been well-studied, we present an innovative approach to this problem: a unified, reduced-order model that combines three key components to describe “work” in a distributed system: computation, communication, and memory. Our model enables an optimizer to explore complex tradeoffs in task placement, such as augmented parallelism, at the expense of data replication increasing memory usage. We propose a fully distributed, heuristic-based load balancing optimization algorithm, and demonstrate that it quickly finds close-to-optimal solutions. We formalize the complex optimization problem as a mixed-integer linear program, and compare it to our strategy. Finally, we show that when applied to an electromagnetics code, our approach obtains up to 2.3x speedups for the imbalanced execution.

97 MATHEMATICS AND COMPUTING↗

A framework for discrete optimization of stellarator coils

Designing magnets for three-dimensional plasma confinement is a key task for advancing the stellarator as a fusion reactor concept. Stellarator magnets must produce an accurate field while leaving adequate room for other components and being reasonably simple to construct and assemble. In this paper, a framework for coil design and optimization is introduced that enables the attainment of sparse magnet solutions with arbitrary restrictions on where coils may be located. The solution space is formulated as a 'wireframe' consisting of a mesh of interconnected wire segments enclosing the plasma. Two methods are developed for optimizing the current distribution on a wireframe: Regularized Constrained Least Squares, which uses a linear least-squares approach to optimize the currents in each segment, and Greedy Stellarator Coil Optimization, a fully discrete procedure in which loops of current are added to the mesh one by one to achieve the desired magnetic field on the plasma boundary. Examples are presented of solutions obtainable with each method, some of which achieve high field accuracy while obeying spatial constraints that permit easy assembly.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Polynomial Chaos Surrogate Construction for Random Fields with Parametric Uncertainty

Engineering and applied science rely on computational experiments to rigorously study physical systems. The mathematical models used to probe these systems are highly complex, and sampling-intensive studies often require prohibitively many simulations for acceptable accuracy. Surrogate models provide a means of circumventing the high computational expense of sampling such complex models. In particular, polynomial chaos expansions (PCEs) have been successfully used for uncertainty quantification studies of deterministic models where the dominant source of uncertainty is parametric. We discuss an extension to conventional PCE surrogate modeling to enable surrogate construction for stochastic computational models that have intrinsic noise in addition to parametric uncertainty. We develop a PCE surrogate on a joint space of intrinsic and parametric uncertainty, enabled by Rosenblatt transformations, which are evaluated via kernel density estimation of the associated conditional cumulative distributions. Furthermore, we extend the construction to random field data via the Karhunen–Loève expansion. We then take advantage of closed-form solutions for computing PCE Sobol indices to perform a global sensitivity analysis of the model which quantifies the intrinsic noise contribution to the overall model output variance. Additionally, the resulting joint PCE is generative in the sense that it allows generating random realizations at any input parameter setting that are statistically approximately equivalent to realizations from the underlying stochastic model. The method is demonstrated on a chemical catalysis example model and a synthetic example controlled by a parameter that enables a switch from unimodal to bimodal response distributions.

97 MATHEMATICS AND COMPUTING↗

Fast permeability measurement for tight reservoir cores using only initial data of the one chamber pressure pulse decay test

Here, in this study, a mathematical model for fast determination of the permeabilities of tight rocks using measurements taken from the initial period of the One Chamber Pressure Pulse Decay (OC-PPD) test is presented. The model applies to measurements taken both before and after the pressure pulse front has reached the downstream end of the specimen. The analytical solutions for the pressure decay in the upstream chamber are derived based on a parabolic arc approximation of pore pressure distribution along the test specimen. This approximation allows converting the initial–boundary value problem of fluid diffusion in the specimen, governed by partial differential equations, to a system of ordinary differential equations that can be easily solved by explicit formulae. Thus, an explicit formula for the pressure decay rate is obtained, which enables inverse analysis of the initial experimental data to estimate the rock permeability. The proposed method expedites the pulse decay test as it does not require the system to reach equilibrium. The method is validated with three sets of experimental data of the OC-PPD test using helium as the diffusing fluid, for which the relative error of the permeability is found to be less than 6%. This method is particularly useful if the equilibrium time of the pulse decay test for rock specimens with permeabilities in the range of nano-Darcy takes hours or days.

early-time solution↗

Advancing Grid Resilience through Smart Charge Management: Findings from Maryland’s Pilot

This report presents research findings from a four-year Smart Charge Management (SCM) pilot program conducted by Maryland’s largest electric utilities—Baltimore Gas and Electric (BGE), Potomac Electric Power Company (Pepco), and Delmarva Power & Light (DPL)—to evaluate strategies for optimizing electric vehicle (EV) charging loads and enhancing grid stability. Supported by the U.S. Department of Energy (DOE), Argonne National Laboratory collaborated with all project partners and examined the effectiveness of Time-of-Use (TOU) and Load Balancing (LB) strategies in managing peak demand, deferring costly infrastructure upgrades, and reducing grid constraints at the feeder level. Using charging data from over 4,600 EV drivers, the study analyzed SCM’s impact on the distribution systems of BGE and Pepco, which consists of over 2000 feeders. Unlike prior research that focused on system-wide trends or synthetic feeders, this analysis offers granular, feeder-level insights based on real-world operational data. It highlights how transformer density, load profiles, and infrastructure constraints influence smart charging performance. Results show feeder-level conditions play a crucial role in SCM effectiveness, with most feeders benefiting more from LB, while TOU-based SCM may be sufficient for others. By 2035, LB reduced peak charging loads by 27% on average, compared to 23% under TOU-based SCM, though some feeders saw reductions exceeding 35%, while others experienced minimal impact. Feeders with higher transformer utilization and limited capacity benefited more from LB, which more effectively distributed charging demand during off-peak hours. Beyond reducing grid constraints, SCM offers long-term operational and financial benefits. By shifting EV charging demand strategically, utilities can optimize asset utilization, delay infrastructure investments, and enhance grid performance. In terms of infrastructure upgrade deferrals, at the feeder level, LB consistently reduced peak charging loads and resulting infrastructure upgrade costs, particularly in high EV enrollment areas, decreasing the number of overloaded transformers by up to 35%, while TOU-based SCM achieved 20-30% reductions depending on feeder characteristics. At the system level, LB has the potential to defer total upgrade costs by $\$$186 million for BGE, compared to $\$$159 million under TOU-based SCM. For Pepco, TOU-based SCM performed slightly better, deferring upgrade costs by $\$$30 million, compared to $\$$29 million under LB. Section 4.5 reviews some of the system differences between BGE and Pepco. However, as EV adoption scales, TOU-based SCM will introduce secondary peak charging loads, reinforcing the need for more advanced, adaptive SCM approaches to prevent new grid challenges. As EV adoption continues to grow, feeder-level managed charging strategies will be essential for mitigating grid stress, improving infrastructure efficiency, and maintaining energy affordability for consumers. This report provides critical insights for utilities, Public Utility Commissions (PUCs), and state agencies on the role of feeder-specific smart charging in infrastructure planning, policy development, and grid modernization. The findings underscore the importance of tailored, data-driven SCM solutions that align with local grid conditions, ensuring a resilient, cost-effective transition to increasing EV adoption while safeguarding distribution system performance.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Approach for energy efficient building design during early phase of design process

Energy consumption in the building sector is about 40% of total energy consumed globally and is trending upwards, along with its contribution to greenhouse gas (GHG) emissions. Given the adverse impacts of GHG emissions, it is crucial to integrate energy efficiency into building designs. The most significant opportunities for enhancing energy performance are present during the initial phases of building design, when there is less impact of other design constraints. Various tools exist for simulating different design options and providing feedback in terms of energy consumption and comfort parameters. These simulation outputs must then be analyzed to derive design solutions. This paper presents an innovative approach that utilizes user input parameters, processes them through cloud computing, and outputs easily understandable strategies for energy-efficient building design. The methodology employs Asynchronous Distributed Task Queues (DTQ) - a more scalable and reliable alternative to conventional speedup techniques-for conducting parametric energy simulations in the cloud. The goal of this approach is to assist design teams in identifying, visualizing, and prioritizing energy-saving design strategies from a range of possible solutions for each project. Furthermore, a tool ‘eDOT’ has been developed utilizing the discussed methodology. Unlike existing tools, eDOT leverages artificial intelligence to dynamically generate and provide design strategies during the early phases of design process. By simplifying the simulation process, eDOT enables design teams to make informed, data-driven decisions without needing to interpret complex simulation outputs. A case study simulated for two locations is provided in this paper to demonstrate the effectiveness of eDOT, further underscoring its practical impact on energy-efficient building design.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Reactor System Facility Modification to Detect Compromised Human Machine Interfaces

This study focuses on a multi-layered Industrial Control System (ICS)/Operational Technology (OT) security architecture to aid in the discovery and mitigation of compromised Human Machine Interface (HMI)/Instrumentation & Control (I&C) based systems for modifying a prototypical reactor condition test facility called the Flowing Autoclave System (FAS) at Idaho National Laboratory (INL). This is achieved through a three-layered combination of network security solutions, hash-based algorithms, and blockchain technologies. Hash algorithms are mathematical functions used to generate a predetermined set of fixed-length values. They are widely used in computer security to verify the integrity of system information and data, both on a local network and the wider internet. Even small amounts of unauthorized system modification will cause the hash algorithm to output a set of characters that deviate significantly from its original value. Assisting secure hash functions, blockchain technology is a secure and distributed technology used to provide an immutable set of records replicated on all devices within a decentralized network. Blockchain offers a cost-effective solution to detect system compromise by providing a traceable breadcrumb trail of all network activity and data modification happening on a system. If both are used in conjunction with network monitoring tools, the integration of this three-pronged approach can become an asset in detecting suspected system compromises before any real damage can occur.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Federated Access from DOE Labs to Distributed Storage in the EIC Era of Computing

The Electron Ion Collider (EIC) collaboration and future experiment is a unique scientific ecosystem within Nuclear Physics as the experiment starts right off as a crosscollaboration between Brookhaven National Lab (BNL) & Jefferson Lab (JLab). As a result, this muti-lab computing model tries at best to provide services accessible from anywhere by anyone who is part of the collaboration. While the computing model for the EIC is not finalized, it is anticipated that the computational and storage resources will be made accessible to a wide range of collaborators across the world. The use of federated ID seems to be a critical element to the strategy of providing such services, allowing seamless access to each lab site computing resources. However, providing Federated access to a Federated storage is not a trivial matter and has its share of technical challenges. In this contribution, we focus on the steps we took towards the deployment of a distributed object storage system that integrates with Amazon S3 and Federated ID. We will first cover for and explain the first stage storage solutions provided to the EIC during the detector design phase. Our initial test deployment consisted of Lustre storage using MinIO, hence providing an S3 interface. High Availability load balancers were added later to provide the initial scalability it lacked. Performance of that system will be shown. While this embryonic solution worked well, it had many limitations. Looking ahead, the Ceph object storage is considered a top-of-the-line solution in the storage community - since the Ceph Object Gateway is compatible with the Amazon S3 API out of the box, our next phase will use a native S3 storage. Our Ceph deployment will consist of erasure coded storage nodes to maximize storage potential along with multiple Ceph Object Gateways for redundant access. We will compare performance of our next stage implementations. Finally, we will present how to leverage OpenID Connect with the Ceph Object Gateway’s to enable Federated ID access. We hope this contribution will serve the community needs as we move forward with cross-lab collaborations and the need for Federated ID access to distributed compute facilities.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Antecedent Hydrologic Conditions Reflected in Stream Lithium Isotope Ratios During Storms

Antecedent hydrological conditions are recorded through the evolution of dissolved lithium isotope signatures (δ 7 Li) by juxtaposing two storm events in an upland watershed subject to a Mediterranean climate. Discharge and δ 7 Li are negatively correlated in both events,but mean δ 7 Li ratios and associated ranges of variation are distinct between them. We apply a previously developed reactive transport model (RTM) for the site to these event-scale flow perturbations, but observed shifts in stream δ 7 Li are not reproduced. To reconcile the stability of the subsurface solute weathering profile with our observations of dynamic stream δ 7 Li signatures, we couple the RTM to a distribution of fluid transit times that evolve based on storm hydrographs. The approach guides appropriate flux-weighting of fluid from the RTM over a range of flow path lengths, or equivalently fluid residence times. This flux-weighted RTM approach accurately reproduces dynamic storm δ 7 Li-discharge patterns distinguished by the antecedent conditions of the watershed.

58 GEOSCIENCES↗

Ultrafast x-ray imaging of coherently controlled molecular dynamics in real space and time

Coherent control aims to manipulate chemical processes on the latent length and timescales of atoms and bonds, scales that are intrinsic to molecular dynamics but not directly resolved by most experimental probes. As a result, existing coherent-control experiments, which overwhelmingly rely on spectroscopic observables, leave a crucial blind spot: the direct, real-space recovery of all atomic and molecular rearrangements in response to coherently controlled excitations. Here, we overcome this limitation by integrating a Tannor-Kosloff-Rice pump-control-probe scheme with ultrafast X-ray scattering to capture snapshots of the wavepacket motion in a benchmark molecular system. We demonstrate this by photoexciting diatomic iodine vapor with a visible pump pulse, selectively steering the wavepacket toward ground-state recombination or dissociative pathways with a time-delayed near-IR control pulse, and recording the dynamics with angstrom and femtosecond precision with an ultrashort hard X-ray probe pulse. By comparing these structural observations with numerical solutions of the time-dependent Schrödinger equation, we reveal how coherent control actively reshapes the molecular charge density distribution. Our results pave the way for leveraging structural feedback as a control handle and provide a fundamental microscopic visualization of quantum decoherence and energy redistribution at the atomic level.

Hopper, Thomas R. [SLAC National Accelerator Labor↗

Guest Editorial Special Section on Advanced Medium-Voltage Power Electronics for Grid Interactive Applications

Medium-voltage power electronics (MVPE) plays essential roles in power grid modernization and links the MV distribution grid with low-voltage consumers and prosumers. Various MVPE devices, such as solid-state transformers or circuit breakers, inverter-based resources, power flow controllers, etc., bring the benefits of voltage conversion and power regulation in small footprint, power quality and efficiency improvements, and enhancements of grid controllability, flexibility, stability, and resilience. The MVPE also makes it possible for sustainable energy systems, such as solar/wind farms and energy storage generating facilities, to directly access to MV grids without multistage conversions. With their intrinsic intelligence and communications, MVPE enables many new smart grid functions and applications, e.g., dc interconnections and electric vehicle charging, which were not envisioned by traditional power grids otherwise. In addition, the integration of physical power processing units with cyber components forms a cyber-physical system, which is essential for long-term sustainability, development, and environmental preservation. Nonetheless, technical challenges on MVPE device reliability, scalable and efficient converter topologies, control stability, large-scale modeling and simulation, to name a few, need to be addressed and advanced to the next level. In conclusion, this Special Section on Advanced MV Power Electronics for Grid Interactive Applications in IEEE Transactions on Power Electronics (TPEL) provides an insight on some of the recent advances in MVPE and emerging challenges and potential solutions.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Software Verification of VARPOW

The VARPOW program is a post-processing utility program for DIF3D, specifically DIF3DVARIANT, and it was developed to provide interface files for the thermal analysis program DASSH. The basic methodology of VARPOW is to take the neutron and gamma flux (moments) calculated by DIF3D (or GAMSOR) and combine them with the heating (coefficient) cross sections to calculate the spatial power distributions using the DIF3DVARIANT spatial basis. VARPOW can use the output from GAMSOR (both steady state neutron and gamma flux calculations) or standard DIF3D/REBUS calculations (neutron flux only). The correct approach for defining the power distribution is to use GAMSOR as its purpose was to properly compute the gamma heating throughout the modeled domain. The power densities calculated by VARPOW are broken into fuel, cladding and coolant terms for which isotope-wise categorization is needed. VARPOW has built in options the user can select for the isotope categorization or VARPOW can import a file that details the isotope categorization. VARPOW can export the solution in the polynomial basis of DIF3D-VARIANT or the monomial basis of DIF3D-VARIANT. The purpose of this work is to verify the power distribution results calculated by VARPOW from both the GAMSOR and DIF3D input options and verify that the input and output options are consistent with the manual. Hand calculation and independent numerical calculation are used for this verification work.

97 MATHEMATICS AND COMPUTING↗

Software Verification of VARPOW

The VARPOW program is a post-processing utility program for DIF3D, specifically DIF3DVARIANT, and it was developed to provide interface files for the thermal analysis program DASSH. The basic methodology of VARPOW is to take the neutron and gamma flux (moments) calculated by DIF3D (or GAMSOR) and combine them with the heating (coefficient) cross sections to calculate the spatial power distributions using the DIF3DVARIANT spatial basis. VARPOW can use the output from GAMSOR (both steady state neutron and gamma flux calculations) or standard DIF3D/REBUS calculations (neutron flux only). The correct approach for defining the power distribution is to use GAMSOR as its purpose was to properly compute the gamma heating throughout the modeled domain. The power desnities calculated by VARPOW are broken into fuel, cladding and coolant terms for which isotope-wise categorization is needed. VARPOW has built in options the user can select for the isotope categorization or VARPOW can import a file that details the isotope categorization. VARPOW can export the solution in the polynomial basis of DIF3D-VARIANT or the monomial basis of DIF3D-VARIANT. The purpose of this work is to verify the power distribution results calculated by VARPOW from both the GAMSOR and DIF3D input options and verify that the input and output options are consistent with the manual. Hand calculation and independent numerical calculation are used for this verification work.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Situational awareness-enhancing community-level load mapping with opportunistic machine learning

Motivated by present and forthcoming challenges in the adoption and integration of distributed renewable energy, we develop a machine learning (ML) approach that builds short-fuse mappings connecting the occasionally-unobservable true load in one target community with information-rich signals collected from relatively more instrumented reference communities. Our setting is inspired by and tailored to target communities with significant unobservable behind-the-meter solar generation, where true load (a relatively well-behaved quantity of interest to grid operators) is hard to discern during daytime due to insufficient instrumentation and/or privacy reasons, but that can be related to reference communities with low unobservable distributed variable generation or with sufficient instrumentation. The developed mapping, herein realized with Support Vector Machine regression, is built using nighttime data from all communities, when their distributed generation is low or zero. Our ML algorithm opportunistically learns to correlate signals of interest and then is operationally used the next day to shed light into target community load evolution. The mapping is subsequently rebuilt, rolling its short-fuse scope perpetually forward in time. Here, we demonstrate the efficacy of our approach on nine synthetically generated topologies and associated timeseries stemming from real-world data, on which we observe cumulative error performance that yields lower than 10% and 15% daily-averaged mean absolute percentage errors in target community load estimation on more than about 75% and 90% of days, respectively, in multiple yearly evaluations that shed light on long-term performance also under seasonal and one-off effects. The proposed ML-powered methodology can offer grid operators much-improved visibility into a previously obscure space and can also serve as an additional source of information in broader, multi-modal solar disaggregation solutions.

14 SOLAR ENERGY↗

Implementing Directive-Based Deferred Execution for Effective Network Aggregation

Remote direct memory access technology provides an efficient mechanism for one-sided communication that can be leveraged to implement a distributed shared memory programming model. However, when applications generate large numbers of small, irregular messages, network congestion often arises. Existing solutions address this small message problem by facilitating message aggregation but typically require disruptive code transformations that detract from the algorithmic intent of applications, or can be limited by dependent operations on aggregated data between synchronisation points. A solution is to use a directive-assisted approach that enables compilers to transform code dependent on aggregated communication for deferred execution. This paper presents an algorithm that a compiler can use to implement and optimise deferred execution for code dependent on aggregated data, based on an "aggregation context" extension for the OpenSHMEM partitioned global address space library. This new capability addresses a key challenge of message aggregation, allowing its full potential to reduce network congestion and enhance programmability to be realised.

Welch, Aaron [ORNL]↗

VOLTTRON/volttron-pnnl-aems

The Autonomous Energy Management Software (AEMS) system will continuously optimize the operations of the distributed energy resources in the small and medium size commercial building by minimizing energy consumption and cost, while providing a solution for maximizing decarbonization benefits from electrification of buildings. Initially, AEMS system will manage rooftop air conditioners and heat pumps but it can be extended in the future to manage, hot water heaters, storage (battery and thermal), electric vehicle charging and monitoring solar photovoltaic. AEMS support both energy efficiency and grid service features.

Bleeker, Amelia [Pacific Northwest National Labora↗