Search NASA⌕ Search

SEARCH · Search NASA

Results for “management locality load balance”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

A GPU-based compressible combustion solver for applications exhibiting disparate space and time scales

High-speed chemically active flows pose significant computational challenges due to their disparate space and time scales, with stiff chemistry often dominating simulation time. While modern scientific computing programs achieve exascale performance by leveraging graphics processing units (GPUs), existing GPU-based compressible combustion solvers face critical limitations in memory management, load balancing, and handling the highly localized nature of chemical reactions. To this end, we present a high-performance compressible reacting flow solver built on the AMReX framework and optimized for multi-GPU settings. Here, our approach addresses three GPU performance bottlenecks: memory access patterns through column-major storage optimization, computational workload variability via a bulk-sparse integration strategy for chemical kinetics, and multi-GPU load distribution for adaptive mesh refinement applications. The solver adapts existing matrix-based chemical kinetics formulations to multi-grid contexts. Using representative combustion applications, including 2D and 3D detonations and a 3D jet-in-crossflow configuration, we demonstrate 1.4–5× performance improvements over initial implementations on an in-house cluster of NVIDIA H100 GPUs, and near-ideal weak scaling on the Frontier supercomputer (Oak Ridge Leadership Computing Facility) with up to 1024 AMD Instinct MI250X GPUs. Roofline analysis reveals substantial improvements in arithmetic intensity for both convection (∼ 10 ×) and chemistry (∼ 4 ×) routines, confirming efficient utilization of GPU memory bandwidth and computational resources.

42 ENGINEERING↗

Demonstration of a Novel Technology to Manage Electricity Demand in Grid-Independent Military Microgrids

This research was conducted by the National Renewable Energy Laboratory (NREL) in collaboration with the S&C Electric Inc. through funding provided by the ESTCP. The project demonstrates use of cybersecure Automated Demand Response (ADR) technology to effectively manage microgrid loads during grid-independent, also known as "islanded," operation. When military microgrids become isolated from the main electrical grid, they are required to balance electricity supply and demand locally. Given that local generation may be constrained, the prevailing strategy involves shedding all but the most critical loads by tripping smart circuit breakers, which then necessitate manual resetting. This approach is generally implemented at the building level, which means that the buildings with mission-critical activities are exempt from load management and remain fully powered, whereas those deemed non-critical can experience a complete loss of service. In this research we developed a method that allows building automation systems to selectively control their assets in response to load shedding request from a microgrid controller, avoiding total loss of service in contrast to the conventional control approach. A commercial OpenADR client server by GridFabric is used for communication between the microgrid controller and the building management system (BMS). The microgrid controller monitors both generation capacity and various assets within the microgrid and issues a demand reduction request when necessary. This request is communicated to the OpenADR server via Modbus. Upon receiving the request, the OpenADR server forwards it to the BMS utilizing the OpenADR protocol. The BMS is pre-configured with various levels of load reduction strategies based on the controllable assets available, allowing for a nuanced approach to demand reduction. Both lab and field tests were performed that considered load shedding needed to achieve closed transition into island mode and to accommodate changing loads and power source availability while islanded. A commercial microgrid controller was used for these tests with normal programming within the expected constraints of the system capabilities. That is, the solution did not require any specialized modification to the code base of the controller. Given the latency of the round-trip communication path between the microgrid controller and the various devices involved with the load shed processes, there are certain scenarios for which the demonstrated solution are appropriate and some which are not. The methods described in this report can be used for load shedding/restoration during transitions between islanded and grid-tied modes of operation, as well as accommodating normal variations in load and the need to remove a power source from operation for maintenance. These methods should not be used for scenarios that require load shedding within a second or two such as sudden and unanticipated significant load increases or loss of power sources through equipment faults.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Efficient Parallelization of Irregular Applications on GPU Architectures

With the enlarging computation capacity of general Graphics Processing Units (GPUs), leveraging GPUs to accelerate parallel applications has become a critical topic in academia and industry. However, a wide range of irregular applications with the computation-/memory-intensive nature cannot easily achieve high GPU utilization. The challenges mainly involve the following aspects: first, data dependence leads to coarse-grained kernel and inefficient parallelism; second, heavy GPU memory usage may cause frequent memory evictions and extra overhead of I/O; third, specific computation patterns produce memory redundancies; last, workload balance and data reusability conjunctly benefit the overall performance, but there may exist a dynamic trade-off between them. Targeting these challenges, this dissertation proposes multiple optimizations to accelerate two real-world applications: many-body correlation functions to simulate nuclear physics in a large-scale scientific system; the other is the eALS-based matrix factorization recommendation system. To accelerate the calculations of many-body correlation functions, this dissertation presents three frameworks in GPU memory management and multi-GPU scheduling. Firstly, an optimized systematic GPU memory management framework, MemHC, utilizes a series of new memory reduction designs in GPU memory allocation, CPU/GPU communications, and GPU memory oversubscription. Secondly, an enhanced multi-GPU scheduling framework, MICCO, particularly by taking both data dimension (e.g., data reuse and data eviction) and computation dimension into account. MICCO designs a heuristic scheduling algorithm and a machine learning-based regression model to generate the optimal settings of a proposed new concept to manage the trade-off. Thirdly, a locality-aware multi-GPU scheduling framework. This scheduler leverages pipeline batch generation with a looking-ahead strategy by building local dependency graphs for memory transfer reduction and better data reuse, achieving up to 79.92% memory cost reduction and 1.67x speedup. To parallelize the eALS-based recommendation system, this dissertation proposes an efficient CPU/GPU heterogeneous recommendation system, HEALS. HEALS employs newly designed architecture-adaptive data formats to achieve load balance and good data locality on CPU and GPU. To mitigate the data dependence, HEALS presents a CPU/GPU collaboration model for both task parallelism and data parallelism with multiple kernel computation optimizations. In summary, this dissertation efficiently accelerates two typical irregular applications on GPUs by building four frameworks, including CPU/GPU collaboration, GPU memory management, and multi-GPU scheduling.

Wang, Qihan↗

Integrating a Microgrid Controller with a Local OpenADR Server

When military microgrids isolate themselves from the main electrical grid, they must locally balance electricity supply and demand. Since local generation may be limited, the current strategy is to shed all but the most critical loads by tripping smart circuit breakers, which must then be reset manually (e.g., ESTCP project EW-201350). This strategy is typically applied at the building level, meaning that entire buildings housing mission critical activities must be excluded from any load management, while those considered non-critical may lose service entirely. The remotely controlled switchgear needed to manage load in this way is very expensive ($\$30,000$-$\$50,000$ per building). While effective at shedding load, this strategy disrupts installation operation and risks damaging equipment during both disconnection and re-energization. With the goals of lowering costs, protecting equipment, and enhancing the agility of DoD microgrids, this report demonstrates the use of cybersecure automated demand response (ADR) technology to manage microgrid loads during grid-independent (a.k.a. "islanded") operation. This automated approach achieves load shedding and shifting through communication signals sent to equipment controllers rather than by cutting off the flow of electricity within the microgrid itself. Because it operates only on the base network, with no connection to external entities, this strategy avoids the main cybersecurity concern raised by past applications of ADR on military bases.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Advancing Grid Resilience through Smart Charge Management: Findings from Maryland’s Pilot

This report presents research findings from a four-year Smart Charge Management (SCM) pilot program conducted by Maryland’s largest electric utilities—Baltimore Gas and Electric (BGE), Potomac Electric Power Company (Pepco), and Delmarva Power & Light (DPL)—to evaluate strategies for optimizing electric vehicle (EV) charging loads and enhancing grid stability. Supported by the U.S. Department of Energy (DOE), Argonne National Laboratory collaborated with all project partners and examined the effectiveness of Time-of-Use (TOU) and Load Balancing (LB) strategies in managing peak demand, deferring costly infrastructure upgrades, and reducing grid constraints at the feeder level. Using charging data from over 4,600 EV drivers, the study analyzed SCM’s impact on the distribution systems of BGE and Pepco, which consists of over 2000 feeders. Unlike prior research that focused on system-wide trends or synthetic feeders, this analysis offers granular, feeder-level insights based on real-world operational data. It highlights how transformer density, load profiles, and infrastructure constraints influence smart charging performance. Results show feeder-level conditions play a crucial role in SCM effectiveness, with most feeders benefiting more from LB, while TOU-based SCM may be sufficient for others. By 2035, LB reduced peak charging loads by 27% on average, compared to 23% under TOU-based SCM, though some feeders saw reductions exceeding 35%, while others experienced minimal impact. Feeders with higher transformer utilization and limited capacity benefited more from LB, which more effectively distributed charging demand during off-peak hours. Beyond reducing grid constraints, SCM offers long-term operational and financial benefits. By shifting EV charging demand strategically, utilities can optimize asset utilization, delay infrastructure investments, and enhance grid performance. In terms of infrastructure upgrade deferrals, at the feeder level, LB consistently reduced peak charging loads and resulting infrastructure upgrade costs, particularly in high EV enrollment areas, decreasing the number of overloaded transformers by up to 35%, while TOU-based SCM achieved 20-30% reductions depending on feeder characteristics. At the system level, LB has the potential to defer total upgrade costs by $\$$186 million for BGE, compared to $\$$159 million under TOU-based SCM. For Pepco, TOU-based SCM performed slightly better, deferring upgrade costs by $\$$30 million, compared to $\$$29 million under LB. Section 4.5 reviews some of the system differences between BGE and Pepco. However, as EV adoption scales, TOU-based SCM will introduce secondary peak charging loads, reinforcing the need for more advanced, adaptive SCM approaches to prevent new grid challenges. As EV adoption continues to grow, feeder-level managed charging strategies will be essential for mitigating grid stress, improving infrastructure efficiency, and maintaining energy affordability for consumers. This report provides critical insights for utilities, Public Utility Commissions (PUCs), and state agencies on the role of feeder-specific smart charging in infrastructure planning, policy development, and grid modernization. The findings underscore the importance of tailored, data-driven SCM solutions that align with local grid conditions, ensuring a resilient, cost-effective transition to increasing EV adoption while safeguarding distribution system performance.

24 POWER TRANSMISSION AND DISTRIBUTION↗

A cell-centered AMR-ALE framework for 3D multi-material hydrodynamics. Part I: Lagrangian and indirect Euler AMR algorithms

Many applications of physics and engineering involve wide ranges of time and spatial scales. The numerical simulation of localized small scales such as shock waves and material interfaces requires a large number of computational cells in these regions. For these applications, Lagrangian and Arbitrary-Lagrangian-Eulerian (ALE) related methods are engaging since the moving mesh feature naturally brings mesh cells on shock discontinuities and material interfaces are carefully captured. In addition, Adaptive-Mesh-Refinement (AMR) strategies aim to optimize computational resources by concentrating finer mesh cells only in areas of interest while using coarser cells elsewhere. A key but challenging AMR requirement consists in efficiently distributing the computational effort to achieve high accuracy without the prohibitive computational costs associated with uniformly fine grids. Here, in this document, the coupling of the p4est AMR library with a cell-centered Lagrangian scheme is presented with the goal to perform reliable 3D Lagrangian-AMR and indirect Euler-AMR multi-material simulations. In particular, it is shown that starting from a 3D indirect ALE code, the memory management and load balancing requirements can be delegated to an external library (here the p4est library) to unlock ALE-AMR capabilities. First, we present a strategy to transcribe the octant-based connectivity of the 3D AMR framework with that of an unstructured mesh of polygonal cells used in Lagrangian hydrodynamics. Then, we show how refinement and coarsening operations must be adapted to the particular Lagrangian framework to ensure the conservation of volume during those steps. Finally, several numerical test cases are presented that demonstrate the capabilities of the Lagrangian-AMR and indirect Euler-AMR algorithms.

3D cell-centered Lagrangian numerical scheme↗

Assessment of Cloud-Based Applications Enabling a Scalable Risk-Informed Predictive Maintenance Strategy Across the Nuclear Fleet

The current light water reactor fleet uses time-based or failure-based maintenance strategies to achieve high-capacity factors. But to make nuclear more competitive in the energy market, these reactors could utilize emerging technologies in terms of artificial intelligence (AI) and cloud computing to enable a cost-effective, predictive maintenance strategy. This report examines the feasibility of cloud computing for the nuclear industry’s needs in terms of the cloud’s computing capabilities, feasibility, and regulatory concerns. The technical viability of cloud computing was analyzed using one year worth of data from a boiling water reactor’s safety relief valve. Models were hosted on a local desktop, Idaho National Laboratory’s high-performance computer, and Microsoft Azure. Data was loaded, processed, and two types of models were trained in an A/B fashion. Based on the speed at which these actions were completed, it was used to determined that cloud computing has adequate computing resources. Additionally, the computing power can scale with the demanded load. To enable cloud computing in the existing fleet, additional sensors, networks, and other requirements must be implemented to ensure a smooth transition from current maintenance strategies. However, there is a benefit as the plant no longer needs manage their own servers, software, cybersecurity, and IT support staff. Many of these features can be offloaded on to the cloud provider. A comprehensive analysis was completed that showed the current annual cost of operating is more expensive than using cloud computing resources. Lastly, the regulatory framework does not explicitly address AI or autonomous control. Currently, the NRC and other regulatory bodies are evaluating providing guidance to address gaps rather than new regulations to address the use of AI and ML. But since many of the AI applications are focused on non-safety related applications, such as balance-of-plant components, they will likely have little or no regulatory restrictions or necessary approvals. Demonstrating how AI can improve maintenance and operation of these non-safety related systems seems like the likely path forward for implementing AI and cloud computing resources inside nuclear power plants (NPPs).

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Local Power Impact Experiment Design for a New Fuel Type for use in the Advanced Test Reactor

The Advanced Test Reactor (ATR), and complimentary zero-power ATR Critical (ATRC) reactor, located at Idaho National Labs (INL), are undergoing conversion from Highly Enriched Uranium (HEU) to Low Enriched Uranium (LEU). Both have a variety of testing locations that can receive large variations in flux due to its unique serpentine design, consisting of five lobes (see Figure 1). Initial criticality and power distribution throughout the core are controlled by core-external outer shim control cylinders (OSCCs). Distinct test loops allow for testing at specific temperatures, pressures, and irradiation conditions. The ATR is one of the key nuclear engineering research and testing facilities within the DOE National Laboratory Complex, and the ATRC supports its operation [1]. Currently, the Office of Material Management and Minimization (M3) within the National Nuclear Security Administration of the DOE is working to convert the remaining research reactors, including the ATR, from 93% HEU fuel to 19.75% LEU fuel (LEU) to support non-proliferation [2]. Extensive materials testing at INL and internationally has demonstrated that a high-density uranium molybdenum (U 10Mo) alloy can meet the performance requirements of the remaining high powered research reactors. However, there are many technical challenges to address before the conversion to LEU can be successful, including the accurate characterization of the reactor core physics with LEU fuel. Reactor physics safety evaluations currently use Monte Carlo for the 21st Century (MC21), a continuous-energy Monte Carlo radiation transport code [3]. Existing MC21 models of the ATR and ATRC cores have a validation basis for use in neutronics analyses with HEU fuel. The models are used to support safety analyses that include comparisons to the safety requirements for the reactors. However, the use of the LOWE element in the ATR and ATRC is not currently covered by the current model validation basis. To deploy the new fuel type, extensive computational reactor physics support is necessary to support the use of LOWE in the ATR and ATRC. Therefore, LOWE requires a rigorous validation basis, aligned with that of HEU fuel, that takes advantage of the existing software tools and processes currently used for the ATR and ATRC. The experiment to validate of the MC21 models for determining power, the Power Impact Validation Experiment, will consist of two flux runs in the ATRC, one with fully HEU loading and one with a single LOWE element. Both flux runs will be instrumented with 20 sets of azimuthal fission wires and 3 sets of axial fission wires, as shown in Figure 4. Standard flux run methodology will be used [4]. Power Impact Validation Experiment data will be compared against MC21 calculated data, both for absolute fission rate accuracy and to determine the relative change in fission rates between the two runs. The results of the Power Impact Validation Experiment and subsequent evaluations will provide the validation basis for MC21 for use with LOWE elements. Key features of the Power Impact Validation Experiment include: (1) Two flux runs to allow for LOWE perturbed measurements to be compared to already validated measurements taken from a full core of HEU fuel, (2) Optimization of instrumentation to balance analytical needs with practical considerations (e.g., limited time window to count beta particles from fission products), and (3) Standard ATRC core loading, including both driver positions and flux traps, to minimize cost while remaining representative of typical ATR core loading.

42 ENGINEERING↗

Effect of dust on rainfall over the Red Sea coast based on WRF-Chem model simulations

Water is the single most important element of life. Rainfall plays an important role in the spatial and temporal distribution of this precious natural resource, and it has a direct impact on agricultural production, daily life activities, and human health. One of the important elements that govern rainfall formation and distribution is atmospheric aerosol, which also affects the Earth's radiation balance and climate. Therefore, understanding how dust compositions and distributions affect the regional rainfall pattern is crucial, particularly in regions with high atmospheric dust loads such as the Middle East. Although aerosol and rainfall research has garnered increasing attention as both an independent and interdisciplinary topic in the last few decades, the details of various direct and indirect pathways by which dust affects rainfall are not yet fully understood. Here, we explored the effects of dust on rainfall formation and distribution as well as the physical mechanisms that govern these phenomena, using high-resolution WRF-Chem simulations (~1.5 km × 1.5 km) configured with an advanced double-moment cloud microphysics scheme coupled with a sectional eight-bin aerosol scheme. Our model-simulated results were realistic, as evaluated from multiple perspectives including vertical profiles of aerosol concentrations, aerosol size distributions, vertical profiles of air temperature, diurnal wind cycles, and spatio-temporal rainfall patterns. Rainfall over the Red Sea coast is mainly caused by warm rain processes, which are typically confined within a height of ~6 km over the Sarawat mountains and exhibit a strong diurnal cycle that peaks in the evening at approximately 18:00 local time under the influence of sea breezes. Numerical experiments indicated that dust could both suppress or enhance rainfall. The effect of dust on rainfall was calculated as total, indirect, and direct effects, based on 10-year August-average daily-accumulated rainfall over the study domain covering the eastern Red Sea coast. For extreme rainfall events (domain-average daily-accumulated rainfall of ≥ 1.33 mm), the net effect of dust on rainfall was positive or enhancement (6.05 %), with the indirect effect (4.54 %) and direct effect (1.51 %) both causing rainfall increase. At a 5 % significance level, the total and indirect effects were statistically significant whereas the direct effect was not. For normal rainfall events (domain-average daily-accumulated rainfall < 1.33 mm), the indirect effect enhanced rainfall (4.76 %) whereas the direct effect suppressed rainfall (-5.78 %), resulting in a negative net suppressing effect (-1.02 %), all of which were statistically significant. We investigated the possible physical mechanisms of the effects and found that the rainfall suppression by dust direct effects was mainly caused by the scattering of solar radiation by dust. The surface cooling induced by dust weakens the sea breeze circulation, which decreases the associated landward moisture transport, ultimately suppressing rainfall. For extreme rainfall events, dust causes net rainfall enhancement through indirect effects as the high dust concentration facilitates raindrops to grow when the water vapor is sufficiently available. Our results have broader scientific and environmental implications. Specifically, although dust is considered a problem from an air quality perspective, our results highlight the important role of dust on sea breeze circulation and associated rainfall over the Red Sea coastal regions. Our results also have implications for cloud seeding and water resource management.

54 ENVIRONMENTAL SCIENCES↗