Search NASA⌕ Search

SEARCH · Search NASA

Results for “Implementation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 757 records · Page 42

TRIM: AI Guided Random Number Generation for Resource-Constrained IoT Systems

Random numbers often serve as the backbone for many security solutions in diverse domains such as cryptography, side channel leakage prevention, and moving target defense. However, generating true random numbers requires a physical source of entropy (e.g. hardware, quantum, environmental phenomenon) making it difficult to realize at a large scale and at a low cost. On the flip side, pseudorandom number generators (easy to implement) following a specific distribution (e.g. Gaussian) can be easily compromised given a sufficient amount of traces. In this work, we have developed a machine learning-guided generative approach that can be used to create portable, resource-efficient, and cost-effective random number generators with high throughput and true randomness characteristics. We implement the proposed approach as a highly parameterized framework and perform extensive evaluation for different settings. The framework was able to learn from true random sources such as irrational numbers and environmental audio noise and imitate those sources towards generating new good quality random numbers on demand. We have generated more than 1 billion bits and observed robust performance in terms of true randomness metrics obtained from NIST SP 800-22 and FIPS 140-1 randomness test suites achieving a throughput of up to 142.85 Mbps. Compared to the state-of-the-art (SOTA) technique, the iso-cost setup of our framework can achieve more than 500 Mbps in a distributed setting. We have evaluated the efficacy of running the true randomness imitation AI models on target edge devices such as Raspberry Pi 4 (Model B), Nvidia Jetson Nano, Nvidia Jetson Orin Nano and Nvidia Jetson Xavier. We have also looked at the security of the TRIM framework itself against different adversarial threat models.

Cybersecurity↗

Enhancing Time Synchronization in Smart Grid With White Rabbit: Theory, Architecture, and Challenges

The smart grid aims to provide economically efficient and sustainable power to consumers with high quality and security. However, the increasing integration of distributed renewable energy sources presents challenges for smart grid protection and control systems. Here, to enhance smart grid operations, this article introduces White Rabbit (WR) as a more precise and accurate time synchronization technique. To explore White Rabbit’s potential applications in the smart grid, this study first explains existing time synchronization techniques and their limitations, followed by an analysis of the role and importance of time synchronization in smart grids. A comprehensive survey is then conducted, covering the theories, principles, implementations, performances, and existing application cases of White Rabbit. The findings suggest that White Rabbit is a promising technique for smart grid deployment. However, several challenges remain on the path to large-scale implementation. These challenges are analyzed in detail, highlighting key areas for future research.

Liu, Yu [University of Tennessee, Knoxville, TN (U↗

Direct Power Control of Back to Back Modular Multilevel Converter with Advanced Grid Support Functions for Grid Forming Application

Possibility of using back-to-back modular multilevel converter system with a direct power control architecture has been investigated in this paper. The rectifier side is controlled to be in grid following mode whereas the inverter side is controlled in grid forming mode. Both the sides’ local controllers use Lyapunov energy function based architecture to accomplish the objective of active and reactive power control as well as maintaining a fixed user defined voltage over the dc bus on the rectifier side. Therefore, the proposed direct power control architecture has embedded dc bus control architecture to maintain fixed user defined value of the capacitors. Similar architecture is implemented on the inverter side to accomplish either voltage control during grid forming or power control during grid following. Advanced grid support functionalities based on IEEE-1547-2018 is implemented and utilized to generate the reference values for either the rectifier or the inverter sides. The overall system is modeled based on MATLAB/Simulink and PLECS domain and various important case studies to verify the efficacy of the overall system has been presented.

direct power control (DPC)↗

Thermal Analysis of a 100 kW Polyphase Wireless Power Transfer System

Charging Electric Vehicles (EVs) fast and safely has a crucial role in the future of the EV technology. High-power Wireless Power Transfer (WPT) helps to significantly decrease the charging time. However, when the power transfer levels increase, thermal management becomes a significant challenge. The thermal design of the WPT systems needs more consideration in the design and implementation steps. This paper presents a thermal analysis of a 100 kW high-power WPT system. The thermal performance of the proposed design was evaluated at different power levels by considering the magnetic design and loss analysis. Finite Element Analysis (FEA) of the proposed design was performed and the thermal images of the implemented system were taken to prove the simulation results. The results show that, a liquid cooling design is needed for a high-power WPT systems for the long-time continuous operations of the charging pads.

Aydin, Emrullah↗

Enhancing Photosynthesis Simulation Performance in ESMs with Machine Learning-Assisted Solvers

When simulating vegetation dynamics, photosynthesis accounts for a large fraction of the computational cost in most Earth System Models (ESMs). This is largely since photosynthesis is represented as a system of nonlinear equations, and the solution requires the use of an initial guess followed by many iterations of the numerical solver to obtain a solution. We use machine learning (ML) to replicate the response surface of the model’s numerical solver to improve the choice of initial guess, therefore requiring fewer iterations to obtain a final solution. We implemented this test on the leaf-level calculations as well as at the canopy scale, and for both we observed fewer iterations of the photosynthesis solver when a ML-based initial guess was implemented. The model tested here is the Energy Exascale Earth System Model - Land Model (ELM). The ML-based algorithms used here are trained on simulations from the model itself and used only to improve the initial guess for the solver; therefore, the model maintains its own set of physics to obtain the final solution. This work shows novel ways to utilize ML-based methods to improve the performance of numerical solvers in ESMs.

Massoud, Elias [ORNL] (ORCID:0000000217725361)↗

Exploring the Landscape of Distributed Graph Clustering on Leadership Supercomputers

The rapid growth of large-scale datasets in fields like biology and social networks has driven the need for advanced graph analytics techniques. Community detection, a fundamental task in graph analytics, identifies closely connected groups of nodes within a network, providing valuable insights across various disciplines. This study focuses on two classic community detection methods, the Louvain algorithm and Markov Clustering (MCL), and evaluates the performance of two prominent distributed community detection algorithms: HiPDPL-GPU, our prior implementation, and HipMCL. We conduct experiments on GPU-accelerated heterogeneous HPC systems, Summit and Frontier, to assess their performance under varying conditions. Our objective is to identify the strengths and weaknesses of these algorithms in terms of scalability, and quality of solutions. We evaluate these algorithms on a diverse set of 70+ networks spanning 13 domains, with sizes ranging up to 4.2 billion edges. Our results demonstrate that HiPDPL-GPU consistently outperforms HipMCL, especially for large-scale networks. HiPDPL-GPU achieves significantly faster runtimes (47x to 1439x), higher modularity scores, and improved scalability. These findings highlight HiPDPL-GPU as a promising solution for efficient and effective large-scale graph analytics in diverse application domains, and provide insights into the feasibility of using MCL-based approaches for certain application domains.

Community detection, graph algorithms↗

AXI4MLIR: User-Driven Automatic Host Code Generation for Custom AXI-Based Accelerators

Tensor algebra operations represent an important class of algorithms used across many applications, including machine learning, scientific computing, and data analytics. As a result, the efficient generation of custom accelerators for tensor operations has received increased attention. Previous efforts have produced automated tools enabling users to prototype and explore optimized accelerators. However, little effort has been focused on the host-accelerator interaction in these tools. Efficient use of hardware accelerators requires knowledge about the accelerator's capabilities (operations, data formats, and opcode support), the host CPU microarchitecture (e.g., memory hierarchy), the host-accelerator interface, and the application's features (which code regions should be mapped onto an accelerator). Manually rewriting the original applications to facilitate improved custom accelerator mapping is an error-prone and time-consuming endeavor. To cope with this, we propose AXI4MLIR, a new framework to automatically generate and optimize the communication between the host CPU and arbitrary accelerators that implement linear algebra algorithms. AXI4MLIR extends the MLIR compiler framework to automatically generate efficient host-accelerator driver code for accelerators with AXI-based interfaces. Our compiler extensions enable automatic driver code generation while carefully considering the host's memory hierarchy and target accelerator features. To demonstrate the flexibility and utility of AXI4MLIR, we test it with diverse use cases that include different types of accelerators, tiling scenarios, and dataflow schemes. We compare our experimental results to manual implementations of host-accelerator driver code and find that our approach can reduce CPU cache references by 56% and deliver up to a 1.65x speedup.

Bohm Agostini, Nicolas↗

Design and Analysis of an Integral MPPT Control Law for Wave Energy Conversion Systems

In this paper, we propose an integral-based maximum power point tracking (MPPT) algorithm for point absorber wave energy conversion (WEC) systems. A permanent magnet synchronous generator (PMSG) is coupled to the point absorber and its drive implements the proposed MPPT control. While the rotating frame of reference of the PMSG implements power control, the slower mechanical frequency of the incident waves are processed with an additional transformation that yields another set of dc dynamics. The computed dc power in the new reference frame, which is focused on slow wave dynamics, is input to the MPPT control law that tunes the emulated resistance of the machine drive to extract peak power from the wave absorber device. We derive a stability condition for the proposed controller and validate our control design on a simulated 10kW system.

wave energy, maximum power point tracking (MPPT), ↗

Comparative Evaluation of the Single-Phase Shift and Extended-Phase Shift Control for Isolated DC-DC Converter

Abstract: This paper presents a comparative evaluation of the Single-Phase Shift (SPS) and Extended-Phase Shift (EPS) controls for the isolated DC-DC Converter. The analysis is based on the power regulation range, flexibility of operation, current stress on power devices, system efficiency while focusing on the issue of backflow power, control complexity, design and implementation overhead, and stability of the entire system. It is also examined how the backflow influences power circulation and increases current stress. Here, the Dual Active Bridge (DAB) Converter is employed to assess the performance of the control methods. Compared to SPS control, EPS offers a wider power regulation range, greater flexibility, lower current stress, and better system efficiency. However, EPS has increased control complexity, design and implementation overhead, and often requires sophisticated feedback control loops. The comparison is made through mathematical modeling of power transfer, backflow, and current stress. Simulation results validate the comparative analysis presented.

Amir, Aamir [The University of Alabama (UA)]↗

Deep Learning-Based Dynamic Modeling of Three-Phase Voltage Source Inverters

Inverter-based resource (IBR) models are necessary to analyze modern power system stability and create effective control strategies. Modeling IBRs in converter-rich power systems is crucial, yet challenging due to the lack of commercial information on converter topologies and control parameters. This paper proposes novel convolutional neural network (CNN)–based data-driven techniques for modeling IBRs, addressing adaptability and proprietary concerns without requiring internal system physics knowledge. The proposed method is tested using real grid-tied commercial IBR transient data and demonstrates effectiveness and accuracy. Furthermore, the developed modeling approach is integrated and implemented in the open-source power distribution simulation and analysis tool, GridLAB-D, to illustrate the potentiality of dynamic analysis of large-scale power systems with high IBRs.

deep learning, artificial intelligence↗

Design of High-Power Polyphase PCB Coil Systems for Wireless Power Transfer

Printed circuit board (PCB) coils have been proposed prior for implementation as inductive wireless charging coils to minimize size and cost. Utilization of PCBs can allow for a reduced cost, improved manufacturability, and a wide range of geometric customization options. To circumvent material limitations on insulation and thermal performance, parallel paths can be implemented to divide the current per path accordingly. Within this paper, an unconventional high-power PCB coil is designed employing all possible techniques for wiring with axial and radial parallel paths with equivalent transposition to minimize circulating currents. Design studies are simulated in 3D finite element analysis (FEA) to evaluate imbalance between phases with and without transposition. Two experimental prototype coils were fabricated with measurements for self-inductance and mutual inductance between phases. These measurements were validated to be sufficiently consistent with FEA results. Additionally, a method is proposed for a two-step optimization of coupling coefficient and coil losses.

Lewis, Donovin D.↗

Modified Andronov-Hopf Oscillator-Based Grid-Forming Converter with Emulated Virtual Cable for Enhanced Power Sharing Performance

Nonlinear oscillator-based grid-forming converters offer superior dynamic and steady-state performance, making them an attractive solution for interconnecting renewable resources. This paper proposes a novel modified Andronov-Hopf oscillator to enhance the operating spectrum and facilitate the integration of renewable energy sources. An inner loop controller based on the Lyapunov energy function is implemented to achieve robust stability and performance, while a virtual cable emulation strategy enables seamless parallel operation. Comprehensive modeling and simulation studies validate the effectiveness of the proposed system, demonstrating its capabilities in addressing diverse operating scenarios, including grid faults, renewable energy fluctuations, and parallel operation. The proposed solution exhibits fast transient response, robust stability, and flexible operation, making it a valuable contribution to the field of renewable energy integration. The results of this study can be used to inform the design and implementation of next-generation grid-forming converters, enabling a more sustainable and reliable energy future. Additionally, the proposed system's ability to operate in both grid-connected and islanded modes makes it an ideal candidate for remote and off-grid renewable energy applications. The proposed solution's scalability and modularity also make it suitable for large-scale renewable energy integration. The proposed system is verified through MATLAB/Simulink and PLECS simulations, demonstrating its effectiveness in ensuring robust and efficient operation.

Andronov-Hopf Oscillator (AHO)↗

Enabling Scientific Applications with Performance-Portability and High-Productivity for Multi-GPU Programming with JACC.Multi

This work bridges the gap between multi-GPU computing and high-productivity, performance-portable programming solutions. Our goal is to enhance scientific applications with a productive and portable solution—program once, deploy everywhere—for multi-GPU programming with no cost to programmability. To accomplish this, we implemented JACC.Multi, which is part of the Julia for ACCelerators (JACC) performance-portable framework. JACC. Multi is the only high-level, portable metaprogramming solution that targets multi-GPU environments and is integrated in a readily accessible programming language (e.g., Julia language). With transparent GPU-to-GPU communication, JACC. Multi is optimized for scientific application workloads and is portable for NVIDIA and AMD accelerators. For the evaluation, we use two modern multi-GPU systems: Hudson, which features two NVIDIA H100 Hopper GPUs per node, and Frontier, which features four AMD MI250X GPUs per node, each with two Graphics Compute Dies (GCDs) for a total of eight GCDs per node. Additionally, as part of the evaluation, we use JACC (one GPU), MPI+JACC, and JACC. Multi codes that implement well-known and widely used scientific algorithms/kernels such as the conjugate gradient algorithm and an explicit forward Euler solver that requires GPU-to-GPU communication. Overall, JACC. Multi codes achieve better performance than MPI+JACC codes and significant speedups over JACC (one GPU), with up to 1.9× on Hudson and 6× on Frontier.

Valero Lara, Pedro [ORNL] (ORCID:0000000214794310)↗

Development of a Distribution Optimal Power Flow Federate for Open-Source OEDI-SI Platform

Increasing numbers of distributed generators in the electric power distribution networks require developing a control strategy to optimize solutions in real time. Linearized optimal distribution flow development has seen growth and acceptance in the distribution systems literature for efficiently modeling the \glspl{opf} for distribution systems. This paper examines the implementation and integration procedure for linearized optimal distribution flow federate to \gls{oedisi} platform. Specifically, we discuss i) the usage of the \gls{oedisi} platform, ii) obtaining a tractable solution using developed \gls{opf} federate, and iii) validation of solutions and bench-marking the \gls{oedisi} platform with developed \gls{opf} federate using OpenDSS. In brief, we demonstrate how a general linearized optimal distribution flow federate can be developed and integrated with a co-simulation environment to mimic real-world examples. The efficacy of the proposed method is demonstrated using the IEEE 123-bus test system under different scenarios to obtain a tractable solution and compare its results.

Sadnan, Rabayet↗

High-Penetration Microgrids Providing Grid Stability Using Frequency Watt Control

The U.S. grid is rapidly transitioning towards utilizing inverter-based renewable energy resources such as solar, wind, and batteries, reducing the carbon emission footprint. Inverter-based microgrid control architectures remain a critical focus to address power system stability issues in future high penetration markets lacking spinning generation assets. Idaho National Laboratory (INL) is researching an active layered inverter based frequency-Watt control scheme that provides distribution level stability in high-penetration markets where grid inertia is lacking. Hardware in the loop case study was implemented using INL’s Microgrid Testbed to combat scalable frequency deviations ranging from 60 Hz down to 50 Hz initialized by a hydropower model implementing step loads using a 540-kW grid emulator. Our research findings demonstrate the importance of distribution level, inverter-based active frequency-Watt controls utilizing a battery energy storage system (BESS) to provide adequate frequency support at the point of common coupling without major power infrastructure upgrades.

14 SOLAR ENERGY↗

Dual Channel Dual Staging: Hierarchical and Portable Staging for GPU-Based In-Situ Workflow

In-situ workflows have emerged as an attractive approach for addressing data movement challenges at very large scales. Since GPU-based architectures dominate the HPC landscapes, porting these in-situ workflows, and, specifically, the inter-application data exchange, to GPU-based systems can be challenging. Technologies such as GPUDirect RDMA (GDR), which is typically used for I/O in GPU applications as an optimization that circumvents the CPU overhead, can be leveraged to support bulk data exchanges between GPU applications. However, current GDR design often lacks performance portability across HPC clusters built with different hardware configurations. Furthermore, the local CPU may also be effectively used as an auxiliary communication mechanism to offload data exchanges. In this paper, we present a dual channel dual staging approach for efficient, scalable, and performance-portable inter-application data exchange for in-situ workflows. This approach exploits the data access pattern within in-situ workflows along with the inherent execution asynchrony to accelerate data exchanges and, at the same time, improve performance portability. Specifically, the dual channel dual staging method leverages both the local CPU and the remote data staging server to build a hierarchical joint staging area and uses this staging area to transform blocking inter-application bulk data exchanges into best-effort local data movements between GPU and CPU. The dual channel dual staging is implemented as a portability extension of the Dataspaces-GPU staging framework. We present an experimental evaluation of its performance, portability, and scalability using this implementation on three leadership GPU clusters. The evaluation results demonstrate that the dual channel dual staging method saves up to 75% in data-exchange time compared to host-based, GDR, and alternate portable designs, while maintaining scalability (up to 512 GPUs) and performance portability across the three platforms.

Zhang, Bo [University of Utah]↗

SYCL for Performance Portability: Application Experience with Coupled Cluster Formalism in Quantum Chemistry on Exascale Systems

The exascale computing has brought unprecedented heterogeneity in node architectures, with systems such as Frontier and Aurora featuring diverse GPU accelerators, network connectivity among others. Ensuring performance portability across these platforms is a key challenge. To address this, we employ the SYCL programming model to develop portable, high-performance quantum chemistry workloads. As a representative application, we focus on the non-iterative Triples component of the coupled-cluster CCSD(T) method, a key driver in quantum chemistry. In this work, we report on our experience deploying SYCL-based implementations using both DPC++ and AdaptiveCPP across two flagship exascale platforms: OLCF Frontier with AMD MI250X GPUs and ALCF Aurora with Intel GPUs. Our results demonstrate that SYCL enables efficient, single-source implementations that scale to thousands of nodes, delivering performance on par with vendor-optimized HIP solutions. We highlight key insights into runtime behavior, kernel portability, and scaling characteristics, showing that SYCL offers a viable path for performance-portable computing.

Bagusetty, Abhishek [Argonne National Laboratory (↗

Demystifying Cyberattacks: Potential for Securing Energy Systems With Explainable AI : Preprint

Modernization of energy systems has led to in- creased interactions among multiple critical infrastructures and diverse stakeholders making the challenge of operational decision making more complex and at times beyond cognitive capabilities of human operators. The state-of-the-art machine learning and deep learning approaches show promise of supporting users with complex decision-making challenges, such as those occurring in our rapidly transforming cyber-physical energy systems. However, successful adoption of data-driven decision support technology for critical infrastructure will be dependent on the ability of these technologies to be trustworthy and contextually interpretable. In this paper, we investigate the feasibility of implementing XAI for interpretable detection of cyberattacks in the energy system. Leveraging a proof-of-concept simulation use case of detection of a data falsification attack on a photovoltaic system using XGBoost algorithm, we demonstrate how Local Interpretable Model-Agnostic Explanations (LIME), a flavor XAI approach, can help provide contextual and actionable interpretation of cyberattack detection.

artificial intelligence↗