Search NASA⌕ Search

SEARCH · Search NASA

Results for “Base”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 433 records · Page 24

Dual Channel Dual Staging: Hierarchical and Portable Staging for GPU-Based In-Situ Workflow

In-situ workflows have emerged as an attractive approach for addressing data movement challenges at very large scales. Since GPU-based architectures dominate the HPC landscapes, porting these in-situ workflows, and, specifically, the inter-application data exchange, to GPU-based systems can be challenging. Technologies such as GPUDirect RDMA (GDR), which is typically used for I/O in GPU applications as an optimization that circumvents the CPU overhead, can be leveraged to support bulk data exchanges between GPU applications. However, current GDR design often lacks performance portability across HPC clusters built with different hardware configurations. Furthermore, the local CPU may also be effectively used as an auxiliary communication mechanism to offload data exchanges. In this paper, we present a dual channel dual staging approach for efficient, scalable, and performance-portable inter-application data exchange for in-situ workflows. This approach exploits the data access pattern within in-situ workflows along with the inherent execution asynchrony to accelerate data exchanges and, at the same time, improve performance portability. Specifically, the dual channel dual staging method leverages both the local CPU and the remote data staging server to build a hierarchical joint staging area and uses this staging area to transform blocking inter-application bulk data exchanges into best-effort local data movements between GPU and CPU. The dual channel dual staging is implemented as a portability extension of the Dataspaces-GPU staging framework. We present an experimental evaluation of its performance, portability, and scalability using this implementation on three leadership GPU clusters. The evaluation results demonstrate that the dual channel dual staging method saves up to 75% in data-exchange time compared to host-based, GDR, and alternate portable designs, while maintaining scalability (up to 512 GPUs) and performance portability across the three platforms.

Zhang, Bo [University of Utah]↗

Mixed Delay/Nondelay Embeddings Based Neuromorphic Computing with Patterned Nanomagnet Arrays

Patterned nanomagnet arrays (PNAs) have been shown to exhibit a strong geometrically frustrated dipole interaction. Some PNAs have also shown emergent domain wall dynamics. Previous works have demonstrated methods to physically probe these magnetization dynamics of PNAs to realize neuromorphic reservoir systems that exhibit chaotic dynamical behavior and high-dimensional nonlinearity. These PNA reservoir systems from prior works leverage echo state properties and linear/nonlinear short-term memory of component reservoir nodes to map and preserve the dynamical information of the input time-series data into nondelay spatial embeddings. Such mappings enable these PNA reservoir systems to imitate and predict/forecast the input time series data. However, these prior PNA reservoir systems are based solely on the nondelay spatial embeddings obtained at component reservoir nodes. As a result, they require a massive number of component reservoir nodes, or a very large spatial embedding (i.e., high-dimensional spatial embedding) per reservoir node, or both, to achieve acceptable imitation and prediction accuracy. These requirements reduce the practical feasibility of such PNA reservoir systems. To address this shortcoming, we present a mixed delay/nondelay embeddings-based PNA reservoir system. Our system uses a single PNA reservoir node with the ability to obtain a mixture of delay/nondelay embeddings of the dynamical information of the time-series data applied at the input of a single PNA reservoir node. Our analysis shows that when these mixed delay/nondelay embeddings are used to train a perceptron at the output layer, our reservoir system outperforms existing PNA-based reservoir systems for the imitation of NARMA 2, NARMA 5, NARMA 7, and NARMA 10 time series data, and for the short-term and long-term prediction of the Mackey Glass time series data.

Ti, Changpeng↗

Autoencoder-Based Sensor Drift Detection and Mitigation for Resilient Charging Systems

This work presents an autoencoder-based approach for sensor signal reconstruction and drift detection for charging systems. The proposed strategy is implemented within a Simulink-based system framework and evaluated under multiple operating conditions. An autoencoder with 8 neurons in the bottleneck layer is adopted, achieving accurate reconstruction across 10 variables and strong agreement with the physical sensor readings under normal conditions. In the case of a sensor fault, the autoencoder reconstruction remains closer to the expected true value compared to the corrupted measurement. Furthermore, feeding the autoencoder-reconstructed signal value back into the control framework in place of the faulty sensor signal leads to improved power monitoring. These results highlight the potential of autoencoder-based virtual sensing to extend the concept of resiliency to all components of the charging system, including sensors.

Rezende Da Costa Reis Kimpara, Renata [ORNL] (ORCI↗

Importance Sampling Model-Based Diffusion for Trajectory Optimization

Trajectory optimization for robotic systems remains a challenging problem. This is especially true for robotic systems featuring nonlinear dynamics and many degrees of freedom. Data-based or model-free diffusion has recently been popularized in the fields of artificial intelligence and trajectory optimization. Model-Based Diffusion provides a data-free method of trajectory optimization, trained at runtime on a system dynamics model, suitable for high-dimensional models. This paper examines how importance sampling can enhance the performance of Model-Based Diffusion for trajectory optimization. Here, we quantify the benefits of importance sampling across three long horizon planning tasks. These results show as much as a 13x improvement in sample efficiency depending on environment and optimization parameters.

Golembeski, Seth [Georgia Institute of Technology,↗

NeuroCoreX: An Open-Source FPGA-Based Spiking Neural Network Emulator with On-Chip Learning

Spiking Neural Networks (SNNs) are computational models inspired by the event-driven communication and connectivity patterns of biological neural circuits. They enable high energy efficiency and natural support for diverse architectures ranging from layered networks to small-world and graphstructured topologies. In this work, we introduce NeuroCoreX, an open-source, FPGA-based spiking neural network emulator that provides real-time, on-chip learning and flexible network organization. NeuroCoreX supports both feedforward sensory inputs streamed directly from sensors or PCs via UART and recurrent on-chip connectivity, enabling simultaneous processing and learning from external stimuli and internal network dynamics-capabilities rarely available in existing FPGA SNN platforms. The system implements a Leaky Integrate-and-Fire (LIF) neuron model with current-based synapses and supports pair-based STDP learning on both feedforward and recurrent synapses. A lightweight Python interface enables interactive configuration, live monitoring, weight read-back, and experiment control. Importantly, NeuroCoreX is tightly integrated with the SuperNeuroMAT simulator, allowing SNN models to be transferred seamlessly from software to hardware for hardware-in-the-loop development. By combining real-time plasticity, flexible connectivity, and an open-source VHDL implementation, NeuroCoreX provides an extensible and accessible platform for neuromorphic research, algorithm-hardware co-design, and energy-efficient edge intelligence.

Gautam, Ashish [ORNL]↗

Hardware-in-the-Loop Evaluation for Potential High Limit Estimation-Based PV Plant Active Control

This paper validates the efficacy of an artificial intelligence (AI)-based photovoltaic (PV) plant control and optimization approach in enabling PV plants as accountable grid reliability service providers. The validation is performed in a realistic laboratory controller-hardware-in-the-loop environment, leveraging accurate PV plant modeling and standard industrial communication protocols. Through simulations that account for diverse weather conditions and active control scenarios, the results highlight the superior performance of the AI-based solution in comparison to a state-of-the-art reference-control grouping-based approach. Such a finding contributes to mitigating the risk of overcurtailment and uninstructed deviations of active PV plant controls, and offers practical guidance for its field deployment. Furthermore, it establishes a standardized testing framework for comparing various PV active control strategies.

hardware-in-the-loop↗

Design of Hopfield Networks Based on Superconducting Coupled Oscillators

The global energy shortage has driven the development of many energy-efficient computational platforms beyond Moore's law, among which brain-inspired neuromorphic computing is one of the promising solutions. Associative memory and pattern recognition are important computations solved by brain-inspired Hopfield networks. Classical Hopfield networks store memories via fixed point attractors of their dynamics. In oscillatory Hopfield networks, these attractors are replaced by periodic orbits. Here, we design an oscillatory Hopfield network based on coupled superconducting oscillators. We first employ a mathematical phase reduction approach to map networks of coupled superconducting rapid single flux quantum (RSFQ) ring oscillators to coupled Kuramoto phase-oscillator networks. We use this theory to numerically optimize the hardware's mutual inductances in order to directly match the phase-reduced superconducting oscillators to a model of phase-oscillator-based Hopfield networks. The resulting network can store multiple oscillatory phase-locked memory patterns and recover the patterns based on the initial phase conditions. As different pattern recognition tasks, or learning, require tunable connectivity strengths between the oscillatory nodes, we further employ a coupler circuit that enables tuning the coupling strength between two oscillators by applying an external flux. We demonstrate the functionality of our design through numerical simulations of a small example network with oscillators operating at 86 GHz and recognizing patterns within 10 ns. Our approach enables the learning and retrieval of dynamical memory patterns with a wide range of applications where rhythmic dynamic output is beneficial.

Cheng, Ran↗

A Twin Circuit Theory-Based Framework for Oscillation Event Analysis in Inverter-Dominated Power Systems With Case Study for Kaua‘i System

Here, this paper proposes a real-world oscillation event analysis framework for power systems that include inverter-based resources together with synchronous generators. Specifically, the proposed framework combines both measurement-and model-based techniques to readily identify potential oscillation sources, replay the oscillation event with numerical simulation, unveil the underlying oscillation mechanism, and suggest mitigation methods for a wide range of oscillation events. To strengthen the theoretical foundation of our analysis framework, this paper proposes a twin circuit theory that provides theoretical support for one key utilized but not well-proven measurement-based oscillation source identification method-Dissipating Energy Flow. Our twin circuit theory also shows that adopting well-tuned grid-forming inverters can be a potential mitigation method for oscillation events. Finally, the effectiveness of our proposed oscillation event analysis framework is demonstrated by addressing a real-world 18-20 Hz oscillation event in Kaua‘i's power system on November 21, 2021.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Deep Reinforcement Learning Based Control of Wind Turbines for Fast Frequency Response

In order to fulfill vital auxiliary grid services, such as load regulation, spin and non-spin reserve provision, and frequency support during emergencies, there is often a requirement for certain wind farms to operate in de-loaded modes. Leveraging the swift response capabilities of wind farms, this study demonstrates that reserving power in de-loaded modes can significantly enhance power grid stability and reliability during system contingencies. Controlling wind farms optimally for frequency support is intricate due to the nonlinearity of models and controllers and the complexity of wind farm interactions with power systems. Here, to address this challenge, this paper introduces a novel approach that integrates wind turbines into reinforcement learning-based solutions for frequency response. This innovative methodology utilizes the state-of-the-art reinforcement learning algorithm known as the surrogate-gradient-based evolutionary strategy. The proposed learning-based algorithm provides continuous control of wind farm output to rapidly stabilize system frequency and prevent unnecessary trips of under-frequency load shedding relays. To facilitate efficient training, parallel computing techniques are employed. The proposed methodology is evaluated on a modified IEEE-39 bus system, and simulation results reveal its efficacy in reliably supporting power system frequency and preventing the need for unnecessary load shedding.

Gao, Wei [Argonne National Laboratory (ANL), Argon↗

Deep Learning-Based Failure Prognostic Model for PV Inverter Using Field Measurements

Here, this study presents a novel approach for the precise monitoring and prognosis of photovoltaic (PV) inverter status, which is crucial for the proactive maintenance of PV systems. It addresses the gaps in traditional model-based methods, which tend to neglect the overall reliability of inverters, and the limitations of data-driven approaches that largely depend on simulated data. This research presents a robust solution applicable to real-world scenarios. The proposed data-driven model for PV inverter failure prognosis employs actual inverter measurements, integrating various operational and weather-related factors based on domain knowledge. This approach effectively represents inverter stressors and operational status. Utilizing an Enhanced Siamese Convolutional Neural Network (ESCNN), the model merges operational data with domain knowledge features, redefining the prognosis challenge as a classification task. Furthermore, the paper discusses an ESCNN-based real-time inverter failure monitoring method developed on the well-trained model. The proposed models are rigorously trained and tested with real inverter data and a novel filtering method is included to address accidental failures in practical scenarios. The results validate the model's efficacy, and the directions for future research are also outlined.

42 ENGINEERING↗

Identification Uncertainty in Inverse Material Model Parameter Determination: A Sensitivity‐Based Decision Process for Load Path Selection

This research proposes a sensitivity-based framework for selecting the optimal prescribed loading path for a biaxial cruciform specimen. Optimality here is determined by the direction and magnitude of the prescribed displacement that minimizes the influence of random noise on the material model parameter identification. Using simulated experimental data based on finite element simulation, in this work, we identify the material model parameters of a Ludwik hardening model and plane stress implementation of the Hill-48 yield criterion using finite element model updating (FEMU). Our analysis reveals that the identification (or estimator) uncertainty of model parameters depends on the displacement boundary conditions (i.e., loading sequence) and the ground-truth value of the individual parameters. Optimal experimental design (OED) criteria based on the Fisher information matrix were investigated to mitigate indecision in the choice of optimal load path when the identification uncertainty of different material model parameters optimized at different load paths. The determinant of the Fisher information matrix was chosen here as the more useful metric due to its ability to capture uncertainty of the most influential material model parameters. The proposed framework demonstrates potential for real-time automated load step selection using scalar criteria derived prior to mechanical loading. The framework can be generalized to other geometries, boundary conditions and material models, allowing this procedure to be utilized for different experimental configurations and materials.

Fayad, Samuel S. [University of Illinois at Urbana↗

Evaluation of Mechanical and Thermomechanical Water Vapor Compression Techniques for Enabling High Temperature Lift Hydration-Based Chemical Heat Pumps

Achieving high temperature lifts (>200 K) via a chemical heat pump based on salt hydration/dehydration reactions requires the transport of water vapor from low to high pressure. Alternative compression approaches require condensing of low-pressure water vapor, pumping of liquid water, and subsequent evaporation when the low-side pressure corresponds to sub-ambient water saturation temperatures. Thus, this study compares four steam compression methods for use within a chemical heat pump system based on a reversible calcium oxide hydration/dehydration reaction with a temperature lift from 350 °C heat to >600 °C. Purely mechanical and thermochemical/mechanical compression technologies are considered. A parametric study of maximum allowable temperature, the isentropic efficiency of mechanical compressors, the effectiveness of heat exchangers, and the assumed allowable heat exchanger pressure drop is conducted to determine the mechanical and thermal energy consumed per kilogram of compressed steam. The system complexity in terms of the number of main system components, maximum pressure ratio, and maximum allowable temperature is estimated. Model results show an absorption-based steam compressor has the highest exergetic efficiency for the required chemical heat pump required conditions. As a result, this system configuration was then experimentally demonstrated to illustrate the impact of system performance on component effectiveness.

Advanced reactors↗

Rapid Commissioning of Large Machine Tools Using Finite Element-Based Correction of Geometric Errors

Large computer numerical control (CNC) machine tools derive their stiffness from monolithic cast iron bases or weldments that are sometimes integral to machine motion systems like box ways or guideways. However, the sheer size of castings and even floor flatness deviations result in dimensional errors in these systems, which manifest as machine motion errors. Typical geometric alignment processes rely on an iterative approach, where measurements are taken to assess alignment (straightness, squareness, and parallelism), followed by adjustment of the machine supports (fixators or leveling pads), which can take weeks even for an experienced operator. Conversely, a novel method is proposed to shorten the correction time by eliminating the trial-and-error process in favor of a more deterministic approach guided by a finite element (FE) method. A feasibility study is conducted on a CNC polymer hybrid machine, with a steel weldment frame, supported by six leveling pads. An FE model of the frame is utilized to obtain recommended leveling pad adjustments, based on measurement of machine errors taken using a laser tracker. After a single adjustment cycle, measurements reveal that geometric errors of the machine tool are reduced from 2.22 mm of flatness deviation to 0.32 mm, achieving an 85.6% reduction. Furthermore, the entire process including measurement, adjustment, and assessment is completed in just 6 h by two operators who are not professional service engineers. In conclusion, this methodology demonstrates feasibility for scaling up, especially to large, high-precision CNC machine tools with bases mounted by fixators, offering the capability for bidirectional adjustment.

42 ENGINEERING↗

A self-supervised robotic system for autonomous contact-based spatial mapping of semiconductor properties

Integrating robotically driven contact-based material characterization techniques into self-driving laboratories can enhance measurement quality, reliability, and throughput. While deep learning models support robust autonomy, current methods lack reliable pixel-precision positioning and require extensive labeled data. To overcome these challenges, we propose an approach for building self-supervised autonomy into contact-based robotic systems that teach the robot to follow domain expert measurement principles at high throughputs. We demonstrate the performance of this approach by autonomously driving a 4-DOF robotic probe for 24 hours to characterize semiconductor photoconductivity at 3025 uniquely predicted poses across a gradient of drop-casted perovskite film compositions, achieving throughputs of more than 125 measurements per hour. Spatially mapping photoconductivity onto each drop-casted film reveals compositional trends and regions of inhomogeneity, valuable for identifying manufacturing defects. With this self-supervised neural network–driven robotic system, we enable high-precision and reliable automation of contact-based characterization techniques at high throughputs, thereby allowing measurement of previously inaccessible yet important semiconductor properties for self-driving laboratories.

Science & Technology - Other Topics↗

Score-Based Physics-Informed Neural Networks for High-Dimensional Fokker–Planck Equations

The Fokker-Planck (FP) equation is a foundational partial differential equation (PDE) in stochastic processes involving Brownian motions. However, the curse of dimensionality (CoD) poses a formidable challenge when dealing with high-dimensional FP equations. Although Monte Carlo simulation and (vanilla) Physics-Informed Neural Networks (PINNs) have shown the potential to tackle CoD, both methods exhibit significant numerical errors in high dimensions when dealing with the probability density function (PDF) associated with Brownian motion. The point-wise PDF values tend to decrease exponentially as dimensionality increases, surpassing the precision of numerical simulations and resulting in substantial errors. In addition, due to its massive sampling, Monte Carlo fails to offer fast sampling. Modeling the logarithm likelihood (LL) via vanilla PINNs transforms the FP equation into a notoriously difficult Hamilton-Jacobi-Bellman (HJB) equation, which is impractical for PINN learning, whose error grows rapidly with dimension. To this end, we propose a novel approach utilizing a score-based solver to fit the score function in stochastic differential equations (SDEs). The score function, defined as the gradient of the LL, plays a fundamental role in inferring LL and PDF and enables fast SDE sampling, offering an effective means to overcome the CoD. Three fitting methods, Score Matching (SM), Sliced Score Matching (SSM), and Score-PINN, are introduced, each contributing unique advantages in computational complexity, accuracy, and generality. The proposed score-based SDE solver operates in two stages: first, employing score matching or Score-PINN to acquire the score function; and second, solving the LL via an ordinary differential equation (ODE) using the obtained score function. Comparative evaluations across these methods showcase varying trade-offs. The proposed methodology is evaluated across diverse SDEs, including anisotropic Ornstein-Uhlenbeck processes, geometric Brownian motion, and Brownian motion with varying eigenspace. We also test various distributions, including Gaussian, Log-normal, Laplace, and Cauchy distributions. The numerical results demonstrate the score-based SDE solver’s stability, speed, and performance across different experimental settings, solidifying its potential as a solution to CoD for high-dimensional FP equations.

97 MATHEMATICS AND COMPUTING↗

Agentic AI vs ML-Based Autotuning: A Comparative Study for Loop Reordering Optimization

High Performance Computing (HPC) applications rely heavily on code optimizations to achieve good performance on modern CPU and GPU architectures. Traditional Machine Learning auto-tuning approaches have demonstrated success in exploring high-dimensional spaces, but they often require expensive compile-run evaluations and lack adaptability for large HPC applications. The recent advances in Large Language Models (LLMs) and Agentic AI systems raise intriguing questions about the potential of these approaches to address specific optimization methodologies. This work aims to answer an essential question for the HPC community: “How Agentic AI Systems Compare to Traditional ML Autotuning Techniques?” To address this question, we present a comparative analysis between a traditional ML-based optimization approach and an Agentic AI system, evaluating their respective capabilities and limitations for loop-level optimization. In addition, we introduced a new Agentic AI system named LoopGen-AI using three different Large Language Models: GPT-4.1, Claude 4.0, and Gemini 2.5. A key finding is that LoopGen-AI achieves competitive per-formance with only a few program runs, the reasoning logs from the agents revealed that their decisions rely heavily on the combination of semantic understanding of the target kernel with dynamic feedback from the environment, highlighting a promising new dimension in performance tuning. In contrast, ML-based autotuners focus on statistical exploration, and require orders of magnitude more runs to reach peak performance. Additionally, our analysis shows that prompt engineering, particularly using Persona + Context Manager patterns, significantly impacts the effectiveness of Agentic AI. Our results indicate that while Agentic AI systems are not yet a complete replacement for ML-based autotuners, it can effectively complement traditional methods.

Rosas, Miguel Romero↗

Harnessing Redox-Active Molecules in Alkylammonium Halide-Based Eutectic Solvents for Redox Flow Batteries

Redox flow batteries (RFBs) are promising for large-scale energy storage, however, advancements in performance and cost-effectiveness are critical factors for adoption. Here we report on alkylammonium halide based eutectic solvents (ESs) and found that a ESs comprising of diethylammonium chloride or bromide, in ethylene glycol demonstrated exceptional stability and a wide electrochemical stability window, making it a promising candidate for RFB applications. The electrochemical stability and redox behavior of the alkylammonium halide-based ESs are significantly influenced by hydrogen-bonding interactions, modulated by the alkyl chain length of the cation and the nature of the anion. A redox-active eutectic electrolyte containing 0.45 M methyl viologen dichloride (MV 2+ ) paired with acetylferrocene exhibited reversible redox behavior with a maximum open-circuit voltage of ∼1.34 V. The first redox couple, representing the viologen dication to radical cation transition, exhibited remarkable stability with consistent performance, achieving an energy efficiency near 70% at charge-discharge current densities of 10 mA cm −2 over 160 cycles with about 3% loss of the initial capacity. This study highlights the potential of alkylammonium halide DES systems for implementing eutectic-based RFB technologies in the future.

25 ENERGY STORAGE↗

Round Robin Analysis of Uranium Isotopics on Cotton Swipes Measured Using Microextraction-Based Sampling Versus Conventional Bulk Analysis

A comparison of microextraction sampling methods to directly analyze uranium isotopics on cotton swipes was performed concurrently with traditional bulk-processing methods. For the microextraction sampling approach, two different detection platforms were evaluated, a quadrupole-based inductively coupled plasma mass spectrometer (ICP-MS) and the liquid sampling-atmospheric pressure glow discharge (LS-APGD) coupled to an Orbitrap mass spectrometer. Results presented from this innovative sampling approach (i.e., microextraction) are compared with a more traditional approach employed for analysis of cotton-based environmental swipes, namely bulk ashing/digestion, separation, and subsequent analysis by high-precision multi-collector ICP-MS. Overall, the microextraction approach proved to be a reliable and accurate means to determine isotopic ratios of uranium collected on cotton swipes. The ICP-MS-based detection had relative standard deviations of <0.65% for the major isotopic determinations, whereas the LS-APGD-Orbitrap method had relative standard deviations of <3%. The percent relative difference for the 235 U/ 238 U ratios, in comparison to the expected values, was <1% for ICP-MS and <4% for the microplasma-Orbitrap method. Additionally, the microextraction ICP-MS accurately (<2%) and precisely (<5%) determined the minor isotopic compositions (i.e., 234 U/ 238 U and 236 U/ 238 U).

ICP-MS↗