Search NASA⌕ Search

SEARCH · Search NASA

Results for “Distributed training”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

Colorado Technology Primer for Economists and Social Scientists: Cooperative Research and Development Final Report, CRADA Number CRD-19-00790

NLR will assist the Colorado School of Mines (Mines) in supporting a series of one-week training workshops. This proposed training program will be two, week-long summer school sessions in each of the next two years to help give early career economists and social scientists a solid introduction and grounding on the technical components of the electrical distribution, transmission, and generation systems as well as basic to advanced overview of clean energy and traditional generation systems.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Multi-Sensor Data from A-Train Instruments Brought Together for Atmospheric Research

The A-Train is comprised of a series of instruments, developed independently, that measure highly related atmospheric components along the same flight path. In order to intercompare data from this multitude of sensors, researchers must access, subset, visualize, analyze and correlate distributed atmosphere measurements from the various A-Train instruments. The A-Train Data Depot (ATDD) has been operational for over a year, successfully performing the aforementioned functions on behalf of researchers, thus providing co-registered data from the Cloudsat, CALIOP, AIRS, and MODIS instruments for further intercomparisons. Of late, significant data from OM1 and POLDER are now included in the 'depot'. By specifying the desired spatial and temporal range, the researcher can subset, visualize, co-register, and access multi-sensor A-Train data related to: Cloud, aerosol, atmospheric temperature, and water vapor parameters (vertical profile visualizations); Cloud Pressure, cloud top temperature, water vapor, cloud optical thickness, and aerosol products (horizontal strips subsetted +/- 100km from the profile visualizations), and; Cloud pressure parameters (2-D line plots overlayed on the vertical profiles). All data is plotted using the GIOVANNI data exploration tool. A new feature of GIOVANNI is its ability to have collocated and subsetted data sets as well as PNG image files downloaded to the researcher's computing facility. By providing a convenient way to visualize and acquire multi-sensor data, ATDD affords users more time and effort to further their research.

Smith, Peter M.↗

Normalizing flows for domain adaptation when identifying Λ hyperon events

Here this study focuses on the application of a normalizing flow as a method of domain adaptation when classifying physics data. Normalizing flows offer a way to transform data points between two different distributions. The present study investigates a novel method of transforming latent representations of physics data to a normal distribution and then to a physics distribution again. The final distribution models a simulated distribution. After being transformed, the data can be classified by a neural network trained on labeled simulation data. The present study succeeds in training two normalizing flows that can transform between data (or simulation) and a Gaussian distribution.

47 OTHER INSTRUMENTATION↗

EV Charging Infrastructure Energization An Overview of Approaches for Simplifying and Accelerating Timelines to Processing EV Charging Load Service Requests

The United States has seen significant growth in electric vehicle (EV) adoption, leading to increased demand for EV charging infrastructure. Over the past decade, EV charging infrastructure site developers, site hosts, and electric distribution utilities have navigated the process to integrate chargers onto the electric grid. Site developers and site hosts have raised the alarm that the integration process for high-powered EV charging projects does not meet the needs of the EV market for timeliness or cost. High-powered charging stations typically require a load service request or an agreement with the local utility to connect to the grid. The process of energizing a new high-powered charging site can be complex and time-consuming, often taking up to 2 years. This timeline is the result of current utility energization processes having been designed for construction projects that take longer to build (i.e., buildings). The specific challenges stem from various factors, including compartmentalization in application processes, the integration of EV charging process approvals with other distributed energy resources (DERs), and the need to ensure grid reliability. The energization process needs to evolve to meet the growing demand for high-powered EV charging. This white paper compiles information gathered through various conversations with key stakeholders, including utilities, utility regulators, EV charging operators, site developers, and authorities having jurisdiction (AHJ) as well as through an extensive literature review. This document identifies the challenges and provides potential solutions to streamline the process of connecting EV charging infrastructure to the power grid in the United States, serving as a starting point for future conversations around these solutions. The solutions noted in this white paper require collaborative efforts among utilities, regulators, and EV charging infrastructure developers to streamline the grid connection process for EV charging infrastructure. They are broadly organized into four areas: 1. Increase data access and transparency: Develop automated load service request tools, integrate hosting capacity and load service request analyses, incorporate EV adoption forecasts, and provide transparency on the processing queue. 2. Improve energization processes and timing: Create fast-track options based on prescreening criteria, provide flexibility or phased approvals in the load service request/interconnection process, build internal knowledge within utilities about EV charging technologies, and provide standardized workforce training. 3. Promote economic efficiency: Right size distribution components to accurately reflect the load requirements of EV charging infrastructure, make proactive investments in grid infrastructure based on EV adoption forecasts and growth projections, and consider energy equity and environmental justice factors such as equitable access to EV charging when planning infrastructure. 4. Improve grid reliability and resilience: Use load management/power control systems (PCS) at EV charging stations, adopt and implement harmonized standards for communication protocols and information models between the EV charging and grid control infrastructure, and address cybersecurity considerations by implementing robust security measures and standards for EV charging infrastructure—with particular emphasis on clarifying the security requirements for the interface to the grid. The objective of the solutions proposed in this white paper is to accelerate the timeline and decrease costs associated with connecting EV charging infrastructure to the grid. Electric utilities, utility regulators, EV charging infrastructure developers, and site hosts will first need to understand which solutions are available in their service territory, and if warranted, which combination of solutions would support their specific needs. Through the successful implementations of solutions at scale detailed here, industry will demonstrate a new and innovative ecosystem where timely deployment and energization of EV charging infrastructure with greater grid resiliency and reliability is a reality.

24 POWER TRANSMISSION AND DISTRIBUTION↗

On the Generalizability of Time-of-Flight Convolutional Neural Networks for Noninvasive Acoustic Measurements

Bulk wave acoustic time-of-flight (ToF) measurements in pipes and closed containers can be hindered by guided waves with similar arrival times propagating in the container wall, especially when a low excitation frequency is used to mitigate sound attenuation from the material. Convolutional neural networks (CNNs) have emerged as a new paradigm for obtaining accurate ToF in non-destructive evaluation (NDE) and have been demonstrated for such complicated conditions. However, the generalizability of ToF-CNNs has not been investigated. In this work, we analyze the generalizability of the ToF-CNN for broader applications, given limited training data. We first investigate the CNN performance with respect to training dataset size and different training data and test data parameters (container dimensions and material properties). Furthermore, we perform a series of tests to understand the distribution of data parameters that need to be incorporated in training for enhanced model generalizability. This is investigated by training the model on a set of small- and large-container datasets regardless of the test data. We observe that the quantity of data partitioned for training must be of a good representation of the entire sets and sufficient to span through the input space. The result of the network also shows that the learning model with the training data on small containers delivers a sufficiently stable result on different feature interactions compared to the learning model with the training data on large containers. To check the robustness of the model, we tested the trained model to predict the ToF of different sound speed mediums, which shows excellent accuracy. Furthermore, to mimic real experimental scenarios, data are augmented by adding noise. We envision that the proposed approach will extend the applications of CNNs for ToF prediction in a broader range.

47 OTHER INSTRUMENTATION↗

Joint aircraft loading/structure response statistics of time to service crack initiation

A reliability analysis for predicting the statistical distribution of time to fatigue crack initiation for aircraft structures in service is presented. The present analysis utilizes the statistical data of the specimen fatigue tests, the full-scale structure tests, and the statistical dispersion of aircraft service loads. The statistical distribution of the time to fatigue crack initiation of the full-scale structure under laboratory loading spectrum is assumed to be Weibull. The service loads for gust turbulences are modeled as Poisson processes for transport-type aircraft, while the maneuver loads are modeled as compound Poisson processes for fighter and training aircraft. It is found that the statistical distribution of time to fatigue crack initiation for aircraft structures in service is not Weibull and that the prediction on the basis of the Weibull distribution is unconservative, in particular in the early service time.

Yang, J.-N.↗

SDYN-GANs: Adversarial learning methods for multistep generative models for general order stochastic dynamics

We introduce adversarial learning methods for data-driven generative modeling of dynamics of nth-order stochastic systems. Our approach builds on Generative Adversarial Networks (GANs) with generative model classes based on stable m-step stochastic numerical integrators. From observations of trajectory samples, we introduce methods for learning long-time predictors and stable representations of the dynamics. Our approaches use discriminators based on Maximum Mean Discrepancy (MMD), training protocols using both conditional and marginal distributions, and methods for learning dynamic responses over different time-scales. We show how our approaches can be used for modeling physical systems to learn force-laws, damping coefficients, and noise-related parameters. Our adversarial learning approaches provide methods for obtaining stable generative models for dynamic tasks including long-time prediction and developing simulations for stochastic systems.

• Artificial intelligence (AI) / machine learning ↗

Multi-head physics-informed neural networks for learning functional priors and uncertainty quantification

In numerous applications, the integration of prior knowledge and historical information is essential, particularly for tasks requiring the solution of ordinary or partial differential equations (ODEs/PDEs) in data-sparse or noisy environments. For instance, achieving accurate solutions to time-dependent PDEs with limited initial condition measurements necessitates an effective strategy for embedding prior knowledge. Hard-parameter sharing architectures in neural networks (NNs) have demonstrated success in both traditional and scientific machine learning domains, facilitating the learning of informative representations. Here, in this study, we introduce a novel, yet efficient, method to enhance physics-informed neural networks (PINNs) by incorporating a multi-head structure that enables the learning of functional priors from both empirical data and governing physical laws. This prior information can then be used to address data sparsity and high-level noise in solving ODE/PDE problems with uncertainty quantification (UQ). The approach, termed Multi-Head PINN (MH-PINN), consists of a shared body NN and multiple head NNs, each corresponding to an individual PINN instance. Our framework for functional prior learning is carried out in two stages: (1) training the MH-PINNs to develop a shared body NN alongside multiple head NNs, and (2) employing these trained head NNs to estimate a prior distribution through a normalizing flow-based density estimator. The learned functional prior can then be applied as a regularization mechanism in deterministic contexts or as an informative prior within a Bayesian inference framework, aiding in the resolution of subsequent ODE/PDE tasks. We evaluate the efficacy of MH-PINNs across five benchmark problems, including a high-dimensional parametric PDE, all characterized by data sparsity or substantial noise levels. Our findings reveal that MH-PINNs deliver accurate solutions and robust UQ, demonstrating adaptability across a range of complex and challenging scenarios.

Bayesian inference↗

Optimisation of the Kaplan hydropower system via PID 2 and digital twin

Here, this paper proposes a proportional–integral-double–derivative (PID 2 ) optimisation method for the Kaplan hydropower system by building a digital twin. The study first uses one multilayer perceptron (MLP) to model the hydroturbine dynamic and then adopts three connected MLPs to model the generator dynamic, both in an open-loop fashion. Inspired by stochastic distribution control (SDC) theory, we regard the training of the turbine's neural network model as a process control problem, and we propose minimising entropy loss to update the network parameters. The next step is to build the digital twin by connecting the neural network models with a PID 2 controller and a lead-lag exciter and run the whole model in a closed-loop fashion. After that, a binary search approach is applied to optimise the PID 2 parameters based on the obtained digital twin model. The simulation results show that the proposed method can reduce the mean square tracking error by more than 90%. Furthermore, the method is extended to jointly optimise the PID 2 controller and excitation system gains through multiobjective optimisation, leveraging Pareto frontier analysis to balance active power and voltage tracking performance. Simulation results confirm the effectiveness of the proposed method, achieving a 83.46% reduction in relative mean square error of active power, a 47.13% reduction in terminal voltage tracking error, and an 82.78% improvement in the overall scalarized objective.

Hydropower system↗

Dynamical Sketching for Enhanced Communication Efficiency in Federated Learning

Federated learning (FL) has revolutionized distributed machine learning by enabling collaborative model training without sharing local data. However, communication efficiency and privacy guarantees remain significant challenges. This paper introduces a dynamic sketching mechanism in FL, optimizing the trade-off between communication efficiency and model accuracy. By dynamically selecting the sketch matrix size, our approach adapts to the evolving characteristics of the data and the model, ensuring optimal performance across diverse scenarios. We leverage Bayesian optimization to systematically tune the sketch parameters, achieving an effective balance between resource efficiency and model performance. Experimental results on the MNIST dataset using a convolutional neural network (CNN) architecture validate the proposed method's efficiency and scalability. Our dynamic sketching approach significantly outperforms fixed-size sketching techniques, achieving higher compression ratios (up to 62x) and providing better privacy guarantees while maintaining high model accuracy. These findings highlight the robustness and versatility of our approach and make it a valuable solution for privacy-preserving, communication-efficient federated learning.

Afrose, Sharmin [ORNL]↗

Landsat-D thematic mapper simulation in an urban area using aircraft multispectral scanner data

A simulation of imagery from the Landsat-D thematic mapper was conducted in order to determine its usefulness for urban land-use classification. Aircraft 24-channel multispectral scanner imagery of the Los Angeles area at 7.5-m resolution was processed digitally by means of matrix averaging and image smoothing techniques to simulate the 30-m resolution of the thematic mapper. Mean and standard deviation statistics of training sites for resolutions of 7.5, 15, 30 and 60 m were used to generate final classification maps. Plots of relative standard deviation showed that for larger training sites, as the resolution decreased, the distribution range of density values also decreased, while plots of relative classification accuracies showed that as resolution decreased, classification accuracies for three levels of standard deviation increased. A point of diminishing returns was indicated, however, confirming the utility of the resolution intended for Landsat-D.

Clark, J.↗

Investigations on the downwash behind a tapered wing with fuselage and propeller

The new downwash measurements behind a tapered wing with parallel center section described in the present report can be brought into good agreement with theoretical calculations if made on the basis of not-rolled-up vortex sheet and allowance is made for the lowering of the sheet. The test values are about 1 degree higher than the "upper limit" established for it, as against approximately 0.5 degrees in the earlier tests behind a rectangular and elliptical wing. The measurements on lateral axes, especially if lying below the wing on a level with the vortex train, disclosed in accord with the lift distribution, a marked change in angle over the span of the tail in contrast to the rectangular and elliptical wing.

Muttray, H↗

NASA A-Train and Terra Observations of the 2010 Russian Wildfires

Wildfires raged throughout western Russia and parts of Eastern Europe during a persistent heat wave in the summer of 2010. Anomalously high surface temperatures (35 - 41 C) and low relative humidity (9 - 25 %) from mid- June to mid-August 2010 shown by analysis of radiosonde data from multiple sites in western Russia were ideal conditions for the wildfires to thrive. Measurements of outgoing longwave radiation (OLR) from the Atmospheric Infrared Sounder (AIRS) over western Russian indicate persistent subsidence during the heat wave. Daily three-day back-trajectories initiated over Moscow reveal a persistent anticyclonic circulation for 18 days in August, coincident with the most intense period of fire activity observed by Moderate Resolution Imaging Spectroradiometer (MODIS). This unfortunate meteorological coincidence allowed transport of polluted air from the region of intense fires to Moscow and the surrounding area. We demonstrate that the 2010 Russian wildfires are unique in the record of observations obtained by remote-sensing instruments on-board NASA satellites: Aura and Aqua (part of the A-Train Constellation) and Terra. Analysis of the distribution of MODIS fire products and aerosol optical thickness (AOT), UV aerosol index (AI) and single-scattering albedo (SSA) from Aura's Ozone Monitoring Instrument (OMI), and total column carbon monoxide (CO) from Aqua s Atmospheric Infrared Sounder (AIRS) show that the region in the center of western Russia surrounding Moscow (52-58 deg N, 33 -43 deg E) is most severely impacted by wildfire emissions. Over this area, AIRS CO, OMI AI, and MODIS AOT are significantly enhanced relative to the historical satellite record during the first 18 days in August when the anti-cyclonic circulation persisted. By mid-August, the anti-cyclonic circulation was replaced with westerly transport over Moscow and vicinity. The heat wave

Witte, J. C.↗

A Risk-Informed Approach to Trustworthiness Assessment in Digital Twins-Based Autonomous Control

In autonomous control systems, digital twins (DTs) are used to perform diagnostic and prognostic functions. The trustworthiness of these DTs is dependent on quality and coverage of the training data, model accuracy and integrity of sensor data. This work introduces a methodology to determine the trustworthiness of a DT system given faulty sensor data using a risk informed approach. Bayesian Belief Networks (BBNs) are used to propagate uncertainties and determine the probability of trustable recommendations. The decision to trust the control action provided by the DT is based on the DT output, expert opinion, and severity of problems. The performance of DTs is reliant on the data they are trained on. When they encounter out of distribution data, the trustworthiness of the recommendations decreases. To address this issue, we include an expert component that provides input on sensor degradation. For this, we utilize a generative artificial intelligence (AI) model, such as Generative Pretrained Transformer (GPT). The GPT functions as an expert with broad knowledge. The GPT is fine-tuned to understand and discriminate sensor degradation scenarios using manufactured data. This methodology is demonstrated through a case study on a Nearly Autonomous Management and Control System (NAMAC) during a steady state scenario. Various sensor degradation types with different severity levels are considered. Degraded sensor data is processed by the DT system and the fine-tuned GPT. Finally, using the BBN, we combine the GPT information and the DT output with its sources of uncertainty. This provides an output regarding the trustworthiness of the DT recommendation.

22 - GENERAL STUDIES OF NUCLEAR REACTORS↗

An expert system for a distributed real-time trainer

The problem addressed by this expert system concerns the expansion of capability of a Real Time Trainer for the Spacelab flight crew. As requirements for more models or fidelity are placed upon the system, expansion is necessary. The simulator can be expanded using a larger processor or by going to a distributed system and expand by adding additional processors. The distributed system is preferable because it is more economical and can be expanded in a more incremental manner. An expert system was developed to evaluate modeling and timing capability within a real time training simulator. The expert system is based upon a distributed configuration. Components of the modeled system are control tasks, network tasks, emulator tasks, processors, displays, and a network. The distributed module expert system (DMES) allows the configuring of processors, tasks, display use, keyboard use, and selection of alternate methods to update the data buffer. Modules can be defined with execution occurring in a specific processor on a network. The system consists of a knowledge front end editor to interactively generate or update the knowledge base, an inference engine, a display module, and a recording module.

Purinton, Steven C.↗

Temperature Field Reconstruction of Surfaces Heated Through Radiative Heat Transfer Using Convolutional Neural Networks

Microreactors could play a crucial role in decarbonizing our energy portfolio. However, their development and implementation come with specific challenges, particularly regarding cost. Due to their compact size and the harsh operational environment, collecting real-time data on reactor operation can be challenging. Many probe designs are unable to withstand extreme conditions (e.g., temperature, radiation) in the reactor. In this context, using convolutional neural networks (CNNs) can pave the way for developing a nonintrusive approach that relies solely on ex-core sensors. A well-trained physics-informed CNN can reconstruct the distribution of a given physical quantity over a domain using only a few sensors, allowing us to reconstruct the desired field distribution even in a limited space or complex geometries where a large array of sensors is impractical. In this work, we present the initial steps toward developing a real-time tool for monitoring the thermal behavior of nuclear reactor pressure vessels. Based on an experimental setup, a computational model using the Multiphysics Object-Oriented Simulation Environment (moose) framework was built, where the Ray Tracing and Heat Conduction modules were used to evaluate the temperature distribution over a convex metal surface heated through radiative heat transfer. This metal surface represents a section of a heated nuclear reactor vessel wall. The model also accounts for solid mechanics physics through the moose Solid Mechanics module. In situ experimental data, acquired from a Texas A&M facility, were used to validate the computational model. Part of the data generated by the moose model was used to train the convolutional neural network to reconstruct the vessel wall's outer surface temperature. The CNN generalization was then compared against the experimental and computational data.

Aldeia Machado, Luiz Carlos↗

Transferable predictions of energetic and structural properties for refractory solid solution alloys across chemical compositions

We present a data-efficient approach to train graph neural networks (GNNs) on density functional theory (DFT) data for accurate and transferable predictions of energetic and structural properties of refractory solid solution alloys in the niobium-tantalum-vanadium (Nb-Ta-V) chemical space. We start by training the GNN model only on DFT data that describes refractory binary alloys niobium-tantalum (Nb-Ta), niobium-vanadium (Nb-V), and tantalum-vanadium (Ta-V) to predict formation enthalpy and root mean squared displacement. Once trained, the GNN predictions are tested on DFT data describing refractory ternary alloys Nb-Ta-V. While, unsurprisingly, direct transferability from binary to ternary is not sufficiently accurate, augmenting the training with only 1% of the available ternary data (uniformly distributed across the entire range of chemical compositions) improves significantly the quality of the GNN predictions. For comparison, we assess the transferability in the opposite direction by training GNN models on ternary Nb-Ta-V data and making predictions on binaries Nb-Ta, Nb-V, and Ta-V, which exhibits notably higher predictive errors. The proposed methodology, which favors transferability from lower-component to higher-component alloys, offers an efficient path towards avoiding the curse of dimensionality incurred when collecting DFT data for discovery and design of multi-component disordered alloys.

Density functional theory calculations↗

Exploring the energy landscape of RBMs: reciprocal space insights into bosons, hierarchical learning and symmetry breaking

Deep generative models have become ubiquitous due to their ability to learn and sample from complex distributions. Despite the proliferation of various frameworks, the relationships among these models remain largely unexplored, a gap that hinders the development of a unified theory of AI learning. In this work, we address two central challenges: clarifying the connections between different deep generative models and deepening our understanding of their learning mechanisms. We focus on Restricted Boltzmann Machines (RBMs), a class of generative models known for their universal approximation capabilities for discrete distributions. By introducing a reciprocal space formulation for RBMs, we reveal a connection between these models, diffusion processes, and systems of coupled bosons. Our analysis shows that at initialization, the RBM operates at a saddle point, where the local curvature is determined by the singular values of the weight matrix, whose distribution follows the Marc̆enko-Pastur law and exhibits rotational symmetry. During training, this rotational symmetry is broken due to hierarchical learning, where different degrees of freedom progressively capture features at multiple levels of abstraction. This leads to a symmetry breaking in the energy landscape, reminiscent of Landau’s theory. This symmetry breaking in the energy landscape is characterized by the singular values and the weight matrix eigenvector matrix. We derive the corresponding free energy in a mean-field approximation. We show that in the limit of infinite size RBM, the reciprocal variables are Gaussian distributed. Our findings indicate that in this regime, there will be some modes for which the diffusion process will not converge to the Boltzmann distribution. To illustrate our results, we trained replicas of RBMs with different hidden layer sizes using the MNIST dataset. Our findings not only bridge the gap between disparate generative frameworks but also shed light on the fundamental processes underpinning learning in deep generative models.

97 MATHEMATICS AND COMPUTING↗