Search NASA⌕ Search

SEARCH · Search NASA

Results for “distributed learning”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 559 records · Page 31

Some properties of adiabatic blast waves in preexisting cavities

Cox and Anderson (1982) have conducted an investigation regarding an adiabatic blast wave in a region of uniform density and finite external pressure. In connection with an application of the results of the investigation to a study of interstellar blast waves in the very hot, low-density matrix, it was found that it would be desirable to examine situations with a positive radial density gradient in the ambient medium. Information concerning such situations is needed to learn about the behavior of blast waves occurring within preexisting, presumably supernova-induced cavities in the interstellar mass distribution. The present investigation is concerned with the first steps of a study conducted to obtain the required information. A review is conducted of Sedov's (1959) similarity solutions for the dynamical structure of any explosion in a medium with negligible pressure and power law density dependence on radius.

Cox, D. P.↗

Boosting with Averaged Weight Vectors

AdaBoost is a well-known ensemble learning algorithm that constructs its constituent or base models in sequence. A key step in AdaBoost is constructing a distribution over the training examples to create each base model. This distribution, represented as a vector, is constructed to be orthogonal to the vector of mistakes made by the previous base model in the sequence. The idea is to make the next base model's errors uncorrelated with those of the previous model. Some researchers have pointed out the intuition that it is probably better to construct a distribution that is orthogonal to the mistake vectors of all the previous base models, but that this is not always possible. We present an algorithm that attempts to come as close as possible to this goal in an efficient manner. We present experimental results demonstrating significant improvement over AdaBoost and the Totally Corrective boosting algorithm, which also attempts to satisfy this goal.

Oza, Nikunj C.↗

Markov Chain Monte Carlo Bayesian Learning for Neural Networks

Conventional training methods for neural networks involve starting al a random location in the solution space of the network weights, navigating an error hyper surface to reach a minimum, and sometime stochastic based techniques (e.g., genetic algorithms) to avoid entrapment in a local minimum. It is further typically necessary to preprocess the data (e.g., normalization) to keep the training algorithm on course. Conversely, Bayesian based learning is an epistemological approach concerned with formally updating the plausibility of competing candidate hypotheses thereby obtaining a posterior distribution for the network weights conditioned on the available data and a prior distribution. In this paper, we developed a powerful methodology for estimating the full residual uncertainty in network weights and therefore network predictions by using a modified Jeffery's prior combined with a Metropolis Markov Chain Monte Carlo method.

Goodrich, Michael S.↗

The Zwicky Transient Facility: Data Processing, Products, and Archive

The Zwicky Transient Facility (ZTF) is a new robotic time-domain survey currently in progress using the Palomar 48-inch Schmidt Telescope. ZTF uses a 47 square degree field with a 600 megapixel camera to scan the entire northern visible sky at rates of ∼3760 square degrees/hour to median depths of g ~ 20.8 and r ~ 20.6 mag (AB, 5σ in 30 sec). We describe the Science Data System that is housed at IPAC, Caltech. This comprises the data-processing pipelines, alert production system, data archive, and user interfaces for accessing and analyzing the products. The real-time pipeline employs a novel image-differencing algorithm, optimized for the detection of point-source transient events. These events are vetted for reliability using a machine-learned classifier and combined with contextual information to generate data-rich alert packets. The packets become available for distribution typically within 13 minutes (95th percentile) of observation. Detected events are also linked to generate candidate moving-object tracks using a novel algorithm. Objects that move fast enough to streak in the individual exposures are also extracted and vetted. We present some preliminary results of the calibration performance delivered by the real-time pipeline. The reconstructed astrometric accuracy per science image with respect to Gaia DR1 is typically 45 to 85 milliarcsec. This is the RMS per-axis on the sky for sources extracted with photometric S/N ≥10 and hence corresponds to the typical astrometric uncertainty down to this limit. The derived photometric precision (repeatability) at bright unsaturated fluxes varies between 8 and 25 millimag. The high end of these ranges corresponds to an airmass approaching ∼2—the limit of the public survey. Photometric calibration accuracy with respect to Pan-STARRS1 is generally better than 2%. The products support a broad range of scientific applications: fast and young supernovae; rare flux transients; variable stars; eclipsing binaries; variability from active galactic nuclei; counterparts to gravitational wave sources; a more complete census of Type Ia supernovae; and solar-system objects.

Frank J. Masci↗

torch-einshard v1.0

torch-einshard is a Python library for describing local and distributed PyTorch tensor computations with compact, einsum-like notation. Its expressions name logical axes, specify how they are sharded across a PyTorch DeviceMesh, and represent partial reductions. The library automatically performs contractions, permutations, reshaping, splitting, gathering, reduction, reduce-scatter, and repartitioning while preserving autograd. Additional features include sharding-aware FFTs, tensor rolls, halo exchange, sliding windows, 1D–3D convolutions, uneven-shard handling, parameter initialization and gradient management, and cost-based execution planning. It is designed for scientific machine learning and large-model workloads, including tensor-, sequence-, and spatial-parallel MLPs, attention, convolutions, and spectral operations. Compared with manually combining torch.einsum and distributed collectives, torch-einshard expresses both the mathematical operation and data placement in one readable formula. This reduces boilerplate and synchronization errors, keeps forward and backward communication consistent, and allows the library to select optimized collective strategies without changing model code.

Morozov, Dmitriy [Lawrence Berkeley National Labor↗

Statistical and Machine Learning Approaches to Analyzing Pipeline Incidents in the United States (2010–2024)

This study applies machine learning methods to analyze natural gas pipeline incidents in the United States using the Pipeline and Hazardous Materials Safety Administration (PHMSA) Gas Distribution Incident Dataset (2010–2024). The dataset includes over 600 variables describing incident characteristics, infrastructure attributes, and contributing factors associated with unintentional gas releases. The objective is to assess whether these features can reliably predict the underlying cause of pipeline failures. Multinomial logistic regression and Random Forest models were developed to classify incident causes, including excavation damage, corrosion, equipment failure, and natural forces. Results show that excavation damage is both the most frequent and most predictable cause, with models achieving strong performance for this category. However, when excavation damage is excluded, model accuracy declines significantly, with some models performing near random levels. Across all approaches, severe class imbalance and limited variability in key predictors constrain predictive performance. Pipeline age and diameter emerge as the most influential variables, but they provide insufficient discriminatory power to distinguish among less frequent failure types. These findings indicate that non-excavation-related incidents are rare, heterogeneous, and weakly represented in the dataset, limiting the effectiveness of machine learning classification. Overall, this study highlights the structural limitations of the PHMSA dataset for predictive modeling and underscores the need for improved data balance and feature enrichment. The results reinforce excavation damage prevention as the most impactful strategy for reducing pipeline incidents.

03 NATURAL GAS↗

Trapped particles and waves, and what can be learned from multisatellite experiments

Calculations concerning the pitch-angle diffusion resulting from resonant wave-particle interactions can lead to definitive predictions of equatorial pitch-angle distributions and rates of particle loss as a function of particle energy and L-value. Thus, given simultaneous high-altitude measurements of pitch-angle distributions and low-altitude measurements of precipitating fluxes as a function of energy and L, the importance of proposed wave-particle interactions can be verified or discarded. Since many wave-particle phenomena occur over large spatial and temporal scales, exact simultaneity in longitude and time is not necessary. Simultaneous low and high altitude (preferably nearly equatorial) particle measurements could thus greatly increase our understanding of trapped particles and their effects on the ionosphere. Furthermore, given a verified pitch-angle diffusion mechanism and simultaneous low- and high-altitude measurements, accurate lowto high-altitude mappings of field lines and magnetospheric boundaries (such as the plasmapause) could be obtained.

Lyons, L. R.↗

Attention to quantum complexity

The imminent era of error-corrected quantum computing demands robust methods to characterize quantum state complexity from limited, noisy measurements. We introduce the Quantum Attention Network (QuAN), a classical artificial intelligence (AI) framework leveraging attention mechanisms tailored for learning quantum complexity. Inspired by large language models, QuAN treats measurement snapshots as tokens while respecting permutation invariance. Combined with our parameter-efficient miniset self-attention block, this enables QuAN to access high-order moments of bit-string distributions and preferentially attend to less noisy snapshots. We test QuAN across three quantum simulation settings: driven hard-core Bose-Hubbard model, random quantum circuits, and toric code under coherent and incoherent noise. QuAN directly learns entanglement and state complexity growth from experimental computational basis measurements, including complexity growth in random circuits from noisy data. In regimes inaccessible to existing theory, QuAN unveils the complete phase diagram for noisy toric code data as a function of both noise types, highlighting AI’s transformative potential for assisting quantum hardware.

Kim, Hyejin [Cornell Univ., Ithaca, NY (United Sta↗

Retrieve Methane from IR sounder measurements Using Machine Learning-Enhanced Physical Inversion

The sensitivity of IR sounder measurements to atmospheric CH 4 is often limited due to interferences from signals of other trace gases, insufficient thermal contrast, and cloud blockage. In order to resolve the geographical and vertical distribution of atmospheric CH 4 profiles, accurate scene-dependent a priori information is critically needed to support an optimal estimation method-based physical inversion scheme. Following the principles of indexing, representation, and retrieval, a spectral fingerprinting methodology is developed to address the needs for both accuracy and computational efficiency in sounder-based CH 4 retrieval. Within this framework, a clustering method based on machine learning is first employed to stratify and identify the a priori state within the pre-constructed database, using optimized spectral radiances as predictors. The corresponding radiative kernel is then used to establish the physical inversion scheme for finding the solution. High-quality data from CH 4 data assimilation systems like the Carbon-Tracker and the Copernicus Atmosphere Monitoring Service (CAMS) reanalysis, as well as the state-of-art sounder products are used to build the training database, including radiative kernels. We will demonstrate the results retrieved from CrIS observations and the associated validation work.

Wan Wu↗

Destabilizing high-capacity high entropy hydrides via earth abundant substitutions: From predictions to experimental validation

The vast chemical space of high entropy alloys (HEAs) makes trial-and-error experimental approaches for materials discovery intractable and often necessitates data-driven and/or first principles computational insights to successfully target materials with desired properties. In the context of materials discovery for hydrogen storage applications, a theoretical prediction-experimental validation approach can vastly accelerate the search for substitution strategies to destabilize high-capacity hydrides based on benchmark HEAs, e.g. TiVNbCr alloys. Here, in this study, machine learning predictions, corroborated by density functional theory calculations, predict substantial hydride destabilization with increasing substitution of earth-abundant Fe content in the (TiVNb) 75 Cr 25-x Fe x system. The as-prepared alloys crystallize in a single-phase bcc lattice for limited Fe content x < 7, while larger Fe content favors the formation of a secondary C14 Laves phase intermetallic. Short range order for alloys with x < 7 can be well described by a random distribution of atoms within the bcc lattice without lattice distortion. Hydrogen absorption experiments performed on selected alloys validate the predicted thermodynamic destabilization of the corresponding fcc hydrides and demonstrate promising lifecycle performance through reversible absorption/desorption. This demonstrates the potential of computationally expedited hydride discovery and points to further opportunities for optimizing bcc alloy ↔ fcc hydrides for practical hydrogen storage applications.

36 MATERIALS SCIENCE↗

Unsupervised learning-enabled pulsed infrared thermographic microscopy of subsurface defects in stainless steel

Metallic structures produced with laser powder bed fusion (LPBF) additive manufacturing method (AM) frequently contain microscopic porosity defects, with typical approximate size distribution from one to 100 microns. Presence of such defects could lead to premature failure of the structure. In principle, structural integrity assessment of LPBF metals can be accomplished with nondestructive evaluation (NDE). Pulsed infrared thermography (PIT) is a non-contact, one-sided NDE method that allows for imaging of internal defects in arbitrary size and shape metallic structures using heat transfer. PIT imaging is performed using compact instrumentation consisting of a flash lamp for deposition of a heat pulse, and a fast frame infrared (IR) camera for measuring surface temperature transients. However, limitations of imaging resolution with PIT include blurring due to heat diffusion, sensitivity limit of the IR camera. We demonstrate enhancement of PIT imaging capability with unsupervised learning (UL), which enables PIT microscopy of subsurface defects in high strength corrosion resistant stainless steel 316 alloy. PIT images were processed with UL spatial–temporal separation-based clustering segmentation (STSCS) algorithm, refined by morphology image processing methods to enhance visibility of defects. The STSCS algorithm starts with wavelet decomposition to spatially de-noise thermograms, followed by UL principal component analysis (PCA), fine-tuning optimization, and neural learning-based independent component analysis (ICA) algorithms to temporally compress de-noised thermograms. The compressed thermograms were further processed with UL-based graph thresholding K-means clustering algorithm for defects segmentation. The STSCS algorithm also includes online learning feature for efficient re-training of the model with new data. For this study, metallic specimens with calibrated microscopic flat bottom hole defects, with diameters in the range from 203 to 76 µm, were produced using electro discharge machining (EDM) drilling. While the raw thermograms do not show any material defects, using STSCS algorithm to process PIT images reveals defects as small as 101 µm in diameter. To the best of our knowledge, this is the smallest reported size of a sub-surface defect in a metal imaged with PIT, which demonstrates the PIT capability of detecting defects in the size range relevant to quality control requirements of LPBF-printed high-strength metals.

36 MATERIALS SCIENCE↗

A Probabilistic Reasoner Based on Bayes Risk for Damage Detection in Structural Systems

Structural health monitoring (SHM) systems are used to inform operation of structural systems subject to loads and environments that may affect their integrity. SHM systems rely on continuous monitoring of the structure to determine its health state. These systems are often coupled with a model of the deployed structure to determine the consequences of changes in the system by forecasting the response to future states. These models, which may be thought of as digital twins, need to be updated to reflect the latest state of the structural system. This work makes use of an uncertainty-aware machine learning model that enforces distance preservation of the original input space to determine deviations from the training data input space distributions. This workflow enables domain shift detection to determine whether damage is present in the structure. The uncertainty metrics generated by this network are then used in a Bayes risk framework to design an optimal damage detector given cost and risk considerations. The approach is demonstrated on a computational example with simulated damage.

Najera-Flores, David [ATA Engineering, Inc.]↗

Abstract for CRADA between NETL and Louisiana State University

The National Energy Technology Laboratory (NETL) and Louisiana State University (PARTICIPANT) will collaborate in developing and testing distributed optical fiber CO 2 sensors for CO 2 storage site monitoring. Safe and reliable monitoring of CO 2 storage sites and transport infrastructure is critical for ensuring the long-term effectiveness and integrity of carbon capture and storage (CCS). In this project, the objective is to develop a novel optical fiber-based distributed CO 2 sensor for long-term monitoring of CCS sites and CO 2 pipelines and also implement machine learning to automate leak and structural anomaly detection. This effort aligns with NETL’s mission in reliable and sustainable energy and decarbonization.

47 OTHER INSTRUMENTATION↗

Score-based deterministic density sampling

We propose a deterministic sampling framework using Score-Based Transport Modeling for sampling an unnormalized target density π given only its score ∇ log π. Our method approximates the Wasserstein gradient flow on KL($f_t$∥π) by learning the time-varying score ∇ log $f_t$ on the fly using score matching. While having the same marginal distribution as Langevin dynamics, our method produces smooth deterministic trajectories, resulting in monotone noise-free convergence. We prove that our method dissipates relative entropy at the same rate as the exact gradient flow, provided sufficient training. Numerical experiments validate our theoretical findings: our method converges at the optimal rate, has smooth trajectories, and is often more sample efficient than its stochastic counterpart. Experiments on high-dimensional image data show that our method produces high-quality generations in as few as 15 steps and exhibits natural exploratory behavior. The memory and runtime scale linearly in the sample size.

97 MATHEMATICS AND COMPUTING↗

Building intelligent systems - Artificial intelligence research at NASA Ames Research Center

The basic components that make up the goal of building autonomous intelligent systems are discussed, and ongoing work at the NASA Ames Research Center is described. It is noted that a clear progression of systems can be seen through research settings (both within and external to NASA) to Space Station testbeds to systems which actually fly on the Space Station. The starting point for the discussion is a 'truly' autonomous Space Station intelligent system, responsible for a major portion of Space Station control. Attention is given to research in fiscal 1987, including reasoning under uncertainty, machine learning, causal modeling and simulation, knowledge from design through operations, advanced planning work, validation methodologies, and hierarchical control of and distributed cooperation among multiple knowledge-based systems.

Friedland, Peter↗

Building intelligent systems: Artificial intelligence research at NASA Ames Research Center

The basic components that make up the goal of building autonomous intelligent systems are discussed, and ongoing work at the NASA Ames Research Center is described. It is noted that a clear progression of systems can be seen through research settings (both within and external to NASA) to Space Station testbeds to systems which actually fly on the Space Station. The starting point for the discussion is a truly autonomous Space Station intelligent system, responsible for a major portion of Space Station control. Attention is given to research in fiscal 1987, including reasoning under uncertainty, machine learning, causal modeling and simulation, knowledge from design through operations, advanced planning work, validation methodologies, and hierarchical control of and distributed cooperation among multiple knowledge-based systems.

Friedland, P.↗

The fate of organic carbon in marine sediments - New insights from recent data and analysis

Organic carbon in marine sediments is a critical component of the global carbon cycle, and its degradation influences a wide range of phenomena, including the magnitude of carbon sequestration over geologic timescales, the recycling of inorganic carbon and nutrients, the dissolution and precipitation of carbonates, the production of methane and the nature of the seafloor biosphere. Although much has been learned about the factors that promote and hinder rates of organic carbon degradation in natural systems, the controls on the distribution of organic carbon in modern and ancient sediments are still not fully understood. In this review, we summarize how recent findings are changing entrenched perspectives on organic matter degradation in marine sediments: a shift from a structurally-based chemical reactivity viewpoint towards an emerging acceptance of the role of the ecosystem in organic matter degradation rates. That is, organic carbon has a range of reactivities determined by not only the nature of the organic compounds, but by the biological, geochemical, and physical attributes of its environment. This shift in mindset has gradually come about due to a greater diversity of sample sites, the molecular revolution in biology, discoveries concerning the extent and limits of life, advances in quantitative modeling, investigations of ocean carbon cycling under a variety of extreme paleo-conditions (e.g. greenhouse environments, euxinic/anoxic oceans), the application of novel analytical techniques and interdisciplinary efforts. Adopting this view across scientific disciplines will enable additional progress in understanding how marine sediments influence the global carbon cycle.

Organic carbon↗

Multi-Scale Thermo-Mechanical Modeling of Porous 3D Woven TPS Materials

This work summarizes the process to compute and analyze the thermal conductivity and mechanical properties of 3D woven TPS materials such as 3MDCP (3D Mid-Density Carbon-Phenolic). The PuMA [1,2] software, developed at NASA Ames, was used to characterize different 3MDCP samples from their constituents' data, averaging their thermal conductivity and elastic properties in the three main directions to obtain their effective orthotropic thermo-mechanical properties. This was performed in multiple steps: firstly, TPS samples were digitally reconstructed using micro computed tomography (µCT) and their constituents were segmented; then, the porous matrix phase was analyzed at the micro-scale and these results were used, along with the fibers’ constituents information, to model the tows at the meso-scale; finally, results from the constituents were homogenized and used to model the thermo-mechanical behavior of the unit cell at the macro-scale. Additionally, to gain a better understanding of the tows’ morphology and distribution of the carbon-phenolic blended fibers, microscopy images of the tows’ cross-section were segmented using deep learning techniques and analyzed with PuMA. This workflow can be applied to any TPS material to computationally obtain its thermo-mechanical properties, enabling more informed TPS design and manufacturing choices.

Multi-Scale↗