Search NASA⌕ Search

SEARCH · Search NASA

Results for “Bayesian”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 649 records · Page 36

A Machine Learning based Approach of Estimating Equivalent Circuit Model Parameters at Different SoCs of Li-ion Batteries from Voltage Relaxation

Abstract: In this study, an approach of estimating the equivalent circuit model (ECM) parameters for Li-ion batteries (LIBs) is proposed based on the voltage value at different intervals while relaxing the LIB after discharge. The typical approach for estimating ECM parameters of a LIB is to conduct electrochemical impedance spectroscopy (EIS) measurements at different frequencies and fit them to a predefined circuit model, which requires additional measuring arrangements and specialized devices. The proposed methodology utilizes four different voltages at 0s, 60s, 360s, and 1800s alongside the specific state of charge (SoC) value for a specific constant discharge current value of ~1C until the relaxation stage to train and evaluate three regression-based machine learning models— Support Vector Regression (SVR), Extreme Gradient Boosting (XGBoost), and Gaussian Process Regression (GPR)—for estimating the ECM parameters of the selected model. Bayesian optimization is employed for hyperparameter tuning to achieve optimal performance for all the regressor models, among which, the GPR provided the best performance with the root-mean-squared error (RMSE) of less than 4x10-4 on average for the resistive components and less than 0.27 for capacitive components with excellent R2 scores. The simplicity of the approach enables it to eliminate the need for sophisticated measuring equipment and computation power.

Sagar, Md. Samiul [The University of Alabama (UA)]↗

Dynamical Sketching for Enhanced Communication Efficiency in Federated Learning

Federated learning (FL) has revolutionized distributed machine learning by enabling collaborative model training without sharing local data. However, communication efficiency and privacy guarantees remain significant challenges. This paper introduces a dynamic sketching mechanism in FL, optimizing the trade-off between communication efficiency and model accuracy. By dynamically selecting the sketch matrix size, our approach adapts to the evolving characteristics of the data and the model, ensuring optimal performance across diverse scenarios. We leverage Bayesian optimization to systematically tune the sketch parameters, achieving an effective balance between resource efficiency and model performance. Experimental results on the MNIST dataset using a convolutional neural network (CNN) architecture validate the proposed method's efficiency and scalability. Our dynamic sketching approach significantly outperforms fixed-size sketching techniques, achieving higher compression ratios (up to 62x) and providing better privacy guarantees while maintaining high model accuracy. These findings highlight the robustness and versatility of our approach and make it a valuable solution for privacy-preserving, communication-efficient federated learning.

Afrose, Sharmin [ORNL]↗

Automatic Extraction of Network Configurations for Realistic Simulation and Validation

Popular HPC network interconnection simulators such as SST Macro provide a variety of configurable parameters to explore the design space of hardware components such as network links and switches. While such knobs provide flexibility to explore design trade-offs for novel hardware, manually configuring simulations for existing hardware to focus on topology exploration can be cumbersome and error-prone, leading to widely inaccurate simulations. This challenge is compounded when specifications of various (proprietary) technologies are not readily available or are intentionally omitted. In this work, we provide a methodology to automatically tune the simulation configuration of the multiple network models running within SST Macro using Bayesian optimization. We perform this optimization in the context of multiple messaging regimes (i.e., small to large and latency to bandwidth-bound messages) and provide a detailed analysis of the simulation error for four systems. With our automated framework, we achieve a 5x improvement in accuracy over best-effort configurations based on available hardware specifications.

Suetterlein, Joshua D.↗

Resilience Assessment for Distribution Systems during Hurricanes: A Learning-Based Framework

This paper presents a proactive strategy for hurricane-resilient distribution systems. It proposes a Bayesian Neural Network-based outage prediction model considering various parameters, including electrical components, and weather and environmental factors. Addressing challenges in imbalanced outage datasets, a Bias-Variance Tradeoff method is proposed. A resilience assessment model quantifies resilience indices, providing insights into system weaknesses. The approach identifies weak points and serves as a planning benchmark. Numerical results on the modified IEEE 123-node test system demonstrate effectiveness in realistic hurricane scenarios.

Vahedi, Soroush↗

Q-Cluster: Quantum Error Mitigation Through Noise-Aware Unsupervised Learning

Quantum error mitigation (QEM) is critical in reducing the impact of noise in the pre-fault-tolerant era, and is expected to complement error correction in fault-tolerant quantum computing (FTQC). In this work, we propose a novel QEM approach, Q-Cluster, that uses unsupervised learning (clustering) to reshape the measured bit-string distribution. Our approach starts with a simplified bit-flip noise model. It first performs clustering on noisy measurement results, i.e., bit-strings, based on the Hamming distance. The centroid of each cluster is calculated using a qubit-wise majority vote. Next, the noisy distribution is adjusted with the clustering outcomes and the bitflip error rates using Bayesian inference. Our simulation results show that Q-Cluster can mitigate high noise rates (up to 40% per qubit) with the simple bit-flip noise model. However, real quantum computers do not fit such a simple noise model. To address the problem, we (a) apply Pauli twirling to tailor the complex noise channels to Pauli errors, and (b) employ a machine learning model, ExtraTrees regressor, to estimate an effective bit-flip error rate using a feature vector consisting of machine calibration data (gate & measurement error rates), circuit features (number of qubits, numbers of different types of gates, etc.) and the shape of the noisy distribution (entropy). Our experimental results show that our proposed Q-Cluster scheme improves the fidelity by a factor of 1.46x, on average, compared to the unmitigated output distribution, for a set of low-entropy benchmarks on five different IBM quantum machines. Our approach outperforms the state-of-art QEM approaches RZNE [28], M3 [24], Hammer [35], and QBEEP [33] by 1.26x,1.29x,1.47x, and 2.65 x, respectively.

42 ENGINEERING↗

Integrated Framework of Multisource Data Fusion for Outage Location in Looped Distribution Systems

Accurate outage location is essential for expediting post-outage power restoration, minimizing outage duration, and enhancing the resilience of distribution networks. With the advent of advanced metering infrastructure, data-driven outage location methods have significantly advanced beyond traditional approaches that rely on manual inspections. However, existing methods still face critical challenges, like reliance on single-source data, limited ability to handle partially observable systems or difficulties with loop networks. To the best of our knowledge, no single approach has comprehensively addressed all of these challenges at once. To this end, this paper proposes a comprehensive multisource data fusion framework for outage locations via probabilistic graph networks. The framework consists of three key phases. First, a novel method for reconstituting distribution networks with loops is developed, transforming looped networks into multiple radial subnetworks that retain all outage causalities of the original network. Second, Bayesian network (BN) models are established for each subnetwork, integrating multiple data sources and network structures. Finally, a joint Gibbs sampling mechanism, featuring forward and backward information flow, is designed to merge data from separate BN models and maximize the utilization of limited evidence, ensuring accurate outage location identification. In conclusion, the framework was validated on two modified public test systems, and comparative studies confirmed its effectiveness.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Body size and early marine conditions drive changes in Chinook salmon productivity across northern latitude ecosystems

Disentangling the influences of climate change from other stressors affecting the population dynamics of aquatic species is particularly pressing for northern latitude ecosystems, where climate‐driven warming is occurring faster than the global average. Chinook salmon (Oncorhynchus tshawytscha) in the Yukon‐Kuskokwim (YK) region occupy the northern extent of their species' range and are experiencing prolonged declines in abundance resulting in fisheries closures and impacts to the well‐being of Indigenous people and local communities. These declines have been associated with physical (e.g., temperature, streamflow) and biological (e.g., body size, competition) conditions, but uncertainty remains about the relative influence of these drivers on productivity across populations and how salmon–environment relationships vary across watersheds. To fill these knowledge gaps, we estimated the effects of marine and freshwater environmental indicators, body size, and indices of competition, on the productivity (adult returns‐per‐spawner) of 26 Chinook salmon populations in the YK region using a Bayesian hierarchical stock‐recruitment model. Across most populations, productivity declined with smaller spawner body size and sea surface temperatures that were colder in the winter and warmer in the summer during the first year at sea. Decreased productivity was also associated with above average fall maximum daily streamflow, increased sea ice cover prior to juvenile outmigration, and abundance of marine competitors, but the strength of these effects varied among populations. Maximum daily stream temperature during spawning migration had a nonlinear relationship with productivity, with reduced productivity in years when temperatures exceeded thresholds in main stem rivers. These results demonstrate for the first time that well‐documented declines in body size of YK Chinook salmon were associated with declining population productivity, while taking climate into account.

54 ENVIRONMENTAL SCIENCES↗

Boron Coordination in Multicomponent Glasses: Analytical Models and Machine Learning With Uncertainty

Borosilicate glasses are extensively used in a variety of applications from kitchenware to nuclear waste immobilization due to the strong network formed by the Si-O-B bond that makes it resistant to chemical corrosion and gives it a low thermal expansion. Boron, however, exists in both trigonal BO3 and tetrahedral BO4 bonds in glass systems, which impacts the chemical durability and thermal resistance of the glass, amongst other properties. Boron coordination (N4), or the ratio of the amount of BO4 to BO3 within a glass, may aid in predicting these properties but is difficult to derive without experimental data due to the complexity of impacts from varied glass compositions and processing factors. For this reason, compositional models have been developed to predict boron coordination, but the models typically include a limited number of glass components. To help fill this gap in the models, in this work, a diverse multicomponent glass dataset of 809 glasses is compiled from a literature search, and then a number of analytical and machine learning (ML) models are trained on the dataset. Previously developed modified Bernstein and modified Du Stebbins analytical models were fitted to update parameters with the new dataset. Then, partially Bayesian neural networks, Gaussian process regressor, and heteroskedastic deterministic neural networks were evaluated. The ML models examined all have different strategies to overcome the potential for overfitting as a result of a limited training dataset, and return results that account for model uncertainty, which can be valuable for understanding model reliability. For the first time, cooling rate is introduced as an input parameter for ML models, showing consistent improvements in performance and solidifying the importance of including parameters outside of composition alone for N4 prediction. The machine learning models examined here show promise in accurate predictions of boron coordination in borosilicate glasses, all achieving R2 values of 0.91.

boron coordination↗

Predicting fusion ignition at the National Ignition Facility with physics-informed deep learning

Here, an inertial confinement fusion experiment, carried out at the National Ignition Facility, has achieved ignition by generating fusion energy exceeding the laser energy that drove the experiment. Prior to the experiment, a generative machine learning model that combines radiation hydrodynamics simulations, deep learning, experimental data, and Bayesian statistics was used to predict, with a probability greater than 70%, that ignition was the most likely outcome for this shot.

Spears, Brian K. [Lawrence Livermore National Labo↗

The decay of HIV under anti-retroviral therapy is biphasic even in humanized mice with just T cells

HIV-1 plasma viral load decays in a biphasic manner during antiretroviral therapy (ART). It was hypothesized that this is due to infection of different cell types, namely CD4+ T cells and macrophages. We studied this possibility directly by modeling the decay of HIV-1 in humanized mice. We utilized previously published data from humanized T-cell only mice (TOM) and myeloid-only mice (MOM) infected with HIV-1 and treated with a potent ART regimen. Viral load decay dynamics were modeled using either a single or a biexponential decay fitted using nonlinear mixed effects techniques. Fits were compared using the corrected Bayesian information criterion (BICc). In TOM, the biphasic model was significantly better than a single-phase decay model (ΔBICc ≈ 16) despite additional parameters. In MOM, the biphasic decay was statistically better, but there was substantial uncertainty because the virus goes below detection very fast. The first-phase half-life was consistent between groups (1.2 days in MOM and 1.3 days in TOM) and similar to the half-life estimated in human infection. The second-phase decay in these mice was minimal likely due to low initial viral loads. Additional analyses with mice containing both CD4+ T cells and macrophages or X4-tropic virus-infected MOM mice confirmed the biphasic pattern, demonstrating the robustness of this result. The biphasic decline in HIV-1 occurs, even with only CD4+ T cells, refuting the hypothesis that distinct cell populations (CD4+ T cells and macrophages) drive each decay phase. These findings support an alternative model in which the observed dynamics arise from intrinsic properties of the viral infection lifecycle rather than from cellular compartmentalization.

59 BASIC BIOLOGICAL SCIENCES↗

Multivariate Testing of Sampling Techniques to Address Class Imbalance in Building Use Type Classification

This study addresses the challenges inherent in building use type classification, particularly focusing on the issue of class imbalance in the training datasets for machine learning classifiers. We comprehensively analyze the efficacy of various class-balancing sampling techniques. Employing Monte Carlo simulations and Bayesian optimization, we evaluated the performance of multiple sampling methods, including Random Oversampling, Random Undersampling, SMOTE, Borderline-SMOTE, and ADASYN, across a dataset encompassing nine southeastern coastal states of the United States. Our findings reveal that simple random over- and undersampling techniques outperform more sophisticated methods. Additionally, we show inherent value in creating an imbalance in training data to effectively train a machine learning classifier for distinguishing between residential and nonresidential buildings. This study provides valuable guidance for future research on building use type classification research and lays essential groundwork for developing attribute-rich building stock datasets.

Adams, Daniel↗

ORBIT-2: Scaling Exascale Vision Foundation Models for Weather and Climate Downscaling

Sparse observations and coarse-resolution climate models limit effective regional decision-making, underscoring the need for robust downscaling. However, existing AI methods struggle with generalization across variables and geographies and are constrained by the quadratic complexity of Vision Transformer (ViT) self-attention. We introduce ORBIT-2, a scalable foundation model for global, hyper-resolution climate downscaling. ORBIT-2 incorporates two key innovations: (1) Residual Slim ViT (Reslim), a lightweight architecture with residual learning and Bayesian regularization for efficient, robust prediction; and (2) TILES, a tile-wise sequence scaling algorithm that reduces self-attention complexity from quadratic to linear, enabling long-sequence processing and massive parallelism. ORBIT-2 scales to 10 billion parameters across 65,536 GPUs, achieving up to 4.1 ExaFLOPS sustained throughput and 74–98% strong scaling efficiency. It supports downscaling to 0.9 km global resolution and processes sequences up to 4.2 billion tokens. On 7 km resolution benchmarks, ORBIT-2 achieves high accuracy with R2 scores in range of 0.98–0.99 against observation data.

Wang, Xiao [ORNL] (ORCID:0000000165451943)↗

ezECM

Allows the user to fit a classical event categorization matrix (ECM) model, and a novel Bayesian event categorization matrix model, both of which are used for nuclear detonation detection.

Koermer, Scott↗

popclass

popclass is a lightweight python package that allows fast, probabilistic classification of the lens of a microlensing event given the event's posterior distribution and a model of the Galaxy. popclass provides the bridge between Galactic simulation and lens classification, an interface to common Bayesian inference libraries, and the ability for users to flexibly specify their own Galactic model and classification parameters.

Mcgill, Peter↗

khaos

An R implementation of a modified version of the sparse Bayesian polynomial chaos expansion algorithm of Shao et al., (2017).

Rumsey, Kellin↗

MIST_paper

Code to reproduce results from and implement functionality described in "A Bayesian error model for synthesis and sequencing of oligonucleotides", Marrs, FW, Gratz, D, and Erkkila, TH.

Marrs, Frank↗

pnnl/MCRASTA

McRasta (Markov Chain Rate and State Analysis) was developed to estimate parameter uncertainty in constitutive friction models via Bayesian inverse and Markov Chain Monte Carlo (MCMC) methods.

Fichera, Marissa [Pacific Northwest National Labor↗

Code Release for “Unlocking Extreme Space Weather through Advanced Modeling of Legacy Vela Spacecraft” ER

The software being developed for this project has two mains aims. First, a Bayesian Model Calibration (BMC) procedure is being developed to calibrate a spallation model that simulates protons hitting a spacecraft orbiting earth to real data. The procedure will be developed for general data (there is no data release requested as part of this code release). Second, an inverse physics modeling task is being undertaken to map the number of resulting neutrons observed from this process to the expected number of protons that hit the model. This second task is of statistical interest; to publish on it, the code will need to be open source.

Murph, Alexander↗