Search NASA⌕ Search

SEARCH · Search NASA

Results for “Bayesian Statistics”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

Bayesian learning

In 1983 and 1984, the Infrared Astronomical Satellite (IRAS) detected 5,425 stellar objects and measured their infrared spectra. In 1987 a program called AUTOCLASS used Bayesian inference methods to discover the classes present in these data and determine the most probable class of each object, revealing unknown phenomena in astronomy. AUTOCLASS has rekindled the old debate on the suitability of Bayesian methods, which are computationally intensive, interpret probabilities as plausibility measures rather than frequencies, and appear to depend on a subjective assessment of the probability of a hypothesis before the data were collected. Modern statistical methods have, however, recently been shown to also depend on subjective elements. These debates bring into question the whole tradition of scientific objectivity and offer scientists a new way to take responsibility for their findings and conclusions.

Denning, Peter J.↗

Bayesian Analysis of the Cosmic Microwave Background

There is a wealth of cosmological information encoded in the spatial power spectrum of temperature anisotropies of the cosmic microwave background! Experiments designed to map the microwave sky are returning a flood of data (time streams of instrument response as a beam is swept over the sky) at several different frequencies (from 30 to 900 GHz), all with different resolutions and noise properties. The resulting analysis challenge is to estimate, and quantify our uncertainty in, the spatial power spectrum of the cosmic microwave background given the complexities of "missing data", foreground emission, and complicated instrumental noise. Bayesian formulation of this problem allows consistent treatment of many complexities including complicated instrumental noise and foregrounds, and can be numerically implemented with Gibbs sampling. Gibbs sampling has now been validated as an efficient, statistically exact, and practically useful method for low-resolution (as demonstrated on WMAP 1 and 3 year temperature and polarization data). Continuing development for Planck - the goal is to exploit the unique capabilities of Gibbs sampling to directly propagate uncertainties in both foreground and instrument models to total uncertainty in cosmological parameters.

methods - statistical↗

Automated Monitoring with a BSP Fault-Detection Test

The figure schematically illustrates a method and procedure for automated monitoring of an asset, as well as a hardware- and-software system that implements the method and procedure. As used here, asset could signify an industrial process, power plant, medical instrument, aircraft, or any of a variety of other systems that generate electronic signals (e.g., sensor outputs). In automated monitoring, the signals are digitized and then processed in order to detect faults and otherwise monitor operational status and integrity of the monitored asset. The major distinguishing feature of the present method is that the fault-detection function is implemented by use of a Bayesian sequential probability (BSP) technique. This technique is superior to other techniques for automated monitoring because it affords sensitivity, not only to disturbances in the mean values, but also to very subtle changes in the statistical characteristics (variance, skewness, and bias) of the monitored signals.

Bickford, Randall L.↗

Bayesian Optimization of The Relativistic Heavy Ion Collider Luminosity via s * Control

A state-of-the-art jet detector named sPHENIX was proposed, commissioned, and operated at the Rel ativistic Heavy Ion Collider (RHIC) from 2023 to 2025. This detector featured precision tracking and calorime try that enable high-statistics studies of the Quark Gluon Plasma through jet modification, upsilon suppres sion, and open heavy flavor production. The innermost component of the three sPHENIX tracking systems is the Monolithic-Active-Pixel-Sensor-based Vertex Detec tor (MVTX) (Fig. 1), which has an acceptance within | s | < 0.1m of the interaction point (IP).

43 PARTICLE ACCELERATORS↗

Automatic Generation of Algorithms for the Statistical Analysis of Planetary Nebulae Images

Analyzing data sets collected in experiments or by observations is a Core scientific activity. Typically, experimentd and observational data are &aught with uncertainty, and the analysis is based on a statistical model of the conjectured underlying processes, The large data volumes collected by modern instruments make computer support indispensible for this. Consequently, scientists spend significant amounts of their time with the development and refinement of the data analysis programs. AutoBayes [GF+02, FS03] is a fully automatic synthesis system for generating statistical data analysis programs. Externally, it looks like a compiler: it takes an abstract problem specification and translates it into executable code. Its input is a concise description of a data analysis problem in the form of a statistical model as shown in Figure 1; its output is optimized and fully documented C/C++ code which can be linked dynamically into the Matlab and Octave environments. Internally, however, it is quite different: AutoBayes derives a customized algorithm implementing the given model using a schema-based process, and then further refines and optimizes the algorithm into code. A schema is a parameterized code template with associated semantic constraints which define and restrict the template s applicability. The schema parameters are instantiated in a problem-specific way during synthesis as AutoBayes checks the constraints against the original model or, recursively, against emerging sub-problems. AutoBayes schema library contains problem decomposition operators (which are justified by theorems in a formal logic in the domain of Bayesian networks) as well as machine learning algorithms (e.g., EM, k-Means) and nu- meric optimization methods (e.g., Nelder-Mead simplex, conjugate gradient). AutoBayes augments this schema-based approach by symbolic computation to derive closed-form solutions whenever possible. This is a major advantage over other statistical data analysis systems which use numerical approximations even in cases where closed-form solutions exist. AutoBayes is implemented in Prolog and comprises approximately 75.000 lines of code. In this paper, we take one typical scientific data analysis problem-analyzing planetary nebulae images taken by the Hubble Space Telescope-and show how AutoBayes can be used to automate the implementation of the necessary anal- ysis programs. We initially follow the analysis described by Knuth and Hajian [KHO2] and use AutoBayes to derive code for the published models. We show the details of the code derivation process, including the symbolic computations and automatic integration of library procedures, and compare the results of the automatically generated and manually implemented code. We then go beyond the original analysis and use AutoBayes to derive code for a simple image segmentation procedure based on a mixture model which can be used to automate a manual preproceesing step. Finally, we combine the original approach with the simple segmentation which yields a more detailed analysis. This also demonstrates that AutoBayes makes it easy to combine different aspects of data analysis.

Fischer, Bernd↗

sparse_bias

This is a python package used to fit an unknown function from data that potential contains systematic biases related to metadata. The model fits the unknown function and uses a Bayesian horseshoe prior model to impose sparsity on the bias terms. This code has been generalized from research code developed for AIACHNE into a package that should have more general application in a wider class of statistical models.

Walton, Noah↗

Neural network classification - A Bayesian interpretation

The relationship between minimizing a mean squared error and finding the optimal Bayesian classifier is reviewed. This provides a theoretical interpretation for the process by which neural networks are used in classification. A number of confidence measures are proposed to evaluate the performance of the neural network classifier within a statistical framework.

Wan, Eric A.↗

Solar Activity Relations in Energetic Electron Events Measured By the MESSENGER Mission

Aims. We perform a statistical study of the relations between the properties of solar energetic electron (SEE) events measured by the MESSENGER mission from 2010 to 2015 and the parameters of the respective parent solar activity phenomena in order to identify the potential correlations between them. During the time of analysis, the MESSENGER heliocentric distance varied between 0.31 and 0.47 au. Methods. We used a published list of 61 SEE events measured by MESSENGER, which includes information on the near-relativistic electron peak intensities, the peak-intensity energy spectral indices, and the measured X-ray peak intensity of the flares related to the SEE events. Taking advantage of multi-viewpoint remote-sensing observations, we reconstructed, whenever possible, the associated coronal mass ejections (CMEs) and shock waves; and we determined the three-dimensional (3D) properties (location, speed, and width) of the CMEs and the maximum speed of the 3D CME-driven shocks in the corona. We used different methods (Spearman, Pearson, and a Bayesian approach, namely the Kelly method to linear regression) to estimate the correlation coefficients between the flare intensity, maximum speed at the apex of the CME-driven shock, CME speed at the apex, and CME width with the electron peak intensities and with the energy spectral indices. In this statistical study, we considered and addressed the limitations of the particle instrument on board MESSENGER (elevated background intensity level, anti-Sun pointing). Results. There is an asymmetry to the east in the range of connection angles (CAs) for which the SEE events present the highest peak intensities, where the CA is the longitudinal separation between the footpoint of the magnetic field connecting to the spacecraft and the flare location. Based on this asymmetry, we define a subsample of well-connected events as when −65° ≤ CA ≤ +33°. For the well-connected sample, we find moderate to strong correlations between the near-relativistic electron peak intensity and the 3D CME-driven shock maximum speed at the apex (Spearman: cc = 0.53 ± 0.05; Pearson: cc = 0.65 ± 0.04; Kelly: cc = 0.87 ± 0.20), the flare peak intensity (Spearman: cc = 0.63 ± 0.03; Pearson: cc = 0.59 ± 0.03; Kelly: cc = 0.74 ± 0.30), and the 3D CME speed at the apex (Spearman: cc = 0.50 ± 0.04; Pearson: cc = 0.46 ± 0.03; Kelly: cc = 0.60 ± 0.39). When including poorly connected events (full sample), the relations between the peak intensities and the solar-activity phenomena are blurred, showing lower correlation coefficients. Conclusions. Based on the comparison of the correlation coefficients presented in this study using near 0.4 au data, (1) both flare and shock-related processes may contribute to the acceleration of near relativistic electrons in large SEE events, in agreement with previous studies based on near 1 au data; and (2) the maximum speed of the CME-driven shock is a better parameter to investigate particle-acceleration-related mechanisms than the average CME speed, as suggested by the stronger correlation with the SEE peak intensities.

Sun↗

Hybrid Gibbs Sampling and MCMC for CMB Analysis at Small Angular Scales

A) Gibbs Sampling has now been validated as an efficient, statistically exact, and practically useful method for "low-L" (as demonstrated on WMAP temperature polarization data). B) We are extending Gibbs sampling to directly propagate uncertainties in both foreground and instrument models to total uncertainty in cosmological parameters for the entire range of angular scales relevant for Planck. C) Made possible by inclusion of foreground model parameters in Gibbs sampling and hybrid MCMC and Gibbs sampling for the low signal to noise (high-L) regime. D) Future items to be included in the Bayesian framework include: 1) Integration with Hybrid Likelihood (or posterior) code for cosmological parameters; 2) Include other uncertainties in instrumental systematics? (I.e. beam uncertainties, noise estimation, calibration errors, other).

Gibbs sampling↗

Identifying neutron sources using recoil and time-of-flight spectroscopy

Identification of neutron sources is central to nuclear physics and its applications, from planetary science to nuclear security, yet direct source discrimination from measured neutron spectra remains fundamentally elusive. Here, we introduce a Bayesian protocol that directly infers source ensembles from measured neutron spectra by combining full-spectrum template matching with probabilistic evidence evaluation. Applying this protocol to recoil and time-of-flight spectroscopy, we recover single- and two-source configurations with strong statistical significance (beyond 4⁢𝜎) at event counts as low as ∼10 3 . These results demonstrate that neutron spectral signatures can be leveraged for robust source identification, opening a new observational window for both fundamental research and operationally driven applications.

neutron physics↗

Confidence set inference with a prior quadratic bound

In the uniqueness part of a geophysical inverse problem, the observer wants to predict all likely values of P unknown numerical properties z = (z sub 1,...,z sub p) of the earth from measurement of D other numerical properties y(0)=(y sub 1(0),...,y sub D(0)) knowledge of the statistical distribution of the random errors in y(0). The data space Y containing y(0) is D-dimensional, so when the model space X is infinite-dimensional the linear uniqueness problem usually is insoluble without prior information about the correct earth model x. If that information is a quadratic bound on x (e.g., energy or dissipation rate), Bayesian inference (BI) and stochastic inversion (SI) inject spurious structure into x, implied by neither the data nor the quadratic bound. Confidence set inference (CSI) provides an alternative inversion technique free of this objection. CSI is illustrated in the problem of estimating the geomagnetic field B at the core-mantle boundary (CMB) from components of B measured on or above the earth's surface. Neither the heat flow nor the energy bound is strong enough to permit estimation of B(r) at single points on the CMB, but the heat flow bound permits estimation of uniform averages of B(r) over discs on the CMB, and both bounds permit weighted disc-averages with continous weighting kernels. Both bounds also permit estimation of low-degree Gauss coefficients at the CMB. The heat flow bound resolves them up to degree 8 if the crustal field at satellite altitudes must be treated as a systematic error, but can resolve to degree 11 under the most favorable statistical treatment of the crust. These two limits produce circles of confusion on the CMB with diameters of 25 deg and 19 deg respectively.

Backus, George E.↗

Prime VI

SAND2025-03757O Prime VI is a distribution-of-disease outbreak model calibration code based on variational inference. It accompanies a publication for submission to Statistics in Medicine journal, and the code will be maintained for open-source use on Sandia's GitLab. The software provides methods for calibrating an epidemiological model to measured case-count data for a multitude of correlated spatial regions. The code solves a Bayesian inverse problem for model calibration where the posterior over-model parameters are approximated through a custom implementation of variational inference. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Safta, Cosmin↗

Data Assimilation for Robust UQ Within Agent-Based Simulation on HPC Systems

Agent-based simulation provides a powerful tool for in silico system modeling. However, these simulations do not provide built-in methods for uncertainty quantification (UQ). Within these types of models a typical approach to UQ is to run multiple realizations of the model then compute aggregate statistics. This approach is limited due to the compute time required for a solution. When faced with an emerging biothreat, public health decisions need to be made quickly and solutions for integrating near real-time data with analytic tools are needed. We propose an integrated Bayesian UQ framework for agent-based models based on sequential Monte Carlo sampling. Given streaming or static data about the evolution of an emerging pathogen this Bayesian framework provides a distribution over the parameters governing the spread of a disease through a population. These estimates of the spread of a disease may be provided to public health agencies seeking to abate the spread. By coupling agent-based simulations with Bayesian modeling in a data assimilation, our proposed framework provides a powerful tool for modeling dynamical systems in silico. We propose a method which reduces model error and provides a range of realistic possible outcomes. Moreover, our method addresses two primary limitations of ABMs: the lack of UQ and an inability to assimilate data. Our proposed framework combines the flexibility of an agent-based model with UQ provided by the Bayesian paradigm in a workflow which scales well to HPC systems. We provide algorithmic details and results on a simulated outbreak with both static and streaming data.

Spannaus, Adam [ORNL] (ORCID:0000000225213657)↗

Continuous Habitable Zones: Pairing a GCM and Bayesian Framework to Predict Habitable Zone Evolution

In the near-future, new space telescopes like JWST, LUVOIR, and HabEx will begin attempting to explore the properties of atmospheres of potentially habitable planets. This will require a significant amount of time and resources for even a single planet, which makes it essential to prioritize observations by those most-likely to have detectable life. Here we present a statistical method to estimate the probabilities that specific exoplanets have been continuously in the habitable zone of their host stars for more than 2 billion years, the approximate time it took life on Earth to significantly increase the oxygen content of the atmosphere. We introduce the use of statistics of an ensemble of 3D planetary general circulation models to estimate these probabilities, replacing prior 1D model estimates.

habitable planets↗

Pairing a GCM and Bayesian Framework to Predict Habitable Zone Evolution

In the near-future, new space telescopes like JWST, LUVOIR, and HabEx will begin attempting to explore the properties of atmospheres of potentially habitable planets. This will require a significant amount of time and resources for even a single planet, which makes it essential to prioritize observations by those most-likely to have detectable life. Here we present a statistical method to estimate the probabilities that specific exoplanets have been continuously in the habitable zone of their host stars for more than 2 billion years, the approximate time it took life on Earth to significantly increase the oxygen content of the atmosphere. We introduce the use of statistics of an ensemble of 3D planetary general circulation models to estimate these probabilities, replacing prior 1D model estimates.

habitable planets↗

Accessing the gluon momentum fraction of nucleons through the gradient flow

We calculate the gluon momentum fraction of the nucleon using lattice QCD, with a nonperturbative renormalization technique based on the gradient flow. The gluon momentum fraction is determined on a single Wilson-clover ensemble using 𝑁 𝑓 =2 +1 flavors with pion mass 358 MeV and lattice spacing 0.094 fm. We employ the variational method to reduce excited-state contamination and apply the distillation framework to ensure a large operator basis. To reduce systematic uncertainties, we apply Bayesian model averaging to all fit procedures. We apply matching coefficients to the flow-time dependent lattice results to recover the gluon momentum fraction in the $\overline{MS}$-scheme at 2 GeV. Our final result is ⟨𝑥⟩ 𝑔 ⁢(𝜇 =2 GeV) =0.482⁢(35), where we quote only statistical uncertainties.

Lattice QCD↗

Structural model optimization using statistical evaluation

The results of research in applying statistical methods to the problem of structural dynamic system identification are presented. The study is in three parts: a review of previous approaches by other researchers, a development of various linear estimators which might find application, and the design and development of a computer program which uses a Bayesian estimator. The method is tried on two models and is successful where the predicted stiffness matrix is a proper model, e.g., a bending beam is represented by a bending model. Difficulties are encountered when the model concept varies. There is also evidence that nonlinearity must be handled properly to speed the convergence.

Collins, J. D.↗

A dynamic likelihood approach to filtering transport processes: advection-diffusion dynamics

A Bayesian data assimilation scheme is formulated for advection-dominated advective and diffusive evolutionary problems, based upon the Dynamic Likelihood (DLF) approach to filtering. The DLF was developed specifically for hyperbolic problems –waves–, and in this paper, it is extended via a split step formulation, to handle advection-diffusion problems. In the dynamic likelihood approach, observations and their statistics are used to propagate probabilities along characteristics, evolving the likelihood in time. The estimate posterior thus inherits phase information. For advection-diffusion the advective part of the time evolution is handled on the basis of observations alone, while the diffusive part is informed through the model as well as observations. We expect, and indeed show here, that in advection-dominated problems, the DLF approach produces better estimates than other assimilation approaches, particularly when the observations are sparse and have low uncertainty. The added computational expense of the method is cubic in the total number of observations over time, which is on the same order of magnitude as a standard Kalman filter and can be mitigated by bounding the number of forward propagated observations, discarding the least informative data.

97 MATHEMATICS AND COMPUTING↗