Search NASA⌕ Search

SEARCH · Search NASA

Results for “Bayesian Statistics”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Bayesian estimation of crack initiation times from service data

Lockheed C-130 Hercules aircraft have during their service life been periodically inspected and growing cracks around rivet holes were recorded. This record has recently been used to determine the statistical distributions of crack initiation times and the distribution of initial crack sizes. When crack initiation times are calculated from such cracks, by backward extrapolation of the growth relation, the resulting distribution of crack initiation times will indicate a preponderance of short times to crack initiation. If however, such distributions are combined with the reliability of the inspection procedure, the statistical distribution of missed initiation times can be estimated. The method used is based on Bayes theorem which permits the calculation of the 'prior' distribution (initiation times before inspection) from a knowledge of the 'posterior' distribution (initiation times obtained from the inspection) and a 'likelihood function' (reliability of the inspection) procedure. The results indicate that during an early inspection a large percentage of initiation times will be missed and that the fraction of located initiation times increases during later inspections.

Heller, R. A.↗

Operations on Graphical Models with Plates

This paper explains how graphical models, for instance Bayesian or Markov networks, can be extended to model problems in data analysis and learning. This provides a unified framework that combines lessons learned from the artificial intelligence, statistical and connectionist communities. This also offers a set of principles for developing a software generator for data analysis, whereby a learning or discovery system can be compiled from specifications. Many of the popular learning algorithms can be compiled in this way from graphical specifications. While in a sense this paper is a multidisciplinary review of learning, the main contribution here is the presentation of the material within the unifying framework of graphical models, and the observation that, as a result, the process of developing learning algorithms can be partly automated.

Buntine, Wray L.↗

Sparse Superpixel Unmixing for Hyperspectral Image Analysis

Software was developed that automatically detects minerals that are present in each pixel of a hyperspectral image. An algorithm based on sparse spectral unmixing with Bayesian Positive Source Separation is used to produce mineral abundance maps from hyperspectral images. A superpixel segmentation strategy enables efficient unmixing in an interactive session. The algorithm computes statistically likely combinations of constituents based on a set of possible constituent minerals whose abundances are uncertain. A library of source spectra from laboratory experiments or previous remote observations is used. A superpixel segmentation strategy improves analysis time by orders of magnitude, permitting incorporation into an interactive user session (see figure). Mineralogical search strategies can be categorized as supervised or unsupervised. Supervised methods use a detection function, developed on previous data by hand or statistical techniques, to identify one or more specific target signals. Purely unsupervised results are not always physically meaningful, and may ignore subtle or localized mineralogy since they aim to minimize reconstruction error over the entire image. This algorithm offers advantages of both methods, providing meaningful physical interpretations and sensitivity to subtle or unexpected minerals.

Castano, Rebecca↗

Iterative Bayesian Classification In Polarimetric SAR

In improved scheme for Bayesian classification of picture elements in polarimetric synthetic-aperture radar image of terrain, priori probability that given picture element belongs to given class, adjusted according to spatial variation of statistical properties of image data. Accuracy increases dramatically in first few iterations. Scheme involves sequence of classifications. In first, a priori probability that element belongs to class taken to be constant over the whole image. In subsequent classifications, adaptive a priori probabilities calculated for each picture element.

Van Zyl, Jakob J.↗

Under-Constrained SEE Data: Implications for Estimating and Bounding SEE Rates

Increasingly scarce SEE testing resources and rapid growth of the New Space sector have increased the prevalence of under-constrained SEE data. We develop Monte Carlo tools to assess implications for SEE rate estimation. We also show that Bayesian Priors based on large datasets of SEL susceptible parts can augment under-constrained data and improve bounds on SEL rates. The resulting Bayesian Priors are also useful for bounding system SEL risk.

Single-event effect↗

Predicting Time Series Outputs and Time-to-Failure for an Aircraft Controller Using Bayesian Modeling

Safety of unmanned aerial systems (UAS) is paramount, but the large number of dynamically changing controller parameters makes it hard to determine if the system is currently stable, and the time before loss of control if not. We propose a hierarchical statistical model using Treed Gaussian Processes to predict (i) whether a flight will be stable (success) or become unstable (failure), (ii) the time-to-failure if unstable, and (iii) time series outputs for flight variables. We first classify the current flight input into success or failure types, and then use separate models for each class to predict the time-to-failure and time series outputs. As different inputs may cause failures at different times, we have to model variable length output curves. We use a basis representation for curves and learn the mappings from input to basis coefficients. We demonstrate the effectiveness of our prediction methods on a NASA neuro-adaptive flight control system.

Statistics↗

Interpolating Fields of Carbon Monoxide Data Using a Hybrid Statistical-Physical Model

Atmospheric Carbon Monoxide (CO) is a pollutant gas of which the US congress has mandated regular monitoring, and satellite sensors can be used to retrieve regional concentrations of CO over several vertical layers. However, CO at cloudy locations cannot be observed and have to be estimated from the observed data set, resulting in an interpolation problem. The current state-of-the-art solution is to combine prior information, computed by a deterministic physical model, with observations. However, the deterministic model may introduce uncertainties that do not derive from the data. While sharing certain features with the physical model, this paper presents a Bayesian hierarchical model for interpolating CO on a 3-dimensional spatial grid, across time. To our knowledge such a model has not been considered before. The model is applied to a hypothetical air-quality monitoring scenario, and is compared to existing interpolation methods. The results provide motivation for the use of the statistical model for regional to local applications.

Arellano, A. A.↗

R2U2: Monitoring and Diagnosis of Security Threats for Unmanned Aerial Systems

We present R2U2, a novel framework for runtime monitoring of security properties and diagnosing of security threats on-board Unmanned Aerial Systems (UAS). R2U2, implemented in FPGA hardware, is a real-time, REALIZABLE, RESPONSIVE, UNOBTRUSIVE Unit for security threat detection. R2U2 is designed to continuously monitor inputs from the GPS and the ground control station, sensor readings, actuator outputs, and flight software status. By simultaneously monitoring and performing statistical reasoning, attack patterns and post-attack discrepancies in the UAS behavior can be detected. R2U2 uses runtime observer pairs for linear and metric temporal logics for property monitoring and Bayesian networks for diagnosis of security threats. We discuss the design and implementation that now enables R2U2 to handle security threats and present simulation results of several attack scenarios on the NASA DragonEye UAS.

Formal Methods↗

New giant planet beyond the snow line for an extended MOA exoplanet microlens sample

Characterizing a planet detected by microlensing is hard if the planetary signal is weak or the lens-source relative trajectory is far from caustics. However, statistical analyses of planet demography must include those planets to accurately determine occurrence rates. As part of a systematic modelling effort in the context of a >10-yr retrospective analysis of MOA’s survey observations to build an extended MOA statistical sample, we analyse the light curve of the planetary microlensing event MOA-2014-BLG-472. This event provides weak constraints on the physical parameters of the lens, as a result of a planetary anomaly occurring at low magnification in the light curve. We use a Bayesian analysis to estimate the properties of the planet, based on a refined Galactic model and the assumption that all Milky Way’s stars have an equal planet-hosting probability. We find that a lens consisting of a 1.9(+2.2,−1.2)M(J) giant planet orbiting a 0.31(+0.36,−0.19)Mꙩ host at a projected separation of 0.75±0.24au is consistent with the observations and is most likely, based on the Galactic priors. The lens most probably lies in the Galactic bulge, at 7.2(+0.6,−1.7)kpc from Earth. The accurate measurement of the measured planet-to-host star mass ratio will be included in the next statistical analysis of cold planet demography detected by microlensing.

Clément Ranc↗

Sensor Selection and Data Validation for Reliable Integrated System Health Management

For new access to space systems with challenging mission requirements, effective implementation of integrated system health management (ISHM) must be available early in the program to support the design of systems that are safe, reliable, highly autonomous. Early ISHM availability is also needed to promote design for affordable operations; increased knowledge of functional health provided by ISHM supports construction of more efficient operations infrastructure. Lack of early ISHM inclusion in the system design process could result in retrofitting health management systems to augment and expand operational and safety requirements; thereby increasing program cost and risk due to increased instrumentation and computational complexity. Having the right sensors generating the required data to perform condition assessment, such as fault detection and isolation, with a high degree of confidence is critical to reliable operation of ISHM. Also, the data being generated by the sensors needs to be qualified to ensure that the assessments made by the ISHM is not based on faulty data. NASA Glenn Research Center has been developing technologies for sensor selection and data validation as part of the FDDR (Fault Detection, Diagnosis, and Response) element of the Upper Stage project of the Ares 1 launch vehicle development. This presentation will provide an overview of the GRC approach to sensor selection and data quality validation and will present recent results from applications that are representative of the complexity of propulsion systems for access to space vehicles. A brief overview of the sensor selection and data quality validation approaches is provided below. The NASA GRC developed Systematic Sensor Selection Strategy (S4) is a model-based procedure for systematically and quantitatively selecting an optimal sensor suite to provide overall health assessment of a host system. S4 can be logically partitioned into three major subdivisions: the knowledge base, the down-select iteration, and the final selection analysis. The knowledge base required for productive use of S4 consists of system design information and heritage experience together with a focus on components with health implications. The sensor suite down-selection is an iterative process for identifying a group of sensors that provide good fault detection and isolation for targeted fault scenarios. In the final selection analysis, a statistical evaluation algorithm provides the final robustness test for each down-selected sensor suite. NASA GRC has developed an approach to sensor data qualification that applies empirical relationships, threshold detection techniques, and Bayesian belief theory to a network of sensors related by physics (i.e., analytical redundancy) in order to identify the failure of a given sensor within the network. This data quality validation approach extends the state-of-the-art, from red-lines and reasonableness checks that flag a sensor after it fails, to include analytical redundancy-based methods that can identify a sensor in the process of failing. The focus of this effort is on understanding the proper application of analytical redundancy-based data qualification methods for onboard use in monitoring Upper Stage sensors.

Garg, Sanjay↗

Comparing hard and soft prior bounds in geophysical inverse problems

In linear inversion of a finite-dimensional data vector y to estimate a finite-dimensional prediction vector z, prior information about X sub E is essential if y is to supply useful limits for z. The one exception occurs when all the prediction functionals are linear combinations of the data functionals. Two forms of prior information are compared: a soft bound on X sub E is a probability distribution p sub x on X which describeds the observer's opinion about where X sub E is likely to be in X; a hard bound on X sub E is an inequality Q sub x(X sub E, X sub E) is equal to or less than 1, where Q sub x is a positive definite quadratic form on X. A hard bound Q sub x can be softened to many different probability distributions p sub x, but all these p sub x's carry much new information about X sub E which is absent from Q sub x, and some information which contradicts Q sub x. Both stochastic inversion (SI) and Bayesian inference (BI) estimate z from y and a soft prior bound p sub x. If that probability distribution was obtained by softening a hard prior bound Q sub x, rather than by objective statistical inference independent of y, then p sub x contains so much unsupported new information absent from Q sub x that conclusions about z obtained with SI or BI would seen to be suspect.

Backus, George E.↗

Comparing hard and soft prior bounds in geophysical inverse problems

In linear inversion of a finite-dimensional data vector y to estimate a finite-dimensional prediction vector z, prior information about X sub E is essential if y is to supply useful limits for z. The one exception occurs when all the prediction functionals are linear combinations of the data functionals. Two forms of prior information are compared: a soft bound on X sub E is a probability distribution p sub x on X which describes the observer's opinion about where X sub E is likely to be in X; a hard bound on X sub E is an inequality Q sub x(X sub E, X sub E) is equal to or less than 1, where Q sub x is a positive definite quadratic form on X. A hard bound Q sub x can be softened to many different probability distributions p sub x, but all these p sub x's carry much new information about X sub E which is absent from Q sub x, and some information which contradicts Q sub x. Both stochastic inversion (SI) and Bayesian inference (BI) estimate z from y and a soft prior bound p sub x. If that probability distribution was obtained by softening a hard prior bound Q sub x, rather than by objective statistical inference independent of y, then p sub x contains so much unsupported new information absent from Q sub x that conclusions about z obtained with SI or BI would seen to be suspect.

Backus, George E.↗

Bayesian learning

In 1983 and 1984, the Infrared Astronomical Satellite (IRAS) detected 5,425 stellar objects and measured their infrared spectra. In 1987 a program called AUTOCLASS used Bayesian inference methods to discover the classes present in these data and determine the most probable class of each object, revealing unknown phenomena in astronomy. AUTOCLASS has rekindled the old debate on the suitability of Bayesian methods, which are computationally intensive, interpret probabilities as plausibility measures rather than frequencies, and appear to depend on a subjective assessment of the probability of a hypothesis before the data were collected. Modern statistical methods have, however, recently been shown to also depend on subjective elements. These debates bring into question the whole tradition of scientific objectivity and offer scientists a new way to take responsibility for their findings and conclusions.

Denning, Peter J.↗

Bayesian Analysis of the Cosmic Microwave Background

There is a wealth of cosmological information encoded in the spatial power spectrum of temperature anisotropies of the cosmic microwave background! Experiments designed to map the microwave sky are returning a flood of data (time streams of instrument response as a beam is swept over the sky) at several different frequencies (from 30 to 900 GHz), all with different resolutions and noise properties. The resulting analysis challenge is to estimate, and quantify our uncertainty in, the spatial power spectrum of the cosmic microwave background given the complexities of "missing data", foreground emission, and complicated instrumental noise. Bayesian formulation of this problem allows consistent treatment of many complexities including complicated instrumental noise and foregrounds, and can be numerically implemented with Gibbs sampling. Gibbs sampling has now been validated as an efficient, statistically exact, and practically useful method for low-resolution (as demonstrated on WMAP 1 and 3 year temperature and polarization data). Continuing development for Planck - the goal is to exploit the unique capabilities of Gibbs sampling to directly propagate uncertainties in both foreground and instrument models to total uncertainty in cosmological parameters.

methods - statistical↗

Automated Monitoring with a BSP Fault-Detection Test

The figure schematically illustrates a method and procedure for automated monitoring of an asset, as well as a hardware- and-software system that implements the method and procedure. As used here, asset could signify an industrial process, power plant, medical instrument, aircraft, or any of a variety of other systems that generate electronic signals (e.g., sensor outputs). In automated monitoring, the signals are digitized and then processed in order to detect faults and otherwise monitor operational status and integrity of the monitored asset. The major distinguishing feature of the present method is that the fault-detection function is implemented by use of a Bayesian sequential probability (BSP) technique. This technique is superior to other techniques for automated monitoring because it affords sensitivity, not only to disturbances in the mean values, but also to very subtle changes in the statistical characteristics (variance, skewness, and bias) of the monitored signals.

Bickford, Randall L.↗

Automatic Generation of Algorithms for the Statistical Analysis of Planetary Nebulae Images

Analyzing data sets collected in experiments or by observations is a Core scientific activity. Typically, experimentd and observational data are &aught with uncertainty, and the analysis is based on a statistical model of the conjectured underlying processes, The large data volumes collected by modern instruments make computer support indispensible for this. Consequently, scientists spend significant amounts of their time with the development and refinement of the data analysis programs. AutoBayes [GF+02, FS03] is a fully automatic synthesis system for generating statistical data analysis programs. Externally, it looks like a compiler: it takes an abstract problem specification and translates it into executable code. Its input is a concise description of a data analysis problem in the form of a statistical model as shown in Figure 1; its output is optimized and fully documented C/C++ code which can be linked dynamically into the Matlab and Octave environments. Internally, however, it is quite different: AutoBayes derives a customized algorithm implementing the given model using a schema-based process, and then further refines and optimizes the algorithm into code. A schema is a parameterized code template with associated semantic constraints which define and restrict the template s applicability. The schema parameters are instantiated in a problem-specific way during synthesis as AutoBayes checks the constraints against the original model or, recursively, against emerging sub-problems. AutoBayes schema library contains problem decomposition operators (which are justified by theorems in a formal logic in the domain of Bayesian networks) as well as machine learning algorithms (e.g., EM, k-Means) and nu- meric optimization methods (e.g., Nelder-Mead simplex, conjugate gradient). AutoBayes augments this schema-based approach by symbolic computation to derive closed-form solutions whenever possible. This is a major advantage over other statistical data analysis systems which use numerical approximations even in cases where closed-form solutions exist. AutoBayes is implemented in Prolog and comprises approximately 75.000 lines of code. In this paper, we take one typical scientific data analysis problem-analyzing planetary nebulae images taken by the Hubble Space Telescope-and show how AutoBayes can be used to automate the implementation of the necessary anal- ysis programs. We initially follow the analysis described by Knuth and Hajian [KHO2] and use AutoBayes to derive code for the published models. We show the details of the code derivation process, including the symbolic computations and automatic integration of library procedures, and compare the results of the automatically generated and manually implemented code. We then go beyond the original analysis and use AutoBayes to derive code for a simple image segmentation procedure based on a mixture model which can be used to automate a manual preproceesing step. Finally, we combine the original approach with the simple segmentation which yields a more detailed analysis. This also demonstrates that AutoBayes makes it easy to combine different aspects of data analysis.

Fischer, Bernd↗

Neural network classification - A Bayesian interpretation

The relationship between minimizing a mean squared error and finding the optimal Bayesian classifier is reviewed. This provides a theoretical interpretation for the process by which neural networks are used in classification. A number of confidence measures are proposed to evaluate the performance of the neural network classifier within a statistical framework.

Wan, Eric A.↗

Solar Activity Relations in Energetic Electron Events Measured By the MESSENGER Mission

Aims. We perform a statistical study of the relations between the properties of solar energetic electron (SEE) events measured by the MESSENGER mission from 2010 to 2015 and the parameters of the respective parent solar activity phenomena in order to identify the potential correlations between them. During the time of analysis, the MESSENGER heliocentric distance varied between 0.31 and 0.47 au. Methods. We used a published list of 61 SEE events measured by MESSENGER, which includes information on the near-relativistic electron peak intensities, the peak-intensity energy spectral indices, and the measured X-ray peak intensity of the flares related to the SEE events. Taking advantage of multi-viewpoint remote-sensing observations, we reconstructed, whenever possible, the associated coronal mass ejections (CMEs) and shock waves; and we determined the three-dimensional (3D) properties (location, speed, and width) of the CMEs and the maximum speed of the 3D CME-driven shocks in the corona. We used different methods (Spearman, Pearson, and a Bayesian approach, namely the Kelly method to linear regression) to estimate the correlation coefficients between the flare intensity, maximum speed at the apex of the CME-driven shock, CME speed at the apex, and CME width with the electron peak intensities and with the energy spectral indices. In this statistical study, we considered and addressed the limitations of the particle instrument on board MESSENGER (elevated background intensity level, anti-Sun pointing). Results. There is an asymmetry to the east in the range of connection angles (CAs) for which the SEE events present the highest peak intensities, where the CA is the longitudinal separation between the footpoint of the magnetic field connecting to the spacecraft and the flare location. Based on this asymmetry, we define a subsample of well-connected events as when −65° ≤ CA ≤ +33°. For the well-connected sample, we find moderate to strong correlations between the near-relativistic electron peak intensity and the 3D CME-driven shock maximum speed at the apex (Spearman: cc = 0.53 ± 0.05; Pearson: cc = 0.65 ± 0.04; Kelly: cc = 0.87 ± 0.20), the flare peak intensity (Spearman: cc = 0.63 ± 0.03; Pearson: cc = 0.59 ± 0.03; Kelly: cc = 0.74 ± 0.30), and the 3D CME speed at the apex (Spearman: cc = 0.50 ± 0.04; Pearson: cc = 0.46 ± 0.03; Kelly: cc = 0.60 ± 0.39). When including poorly connected events (full sample), the relations between the peak intensities and the solar-activity phenomena are blurred, showing lower correlation coefficients. Conclusions. Based on the comparison of the correlation coefficients presented in this study using near 0.4 au data, (1) both flare and shock-related processes may contribute to the acceleration of near relativistic electrons in large SEE events, in agreement with previous studies based on near 1 au data; and (2) the maximum speed of the CME-driven shock is a better parameter to investigate particle-acceleration-related mechanisms than the average CME speed, as suggested by the stronger correlation with the SEE peak intensities.

Sun↗