Search NASA⌕ Search

SEARCH · Search NASA

Results for “Bayesian statistics”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

190 records · Page 11

Comparison of Likelihood Methods for Generalized Linear Mixed Models with Application to Quiet Supersonic Flights 2018 Data

Repeated measurement will be a feature of the survey data collected during the Quesst missionX-59 community response tests (CRT). Since each participant will report his or her categorical level of annoyance in response to multiple events, the responses from any single individual may be correlated with one another. Several models within the class of generalized linear mixed models (GLMM) are pertinent to the analysis of correlated categorical outcomes; the random intercept logistic regression model is one example. Both Bayesian and frequentist methods for fitting these models are available, with frequentist methods relying on some form of approximation (of either an integral or the integrand) that appears in the marginal likelihood function. Given several anticipated similarities of the X-59 CRT data to data collected during a past risk reduction, Quiet Supersonic Flights 2018 (QSF18), this short note is intended to create awareness. It documents an instance in which a reported population average dose-response relationship derived from QSF18 single event data was distorted by the integral approximation applied in likelihood-based methods. We review some of the available literature on the topic, compare the outputs of several different computational approaches implemented in available statistical software, and present simple corrective actions that may be useful during the Quesst mission.

dose-response model↗

Topics in inference and decision-making with partial knowledge

Two essential elements needed in the process of inference and decision-making are prior probabilities and likelihood functions. When both of these components are known accurately and precisely, the Bayesian approach provides a consistent and coherent solution to the problems of inference and decision-making. In many situations, however, either one or both of the above components may not be known, or at least may not be known precisely. This problem of partial knowledge about prior probabilities and likelihood functions is addressed. There are at least two ways to cope with this lack of precise knowledge: robust methods, and interval-valued methods. First, ways of modeling imprecision and indeterminacies in prior probabilities and likelihood functions are examined; then how imprecision in the above components carries over to the posterior probabilities is examined. Finally, the problem of decision making with imprecise posterior probabilities and the consequences of such actions are addressed. Application areas where the above problems may occur are in statistical pattern recognition problems, for example, the problem of classification of high-dimensional multispectral remote sensing image data.

Safavian, S. Rasoul↗

Modeling Increased Complexity and the Reliance on Automation: FLightdeck Automation Problems (FLAP) Model

This paper highlights the development of a model that is focused on the safety issue of increasing complexity and reliance on automation systems in transport category aircraft. Recent statistics show an increase in mishaps related to manual handling and automation errors due to pilot complacency and over-reliance on automation, loss of situational awareness, automation system failures and/or pilot deficiencies. Consequently, the aircraft can enter a state outside the flight envelope and/or air traffic safety margins which potentially can lead to loss-of-control (LOC), controlled-flight-into-terrain (CFIT), or runway excursion/confusion accidents, etc. The goal of this modeling effort is to provide NASA's Aviation Safety Program (AvSP) with a platform capable of assessing the impacts of AvSP technologies and products towards reducing the relative risk of automation related accidents and incidents. In order to do so, a generic framework, capable of mapping both latent and active causal factors leading to automation errors, is developed. Next, the framework is converted into a Bayesian Belief Network model and populated with data gathered from Subject Matter Experts (SMEs). With the insertion of technologies and products, the model provides individual and collective risk reduction acquired by technologies and methodologies developed within AvSP.

Ancel, Ersin↗

Atacama Cosmology Telescope measurements of a large sample of candidates from the Massive and Distant Clusters of WISE Survey: Sunyaev-Zeldovich effect confirmation of MaDCoWS candidates using ACT

Context. Galaxy clusters are an important tool for cosmology, and their detection and characterization are key goals for current and future surveys. Using data from the Wide-field Infrared Survey Explorer (WISE), the Massive and Distant Clusters of WISE Survey (MaDCoWS) located 2839 significant galaxy overdensities at redshifts 0.7 . z . 1.5, which included extensive follow-up imaging from the Spitzer Space Telescope to determine cluster richnesses. Concurrently, the Atacama Cosmology Telescope (ACT) has produced large area millimeter-wave maps in three frequency bands along with a large catalog of Sunyaev-Zeldovich (SZ)-selected clusters as part of its Data Release 5 (DR5). Aims. We aim to verify and characterize MaDCoWS clusters using measurements of, or limits on, their thermal SZ effect signatures. We also use these detections to establish the scaling relation between SZ mass and the MaDCoWS-defined richness. Methods. Using the maps and cluster catalog from DR5, we explore the scaling between SZ mass and cluster richness. We do this by comparing cataloged detections and extracting individual and stacked SZ signals from the MaDCoWS cluster locations. We use complementary radio survey data from the Very Large Array, submillimeter data from Herschel, and ACT 224 GHz data to assess the impact of contaminating sources on the SZ signals from both ACT and MaDCoWS clusters. We use a hierarchical Bayesian model to fit the mass-richness scaling relation, allowing for clusters to be drawn from two populations: one, a Gaussian centered on the mass-richness relation, and the other, a Gaussian centered on zero SZ signal. Results. We find that MaDCoWS clusters have submillimeter contamination that is consistent with a gray-body spectrum, while the ACT clusters are consistent with no submillimeter emission on average. Additionally, the intrinsic radio intensities of ACT clusters are lower than those of MaDCoWS clusters, even when the ACT clusters are restricted to the same redshift range as the MaDCoWS clusters. We find the best-fit ACT SZ mass versus MaDCoWS richness scaling relation has a slope of p1 = 1.84+0.15 −0.14, where the slope is defined as M ∝ λ p1 15 and λ15 is the richness. We also find that the ACT SZ signals for a significant fraction (∼57%) of the MaDCoWS sample can statistically be described as being drawn from a noise-like distribution, indicating that the candidates are possibly dominated by low-mass and unvirialized systems that are below the mass limit of the ACT sample. Further, we note that a large portion of the optically confirmed ACT clusters located in the same volume of the sky as MaDCoWS are not selected by MaDCoWS, indicating that the MaDCoWS sample is not complete with respect to SZ selection. Finally, we find that the radio loud fraction of MaDCoWS clusters increases with richness, while we find no evidence that the submillimeter emission of the MaDCoWS clusters evolves with richness. Conclusions. We conclude that the original MaDCoWS selection function is not well defined and, as such, reiterate the MaDCoWS collaboration’s recommendation that the sample is suited for probing cluster and galaxy evolution, but not cosmological analyses. We find a best-fit mass-richness relation slope that agrees with the published MaDCoWS preliminary results. Additionally, we find that while the approximate level of infill of the ACT and MaDCoWS cluster SZ signals (1–2%) is subdominant to other sources of uncertainty for current generation experiments, characterizing and removing this bias will be critical for next-generation experiments hoping to constrain cluster masses at the sub-percent level.

large↗

Setting the Bar for the Replacement of the Probability of Collision Metric

To date, satellite conjunction assessment (CA) risk analysis has largely embraced the probability of collision (Pc) as the omnibus metric to evaluate collision likelihood, and its use in such assessments has mostly been straightforward: at the point at which a conjunction mitigation decision is required, the calculated Pc is compared to a threshold; and if the calculated Pc exceeds that threshold, then a mitigation action is warranted. With only minor variation, this approach is employed by major CA risk assessment centers (e.g., NASA, EUSST, CNES, JAXA) and is advanced as the preferred method in the published CA best practices handbooks. Despite this near unanimity of operational practice, there is a major strain of secondary literature critical of the Pc and willing to propose alternatives. Alfano (2005) pointed out the ability of the Pc to underrepresent the risk in certain situations and counselled a maximum Pc construct. Carpenter (2017, 2019) reiterated this criticism and proposed using instead a confidence interval on the miss distance. Balch et al. (2019) identified what they argued was a defect in the entire Bayesian Pc construct and believed that the use of a more conservative methodology based on covariance ellipsoid overlap was necessary. Delande (2022) introduced the framework of collision “plausibility” to the risk assessment process and sketched out how this might be used operationally. Elkantassi (2022) published a full development of the miss distance confidence interval approach and applied it to several worked examples. While these different approaches to collision risk assessment do differ in their details, they all converge on two central points: first, the Pc’s failure to give an adequate expression of the risk in dilution region situations is a fatal flaw; and second, a conjunction should be presumed risky and in need of mitigation until the evidence of the situation can establish otherwise. These criticisms, if correct, would counsel a number of modifications to current CA operational practice; as such, they force a re-examination of fundamental aspects of the CA problem, including the following: 1. Is the CA risk assessment a probability problem, a statistics problem, or something else? 2. If it is a statistics problem, does it lend itself naturally to a hypothesis test construction? 3. If it can be construed as a hypothesis test, what form should the null hypothesis take, to wit: what constraints exist on the choice of the null hypothesis, what selections are in best alignment with all of the attendant parameters of the problem, and what is implied philosophically by different choices? 4. What are the implications of using the different proposed risk assessment parameters for CA? This question should be answered both in determining how frequently the dilution region situation cited by the critics of the Pc actually appears in an operationally significant manner and the missed detection and false alarm rates of all of the proposed risk assessment metrics, compared both to the Pc and to each other. This paper explores and offers preliminary answers to the above questions, presenting a researched treatment of the philosophical nature of the CA problem and the null hypothesis choice that achieves the greatest consistency with all of the different aspects of operational CA conduct. It then profiles all of the different proposed risk assessment metrics enumerated in the earlier paragraph against an extremely large database of conjunction events at both the 550km and 700km altitudes. The combination of the philosophical exploration of the CA problem and the results of the profiling activity articulates what a risk assessment metric will need to demonstrate, in terms of both innate construction and performance, in order to be a true competitor to the Pc.

conjunction assessment↗

A Sub-Neptune and a Non-Transiting Neptune-Mass Companion Unveiled by ESPRESSO Around the Bright Late-F Dwarf HD 5278 (TOI-130)

Context. Transiting sub-Neptune-type planets, with radii approximately between 2 and 4R⊕, are of particular interest as their study allows us to gain insight into the formation and evolution of a class of planets that are not found in our Solar System. Aims. We exploit the extreme radial velocity (RV) precision of the ultra-stable echelle spectrograph ESPRESSO on the VLT to unveil the physical properties of the transiting sub-Neptune TOI-130 b, uncovered by the TESS mission orbiting the nearby, bright, late F-typestar HD 5278 (TOI-130) with a period of Pb=14.3 days. Methods. We used 43 ESPRESSO high-resolution spectra and broad-band photometry information to derive accurate stellar atmospheric and physical parameters of HD 5278. We exploited the TESS light curve and spectroscopic diagnostics to gauge the impact of stellar activity on the ESPRESSO RVs. We performed separate as well as joint analyses of the TESS photometry and the ESPRESSORVs using fully Bayesian frameworks to determine the system parameters. Results. Based on the ESPRESSO spectra, the updated stellar parameters of HD 5278 are Teff=6203±64K, logg=4.50±0.11dex, [Fe/H] =−0.12±0.04dex,M?=1.126+0.036−0.035M, and R?=1.194+0.017−0.016R. We determine HD 5278 b’s mass and radius to be Mb=7.8+1.5−1.4M⊕ and Rb=2.45±0.05R⊕. The derived mean density, %b=2.9+0.6−0.5g cm−3, is consistent with the bulk composition of a sub-Neptune with a substantial (∼30%) water mass fraction and with a gas envelope comprising ∼17% of the measured radius. Given the host brightness and irradiation levels, HD 5278 b is one of the best targets orbiting G-F primaries for follow-up atmospheric characterization measurements with HST and JWST. We discover a second, non-transiting companion in the system, with a period of Pc=40.87+0.18−0.17days and a minimum mass of Mcsinic=18.4+1.8−1.9M⊕. We study emerging trends in parameters space (e.g., mass, radius, stellar insolation, and mean density) of the growing population of transiting sub-Neptunes, and provide statistical evidence for a low occurrence of close-in,10−15M⊕companions around G-F primaries withTeff&5500K.

planetary systems↗

Bayesian Approach for Determining Microlens System Properties with High-angular resolution Follow-up Imaging

We present the details of the Bayesian analysis of the planetary microlensing event MOA-2016-BLG-227, whose excess flux is likely due to a source/lens companion or an unrelated ambient star, as well as of the assumed prior distributions. Furthermore, we apply this method to four reported planetary events, MOA-2008-BLG-310, MOA2011-BLG-293, OGLE-2012-BLG-0527, and OGLE-2012-BLG-0950, where adaptive optics observations have detected excess flux at the source star positions. For events with small angular Einstein radii, our lens mass estimates are more uncertain than those of previous analyses, which assumed that the excess was due to the lens. Our predictions for MOA-2008-BLG-310 and OGLE-2012-BLG-0950 are consistent with recent results on these events obtained via Keck and Hubble Space Telescope observations when the source star is resolvable from the lens star. For events with small angular Einstein radii, we find that it is generally difficult to conclude whether the excess flux comes from the host star. Therefore, it is necessary to identify the lens star by measuring its proper motion relative to the source star to determine whether the excess flux comes from the lens star. Even without such measurements, our method can be used to statistically test the dependence of the planet-hosting probability on the stellar mass.

Naoki Koshimoto↗

Rao-Blackwellization for Adaptive Gaussian Sum Nonlinear Model Propagation

When dealing with imperfect data and general models of dynamic systems, the best estimate is always sought in the presence of uncertainty or unknown parameters. In many cases, as the first attempt, the Extended Kalman filter (EKF) provides sufficient solutions to handling issues arising from nonlinear and non-Gaussian estimation problems. But these issues may lead unacceptable performance and even divergence. In order to accurately capture the nonlinearities of most real-world dynamic systems, advanced filtering methods have been created to reduce filter divergence while enhancing performance. Approaches, such as Gaussian sum filtering, grid based Bayesian methods and particle filters are well-known examples of advanced methods used to represent and recursively reproduce an approximation to the state probability density function (pdf). Some of these filtering methods were conceptually developed years before their widespread uses were realized. Advanced nonlinear filtering methods currently benefit from the computing advancements in computational speeds, memory, and parallel processing. Grid based methods, multiple-model approaches and Gaussian sum filtering are numerical solutions that take advantage of different state coordinates or multiple-model methods that reduced the amount of approximations used. Choosing an efficient grid is very difficult for multi-dimensional state spaces, and oftentimes expensive computations must be done at each point. For the original Gaussian sum filter, a weighted sum of Gaussian density functions approximates the pdf but suffers at the update step for the individual component weight selections. In order to improve upon the original Gaussian sum filter, Ref. [2] introduces a weight update approach at the filter propagation stage instead of the measurement update stage. This weight update is performed by minimizing the integral square difference between the true forecast pdf and its Gaussian sum approximation. By adaptively updating each component weight during the nonlinear propagation stage an approximation of the true pdf can be successfully reconstructed. Particle filtering (PF) methods have gained popularity recently for solving nonlinear estimation problems due to their straightforward approach and the processing capabilities mentioned above. The basic concept behind PF is to represent any pdf as a set of random samples. As the number of samples increases, they will theoretically converge to the exact, equivalent representation of the desired pdf. When the estimated qth moment is needed, the samples are used for its construction allowing further analysis of the pdf characteristics. However, filter performance deteriorates as the dimension of the state vector increases. To overcome this problem Ref. [5] applies a marginalization technique for PF methods, decreasing complexity of the system to one linear and another nonlinear state estimation problem. The marginalization theory was originally developed by Rao and Blackwell independently. According to Ref. [6] it improves any given estimator under every convex loss function. The improvement comes from calculating a conditional expected value, often involving integrating out a supportive statistic. In other words, Rao-Blackwellization allows for smaller but separate computations to be carried out while reaching the main objective of the estimator. In the case of improving an estimator's variance, any supporting statistic can be removed and its variance determined. Next, any other information that dependents on the supporting statistic is found along with its respective variance. A new approach is developed here by utilizing the strengths of the adaptive Gaussian sum propagation in Ref. [2] and a marginalization approach used for PF methods found in Ref. [7]. In the following sections a modified filtering approach is presented based on a special state-space model within nonlinear systems to reduce the dimensionality of the optimization problem in Ref. [2]. First, the adaptive Gaussian sum propagation is explained and then the new marginalized adaptive Gaussian sum propagation is derived. Finally, an example simulation is presented.

state estimation↗

A Bayesian Analysis of SDSS J0914+0853, a Low-mass Dual AGN Candidate

We present the first results from Bayesian AnalYsis of Multiple AGN in X-rays (BAYMAX), a tool that uses a Bayesian framework to quantitatively evaluate whether a given Chandra observation is more likely a single or dual point source. Although the most robust method of determining the presence of dual active galactic nuclei (AGNs) is to use X-ray observations, only sources that are widely separated relative to the instrumentʼs point-spread function are easy to identify. It becomes increasingly difficult to distinguish dual AGNs from single AGNs when the separation is on the order of Chandraʼs angular resolution (<1″). Using likelihood models for single and dual point sources, BAYMAX quantitatively evaluates the likelihood of an AGN for a given source. Specifically, we present results from BAYMAX analyzing the lowest-mass dual AGN candidate to date, SDSS J0914+0853, where archival Chandra data shows a possible secondary AGN ∼ 0"3 from the primary. Analyzing a new 50 ks Chandra observation, results from BAYMAX shows that SDSS J0914+0853 is most likely a single AGN with a Bayes factor of 13.5 in favor of a single point source model. Further, posterior distributions from the dual point source model are consistent with emission from a single AGN. We find a very low probability of SDSS J0914+0853 being a dual AGN system with a flux ratio f>0.3 and separation r>0"3. Overall, BAYMAX will be an important tool for correctly classifying candidate dual AGNs in the literature, as well as studying the dual AGN population where past spatial resolution limits have prevented systematic analyses.

Active galaxies↗

Intelligent machines in the twenty-first century: foundations of inference and inquiry

The last century saw the application of Boolean algebra to the construction of computing machines, which work by applying logical transformations to information contained in their memory. The development of information theory and the generalization of Boolean algebra to Bayesian inference have enabled these computing machines, in the last quarter of the twentieth century, to be endowed with the ability to learn by making inferences from data. This revolution is just beginning as new computational techniques continue to make difficult problems more accessible. Recent advances in our understanding of the foundations of probability theory have revealed implications for areas other than logic. Of relevance to intelligent machines, we recently identified the algebra of questions as the free distributive algebra, which will now allow us to work with questions in a way analogous to that which Boolean algebra enables us to work with logical statements. In this paper, we examine the foundations of inference and inquiry. We begin with a history of inferential reasoning, highlighting key concepts that have led to the automation of inference in modern machine-learning systems. We then discuss the foundations of inference in more detail using a modern viewpoint that relies on the mathematics of partially ordered sets and the scaffolding of lattice theory. This new viewpoint allows us to develop the logic of inquiry and introduce a measure describing the relevance of a proposed question to an unresolved issue. Last, we will demonstrate the automation of inference, and discuss how this new logic of inquiry will enable intelligent machines to ask questions. Automation of both inference and inquiry promises to allow robots to perform science in the far reaches of our solar system and in other star systems by enabling them not only to make inferences from data, but also to decide which question to ask, which experiment to perform, or which measurement to take given what they have learned and what they are designed to understand.

Review↗