Search NASA⌕ Search

SEARCH · Search NASA

Results for “Statistical sampling techniques”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

CROP type analysis using Landsat digital data

Classification and statistical sampling techniques for crop type discrimination using Landsat digital data have been developed by the University of California in cooperation with NASA and the California Department of Water Resources. Ratioed bands (MSS 7/5 and 5/4) and a sun-angle corrected Euclidean albedo band were prepared from data for the Sacramento Valley for five different dates. The test area was stratified into general crop groupings based on the particular patterns of irrigation timing for each crop. Data classified within each stratum were used to produce a crop type map. Comparison with ground data indicates that certain crops and crop groups are discernable. Small grains and rice are easily identifiable, as are deciduous fruit varieties as a group. However, it is not feasible to separate various fruit and nut varieties, or separate vegetable crops with these techniques at present.

Brown, C. E.↗

Forward modeling fluctuations in the DESI LRGs target sample using image simulations

We use the forward modeling pipeline, Obiwan, to study the imaging systematics of the Luminous Red Galaxies (LRGs) targeted by the Dark Energy Spectroscopic Instrument (DESI). Imaging systematics refers to the false fluctuation of galaxy densities due to varying observing conditions and astrophysical foregrounds corresponding to the imaging surveys from which DESI LRG target galaxies are selected. We update the Obiwan pipeline, which we previously developed to simulate the optical images used to target DESI data, to further simulate WISE images in the infrared. This addition allows simulating the DESI LRGs sample, which utilizes WISE data in the target selection. Deep DESI imaging data combined with a method to account for biases in their shapes is used to define a truth sample of potential LRG targets. We inject these data evenly throughout the DESI Legacy Imaging Survey footprint at declinations between -30 and 32.375 degrees. We simulate a total of 15 million galaxies to obtain a simulated LRG sample (Obiwan LRGs) that predicts the variations in target density due to imaging properties. We find that the simulations predict the trends with depth observed in the data, including how they depend on the intrinsic brightness of the galaxies. We observe that faint LRGs are the main contributing source of the imaging systematics trend induced by depth. We also find significant trends in the data against Galactic extinction that are not predicted by Obiwan. These trends depend strongly on the particular map of Galactic extinction chosen to test against, implying systematic contamination in the Galactic extinction maps is a likely root cause (e.g., Cosmic-Infrared Background, dust temperature correction). We additionally observe a morphological change of the DESI LRGs population evidenced by a correlation between OII emission line average intensity and the size of the z-band PSF. This effect most likely results from uncertainties in background subtraction. The detailed findings we present should be used to guide any observational systematics mitigation treatment for the clustering of the DESI LRGs sample.

79 ASTRONOMY AND ASTROPHYSICS↗

Statistical Analysis Techniques for Small Sample Sizes

The small sample sizes problem which is encountered when dealing with analysis of space-flight data is examined. Because of such a amount of data available, careful analyses are essential to extract the maximum amount of information with acceptable accuracy. Statistical analysis of small samples is described. The background material necessary for understanding statistical hypothesis testing is outlined and the various tests which can be done on small samples are explained. Emphasis is on the underlying assumptions of each test and on considerations needed to choose the most appropriate test for a given type of analysis.

Navard, S. E.↗

Experimental analysis of computer system dependability

This paper reviews an area which has evolved over the past 15 years: experimental analysis of computer system dependability. Methodologies and advances are discussed for three basic approaches used in the area: simulated fault injection, physical fault injection, and measurement-based analysis. The three approaches are suited, respectively, to dependability evaluation in the three phases of a system's life: design phase, prototype phase, and operational phase. Before the discussion of these phases, several statistical techniques used in the area are introduced. For each phase, a classification of research methods or study topics is outlined, followed by discussion of these methods or topics as well as representative studies. The statistical techniques introduced include the estimation of parameters and confidence intervals, probability distribution characterization, and several multivariate analysis methods. Importance sampling, a statistical technique used to accelerate Monte Carlo simulation, is also introduced. The discussion of simulated fault injection covers electrical-level, logic-level, and function-level fault injection methods as well as representative simulation environments such as FOCUS and DEPEND. The discussion of physical fault injection covers hardware, software, and radiation fault injection methods as well as several software and hybrid tools including FIAT, FERARI, HYBRID, and FINE. The discussion of measurement-based analysis covers measurement and data processing techniques, basic error characterization, dependency analysis, Markov reward modeling, software-dependability, and fault diagnosis. The discussion involves several important issues studies in the area, including fault models, fast simulation techniques, workload/failure dependency, correlated failures, and software fault tolerance.

Iyer, Ravishankar, K.↗

Statistical Symbolic Execution with Informed Sampling

Symbolic execution techniques have been proposed recently for the probabilistic analysis of programs. These techniques seek to quantify the likelihood of reaching program events of interest, e.g., assert violations. They have many promising applications but have scalability issues due to high computational demand. To address this challenge, we propose a statistical symbolic execution technique that performs Monte Carlo sampling of the symbolic program paths and uses the obtained information for Bayesian estimation and hypothesis testing with respect to the probability of reaching the target events. To speed up the convergence of the statistical analysis, we propose Informed Sampling, an iterative symbolic execution that first explores the paths that have high statistical significance, prunes them from the state space and guides the execution towards less likely paths. The technique combines Bayesian estimation with a partial exact analysis for the pruned paths leading to provably improved convergence of the statistical analysis. We have implemented statistical symbolic execution with in- formed sampling in the Symbolic PathFinder tool. We show experimentally that the informed sampling obtains more precise results and converges faster than a purely statistical analysis and may also be more efficient than an exact symbolic analysis. When the latter does not terminate symbolic execution with informed sampling can give meaningful results under the same time and memory limits.

Reliability↗

Galaxy cluster profiles: a Gaussian mixture model approach to halo miscentering

Measurements of the galaxy density and weak-lensing profiles of galaxy clusters typically rely on an assumed cluster center, which is taken to be the brightest cluster galaxy or other proxies for the true halo center defined as the minimum in the potential well. Departure of the assumed cluster center from the true halo center bias the resultant profile measurements, an effect known as miscentering bias. Currently, miscentering is typically modeled in stacked profiles of clusters with a two parameter model. We use an alternate approach in which the profiles of individual clusters are used with the corresponding likelihood computed using a Gaussian mixture model. We test the approach using halos and the corresponding subhalo profiles from the IllustrisTNG hydrodynamic simulations. We obtain significantly improved estimates of the miscentering parameters for both 3D and projected 2D profiles relevant for imaging surveys. We discuss applications to upcoming cosmological surveys. Our Python package for the Gaussian mixture model is publicly available at https://github.com/KyleMiller1/Halo-Miscentering-Mixture-Model.

Bayesian reasoning↗

First astrometric constraints on parity-violation in the gravitational wave background

Astrometry, the precise measurement of stellar positions and velocities, offers a promising approach to probing the low-frequency stochastic gravitational wave background (SGWB). Notably, astrometric vector sky maps are sensitive to parity-violating SGWB signals, which cannot be distinguished using pulsar timing array observations in an isotropic SGWB. We present the first astrometric constraints on parity-violating SGWB using quasar catalogs from Gaia DR3 and VLBA data. By analyzing the EB correlation in the two-point correlation function of the proper motions of the quasars, we find 2σ constraints on the parity-violating SGWB amplitude h 70 2 Ω V = -0.020 ± 0.025 from Gaia DR3 and h 70 2 Ω V = -0.004 ± 0.010 from VLBA. These constraints are valid in the frequency range 4.2 × 10 -18 Hz < f < 1.1 × 10 -8 Hz. Although not currently a tight constraint on theoretical models, this first attempt lays the groundwork for future investigations using more precise astrometric data.

Gravitational waves in GR and beyond: theory↗

Quantifying bias due to non-Gaussian foregrounds in an optimal reconstruction of CMB lensing and temperature power spectra

We estimate the magnitude of the bias due to non-Gaussian extragalactic foregrounds on the optimal reconstruction of the cosmic microwave background (CMB) lensing potential and temperature power spectra. The reconstruction is performed using a Bayesian inference method known as the marginal unbiased score expansion (MUSE). We apply MUSE to a minimum variance combination of multifrequency maps drawn from the Agora publicly available simulations of the lensed CMB and correlated extragalactic foreground emission. Taking noise levels appropriate to the SPT-3G D1 release, we find non-Gaussian foregrounds may bias the MUSE reconstruction of the lensing potential amplitude at the level of (0.7 ± 0.3)σ when using modes up to ℓ max = 3500. We do not detect a statistically significant bias, finding a value of (-0.4 ± 0.3)σ, when restricted to lower angular multipoles, ℓ max = 3000. This work is a first step toward understanding the impact of extragalactic foregrounds on optimal reconstructions of CMB temperature and lensing potential power spectra.

Statistical sampling techniques↗

Field-level reconstruction from foreground-contaminated 21-cm maps

Current and upcoming 21-cm experiments will soon be able to map 21-cm spatial fluctuations in three dimensions for a wide range of redshifts. However, bright foreground contamination and the nature of radio interferometry create significant challenges, making it difficult to access rich cosmological information from the Fourier modes that lie within the “foreground wedge”. Here, in this work, we introduce two approaches aiming to reconstruct the full 21-cm density field, including the missing modes in the wedge: (a) a field-level inference under an effective field theory (EFT) framework; (b) a diffusion-based deep generative model trained on simulations. Under the EFT framework, we implement a fully differentiable forward model that maps the initial conditions of matter fluctuations to the observed, foreground-filtered 21-cm maps. This enables a gradient-based sampler to simultaneously sample the initial conditions and bias parameters, allowing a physically motivated mode reconstruction. Alternatively, we apply a variational diffusion model to perform 21-cm density reconstruction at the map level. Our model is trained on semi-numerical simulations over a wide range of astrophysical parameters. Our results from both approaches should provide improved cosmological constraints from the field level and also enable cross-correlation between experiments that have little or no overlapping modes.

cosmological perturbation theory↗

The Atacama Cosmology Telescope: Map-Based Noise Simulations for DR6

The increasing statistical power of cosmic microwave background (CMB) datasets requires a commensurate effort in understanding their noise properties. The noise in maps from ground-based instruments is dominated by large-scale correlations, which poses a modeling challenge. This paper develops novel models of the complex noise covariance structure in the Atacama Cosmology Telescope Data Release 6 (ACT DR6) maps. We first enumerate the noise properties that arise from the combination of the atmosphere and the ACT scan strategy. We then prescribe a class of Gaussian, map-based noise models, including a new wavelet-based approach that uses directional wavelet kernels for modeling correlated instrumental noise. The models are empirical, whose only inputs are a small number of independent realizations of the same region of sky. We evaluate the performance of these models against the ACT DR6 data by drawing ensembles of noise realizations. Applying these simulations to the ACT DR6 power spectrum pipeline reveals a ≥ 20% excess in the covariance matrix diagonal when compared to an analytic expression that assumes noise properties are uniquely described by their power spectrum. Along with our public code, mnms, this work establishes a necessary element in the science pipelines of both ACT DR6 and future ground-based CMB experiments such as the Simons Observatory (SO).

CMBR experiments↗

Analysis of the Einstein sample of early-type galaxies

The EINSTEIN galaxy catalog contains x-ray data for 148 early-type (E and SO) galaxies. A detailed analysis of the global properties of this sample are studied. By comparing the x-ray properties with other tracers of the ISM, as well as with observables related to the stellar dynamics and populations of the sample, we expect to determine more clearly the physical relationships that determine the evolution of early-type galaxies. Previous studies with smaller samples have explored the relationships between x-ray luminosity (L(sub x)) and luminosities in other bands. Using our larger sample and the statistical techniques of survival analysis, a number of these earlier analyses were repeated. For our full sample, a strong statistical correlation is found between L(sub X) and L(sub B) (the probability that the null hypothesis is upheld is P less than 10(exp -4) from a variety of rank correlation tests. Regressions with several algorithms yield consistent results.

Eskridge, Paul B.↗

Statistical classification techniques for engineering and climatic data samples

Fisher's sample linear discriminant function is modified through an appropriate alteration of the common sample variance-covariance matrix. The alteration consists of adding nonnegative values to the eigenvalues of the sample variance covariance matrix. The desired results of this modification is to increase the number of correct classifications by the new linear discriminant function over Fisher's function. This study is limited to the two-group discriminant problem.

Temple, E. C.↗

Space shuttle solid rocket booster recovery system definition, volume 1

The performance requirements, preliminary designs, and development program plans for an airborne recovery system for the space shuttle solid rocket booster are discussed. The analyses performed during the study phase of the program are presented. The basic considerations which established the system configuration are defined. A Monte Carlo statistical technique using random sampling of the probability distribution for the critical water impact parameters was used to determine the failure probability of each solid rocket booster component as functions of impact velocity and component strength capability.

Source record↗

Growth of GaAs crystals

Study on effects of melt and growth on solute segregation and crystal quality uses statistical techniques to reduce sample numbers and experimental costs.

Li, C.↗

Investigation of Error Patterns in Geographical Databases

The objective of the research conducted in this project is to develop a methodology to investigate the accuracy of Airport Safety Modeling Data (ASMD) using statistical, visualization, and Artificial Neural Network (ANN) techniques. Such a methodology can contribute to answering the following research questions: Over a representative sampling of ASMD databases, can statistical error analysis techniques be accurately learned and replicated by ANN modeling techniques? This representative ASMD sample should include numerous airports and a variety of terrain characterizations. Is it possible to identify and automate the recognition of patterns of error related to geographical features? Do such patterns of error relate to specific geographical features, such as elevation or terrain slope? Is it possible to combine the errors in small regions into an error prediction for a larger region? What are the data density reduction implications of this work? ASMD may be used as the source of terrain data for a synthetic visual system to be used in the cockpit of aircraft when visual reference to ground features is not possible during conditions of marginal weather or reduced visibility. In this research, United States Geologic Survey (USGS) digital elevation model (DEM) data has been selected as the benchmark. Artificial Neural Networks (ANNS) have been used and tested as alternate methods in place of the statistical methods in similar problems. They often perform better in pattern recognition, prediction and classification and categorization problems. Many studies show that when the data is complex and noisy, the accuracy of ANN models is generally higher than those of comparable traditional methods.

Dryer, David↗

Chi-squared and C statistic minimization for low count per bin data

Results are presented from a computer simulation comparing two statistical fitting techniques on data samples with large and small counts per bin; the results are then related specifically to X-ray astronomy. The Marquardt and Powell minimization techniques are compared by using both to minimize the chi-squared statistic. In addition, Cash's C statistic is applied, with Powell's method, and it is shown that the C statistic produces better fits in the low-count regime than chi-squared.

Nousek, John A.↗

Convergence of oscillator spectral estimators for counted-frequency measurements.

A common intermediary connecting frequency-noise calibration or testing of an oscillator to useful applications is the spectral density of the frequency-deviating process. In attempting to turn test data into predicts of performance characteristics, one is naturally led to estimation of statistical values by sample-mean and sample-variance techniques. However, sample means and sample variances themselves are statistical quantities that do not necessarily converge (in the mean-square sense) to actual ensemble-average means and variances, except perhaps for excessively large sample sizes. This is especially true for the flicker noise component of oscillators. This article shows, for the various types of noises found in oscillators, how sample averages converge (or do not converge) to their statistical counterparts. The convergence rate is shown to be the same for all oscillators of a given spectral type.

Tausworthe, R. C.↗