Search NASA⌕ Search

SEARCH · Search NASA

Results for “Statistical sampling techniques”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

CROP type analysis using Landsat digital data

Classification and statistical sampling techniques for crop type discrimination using Landsat digital data have been developed by the University of California in cooperation with NASA and the California Department of Water Resources. Ratioed bands (MSS 7/5 and 5/4) and a sun-angle corrected Euclidean albedo band were prepared from data for the Sacramento Valley for five different dates. The test area was stratified into general crop groupings based on the particular patterns of irrigation timing for each crop. Data classified within each stratum were used to produce a crop type map. Comparison with ground data indicates that certain crops and crop groups are discernable. Small grains and rice are easily identifiable, as are deciduous fruit varieties as a group. However, it is not feasible to separate various fruit and nut varieties, or separate vegetable crops with these techniques at present.

Brown, C. E.↗

Statistical Analysis Techniques for Small Sample Sizes

The small sample sizes problem which is encountered when dealing with analysis of space-flight data is examined. Because of such a amount of data available, careful analyses are essential to extract the maximum amount of information with acceptable accuracy. Statistical analysis of small samples is described. The background material necessary for understanding statistical hypothesis testing is outlined and the various tests which can be done on small samples are explained. Emphasis is on the underlying assumptions of each test and on considerations needed to choose the most appropriate test for a given type of analysis.

Navard, S. E.↗

Experimental analysis of computer system dependability

This paper reviews an area which has evolved over the past 15 years: experimental analysis of computer system dependability. Methodologies and advances are discussed for three basic approaches used in the area: simulated fault injection, physical fault injection, and measurement-based analysis. The three approaches are suited, respectively, to dependability evaluation in the three phases of a system's life: design phase, prototype phase, and operational phase. Before the discussion of these phases, several statistical techniques used in the area are introduced. For each phase, a classification of research methods or study topics is outlined, followed by discussion of these methods or topics as well as representative studies. The statistical techniques introduced include the estimation of parameters and confidence intervals, probability distribution characterization, and several multivariate analysis methods. Importance sampling, a statistical technique used to accelerate Monte Carlo simulation, is also introduced. The discussion of simulated fault injection covers electrical-level, logic-level, and function-level fault injection methods as well as representative simulation environments such as FOCUS and DEPEND. The discussion of physical fault injection covers hardware, software, and radiation fault injection methods as well as several software and hybrid tools including FIAT, FERARI, HYBRID, and FINE. The discussion of measurement-based analysis covers measurement and data processing techniques, basic error characterization, dependency analysis, Markov reward modeling, software-dependability, and fault diagnosis. The discussion involves several important issues studies in the area, including fault models, fast simulation techniques, workload/failure dependency, correlated failures, and software fault tolerance.

Iyer, Ravishankar, K.↗

Statistical Symbolic Execution with Informed Sampling

Symbolic execution techniques have been proposed recently for the probabilistic analysis of programs. These techniques seek to quantify the likelihood of reaching program events of interest, e.g., assert violations. They have many promising applications but have scalability issues due to high computational demand. To address this challenge, we propose a statistical symbolic execution technique that performs Monte Carlo sampling of the symbolic program paths and uses the obtained information for Bayesian estimation and hypothesis testing with respect to the probability of reaching the target events. To speed up the convergence of the statistical analysis, we propose Informed Sampling, an iterative symbolic execution that first explores the paths that have high statistical significance, prunes them from the state space and guides the execution towards less likely paths. The technique combines Bayesian estimation with a partial exact analysis for the pruned paths leading to provably improved convergence of the statistical analysis. We have implemented statistical symbolic execution with in- formed sampling in the Symbolic PathFinder tool. We show experimentally that the informed sampling obtains more precise results and converges faster than a purely statistical analysis and may also be more efficient than an exact symbolic analysis. When the latter does not terminate symbolic execution with informed sampling can give meaningful results under the same time and memory limits.

Reliability↗

The Atacama Cosmology Telescope: Map-Based Noise Simulations for DR6

The increasing statistical power of cosmic microwave background (CMB) datasets requires a commensurate effort in understanding their noise properties. The noise in maps from ground-based instruments is dominated by large-scale correlations, which poses a modeling challenge. This paper develops novel models of the complex noise covariance structure in the Atacama Cosmology Telescope Data Release 6 (ACT DR6) maps. We first enumerate the noise properties that arise from the combination of the atmosphere and the ACT scan strategy. We then prescribe a class of Gaussian, map-based noise models, including a new wavelet-based approach that uses directional wavelet kernels for modeling correlated instrumental noise. The models are empirical, whose only inputs are a small number of independent realizations of the same region of sky. We evaluate the performance of these models against the ACT DR6 data by drawing ensembles of noise realizations. Applying these simulations to the ACT DR6 power spectrum pipeline reveals a ≥ 20% excess in the covariance matrix diagonal when compared to an analytic expression that assumes noise properties are uniquely described by their power spectrum. Along with our public code, mnms, this work establishes a necessary element in the science pipelines of both ACT DR6 and future ground-based CMB experiments such as the Simons Observatory (SO).

CMBR experiments↗

Analysis of the Einstein sample of early-type galaxies

The EINSTEIN galaxy catalog contains x-ray data for 148 early-type (E and SO) galaxies. A detailed analysis of the global properties of this sample are studied. By comparing the x-ray properties with other tracers of the ISM, as well as with observables related to the stellar dynamics and populations of the sample, we expect to determine more clearly the physical relationships that determine the evolution of early-type galaxies. Previous studies with smaller samples have explored the relationships between x-ray luminosity (L(sub x)) and luminosities in other bands. Using our larger sample and the statistical techniques of survival analysis, a number of these earlier analyses were repeated. For our full sample, a strong statistical correlation is found between L(sub X) and L(sub B) (the probability that the null hypothesis is upheld is P less than 10(exp -4) from a variety of rank correlation tests. Regressions with several algorithms yield consistent results.

Eskridge, Paul B.↗

Statistical classification techniques for engineering and climatic data samples

Fisher's sample linear discriminant function is modified through an appropriate alteration of the common sample variance-covariance matrix. The alteration consists of adding nonnegative values to the eigenvalues of the sample variance covariance matrix. The desired results of this modification is to increase the number of correct classifications by the new linear discriminant function over Fisher's function. This study is limited to the two-group discriminant problem.

Temple, E. C.↗

Space shuttle solid rocket booster recovery system definition, volume 1

The performance requirements, preliminary designs, and development program plans for an airborne recovery system for the space shuttle solid rocket booster are discussed. The analyses performed during the study phase of the program are presented. The basic considerations which established the system configuration are defined. A Monte Carlo statistical technique using random sampling of the probability distribution for the critical water impact parameters was used to determine the failure probability of each solid rocket booster component as functions of impact velocity and component strength capability.

Source record↗

Growth of GaAs crystals

Study on effects of melt and growth on solute segregation and crystal quality uses statistical techniques to reduce sample numbers and experimental costs.

Li, C.↗

Investigation of Error Patterns in Geographical Databases

The objective of the research conducted in this project is to develop a methodology to investigate the accuracy of Airport Safety Modeling Data (ASMD) using statistical, visualization, and Artificial Neural Network (ANN) techniques. Such a methodology can contribute to answering the following research questions: Over a representative sampling of ASMD databases, can statistical error analysis techniques be accurately learned and replicated by ANN modeling techniques? This representative ASMD sample should include numerous airports and a variety of terrain characterizations. Is it possible to identify and automate the recognition of patterns of error related to geographical features? Do such patterns of error relate to specific geographical features, such as elevation or terrain slope? Is it possible to combine the errors in small regions into an error prediction for a larger region? What are the data density reduction implications of this work? ASMD may be used as the source of terrain data for a synthetic visual system to be used in the cockpit of aircraft when visual reference to ground features is not possible during conditions of marginal weather or reduced visibility. In this research, United States Geologic Survey (USGS) digital elevation model (DEM) data has been selected as the benchmark. Artificial Neural Networks (ANNS) have been used and tested as alternate methods in place of the statistical methods in similar problems. They often perform better in pattern recognition, prediction and classification and categorization problems. Many studies show that when the data is complex and noisy, the accuracy of ANN models is generally higher than those of comparable traditional methods.

Dryer, David↗

Chi-squared and C statistic minimization for low count per bin data

Results are presented from a computer simulation comparing two statistical fitting techniques on data samples with large and small counts per bin; the results are then related specifically to X-ray astronomy. The Marquardt and Powell minimization techniques are compared by using both to minimize the chi-squared statistic. In addition, Cash's C statistic is applied, with Powell's method, and it is shown that the C statistic produces better fits in the low-count regime than chi-squared.

Nousek, John A.↗

A comparison of Landsat point and rectangular field training sets for land-use classification

Rectangular training fields of homogeneous spectroreflectance are commonly used in supervised pattern recognition efforts. Trial image classification with manually selected training sets gives irregular and misleading results due to statistical bias. A self-verifying, grid-sampled training point approach is proposed as a more statistically valid feature extraction technique. A systematic pixel sampling network of every ninth row and ninth column efficiently replaced the full image scene with smaller statistical vectors which preserved the necessary characteristics for classification. The composite second- and third-order average classification accuracy of 50.1 percent for 331,776 pixels in the full image substantially agreed with the 51 percent value predicted by the grid-sampled, 4,100-point training set.

Tom, C. H.↗

Convergence of oscillator spectral estimators for counted-frequency measurements.

A common intermediary connecting frequency-noise calibration or testing of an oscillator to useful applications is the spectral density of the frequency-deviating process. In attempting to turn test data into predicts of performance characteristics, one is naturally led to estimation of statistical values by sample-mean and sample-variance techniques. However, sample means and sample variances themselves are statistical quantities that do not necessarily converge (in the mean-square sense) to actual ensemble-average means and variances, except perhaps for excessively large sample sizes. This is especially true for the flicker noise component of oscillators. This article shows, for the various types of noises found in oscillators, how sample averages converge (or do not converge) to their statistical counterparts. The convergence rate is shown to be the same for all oscillators of a given spectral type.

Tausworthe, R. C.↗

Statistical Sampling of Tide Heights Study

The goal of the study was to determine if it was possible to reduce the cost of verifying computational models of tidal waves and currents. Statistical techniques were used to determine the least number of samples required, in a given situation, to remain statistically significant, and thereby reduce overall project costs. Commercial, academic, and Federal agencies could benefit by applying these techniques, without the need to 'touch' every item in the population. For example, the requirement of this project was to measure the heights and times of high and low tides at 8,000 locations for verification of computational models of tidal waves and currents. The application of the statistical techniques began with observations to determine the correctness of submitted measurement data, followed by some assumptions based on the observations. Among the assumptions were that the data were representative of data-collection techniques used at the measurement locations, that time measurements could be ignored (that is, height measurements alone would suffice), and that the height measurements were from a statistically normal distribution. Sample means and standard deviations were determined for all locations. Interval limits were determined for confidence levels of 95, 98, and 99 percent. It was found that the numbers of measurement locations needed to attain these confidence levels were 55, 78, and 96, respectively.

Source record↗

Lagged average forecasting, an alternative to Monte Carlo forecasting

A 'lagged average forecast' (LAF) model is developed for stochastic dynamic weather forecasting and used for predictions in comparison with the results of a Monte Carlo forecast (MCF). The technique involves the calculation of sample statistics from an ensemble of forecasts, with each ensemble member being an ordinary dynamical forecast (ODF). Initial conditions at a time lagging the start of the forecast period are used, with varying amounts of time for the lags. Forcing by asymmetric Newtonian heating of the lower layer is used in a two-layer, f-plane, highly truncated spectral model in a test forecasting run. Both the LAF and MCF are found to be more accurate than the ODF due to ensemble averaging with the MCF and the LAF. When a regression filter is introduced, all models become more accurate, with the LAF model giving the best results. The possibility of generating monthly or seasonal forecasts with the LAF is discussed.

Hoffman, R. N.↗

Improved Reference Sampling and Subtraction: A Technique for Reducing the Read Noise of Near-Infrared Detector Systems

Near-infrared array detectors, like the James Webb Space Telescope (JWST) NIRSpec's Teledyne's H2RGs, often provide reference pixels and a reference output. These are used to remove correlated noise. Improved reference sampling and subtraction (IRS(exp 2)) is a statistical technique for using this reference information optimally in a leastsquares sense. Compared with the traditional H2RG readout, IRS(exp 2) uses a different clocking pattern to interleave many more reference pixels into the data than is otherwise possible. Compared with standard reference correction techniques, IRS(exp 2) subtracts the reference pixels and reference output using a statistically optimized set of frequency dependent weights. The benefits include somewhat lower noise variance and much less obvious correlated noise. NIRSpec's IRS(exp 2) images are cosmetically clean, with less 1/f banding than in traditional data from the same system. This article describes the IRS(exp 2) clocking pattern and presents the equations needed to use IRS(exp 2) in systems other than NIRSpec. For NIRSpec, applying these equations is already an option in the calibration pipeline. As an aid to instrument builders, we provide our prototype IRS(exp 2) calibration software and sample JWST NIRSpec data. The same techniques are applicable to other detector systems, including those based on Teledyne's H4RG arrays. The H4RG's interleaved reference pixel readout mode is effectively one IRS(exp 2) pattern.

Methods-statistical↗

Chemical studies of H chondrites. 5: Temporal variations of sources

We report Cl-36 (301-kyr half-life) data obtained by accelerator mass spectrometry allowing nominal terrestrial ages to be determined for 39 Antarctic H4-6 chondrites for which contents of volatile trace elements are known. The compositional difference between these Antarctic meteorites and 58 non-Antarctic falls increases with terrestrial age and, using multivariate statistical techniques, becomes highly significant for Antarctic samples with ages greater than 50 kyr. The compositional difference is inconsistent with trivial causes such as weathering and seems to reflect differences in thermal histories of parent sources. Temporal source variations for the H chondrite flux on Earth thus exist not only on a short-term, 40 years, basis (Dodd et al., 1993) but also on a long-term, greater than 50 kyr, basis.

Michlovich, Edward S.↗