Statistical algorithms and computer programs for analysis of multi-spectral observations Final report
Statistical algorithms and computer programs for remote sensor multispectral data analysis
SEARCH · Search NASA
Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.
Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.
Statistical algorithms and computer programs for remote sensor multispectral data analysis
The construction of the Space Telescope Guide Star Catalog from digitized Schmidt survey plates covering the entire sky is described. In order to provide sufficient pointing information, the Guide Star Selection System has to contain a catalog of 20 million guide star candidates in the range of 9.0 to 14.5 visual magnitudes. An image inventory process extracts the relevant object data; after photometric and astrometric calibration the data are screened to arrive at suitable guide star candidates. The selection process for a given target takes into account geometric and photometric parameters of the field, scheduling information, and acquisition probabilities for each guide star pair.
The Algorithm Simulation Test and Evaluation Program (ASTEP) is a modular computer program developed for the purpose of testing and evaluating methods of processing remotely sensed multispectral scanner earth resources data. ASTEP is written in FORTRAND V on the UNIVAC 1110 under the EXEC 8 operating system and may be operated in either a batch or interactive mode. The program currently contains over one hundred subroutines consisting of data classification and display algorithms, statistical analysis algorithms, utility support routines, and feature selection capability. The current program can accept data in LARSC1, LARSC2, ERTS, and Universal formats, and can output processed image or data tapes in Universal format.
The optical and microphysical structure of warm boundary layer marine clouds is of fundamental importance for understanding a variety of cloud radiation and precipitation processes. With the advent of MODIS (Moderate Resolution Imaging Spectroradiometer) on the NASA EOS Terra and Aqua platforms, simultaneous global/daily 1km retrievals of cloud optical thickness and effective particle size are provided, as well as the derived water path. In addition, the cloud product (MOD06/MYD06 for MODIS Terra and Aqua, respectively) provides separate effective radii results using the l.6, 2.1, and 3.7 ~m spectral channels. Cloud retrieval statistics are highly sensitive to how a pixel identified as being "notclear" by a cloud mask (e.g., the MOD35/MYD35 product) is determined to be useful for an optical retrieval based on a 1-D cloud model. The Collection 5 MODIS retrieval algorithm removed pixels associated with cloud'edges as well as ocean pixels with partly cloudy elements in the 250m MODIS cloud mask - part of the so-called Clear Sky Restoral (CSR) algorithm. Collection 6 attempts retrievals for those two pixel populations, but allows a user to isolate or filter out the populations via CSR pixel-level Quality Assessment (QA) assignments. In this paper, using the preliminary Collection 6 MOD06 product, we present global and regional statistical results of marine warm cloud retrieval sensitivities to the cloud edge and 250m partly cloudy pixel populations. As expected, retrievals for these pixels are generally consistent with a breakdown of the ID cloud model. While optical thickness for these suspect pixel populations may have some utility for radiative studies, the retrievals should be used with extreme caution for process and microphysical studies.
A method for predicting the audibility of an arbitrary time-varying noise (signal) in the presence of masking noise has been developed. The statistical audibility prediction (SAP) method relies on the specific loudness, or loudness perceived through the individual auditory filters, for accurate statistical estimation of audibility vs. time. More recent development has focused on derivation and inclusion of frequency-dependent correction factors in SAP’s model to account for the ability to hear signals below the level of the masking noise. Audibility prediction vs. time is intuitive since it captures changes in audibility with time as it occurs, which is critical for the study of human response to noise. Concurrently time-frequency prediction of audibility may also provide valuable information about the root cause(s) for audibility useful for the design and operation of sources of noise. Empirical data, gathered under a three-alternative forced-choice (3AFC) test paradigm for low-frequency sound, has been used to examine the accuracy of SAP.
The mortality related to cervical cancer can be substantially reduced through early detection and treatment. However, current detection techniques, such as Pap smear and colposcopy, fail to achieve a concurrently high sensitivity and specificity. In vivo fluorescence spectroscopy is a technique which quickly, noninvasively and quantitatively probes the biochemical and morphological changes that occur in precancerous tissue. A multivariate statistical algorithm was used to extract clinically useful information from tissue spectra acquired from 361 cervical sites from 95 patients at 337-, 380-, and 460-nm excitation wavelengths. The multivariate statistical analysis was also employed to reduce the number of fluorescence excitation-emission wavelength pairs required to discriminate healthy tissue samples from precancerous tissue samples. The use of connectionist methods such as multilayered perceptrons, radial basis function (RBF) networks, and ensembles of such networks was investigated. RBF ensemble algorithms based on fluorescence spectra potentially provide automated and near real-time implementation of precancer detection in the hands of nonexperts. The results are more reliable, direct, and accurate than those achieved by either human experts or multivariate statistical algorithms.
In this paper, we propose a new heuristic scheduling algorithm based on the statistical analysis of the cumulative frequency distribution of operations among control steps. It has a tendency of escaping from local minima and therefore reaching a globally optimal solution. The presented algorithm considers the real world constraints such as chained operations, multicycle operations, and pipelined data paths. The result of the experiment shows that it gives optimal solutions, even though it is greedy in nature.
Parameters of spectral envelope obtained statistically. Algorithm for digital compression of speech signals encodes power spectral density of each short interval of speech by use of quantiles or order statistics. Purpose to reduce bit rate and bandwidth required for transmission. When fully developed, quantile vocoder - speech-encoding system based on new algorithm - expected moderately complicated compared with other speech-encoding systems and reproduce high-quality speech from code transmitted at relatively low bit rates. Speech signal treated mathematically as though amplitude spectrum stationary during short coding intervals, called "windows". Window duration of 20 or 35 ms, chosen as compromise between frequency resolution and time resolution. During each window, short-time amplitude and power spectra found by sampling at high rate (typically 10 kHz) and taking fast Fourier transforms (FFT's).
The goal of this project is to develop reliable statistical algorithms for on-line analysis of physiologic and neurobehavioral variables monitored during long-duration space missions. Maintenance of physiologic and neurobehavioral homeostasis during long-duration space missions is crucial for ensuring optimal crew performance. If countermeasures are not applied, alterations in homeostasis will occur in nearly all-physiologic systems. During such missions data from most of these systems will be either continually and/or continuously monitored. Therefore, if these data can be analyzed as they are acquired and the status of these systems can be continually assessed, then once alterations are detected, appropriate countermeasures can be applied to correct them. One of the most important physiologic systems in which to maintain homeostasis during long-duration missions is the circadian system. To detect and treat alterations in circadian physiology during long duration space missions requires development of: 1) a ground-based protocol to assess the status of the circadian system under the light-dark environment in which crews in space will typically work; and 2) appropriate statistical methods to make this assessment. The protocol in Project 1, Circadian Entrainment, Sleep-Wake Regulation and Neurobehavioral will study human volunteers under the simulated light-dark environment of long-duration space missions. Therefore, we propose to develop statistical models to characterize in near real time circadian and neurobehavioral physiology under these conditions. The specific aims of this project are to test the hypotheses that: 1) Dynamic statistical methods based on the Kronauer model of the human circadian system can be developed to estimate circadian phase, period, amplitude from core-temperature data collected under simulated light- dark conditions of long-duration space missions. 2) Analytic formulae and numerical algorithms can be developed to compute the error in the estimates of circadian phase, period and amplitude determined from the data in Specific Aim 1. 3) Statistical models can detect reliably in near real- time (daily) significant alternations in the circadian physiology of individual subjects by analyzing the circadian and neurobehavioral data collected in Project 1. 4) Criteria can be developed using the Kronauer model and the recently developed Jewett model of cognitive -performance and subjective alertness to define altered circadian and neurobehavioral physiology and to set conditions for immediate administration of countermeasures.
The Launch Services Program at NASA’s Kennedy Space Center (KSC) is the primary gate for acquiring commercial vehicles to provide a cost effective ride to space for NASA spacecraft. With the lunar Gateway, more human tended elements are planned for launch. One challenge facing the space industry is the proliferation of communication and science transmitters at frequencies beyond the qualification of space avionics and instruments. Studies have ensued to examine the intricacies of performing radiated susceptibility testing and analysis above 18 GHz. Changes in the launch vehicle communications interface to the range have also lead to new launch vehicle antenna systems and more reliance on GPS and telemetry systems. Finally, research initiated at KSC in the area of predicted electric field distributions in launch vehicle payload fairings have spawned Small Business Technology Transfer initiatives for industry to investigate statistical algorithm and computational improvements in large payload fairing modeling of transmitters at frequencies in the GHz range. These topics, along with electromagnetic compatibility testing for launch vehicles will be discussed.
We have developed an algorithm that retrieves wind speed under rain using C-hand and X-band channels of passive microwave satellite radiometers. The spectral difference of the brightness temperature signals due to wind or rain allows to find channel combinations that are sufficiently sensitive to wind speed but little or not sensitive to rain. We &ve trained a statistical algorithm that applies under hurricane conditions and is able to measure wind speeds in hurricanes to an estimated accuracy of about 2 m/s. We have also developed a global algorithm, that is less accurate but can be applied under all conditions. Its estimated accuracy is between 2 and 5 mls, depending on wind speed and rain rate. We also extend the wind speed region in our model for the wind induced sea surface emissivity from currently 20 m/s to 40 mls. The data indicate that the signal starts to saturate above 30 mls. Finally, we make an assessment of the performance of wind direction retrievals from polarimetric radiometers as function of wind speed and rain rate
Signal/Noise Ratio Meter measures ratio of signal power to noise power in input that contains both signal and noise. Signal and noise first filtered and normalized in analog circuitry, then digitized and sampled. Performance of SNR meter determined by statistical algorithm chosen for analysis of samples.
In 13 years of operation, IUE has gathered approximately 5000 spectra of almost 600 Active Galactic Nuclei (AGN). In order to undertake AGN studies which require large amounts of data, we are consistently reducing this entire archive and creating a homogeneous, easy-to-use database. First, the spectra are extracted using the Optimal extraction algorithm. Continuum fluxes are then measured across predefined bands, and line fluxes are measured with a multi-component fit. These results, along with source information such as redshifts and positions, are placed in the IUEAGN relational database. Analysis algorithms, statistical tests, and plotting packages run within the structure, and this flexible database can accommodate future data when they are released. This archival approach has already been used to survey line and continuum variability in six bright Seyfert 1s and rapid continuum variability in 14 blazars. Among the results that could only be obtained using a large archival study is evidence that blazars show a positive correlation between degree of variability and apparent luminosity, while Seyfert 1s show an anti-correlation. This suggests that beaming dominates the ultraviolet properties for blazars, while thermal emission from an accretion disk dominates for Seyfert 1s. Our future plans include a survey of line ratios in Seyfert 1s, to be fitted with photoionization models to test the models and determine the range of temperatures, densities and ionization parameters. We will also include data from IRAS, Einstein, EXOSAT, and ground-based telescopes to measure multi-wavelength correlations and broadband spectral energy distributions.
Implementation of computer programs based on multivariate statistical algorithms makes possible obtaining reliable information from long data vectors that contain large amounts of extraneous information, for example, noise and/or analytes that we do not wish to control. Three examples are described. Each of these applications requires the use of techniques characteristic of modern analytical chemistry. The first example, using a quantitative or analytical model, describes the determination of the acid dissociation constant for 2,2'-pyridyl thiophene using archived data. The second example describes an investigation to determine the active biocidal species of iodine in aqueous solutions. The third example is taken from a research program directed toward advanced fiber-optic chemical sensors. The second and third examples require heuristic or empirical models.
Given the substantial radiative effects of cirrus clouds and the need to validate cirrus cloud mass in climate models, it is important to measure the global distribution of cirrus properties with satellite remote sensing. Existing cirrus remote sensing techniques, such as solar reflectance methods, measure cirrus ice water path (IWP) rather indirectly and with limited accuracy. Submillimeter/wave radiometry is an independent method of cirrus remote sensing based on ice particles scattering the upwelling radiance emitted by the lower atmosphere. A new aircraft instrument, the Far Infrared Sensor for Cirrus (FIRSC), is described. The FIRSC employs a Fourier Transform Spectrometer (FTS). which measures the upwelling radiance across the whole submillimeter region (0.1 1.0-mm wavelength). This wide spectral coverage gives high sensitivity to most cirrus particle sizes and allows accurate determination of the characteristic particle size. Radiative transfer modeling is performed to analyze the capabilities of the submillimeter FTS technique. A linear inversion analysis is done to show that cirrus IWP, particle size, and upper-tropospheric temperature and water vapor may be accurately measured, A nonlinear statistical algorithm is developed using a database of 20000 spectra simulated by randomly varying most relevant cirrus and atmospheric parameters. An empirical orthogonal function analysis reduces the 500-point spectrum (20 - 70/cm) to 15 "pseudo-channels" that are then input to a neural network to retrieve cirrus IWP and median particle diameter. A Monte Carlo accuracy study is performed with simulated spectra having realistic noise. The retrieval errors are low for IWP (rms less than a factor of 1.5) and for particle sizes (rins less than 30%) for IWP greater than 5 g/sq m and a wide range of median particle sizes. This detailed modeling indicates that there is good potential to accurately measure cirrus properties with a submillimeter FTS.
The accurate determination of upper ocean apparent optical properties (AOPs) is essential for the vicarious calibration of the Sea-viewing Wide Field-of-view Sensor (SeaWiFS) instrument and the validation of the derived data products. To evaluate the importance of data analysis methods upon derived AOP values, the Second Data Analysis Round Robin (DARR-00) activity was planned during the latter half of 1999 and executed during March 2000. The focus of the study was the intercomparison of several standard AOP parameters: (1) the upwelled radiance immediately below the sea surface, L(sub u)(0(-),lambda); (2) the downward irradiance immediately below the sea surface, E(sub d)(0(-),lambda); (3) the diffuse attenuation coefficients from the upwelling radiance and the downward irradiance profiles, L(sub L)(lambda) and K(sub d)(lambda), respectively; (4) the incident solar irradiance immediately above the sea surface, E(sub d)(0(+),lambda); (5) the remote sensing reflectance, R(sub rs)(lambda); (6) the normalized water-leaving radiance, [L(sub W)(lambda)](sub N); (7) the upward irradiance immediately below the sea surface, E(sub u)(0(-)), which is used with the upwelled radiance to derive the nadir Q-factor immediately below the sea surface, Q(sub n)(0(-),lambda); and (8) ancillary parameters like the solar zenith angle, theta, and the total chlorophyll concentration, C(sub Ta), derived from the optical data through statistical algorithms. In the results reported here, different methodologies from three research groups were applied to an identical set of 40 multispectral casts in order to evaluate the degree to which differences in data analysis methods influence AOP estimation, and whether any general improvements can be made. The overall results of DARR-00 are presented in Chapter 1 and the individual methods used by the three groups and their data processors are presented in Chapters 2-4.
We report on several projects in the field of computational astrobiology, which is devoted to advancing our understanding of the origin, evolution and distribution of life in the Universe using theoretical and computational tools. Research projects included modifying existing computer simulation codes to use efficient, multiple time step algorithms, statistical methods for analysis of astrophysical data via optimal partitioning methods, electronic structure calculations on water-nuclei acid complexes, incorporation of structural information into genomic sequence analysis methods and calculations of shock-induced formation of polycylic aromatic hydrocarbon compounds.
Interpreting data from large-scale protein interaction experiments has been a challenging task because of the widespread presence of random false positives. Here, we present a network-based statistical algorithm that overcomes this difficulty and allows us to derive functions of unannotated proteins from large-scale interaction data. Our algorithm uses the insight that if two proteins share significantly larger number of common interaction partners than random, they have close functional associations. Analysis of publicly available data from Saccharomyces cerevisiae reveals >2,800 reliable functional associations, 29% of which involve at least one unannotated protein. By further analyzing these associations, we derive tentative functions for 81 unannotated proteins with high certainty. Our method is not overly sensitive to the false positives present in the data. Even after adding 50% randomly generated interactions to the measured data set, we are able to recover almost all (approximately 89%) of the original associations.