Search NASASearch

SEARCH · Search NASA

Results for “Statistical Algorithms”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

Design and Calibration of Autonomous Coherent Doppler Lidar for Space Missions

Developed a new algorithm for the simulation of three dimensional homogeneous turbulent velocity fields. For typical atmospheric conditions it is impossible to produce a simulated velocity field that simultaneously satisfy a given spatial correlation and the corresponding spatial spectrum because of spectral aliasing. The new algorithms produce a turbulent velocity field which has accurate spatial correlations which is required for performance predictions from space-based systems. Developed a new algorithm for extracting the spatial statistics of the atmospheric velocity field using coherent Doppler lidar. The performance of the algorithm was compared with past methods and the new algorithm produces useful results for space-based data, which was not possible before. Developed new methods for verification of the errors in ground-based and space-based Doppler lidar wind measurements. These new methods do not require independent in situ data. This is an important issue for the verification of space-based Doppler lidar measurements of the global wind field. The performance of the new algorithm was compared with past results for both space-based and ground-based operation. The new algorithm has the best performance and is the only algorithm that performed satisfactory for spacebased operation. The performance of coherent Doppler lidar for a space missions with various scanning geometries was determined using computer simulation which contained the effects of random instrumental velocity errors, wind shear, wind variability along the range-gate and from shot-to-shot, and random variations in atmospheric aerosol backscatter over the measurement volume. The bias in the velocity estimates was small and the accuracy in the is typically less than 0.5 m/s for high signal conditions. For a large number of shot per velocity estimate, the threshold signal level for acceptable estimates is proportional to the number of shots to the minus one half power. This agrees with previous results determined for ground-based measurements. The use of multi-element optical detectors for autonomous operation of coherent Doppler lidar was shown to be a very promising technique. Optimal detector geometries were determined by computer simulation of performance: for ground-based testing with a fixed calibration target and for space-based operation using the random surface returns. The effects of refractive turbulence on ground-based calibration of coherent Doppler lidar was determined by computer simulations and compared with theoretical predictions. New techniques were required to correctly predict performance for the focused beam geometry commonly used for verification of space-based operation. An improved velocity estimator was evaluated for space-based applications were signal shot measurements are used to produce vector wind measurements. This permits more accurate measurements when the signal level is not known a priori or not available from multiple shot measurements. The average Doppler lidar signal spectrum including the effects of velocity turbulence was derived and calculated. This permits new estimation algorithms for turbulence based on spectral estimates. In situ atmospheric measurements were conducted and analyzed using an instrumented kite-platform. This work helps provide the required in situ data for verification of Doppler lidar velocity statistics.

Frehlich, Rod G.

Land Surface Temperature Measurements form EOS MODIS Data

We have developed a physics-based land-surface temperature (LST) algorithm for simultaneously retrieving surface band-averaged emissivities and temperatures from day/night pairs of MODIS (Moderate Resolution Imaging Spectroradiometer) data in seven thermal infrared bands. The set of 14 nonlinear equations in the algorithm is solved with the statistical regression method and the least-squares fit method. This new LST algorithm was tested with simulated MODIS data for 80 sets of band-averaged emissivities calculated from published spectral data of terrestrial materials in wide ranges of atmospheric and surface temperature conditions. Comprehensive sensitivity and error analysis has been made to evaluate the performance of the new LST algorithm and its dependence on variations in surface emissivity and temperature, upon atmospheric conditions, as well as the noise-equivalent temperature difference (NE(Delta)T) and calibration accuracy specifications of the MODIS instrument. In cases with a systematic calibration error of 0.5%, the standard deviations of errors in retrieved surface daytime and nighttime temperatures fall between 0.4-0.5 K over a wide range of surface temperatures for mid-latitude summer conditions. The standard deviations of errors in retrieved emissivities in bands 31 and 32 (in the 10-12.5 micrometer IR spectral window region) are 0.009, and the maximum error in retrieved LST values falls between 2-3 K. Several issues related to the day/night LST algorithm (uncertainties in the day/night registration and in surface emissivity changes caused by dew occurrence, and the cloud cover) have been investigated. The LST algorithms have been validated with MODIS Airborne Simulator (MAS) dada and ground-based measurement data in two field campaigns conducted in Railroad Valley playa, NV in 1995 and 1996. The MODIS LST version 1 software has been delivered.

Wan, Zhengming

SeaWiFS technical report series. Volume 32: Level-3 SeaWiFS data products. Spatial and temporal binning algorithms

The level-3 data products from the Sea-viewing Wide Field-of-view Sensor (SeaWiFS) are statistical data sets derived from level-2 data. Each data set will be based on a fixed global grid of equal-area bins that are approximately 9 x 9 sq km. Statistics available for each bin include the sum and sum of squares of the natural logarithm of derived level-2 geophysical variables where sums are accumulated over a binning period. Operationally, products with binning periods of 1 day, 8 days, 1 month, and 1 year will be produced and archived. From these accumulated values and for each bin, estimates of the mean, standard deviation, median, and mode may be derived for each geophysical variable. This report contains two major parts: the first (Section 2) is intended as a users' guide for level-3 SeaWiFS data products. It contains an overview of level-0 to level-3 data processing, a discussion of important statistical considerations when using level-3 data, and details of how to use the level-3 data. The second part (Section 3) presents a comparative statistical study of several binning algorithms based on CZCS and moored fluorometer data. The operational binning algorithms were selected based on the results of this study.

Hooker, Stanford B.

High-Throughput Discovery Illuminates Design Principles and Limits for Long-Lived Charged Species in Organic Electrolytes

The chemical stability of charged molecules in all-organic redox flow batteries (RFBs) is required for the prolonged operation of these devices. Molecular engineering and electrolyte optimization are used to mitigate parasitic reactions and extend the lifetimes of the charge carriers. However, how much can structural variation extend the lifetime? To probe this query, we designed a high-throughput kinetic study of the radical cation of N-methylphenothiazinium, guided by statistical sampling and learning algorithms. Using Argonne’s autonomous discovery facility, we conducted over 6,000 kinetic experiments with robotic sample preparation, parallel kinetic measurements, and machine learning inputs, testing 188 solvent molecules selected from a space of over 540 candidates from 11 chemical classes. Algorithmic selections guided us to stable solvent candidates, which were further tested in high concentration with and without supporting electrolyte. Our findings reveal the inherent difficulty of exceeding the current state of the art through solvent variation. The desired stability is statistically rare and poorly predictable. Among the many tested, only three solvents significantly outperformed our baseline, acetonitrile─and none by more than a factor of 3─suggesting a general challenge in achieving the necessary techno-economic targets. Furthermore, we suggest that self-discharge through solvent homolysis is the cause of the observed limitations. Several structural motifs contribute to >1,000 h half-life stability including molecular simplicity, symmetry, oxidation complement, and strategic fluorination. Importantly, this workflow establishes effective assays for diagnosing and predicting oxidative stress for highly stable liquid electrolytes in all batteries.

Batteries

Study of photon correlation techniques for processing of laser velocimeter signals

The objective was to provide the theory and a system design for a new type of photon counting processor for low level dual scatter laser velocimeter (LV) signals which would be capable of both the first order measurements of mean flow and turbulence intensity and also the second order time statistics: cross correlation auto correlation, and related spectra. A general Poisson process model for low level LV signals and noise which is valid from the photon-resolved regime all the way to the limiting case of nonstationary Gaussian noise was used. Computer simulation algorithms and higher order statistical moment analysis of Poisson processes were derived and applied to the analysis of photon correlation techniques. A system design using a unique dual correlate and subtract frequency discriminator technique is postulated and analyzed. Expectation analysis indicates that the objective measurements are feasible.

Mayo, W. T., Jr.

SeaWiFS technical report series. Volume 4: An analysis of GAC sampling algorithms. A case study

The Sea-viewing Wide Field-of-view Sensor (SeaWiFS) instrument will sample at approximately a 1 km resolution at nadir which will be broadcast for reception by realtime ground stations. However, the global data set will be comprised of coarser four kilometer data which will be recorded and broadcast to the SeaWiFS Project for processing. Several algorithms for degrading the one kilometer data to four kilometer data are examined using imagery from the Coastal Zone Color Scanner (CZCS) in an effort to determine which algorithm would best preserve the statistical characteristics of the derived products generated from the one kilometer data. Of the algorithms tested, subsampling based on a fixed pixel within a 4 x 4 pixel array is judged to yield the most consistent results when compared to the one kilometer data products.

Yeh, Eueng-Nan

Detection and Modeling of High-Dimensional Thresholds for Fault Detection and Diagnosis

Many Fault Detection and Diagnosis (FDD) systems use discrete models for detection and reasoning. To obtain categorical values like oil pressure too high, analog sensor values need to be discretized using a suitablethreshold. Time series of analog and discrete sensor readings are processed and discretized as they come in. This task isusually performed by the wrapper code'' of the FDD system, together with signal preprocessing and filtering. In practice,selecting the right threshold is very difficult, because it heavily influences the quality of diagnosis. If a threshold causesthe alarm trigger even in nominal situations, false alarms will be the consequence. On the other hand, if threshold settingdoes not trigger in case of an off-nominal condition, important alarms might be missed, potentially causing hazardoussituations. In this paper, we will in detail describe the underlying statistical modeling techniques and algorithm as well as the Bayesian method for selecting the most likely shape and its parameters. Our approach will be illustrated by several examples from the Aerospace domain.

Computer Systems

AutoBayes Program Synthesis System System Internals

This lecture combines the theoretical background of schema based program synthesis with the hands-on study of a powerful, open-source program synthesis system (Auto-Bayes). Schema-based program synthesis is a popular approach toward program synthesis. The lecture will provide an introduction into this topic and discuss how this technology can be used to generate customized algorithms. The synthesis of advanced numerical algorithms requires the availability of a powerful symbolic (algebra) system. Its task is to symbolically solve equations, simplify expressions, or to symbolically calculate derivatives (among others) such that the synthesized algorithms become as efficient as possible. We will discuss the use and importance of the symbolic system for synthesis. Any synthesis system is a large and complex piece of code. In this lecture, we will study Autobayes in detail. AutoBayes has been developed at NASA Ames and has been made open source. It takes a compact statistical specification and generates a customized data analysis algorithm (in C/C++) from it. AutoBayes is written in SWI Prolog and many concepts from rewriting, logic, functional, and symbolic programming. We will discuss the system architecture, the schema libary and the extensive support infra-structure. Practical hands-on experiments and exercises will enable the student to get insight into a realistic program synthesis system and provides knowledge to use, modify, and extend Autobayes.

Statistical Algorithms

Computing approximate random Delta v magnitude probability densities

This paper describes the development and use of an algorithm to compute approximate statistics of the magnitude of a single random trajectory correction maneuver (TCM) Delta v vector. The TCM Delta v vector is modeled as a three component Cartesian vector each of whose components is a random variable having a normal (Gaussian) distribution with zero mean and possibly unequal standard deviations. The algorithm uses these standard deviations as input to produce approximations to (1) the mean and standard deviation of the magnitude of Delta v, (2) points of the probability density function of the magnitude of Delta v, and (3) points of the cumulative and inverse cumulative distribution functions of Delta v. The approximates are based on Monte Carlo techniques developed in a previous paper by the author and extended here. The algorithm described is expected to be useful in both pre-flight planning and in-flight analysis of maneuver propellant requirements for space missions.

Chadwick, C.

A Coulomb collision algorithm for weighted particle simulations

A binary Coulomb collision algorithm is developed for weighted particle simulations employing Monte Carlo techniques. Charged particles within a given spatial grid cell are pair-wise scattered, explicitly conserving momentum and implicitly conserving energy. A similar algorithm developed by Takizuka and Abe (1977) conserves momentum and energy provided the particles are unweighted (each particle representing equal fractions of the total particle density). If applied as is to simulations incorporating weighted particles, the plasma temperatures equilibrate to an incorrect temperature, as compared to theory. Using the appropriate pairing statistics, a Coulomb collision algorithm is developed for weighted particles. The algorithm conserves energy and momentum and produces the appropriate relaxation time scales as compared to theoretical predictions. Such an algorithm is necessary for future work studying self-consistent multi-species kinetic transport.

Miller, Ronald H.

Context distribution estimation for contextual classification of multispectral image data

A classification algorithm incorporating contextual information in a general, statistical manner is presented. Methods are investigated for obtaining adequate estimates of the context distribution (a statistical characterization of context) upon which the classification algorithm depends. Finally, a method of estimating optimal algorithm parameters prior to performing preliminary classifications is explored.

Tilton, J. C.

A new algorithm for attitude-independent magnetometer calibration

A new algorithm is developed for inflight magnetometer bias determination without knowledge of the attitude. This algorithm combines the fast convergence of a heuristic algorithm currently in use with the correct treatment of the statistics and without discarding data. The algorithm performance is examined using simulated data and compared with previous algorithms.

Alonso, Roberto

A passive microwave technique for estimating rainfall and vertical structure information from space. Part 1: Algorithm description

This paper describes a multichannel physical approach for retrieving rainfall and vertical structure information from satellite-based passive microwave observations. The algorithm makes use of statistical inversion techniques based upon theoretically calculated relations between rainfall rates and brightness temperatures. Potential errors introduced into the theoretical calculations by the unknown vertical distribution of hydrometeors are overcome by explicity accounting for diverse hydrometeor profiles. This is accomplished by allowing for a number of different vertical distributions in the theoretical brightness temperature calculations and requiring consistency between the observed and calculated brightness temperatures. This paper will focus primarily on the theoretical aspects of the retrieval algorithm, which includes a procedure used to account for inhomogeneities of the rainfall within the satellite field of view as well as a detailed description of the algorithm as it is applied over both ocean and land surfaces. The residual error between observed and calculated brightness temperatures is found to be an important quantity in assessing the uniqueness of the solution. It is further found that the residual error is a meaningful quantity that can be used to derive expected accuracies from this retrieval technique. Examples comparing the retrieved results as well as the detailed analysis of the algorithm performance under various circumstances are the subject of a companion paper.

Kummerow, Christian

ERBE Geographic Scene and Monthly Snow Data

The Earth Radiation Budget Experiment (ERBE) is a multisatellite system designed to measure the Earth's radiation budget. The ERBE data processing system consists of several software packages or sub-systems, each designed to perform a particular task. The primary task of the Inversion Subsystem is to reduce satellite altitude radiances to fluxes at the top of the Earth's atmosphere. To accomplish this, angular distribution models (ADM's) are required. These ADM's are a function of viewing and solar geometry and of the scene type as determined by the ERBE scene identification algorithm which is a part of the Inversion Subsystem. The Inversion Subsystem utilizes 12 scene types which are determined by the ERBE scene identification algorithm. The scene type is found by combining the most probable cloud cover, which is determined statistically by the scene identification algorithm, with the underlying geographic scene type. This Contractor Report describes how the geographic scene type is determined on a monthly basis.

Coleman, Lisa H.

The Precipitation Rate Retrieval Algorithms for the GPM Dual-frequency Precipitation Radar

In this paper, precipitation rate retrieval algorithms for the Global Precipitation Measurement mission's Dual-frequency Precipitation Radar (DPR) are developed. The DPR consists of a Ku-band radar (KuPR; 13.6 GHz) and a Ka-band radar (KaPR; 35.5 GHz). For the KuPR, an algorithm similar to the Tropical Rainfall Measuring Mission's Precipitation Radar algorithm is developed, with the relation between precipitation rate R and massweighted mean diameter D m (R−D m relation) replacing the relation between the specific attenuation k and effective radar reflectivity factor Z e . The R−D m relation can also be applied to the KaPR and dual-frequency algorithms. In both the single-frequency and dual-frequency algorithms, the forward retrieval method is applied with an assumed adjustment factor for the R−D m relation (ε) and the results are evaluated to select the best value of ε. The advantages of the dual-frequency algorithm are the availability of the dual-frequency surface reference technique and the ZfKa method, which is a method to use the attenuation-corrected radar reflectivity factor Z f of KaPR, to select ε as well as the possibility to selectively use measurements from KuPR or KaPR. This paper also describes the derivation of the scattering table and the R−D m relation as well as the procedure for non-uniform beam filling correction in detail. The outputs are then statistically analyzed to demonstrate algorithm performance.

precipitation radar

G-Mapper: Learning a Cover in the Mapper Construction

The Mapper algorithm is a visualization technique in topological data analysis (TDA) that outputs a graph reflecting the structure of a given dataset. However, the Mapper algorithm requires tuning several parameters in order to generate a “nice” Mapper graph. This paper focuses on selecting the cover parameter. We present an algorithm that optimizes the cover of a Mapper graph by splitting a cover repeatedly according to a statistical test for normality. Our algorithm is based on G-means clustering, which searches for the optimal number of clusters in 𝑘-means by iteratively applying the Anderson–Darling test. Our splitting procedure employs a Gaussian mixture model to carefully choose the cover according to the distribution of the given data. In conclusion, experiments for synthetic and real-world datasets demonstrate that our algorithm generates covers so that the Mapper graphs retain the essence of the datasets, while also running significantly faster than a previous iterative method.

G-means clustering

Measurement of top-quark pair production in association with charm quarks in proton–proton collisions at √s = 13 TeV with the ATLAS detector

Inclusive cross-sections or top-quark pair production in association with charm quarks are measured with proton-proton collision data at a center-of-mass energy of 13 TeV corresponding to an integrated luminosity of 140 fb -1 , collected with the ATLAS experiment at LHC between 2015 and 2018. The measurements are performed by requiring one or two charged leptons (electrons and muons), two b-tagged jets, and at least one additional jet in the final state. A custom flavor-tagging algorithm is employed for the simultaneous identification of b-jets and c-jets. In a fiducial phase space that replicates the acceptance of the ATLAS detector, the cross-sections for $t\bar{t}$ + ≥ 2c and $t\bar{t}$ + 1c production are measured to be $1.28^{+0.27}_{-0.24}$ pb and $6.4^{+1.0}_{-0.9}$ pb, respectively. The measurements are primarily limited by uncertainties in the modeling of inclusive $t\bar{t}$ and $t\bar{t}$ + $b\bar{b}$ production, in the calibration of the flavor-tagging algorithm, and by data statistics. Cross-section predictions from various $t\bar{t}$ simulations are largely consistent with the measured cross-section values, though all underpredict the observed values by 0.5 to 2.0 standard deviations. In a phase-space volume without requirements on the $t\bar{t}$ decay products and the jet multiplicity, the cross-section ratios of $t\bar{t}$ + ≥ 2c and $t\bar{t}$ + 1c to total $t\bar{t}$ + jets production are determined to be (1.23 ± 0.25)% and (8.8 ± 1.3)%.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS