Search NASA⌕ Search

SEARCH · Search NASA

Results for “High dimensional data,”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 163 records · Page 9

Efficient Methods to Assimilate Satellite Retrievals Based on Information Content: Suboptimal Retrieval Assimilation - Part 2

One of the outstanding problems in data assimilation has been and continues to be how best to utilize satellite data while balancing the tradeoff between accuracy and computational cost. A number of weather prediction centers have recently achieved remarkable success in improving their forecast skill by changing the method by which satellite data are assimilated into the forecast model from the traditional approach of assimilating retrievals to the direct assimilation of radiances in a variational framework. The operational implementation of such a substantial change in methodology involves a great number of technical details, e.g., pertaining to quality control procedures, systematic error correction techniques, and tuning of the statistical parameters in the analysis algorithm. Although there are clear theoretical advantages to the direct radiance assimilation approach, it is not obvious at all to what extent the improvements that have been obtained so far can be attributed to the change in methodology, or to various technical aspects of the implementation. The issue is of interest because retrieval assimilation retains many practical and logistical advantages which may become even more significant in the near future when increasingly high-volume data sources become available. The central question we address here is: how much improvement can we expect from assimilating radiances rather than retrievals, all other things being equal? We compare the two approaches in a simplified one-dimensional theoretical framework, in which problems related to quality control and systematic error correction are conveniently absent. By assuming a perfect radiative transfer model and perfect knowledge of radiance and background error covariances, we are able to formulate a nonlinear local error analysis for each assimilation method. Direct radiance assimilation is optimal in this idealized context, while the traditional method of assimilating retrievals is suboptimal because it ignores the cross-covariances between background errors and retrieval errors. We show that interactive retrieval assimilation (where the same background used for assimilation is also used in the retrieval step) is equivalent to direct assimilation of radiances with suboptimal analysis weights. We illustrate and extend these theoretical arguments with several one-dimensional assimilation experiments, where we estimate vertical atmospheric profiles using simulated data from both the High-resolution InfraRed Sounder 2 (HIRS2) and the future Atmospheric InfraRed Sounder (AIRS).

Joiner, J.↗

Putting Priors in Mixture Density Mercer Kernels

This paper presents a new methodology for automatic knowledge driven data mining based on the theory of Mercer Kernels, which are highly nonlinear symmetric positive definite mappings from the original image space to a very high, possibly infinite dimensional feature space. We describe a new method called Mixture Density Mercer Kernels to learn kernel function directly from data, rather than using predefined kernels. These data adaptive kernels can en- code prior knowledge in the kernel using a Bayesian formulation, thus allowing for physical information to be encoded in the model. We compare the results with existing algorithms on data from the Sloan Digital Sky Survey (SDSS). The code for these experiments has been generated with the AUTOBAYES tool, which automatically generates efficient and documented C/C++ code from abstract statistical model specifications. The core of the system is a schema library which contains template for learning and knowledge discovery algorithms like different versions of EM, or numeric optimization methods like conjugate gradient methods. The template instantiation is supported by symbolic- algebraic computations, which allows AUTOBAYES to find closed-form solutions and, where possible, to integrate them into the code. The results show that the Mixture Density Mercer-Kernel described here outperforms tree-based classification in distinguishing high-redshift galaxies from low- redshift galaxies by approximately 16% on test data, bagged trees by approximately 7%, and bagged trees built on a much larger sample of data by approximately 2%.

Srivastava, Ashok N.↗

A functional video-based anthropometric measuring system

A high-speed anthropometric three dimensional measurement system using the Selcom Selspot motion tracking instrument for visual data acquisition is discussed. A three-dimensional scanning system was created which collects video, audio, and performance data on a single standard video cassette recorder. Recording rates of 1 megabit per second for periods of up to two hours are possible with the system design. A high-speed off-the-shelf motion analysis system for collecting optical information as used. The video recording adapter (VRA) is interfaced to the Selspot data acquisition system.

Nixon, J. H.↗

Discovering System Health Anomalies Using Data Mining Techniques

We present a data mining framework for the analysis and discovery of anomalies in high-dimensional time series of sensor measurements that would be found in an Integrated System Health Monitoring system. We specifically treat the problem of discovering anomalous features in the time series that may be indicative of a system anomaly, or in the case of a manned system, an anomaly due to the human. Identification of these anomalies is crucial to building stable, reusable, and cost-efficient systems. The framework consists of an analysis platform and new algorithms that can scale to thousands of sensor streams to discovers temporal anomalies. We discuss the mathematical framework that underlies the system and also describe in detail how this framework is general enough to encompass both discrete and continuous sensor measurements. We also describe a new set of data mining algorithms based on kernel methods and hidden Markov models that allow for the rapid assimilation, analysis, and discovery of system anomalies. We then describe the performance of the system on a real-world problem in the aircraft domain where we analyze the cockpit data from aircraft as well as data from the aircraft propulsion, control, and guidance systems. These data are discrete and continuous sensor measurements and are dealt with seamlessly in order to discover anomalous flights. We conclude with recommendations that describe the tradeoffs in building an integrated scalable platform for robust anomaly detection in ISHM applications.

Sriastava, Ashok, N.↗

Design of QMF (Quadrature Mirror Filter) in spatial domain and edge encoding

Simoncelli and Adelson have extended the one dimensional Quadrature Mirror Filter (QMF) to two dimensions with hexagon symmetry and three dimensional spatio-temporal extensions with rhombic-duodecahedray symmetry. Jain and Crochiere presented an excellent QMF design technique in the time domain. It is proposed to extend the design of a two dimensional QMF over a rectangular lattice in the spatial domain based primarily on the extension of the idea of Jain and Crochiere. In addition, the design will investigate the use of two dimensional Z-transformations. Since this proposed QMF is intended for the applications in image processing, all the important and interesting engineering issues will be addressed throughout the development phase. The design of a two dimensional QMF is discussed. The motivation is to achieve an extremely high data compression ratio. It is entirely possible to achieve dramatic results when pattern recognition techniques are employed. The final goal is the demonstration of extremely high data compression ratios using NASA pictures.

Wang, Paul P.↗

Free-Stream Turbulence Intensity in the Langley 14- by 22-Foot Subsonic Tunnel

An investigation was conducted using hot-wire anemometry to determine the turbulence intensity levels in the test section of the Langley 14- by 22-Foot Subsonic Tunnel in the closed or walls-down configuration. This study was one component of the three-dimensional High-Lift Flow Physics experiment designed to provide code validation data. Turbulence intensities were measured during two stages of the study. In the first stage, the free-stream turbulence levels were measured before and after a change was made to the floor suction surface of the wind tunnel s boundary layer removal system. The results indicated that the new suction surface at the entrance to the test section had little impact on the turbulence intensities. The second stage was an overall flow quality survey of the empty tunnel including measurements of the turbulence levels at several vertical and streamwise locations. Results indicated that the turbulence intensity is a function of tunnel dynamic pressure and the location in the test section. The general shape of the frequency spectrum is fairly consistent throughout the wind tunnel, changing mostly in amplitude (also slightly with frequency) with change in condition and location.

Neuhart, Dan H.↗

Aerodynamic Simulation of Ice Accretion on Airfoils

This report describes recent improvements in aerodynamic scaling and simulation of ice accretion on airfoils. Ice accretions were classified into four types on the basis of aerodynamic effects: roughness, horn, streamwise, and spanwise ridge. The NASA Icing Research Tunnel (IRT) was used to generate ice accretions within these four types using both subscale and full-scale models. Large-scale, pressurized windtunnel testing was performed using a 72-in.- (1.83-m-) chord, NACA 23012 airfoil model with high-fidelity, three-dimensional castings of the IRT ice accretions. Performance data were recorded over Reynolds numbers from 4.5 x 10(exp 6) to 15.9 x 10(exp 6) and Mach numbers from 0.10 to 0.28. Lower fidelity ice-accretion simulation methods were developed and tested on an 18-in.- (0.46-m-) chord NACA 23012 airfoil model in a small-scale wind tunnel at a lower Reynolds number. The aerodynamic accuracy of the lower fidelity, subscale ice simulations was validated against the full-scale results for a factor of 4 reduction in model scale and a factor of 8 reduction in Reynolds number. This research has defined the level of geometric fidelity required for artificial ice shapes to yield aerodynamic performance results to within a known level of uncertainty and has culminated in a proposed methodology for subscale iced-airfoil aerodynamic simulation.

Broeren, Andy P.↗

Improved Interactive Medical-Imaging System

An improved computational-simulation system for interactive medical imaging has been invented. The system displays high-resolution, three-dimensional-appearing images of anatomical objects based on data acquired by such techniques as computed tomography (CT) and magnetic-resonance imaging (MRI). The system enables users to manipulate the data to obtain a variety of views for example, to display cross sections in specified planes or to rotate images about specified axes. Relative to prior such systems, this system offers enhanced capabilities for synthesizing images of surgical cuts and for collaboration by users at multiple, remote computing sites.

Ross, Muriel D.↗

An Ensemble Approach to Building Mercer Kernels with Prior Information

This paper presents a new methodology for automatic knowledge driven data mining based on the theory of Mercer Kernels, which are highly nonlinear symmetric positive definite mappings from the original image space to a very high, possibly dimensional feature space. we describe a new method called Mixture Density Mercer Kernels to learn kernel function directly from data, rather than using pre-defined kernels. These data adaptive kernels can encode prior knowledge in the kernel using a Bayesian formulation, thus allowing for physical information to be encoded in the model. Specifically, we demonstrate the use of the algorithm in situations with extremely small samples of data. We compare the results with existing algorithms on data from the Sloan Digital Sky Survey (SDSS) and demonstrate the method's superior performance against standard methods. The code for these experiments has been generated with the AUTOBAYES tool, which automatically generates efficient and documented C/C++ code from abstract statistical model specifications. The core of the system is a schema library which contains templates for learning and knowledge discovery algorithms like different versions of EM, or numeric optimization methods like conjugate gradient methods. The template instantiation is supported by symbolic-algebraic computations, which allows AUTOBAYES to find closed-form solutions and, where possible, to integrate them into the code.

Srivastava, Ashok N.↗

Convective dynamics - Panel report

Aspects of highly organized forms of deep convection at midlatitudes are reviewed. Past emphasis in field work and cloud modeling has been directed toward severe weather as evidenced by research on tornadoes, hail, and strong surface winds. A number of specific issues concerning future thrusts, tactics, and techniques in convective dynamics are presented. These subjects include; convective modes and parameterization, global structure and scale interaction, convective energetics, transport studies, anvils and scale interaction, and scale selection. Also discussed are analysis workshops, four-dimensional data assimilation, matching models with observations, network Doppler analyses, mesoscale variability, and high-resolution/high-performance Doppler. It is also noted, that, classical surface measurements and soundings, flight-level research aircraft data, passive satellite data, and traditional photogrammetric studies are examples of datasets that require assimilation and integration.

Carbone, Richard↗

A new method for mapping multidimensional data to lower dimensions

A multispectral mapping method is proposed which is based on the new concept of BEND (Bidimensional Effective Normalised Difference). The method, which involves taking one sample point at a time and finding the interrelationships between its features, is found very economical from the point of view of storage and processing time. It has good dimensionality reduction and clustering properties, and is highly suitable for computer analysis of large amounts of data. The transformed values obtained by this procedure are suitable for either a planar 2-space mapping of geological sample points or for making grayscale and color images of geo-terrains. A few examples are given to justify the efficacy of the proposed procedure.

Gowda, K. C.↗

Application of Machine Learning Algorithms to the Study of Noise Artifacts in Gravitational-Wave Data

The sensitivity of searches for astrophysical transients in data from the Laser Interferometer Gravitationalwave Observatory (LIGO) is generally limited by the presence of transient, non-Gaussian noise artifacts, which occur at a high-enough rate such that accidental coincidence across multiple detectors is non-negligible. Furthermore, non-Gaussian noise artifacts typically dominate over the background contributed from stationary noise. These "glitches" can easily be confused for transient gravitational-wave signals, and their robust identification and removal will help any search for astrophysical gravitational-waves. We apply Machine Learning Algorithms (MLAs) to the problem, using data from auxiliary channels within the LIGO detectors that monitor degrees of freedom unaffected by astrophysical signals. Terrestrial noise sources may manifest characteristic disturbances in these auxiliary channels, inducing non-trivial correlations with glitches in the gravitational-wave data. The number of auxiliary-channel parameters describing these disturbances may also be extremely large; high dimensionality is an area where MLAs are particularly well-suited. We demonstrate the feasibility and applicability of three very different MLAs: Artificial Neural Networks, Support Vector Machines, and Random Forests. These classifiers identify and remove a substantial fraction of the glitches present in two very different data sets: four weeks of LIGO's fourth science run and one week of LIGO's sixth science run. We observe that all three algorithms agree on which events are glitches to within 10% for the sixth science run data, and support this by showing that the different optimization criteria used by each classifier generate the same decision surface, based on a likelihood-ratio statistic. Furthermore, we find that all classifiers obtain similar limiting performance, suggesting that most of the useful information currently contained in the auxiliary channel parameters we extract is already being used. Future performance gains are thus likely to involve additional sources of information, rather than improvements in the MLAs themselves.

gravitational-wave data↗

Ares I and Ares I-X Stage Separation Aerodynamic Testing

The aerodynamics of the Ares I crew launch vehicle (CLV) and Ares I-X flight test vehicle (FTV) during stage separation was characterized by testing 1%-scale models at the Arnold Engineering Development Center s (AEDC) von Karman Gas Dynamics Facility (VKF) Tunnel A at Mach numbers of 4.5 and 5.5. To fill a large matrix of data points in an efficient manner, an injection system supported the upper stage and a captive trajectory system (CTS) was utilized as a support system for the first stage located downstream of the upper stage. In an overall extremely successful test, this complex experimental setup associated with advanced postprocessing of the wind tunnel data has enabled the construction of a multi-dimensional aerodynamic database for the analysis and simulation of the critical phase of stage separation at high supersonic Mach numbers. Additionally, an extensive set of data from repeated wind tunnel runs was gathered purposefully to ensure that the experimental uncertainty would be accurately quantified in this type of flow where few historical data is available for comparison on this type of vehicle and where Reynolds-averaged Navier-Stokes (RANS) computational simulations remain far from being a reliable source of static aerodynamic data.

Pinier, Jeremy T.↗

Improved Plate and Beam Models for Thermoviscoelastic Constitutive Modeling of Composites

The effective properties of composites are influenced by the time-dependent behavior of polymer matrices very sensitive to changes in temperature. Improved plate and beam models are required to efficiently design, and simulate composite structures when the long-term performance of large anisotropic composite structures is the matter of interest. In this work, mechanics of structure genome (MSG) is used to con-struct linear thermoviscoelastic plate and beam models that can homogenize three-dimensional heterogeneous materials made of constituents with time- and temperature-dependent behavior. The formulation derives the transient strain energy based on integral formulation for thermorheologically simple materials subject to finite temperature changes with the restriction that the strain is small. The reduced time parameter is introduced to relate the time-temperature dependency of the anisotropic material by means of master curves at reference conditions. The new formulation has been implemented in SwiftCompTM, a general-purpose multiscale constitutive modeling code based on MSG. Experimental data and three-dimensional direct numerical simulations of thin-ply high-strain composites (TP-HSC) using a commercial finite element analysis (FEA) package are conducted to verify the accuracy of SwiftCompTM results. The paper also analyzes the relationship between the shift factor of the polymer matrix and the temperature dependencies of the effective beam properties.

Finite element analysis↗

An Efficient GPU-Accelerated Multi-Source Global Fit Pipeline for LISA Data Analysis

The large-scale analysis task of deciphering gravitational wave signals in the LISA data stream will be difficult, requiring a large amount of computational resources and extensive development of computational methods. Its high dimensionality, multiple model types, and complicated noise profile require a global fit to all parameters and input models simultaneously. In this work, we detail our global fit algorithm, called “Erebor,” designed to accomplish this challenging task. It is capable of analysing current state-of-the-art datasets and then growing into the future as more pieces of the pipeline are completed and added. We describe our pipeline strategy, the algorithmic setup, and the results from our analysis of the LDC2A Sangria dataset, which contains Massive Black Hole Binaries, compact Galactic Binaries, and a parameterized noise spectrum whose parameters are unknown to the user. The Erebor algorithm includes three unique and very useful contributions: GPU acceleration for enhanced computational efficiency; ensemble MCMC sampling with multiple MCMC walkers per temperature for better mixing and parallelized sample creation; and special online updates to reversible-jump (or trans-dimensional) sampling distributions to ensure sampler mixing and accurate initial estimates for detectable sources in the data. We recover posterior distributions for all 15 (6) of the injected MBHBs in the LDC2A training (hidden) dataset. We catalog ∼12000 Galactic Binaries (∼8000 as high confidence detections) for both the training and hidden datasets. All of the sources and their posterior distributions are provided in publicly available catalogs.

LISA global fit↗

Establishing Calibration Standards for Remote Sensing Retrievals of Greenhouse Gases

While in situ greenhouse gas measurements have a concrete calibration standard, establishing a similar standard for remote sensing retrievals remains challenging. As the constellation of greenhouse gas observing satellites grows, such a standard is essential to ensure data are of the quality necessary to support scientific and policy applications. A primary goal of NASA's Atmospheric Carbon and Transport (ACT) - America aircraft campaign is to evaluate column CO2 retrievals from the Orbiting Carbon Observatory 2 (OCO-2) through coordinated underflights. The campaign also includes airborne lidar instruments that measure the amount of CO2 and CH4 below the aircraft. This presentation introduces a system that establishes calibration standards for OCO-2 and lidar retrievals based on in situ data from the ACT-America campaign. The system assimilates the in situ data into NASA's Goddard Earth Observing System (GEOS) to produce high-resolution, two-dimensional transects of CO2 along the flight path which we refer to as curtains. Excluding the ability to sample the entire atmosphere at once, any such analysis must make assumptions about the connection of measurements at different places and times to a given retrieval. We chose to use the GEOS general circulation model forced by meteorology from its data assimilation system because their scientific merits are extensively documented. Furthermore, in areas rich in data, the assimilated curtains approach a field constrained by data alone. Where data are lacking, e.g., the stratosphere, age of air and other transport diagnostics can be used to quantify the uncertainty introduced by the model. Given these uncertainties, we can determine the uncertainties of the curtains and thus our ability to evaluate remote sensing instruments. Here, we demonstrate this for several flights over North America and discuss possible applications to upcoming missions, e.g., GeoCarb.

Weir, B.↗

The NATA code; theory and analysis. Volume 2: User's manual

The NATA code is a computer program for calculating quasi-one-dimensional gas flow in axisymmetric nozzles and rectangular channels, primarily to describe conditions in electric archeated wind tunnels. The program provides solutions based on frozen chemistry, chemical equilibrium, and nonequilibrium flow with finite reaction rates. The shear and heat flux on the nozzle wall are calculated and boundary layer displacement effects on the inviscid flow are taken into account. The program contains compiled-in thermochemical, chemical kinetic and transport cross section data for high-temperature air, CO2-N2-Ar mixtures, helium, and argon. It calculates stagnation conditions on axisymmetric or two-dimensional models and conditions on the flat surface of a blunt wedge. Included in the report are: definitions of the inputs and outputs; precoded data on gas models, reactions, thermodynamic and transport properties of species, and nozzle geometries; explanations of diagnostic outputs and code abort conditions; test problems; and a user's manual for an auxiliary program (NOZFIT) used to set up analytical curvefits to nozzle profiles.

Bade, W. L.↗