Search NASA⌕ Search

SEARCH · Search NASA

Results for “heterogeneous data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

ROSAT observations of z greater than 3 quasars

Successful pointed observations using the Roentgen Satellite (ROSAT) Position Sensitive Proportional Counter (PSPC) were made of seven z greater than 3 optically selected quasars from the Large Bright Quasar Survey (LBQS). Four detections and three 3 sigma upper limits resulted. Combining these data with the heterogeneous sample of Avni & Tananbaum (1986) confirms their conclusion that the ratio of x-ray to optical luminosity is correlated with optical luminosity and probably not correlated with redshift. This suggests that x-ray luminosity evolves more slowly than optical luminosity. These results are then used in conjunction with the LBQS database to estimate the contribution to the 2 keV x-ray background of bright, optically selected quasars with m(sub B(sub J)) less than 18.85; the result is about 5%.

Pickering, T. E.↗

The Web Measurement Environment (WebME): A Tool for Combining and Modeling Distributed Data

Many organizations have incorporated data collection into their software processes for the purpose of process improvement. However, in order to improve, interpreting the data is just as important as the collection of data. With the increased presence of the Internet and the ubiquity of the World Wide Web, the potential for software processes being distributed among several physically separated locations has also grown. Because project data may be stored in multiple locations and in differing formats, obtaining and interpreting data from this type of environment becomes even more complicated. The Web Measurement Environment (WebME), a Web-based data visualization tool, is being developed to facilitate the understanding of collected data in a distributed environment. The WebME system will permit the analysis of development data in distributed, heterogeneous environments. This paper provides an overview of the system and its capabilities.

Tesoriero, Roseanne↗

CTE Measurements of Carbon Fibers in ESEM

The strong appeal of carbon fiber reinforced ceramic matrix and carbon-carbon composites arises from their promising high temperature performance within laboratory-simulated nozzle, leading edge and combustor environments. The main drawback in these systems are two fold; namely oxidation affinity and CTE mismatch. To date the emphasis has been on axial CTE mismatch with the push on estimating (component) laminate level data. However, in heterogeneous and anisotropic systems, all failure mechanisms are at micro or nano scales. This paper will report plans and progress on single carbon fibers efforts in order to develop its temperature dependent, anisotropic thermophysical and thermomechanical properties using combination of x-ray spectroscopy and environmental scanning electron microscopy to capture circumferential expansion from 20-900 C in within the SEM chamber. Surface roughness measurements were also determined. Crystallinity was also determined by the quality of the electron diffraction (ED) patterns.

Effinger, Michael↗

Using High Spatial Resolution Satellite Imagery to Map Forest Burn Severity Across Spatial Scales in a Pine Barrens Ecosystem

As a primary disturbance agent, fire significantly influences local processes and services of forest ecosystems. Although a variety of remote sensing based approaches have been developed and applied to Landsat mission imagery to infer burn severity at 30 m spatial resolution, forest burn severity have still been seldom assessed at fine spatial scales (less than or equal to 5 m) from very-high-resolution (VHR) data. We assessed a 432 ha forest fire that occurred in April 2012 on Long Island, New York, within the Pine Barrens region, a unique but imperiled fire-dependent ecosystem in the northeastern United States. The mapping of forest burn severity was explored here at fine spatial scales, for the first time using remotely sensed spectral indices and a set of Multiple Endmember Spectral Mixture Analysis (MESMA) fraction images from bi-temporal - pre- and post-fire event - WorldView-2 (WV-2) imagery at 2 m spatial resolution. We first evaluated our approach using 1 m by 1 m validation points at the sub-crown scale per severity class (i.e. unburned, low, moderate, and high severity) from the post-fire 0.10 m color aerial ortho-photos; then, we validated the burn severity mapping of geo-referenced dominant tree crowns (crown scale) and 15 m by 15 m fixed-area plots (inter-crown scale) with the post-fire 0.10 m aerial ortho-photos and measured crown information of twenty forest inventory plots. Our approach can accurately assess forest burn severity at the sub-crown (overall accuracy is 84% with a Kappa value of 0.77), crown (overall accuracy is 82% with a Kappa value of 0.76), and inter-crown scales (89% of the variation in estimated burn severity ratings (i.e. Geo-Composite Burn Index (CBI)). This work highlights that forest burn severity mapping from VHR data can capture heterogeneous fire patterns at fine spatial scales over the large spatial extents. This is important since most ecological processes associated with fire effects vary at the less than 30 m scale and VHR approaches could significantly advance our ability to characterize fire effects on forest ecosystems.

Meng, Ran↗

Transcriptomics-based Machine Learning (ML) Analysis Predicts Space-Exposed Murine Livers

NASA has employed high-throughput molecular assays to identify sub-cellular changes impacting human physiology during spaceflight. Machine learning (ML) methods hold the promise to improve our ability to identify important signals within highly dimensional molecular data. However, the inherent limitation of study subject numbers within a spaceflight mission minimizes the utility of ML approaches. To overcome the sample power limitations, data from multiple spaceflight missions must be aggregated while appropriately addressing intra- and inter-study variabilities. Here we describe an approach to log transform, scale and normalize data from six heterogeneous, mouse liver derived transcriptomics datasets (ntotal=137) which enabled ML-methods to perform well (AUC ≥ 0.87) in classifying spaceflown vs ground control animals rather than mission-of-origin. Concordance was found between liver-specific biological processes identified from harmonized ML-based analysis and study-by-study classical omics analysis. This work demonstrates the feasibility of applying ML methods on integrated, heterogeneous datasets of small sample size.

Machine Learning↗

IRIS-MEMFLOW: Data Flow-Enabled Portable Memory Orchestration in IRIS Runtime for Diverse Heterogeneity

Task-based programming models and execution paradigms provide a means to decompose a computation by expressing it as a graph in which each node represents a specific computation operating on memory objects and the edges define the dependencies in the execution flow. In this execution model, independent nodes in the graph can be executed concurrently in different computing devices, making it suitable for heterogeneous systems in which computing devices with different architectures coexist. However, careful memory orchestration across heterogeneous devices is needed because copies of the same memory object may reside in multiple devices during execution. Manually ensuring such an orchestration is quite challenging. Not only must an application developer guard against race conditions, but they must also optimize data movement between the host and devices because unnecessary data movement significantly impacts performance. To mitigate these challenges, we enhance the IRIS heterogeneous runtime and introduce IRIS-MEMFLOW–a data flow–enabled portable memory abstraction for seamlessly orchestrating memory in diverse heterogeneous computing environments. By using data-flow analysis, IRIS-MEMFLOW guards against race conditions while multiple heterogeneous devices access memory objects. IRIS-MEMFLOW also optimizes data movement between the host and devices without manual intervention. As a result, IRIS provides improved programming productivity, performance, and portability for multidevice heterogeneous executions in high-performance computing and cloud systems that run diverse architectures from different vendors. The efficacy of IRIS-MEMFLOW is evaluated through experiments that show its capability in terms of programming productivity, multidevice heterogeneity, portability, and low overhead versus the state of the art.

Monil, M. A. H. [ORNL] (ORCID:0000000334194037)↗

Digital image correlation and infrared thermography data for seven unique geometries of 304L stainless steel

Material Testing 2.0 (MT2.0) is a paradigm that advocates for the use of rich, full-field data, such as from digital image correlation and infrared thermography, for material identification. By employing heterogeneous, multi-axial data in conjunction with sophisticated inverse calibration techniques such as finite element model updating and the virtual fields method, MT2.0 aims to reduce the number of specimens needed for material identification and to increase confidence in the calibration results. To support continued development, improvement, and validation of such inverse methods—specifically for rate-dependent, temperature-dependent, and anisotropic metal plasticity models—we provide here a thorough experimental data set for 304L stainless steel sheet metal. The data set includes full-field displacement, strain, and temperature data for seven unique specimen geometries tested at different strain rates and in different material orientations. Commensurate extensometer strain data from tensile dog bones is provided as well for comparison. We believe this complete data set will be a valuable contribution to the experimental and computational mechanics communities, supporting continued advances in material identification methods.

36 MATERIALS SCIENCE↗

3-D Geological Modeling for Numerical Flow Simulation Studies of Gas Hydrate Reservoirs at the Kuparuk State 7-11-12 Pad in the Prudhoe Bay Unit on the Alaska North Slope

Accurate reservoir evaluation requires reliable three-dimensional (3-D) geological models. Here, this study conducted 3-D geological modeling for numerical flow simulation of the B1 sand gas hydrate reservoir at the Kuparuk State 7-11-12 pad, Prudhoe Bay Unit, Alaska North Slope. The model integrates well logs, core, and seismic data to address spatial heterogeneity in geological structures and reservoir properties. Two modeling types were performed: structural framework modeling and petrophysical property modeling. For structural framework modeling, seismic data and well log markers were used to reproduce subsurface structures characterized by a normal fault system. A volume-based modeling algorithm and stair-stepping grid were applied. The resulting 3-D model comprised 2,640,000 grid cells across 264 layers, including seven fault grids. For petrophysical property modeling, total porosity was initially modeled using sequential Gaussian simulation with collocated cokriging. To reproduce the upward coarsening of the B1 sand, upscaled log-derived total porosity and a three-dimensional (3-D) trend depicting total porosity variation were used as primary and secondary data, respectively. Gas hydrate saturation distribution was modeled similarly, with secondary data from estimated porosity distribution and seismic-derived acoustic impedance map enhancing accuracy. Results indicate higher gas hydrate saturation in the upper part of the B1 sand and areas with higher acoustic impedance. Intrinsic permeability was modeled from the total porosity and clay-bound water volume, and effective permeability was derived from the gas hydrate saturation and intrinsic permeability distributions based on the “Tokyo model”. Effective permeability distributions were influenced by the total porosity, gas hydrate saturation, and intrinsic permeability. Within the same layer, higher gas hydrate saturation leads to decreased effective permeability. In total, 100 sets of multiple scenarios were prepared, providing input data for dynamic flow simulations to evaluate the effects of lateral heterogeneity in reservoir properties and the hydraulic characteristics of faults on production behavior for preassessment before the long-term production test.

58 GEOSCIENCES↗

On learning what to learn: Heterogeneous observations of dynamics and establishing possibly causal relations among them

Abstract Before we attempt to (approximately) learn a function between two sets of observables of a physical process, we must first decide what the inputs and outputs of the desired function are going to be. Here we demonstrate two distinct, data-driven ways of first deciding “the right quantities” to relate through such a function, and then proceeding to learn it. This is accomplished by first processing simultaneous heterogeneous data streams (ensembles of time series) from observations of a physical system: records of multiple observation processes of the system. We determine (i) what subsets of observables are common between the observation processes (and therefore observable from each other, relatable through a function); and (ii) what information is unrelated to these common observables, therefore particular to each observation process, and not contributing to the desired function. Any data-driven technique can subsequently be used to learn the input–output relation—from k-nearest neighbors and Geometric Harmonics to Gaussian Processes and Neural Networks. Two particular “twists” of the approach are discussed. The first has to do with the identifiability of particular quantities of interest from the measurements. We now construct mappings from a single set of observations from one process to entire level sets of measurements of the second process, consistent with this single set. The second attempts to relate our framework to a form of causality: if one of the observation processes measures “now,” while the second observation process measures “in the future,” the function to be learned among what is common across observation processes constitutes a dynamical model for the system evolution.

Sroczynski, David W.↗

Understanding Generative AI Content with Embedding Models

The construction of high-quality numerical features is critical to any quantitative data analysis. Feature engineering has been historically addressed by carefully hand-crafting data representations based on domain expertise. This work views the internal representations of modern deep neural networks (DNNs), called embeddings, as an implicit form of traditional feature engineering. For trained DNNs, we show that these embeddings can reveal interpretable, high-level concepts in unstructured sample data. We use these embeddings in natural language and computer vision tasks to uncover both inherent heterogeneity in the underlying data and human-understandable explanations for it. In particular, we find empirical evidence that there is inherent separability between real data and those generated from AI models.

Vargas, Max↗

A Comparative Analysis of Micrometeorological Determinants of Evapotranspiration Rates Within a Heterogeneous Urban Environment

Variability in micrometeorological conditions and their influence on estimated reference evapotranspiration (RET) rates were evaluated across a heterogeneous urban environment. Micrometeorological data sets (incoming solar radiation, air temperature, relative humidity and wind speed) were collected over a one-year period at six weather stations in New York City, NY (USA). Weather stations are located at four new urban green space monitoring sites and two airports. Reference evapotranspiration (RET) rates were estimated from the micrometeorological data sets for a short reference surface at a daily time-step using the ASCE Standardized Reference Evapotranspiration Equation, a Penman-Monteith based combination equation. Nonparametric comparative statistical analyses (Kruskal-Wallis) revealed statistically significant differences (at significance level α = 0.05) in micrometeorological conditions and estimated RET rates between the six sites. On a cumulative annual basis, estimated RET varied by up to 40 percent between the sites. A new technique for adjusting weather data collected at one location (e.g. regional airports) for use at another location (e.g. interior engineered urban green spaces) was evaluated. The study highlights the importance, for accurate estimation of ET, of onsite micrometeorological data sets, but concludes that additional research is needed to more thoroughly characterize micrometeorological variability across heterogeneous urban environments, and also to evaluate the influence of non-meteorological determinants, e.g. vegetation type, soil/media type, media moisture conditions and anthropogenic heat fluxes, on urban ET.

Urban environment↗

Estimation of canopy parameters for inhomogeneous vegetation canopies from reflectance data. III - TRIM: A model for radiative transfer in heterogeneous three-dimensional canopies

A model for radiative transfer in heterogeneous three-dimensional canopies such as those found in forests is proposed. Its use in estimating important biophysical variables such as leaf area index and canopy architecture from bidirectional canopy reflectance data is discussed. The model and its use in estimating canopy parameters through its inversion are validated with measured canopy reflectance data for corn canopies.

Goel, Narendra S.↗

Assessed Using Satellite Data and Ground Based Sun/Sky Radiometers

The short aerosol lifetime and the variety of sources and atmospheric processes that affect the aerosol concentration and properties, generate a very heterogeneous aerosol field. Satellite data are needed to assess the daily global distribution of aerosol loading, properties and radiative forcing of climate. The Moderate Resolution Imaging Spectroradiometer (MODIS) instrument launched in July 1999 measures the aerosol properties and backscattered radiation in 7 spectral channels across the solar channel with spatial resolution of 500 m. Novel inversion techniques are used to derive the aerosol forcing at the top of the atmosphere in these 7 spectral bands with even higher accuracy than the derivation of the aerosol loading or microphysical properties. But satellites cannot sense the aerosol attenuation of radiation reaching the Earth surface. A network of 100 sun/sky radiometers, members of the AErosol RObotic NEtwork (AERONET) is used to assess, simultaneously, the radiative forcing of aerosol at the surface level. The relevance of MODIS and AERONET to assess the global aerosol forcing, and the inversion techniques will be presented and discussed.

Kaufman, Yoram J.↗

Exploration Clinical Decision Support System: Medical Data Architecture

The Exploration Clinical Decision Support (ECDS) System project is intended to enhance the Exploration Medical Capability (ExMC) Element for extended duration, deep-space mission planning in HRP. A major development guideline is the Risk of "Adverse Health Outcomes & Decrements in Performance due to Limitations of In-flight Medical Conditions". ECDS attempts to mitigate that Risk by providing crew-specific health information, actionable insight, crew guidance and advice based on computational algorithmic analysis. The availability of inflight health diagnostic computational methods has been identified as an essential capability for human exploration missions. Inflight electronic health data sources are often heterogeneous, and thus may be isolated or not examined as an aggregate whole. The ECDS System objective provides both a data architecture that collects and manages disparate health data, and an active knowledge system that analyzes health evidence to deliver case-specific advice. A single, cohesive space-ready decision support capability that considers all exploration clinical measurements is not commercially available at present. Hence, this Task is a newly coordinated development effort by which ECDS and its supporting data infrastructure will demonstrate the feasibility of intelligent data mining and predictive modeling as a biomedical diagnostic support mechanism on manned exploration missions. The initial step towards ground and flight demonstrations has been the research and development of both image and clinical text-based computer-aided patient diagnosis. Human anatomical images displaying abnormal/pathological features have been annotated using controlled terminology templates, marked-up, and then stored in compliance with the AIM standard. These images have been filtered and disease characterized based on machine learning of semantic and quantitative feature vectors. The next phase will evaluate disease treatment response via quantitative linear dimension biomarkers that enable image content-based retrieval and criteria assessment. In addition, a data mining engine (DME) is applied to cross-sectional adult surveys for predicting occurrence of renal calculi, ranked by statistical significance of demographics and specific food ingestion. In addition to this precursor space flight algorithm training, the DME will utilize a feature-engineering capability for unstructured clinical text classification health discovery. The ECDS backbone is a proposed multi-tier modular architecture providing data messaging protocols, storage, management and real-time patient data access. Technology demonstrations and success metrics will be finalized in FY16.

Biomedical support↗

The Unified Phenotype Ontology : a framework for cross-species integrative phenomics

Phenotypic data are critical for understanding biological mechanisms and consequences of genomic variation, and are pivotal for clinical use cases such as disease diagnostics and treatment development. For over a century, vast quantities of phenotype data have been collected in many different contexts covering a variety of organisms. The emerging field of phenomics focuses on integrating and interpreting these data to inform biological hypotheses. A major impediment in phenomics is the wide range of distinct and disconnected approaches to recording the observable characteristics of an organism. Phenotype data are collected and curated using free text, single terms or combinations of terms, using multiple vocabularies, terminologies, or ontologies. Integrating these heterogeneous and often siloed data enables the application of biological knowledge both within and across species. Existing integration efforts are typically limited to mappings between pairs of terminologies; a generic knowledge representation that captures the full range of cross-species phenomics data is much needed. We have developed the Unified Phenotype Ontology (uPheno) framework, a community effort to provide an integration layer over domain-specific phenotype ontologies, as a single, unified, logical representation. uPheno comprises (1) a system for consistent computational definition of phenotype terms using ontology design patterns, maintained as a community library; (2) a hierarchical vocabulary of species-neutral phenotype terms under which their species-specific counterparts are grouped; and (3) mapping tables between species-specific ontologies. This harmonized representation supports use cases such as cross-species integration of genotype-phenotype associations from different organisms and cross-species informed variant prioritization.

59 BASIC BIOLOGICAL SCIENCES↗

Heterogeneity in Permeability and Particulate Organic Carbon Content Controls the Redox Condition of Riverbed Sediments at Different Timescales

Abstract The hydrological and biogeochemical properties of the hyporheic zone in stream and riverine ecosystems have been extensively studied over the past two decades. Although it is widely acknowledged that sediment heterogeneity can influence biogeochemical reactions, little effort has been made to understand the role of heterogeneity on the spatiotemporal variability of riverbed redox conditions under changing flow dynamics at different timescales. Here we integrate a mechanistic model and field data to demonstrate that heterogeneity in permeability plays a vital role in modulating sediment redox conditions at both seasonal (annual) and event (daily‐to‐weekly) timescales, whereas heterogeneity in particulate organic carbon (POC) content only has a comparable influence on redox conditions at the seasonal timescale. These findings underscore the importance of accurately characterizing sediment heterogeneity, in terms of permeability and POC content, in quantifying biogeochemical dynamics in the riverbed and hyporheic zones of riverine ecosystems.

Geology↗

Data and scripts associated with “Allometric scaling of hyporheic respiration across basins in the Pacific Northwest USA"

This data package is associated with the publication “Allometric scaling of hyporheic respiration across basins in the Pacific Northwest USA” submitted to JGR-Biogeosciences (Regier et al. 2025).This study used reach-scale modeled estimates of hyporheic aerobic respiration made by the River Corridor Model (Fang et al. 2020) and watershed characteristics across the Willamette and Yakima River basins to explore potential allometric scaling (i.e., power-law relationships between size and function) of cumulative hyporheic respiration across catchment-to-basin scales. Scaling was explored quantitatively via the R2, slope, and y-intercept of relationships between cumulative hyporheic respiration and watershed area, divided into hyporheic exchange flux (HEF) quantiles. We also explored relationships between allometric scaling and other watershed characteristics through linear regression, spatial patterns, and mutual information analyses. Our results also suggest variability of hyporheic respiration allometry for middle exchange flux quantiles, and in relation to land-cover. Our findings provide initial evidence that allometric scaling may be useful for predicting hyporheic biogeochemical dynamics across watersheds from reach to basin scales. This data package is associated with the GitHub repository found at https://github.com/peterregier/rc_wrb_yrb_scaling. The data package is organized into several key directories. The “data” folder contains multiple CSV files, including landscape heterogeneity, scaling analysis, and watershed boundary data. The “figures” folder has all figure files in both PDF and PNG formats. Core analysis scripts and figure generation scripts are in the “scripts” directory, systematically numbered for sequential execution. The root directory includes essential project files; please see the file ending in “flmd.csv” for a list and description of all files contained in this data package and the file ending in “dd.csv” for data dictionaries used to describe tabular column headers.

54 ENVIRONMENTAL SCIENCES↗

Surface Areas and Morphology of Thin Ice Films

Thin ice films formed by deposition from the vapor phase in a fast flow-tube reactor have been used to simulate polar stratospheric cloud surfaces in order to obtain laboratory data on uptake and heterogeneous reaction rates. Surface areas are determined from BET (Brunauer, Emmett, and Teller) analysis of gas adsorption isotherms. The results for ices prepared at 196 K or 77 K are consistent with previous data on thicker ice films. Environmental scanning electron microscopy is used to obtain particle sizes and shapes, and to investigate the morphology of the ices on borosilicate or silicon windows. In addition, the uptake of HCI on ice films prepared at 196 K is investigated. The results suggest that the layer model we have previously developed for analysis of uptake and heterogeneous reaction rates on ice films is valid. Detailed information will be presented at the conference.

ice films↗