Search NASA⌕ Search

SEARCH · Search NASA

Results for “data integration”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 379 records · Page 21

Utilization of wind tunnel instrumentation with software verifications

Software tools developed for the National Full-Scale Aerodynamic Complex (NFAC) for verifying data integrity and troublehooting problems are discussed. The Hardware Check verifies that the incoming signals are properly connected and are being acquired into the real time data system. The Zero/Cal Check program verifies the reliability of the wind tunnel instrumentation by checking the zero and calibration points. The Power Spectral Density Plots help to identify the frequency components of a signal. Drift Program and Thermal Plots tools are also described.

Silva, Betty W.↗

TPSAS-NF1676L-16833-DND

Semantic Infrastructure is central to realizing the first goal of the ASDC's Strategic Plan: expanding the ASDC's customer base by improving access to ASDC data. ASDC data comprises a widely heterogeneous set of complex products which presents two significant challenges in data access: Helping customers discover, among many available options, the most suitable data products for their purpose; and Guiding customers to easily and appropriately use products. Data products differ significantly in terms of how the data was collected and processed, even with similar subject matter. Understanding differences is critical to using data effectively. To reach a broader customer range, the ASDC must provide prospective users with enough information to quickly and meaningfully compare and evaluate data products. Data formats and structures also differ among products. Applications displaying and analyzing data need access to federated and semantically disambiguated data. Semantic technologies offer functionality for addressing this issue. Ontologies can provide robust, stable domain models serving as common schema for discovering, evaluating, comparing, and integrating data from disparate products. Reasoning engines and triple stores can leverage ontologies to support intelligent search applications allowing users to discover, query, retrieve, and easily reformat data from a broad spectrum of sources.

Beth Huffer↗

Advanced Modeling and Uncertainty Quantification for Flight Dynamics; Interim Results and Challenges

As part of the NASA Vehicle Systems Safety Technologies (VSST), Assuring Safe and Effective Aircraft Control Under Hazardous Conditions (Technical Challenge #3), an effort is underway within Boeing Research and Technology (BR&T) to address Advanced Modeling and Uncertainty Quantification for Flight Dynamics (VSST1-7). The scope of the effort is to develop and evaluate advanced multidisciplinary flight dynamics modeling techniques, including integrated uncertainties, to facilitate higher fidelity response characterization of current and future aircraft configurations approaching and during loss-of-control conditions. This approach is to incorporate multiple flight dynamics modeling methods for aerodynamics, structures, and propulsion, including experimental, computational, and analytical. Also to be included are techniques for data integration and uncertainty characterization and quantification. This research shall introduce new and updated multidisciplinary modeling and simulation technologies designed to improve the ability to characterize airplane response in off-nominal flight conditions. The research shall also introduce new techniques for uncertainty modeling that will provide a unified database model comprised of multiple sources, as well as an uncertainty bounds database for each data source such that a full vehicle uncertainty analysis is possible even when approaching or beyond Loss of Control boundaries. Methodologies developed as part of this research shall be instrumental in predicting and mitigating loss of control precursors and events directly linked to causal and contributing factors, such as stall, failures, damage, or icing. The tasks will include utilizing the BR&T Water Tunnel to collect static and dynamic data to be compared to the GTM extended WT database, characterizing flight dynamics in off-nominal conditions, developing tools for structural load estimation under dynamic conditions, devising methods for integrating various modeling elements into a real-time simulation capability, generating techniques for uncertainty modeling that draw data from multiple modeling sources, and providing a unified database model that includes nominal plus increments for each flight condition. This paper presents status of testing in the BR&T water tunnel and analysis of the resulting data and efforts to characterize these data using alternative modeling methods. Program challenges and issues are also presented.

Hyde, David C.↗

MLSPICE: Machine Learning based SPICE Modeling Platform for Power Magnetics

Electrical power converters are critical to a wide range of applications ranging from renewable integration to transportation electrification, and can be a key factor determining the size, weight, and efficiency of energy conversion systems. Magnetic components are typically the largest and least efficient components in power electronics. While there have been major strides in the modeling and analysis of power semiconductor devices and circuit simulations, the necessary advances in the design of power magnetics have lagged. In this project, we have transformed the modeling and design of power magnetics with machine learning enabled methods and catalyze simultaneous disruptive improvements for ML-based power electronics design tools. A fully automated open-source machine learning based magnetics modeling platform – the MagNet project - with innovations in full stack have been developed to greatly accelerate the design process and provide new insights to magnetic material and geometry design. The ARPA-E funded MagNet platform contains three major building blocks: 1) a ML-Integrated Data Acquisition System (MIDAS): a highly automated data acquisition testbed which is capable of measuring a large number of magnetic cores with a wide range of electrical circuit excitations; 2) a ML-integrated Core Loss Model (MICLM): a machine-learning trained modeling method for modeling the core loss and saturation effects of magnetic materials for arbitrary excitation waveforms; 3) ML-guided Magnetics SPICE Simulation Tool (PMSPICE): a fully integrated CAD tool which can simulate the magnetics in SPICE. It can help the designers to quickly model the linear and non-linear characteristics of magnetic components and evaluate their behavior in SPICE simulations. The developed MagNet system has fully demonstrated the proposed performance target and has been open sourced to the entire power electronics community to advance the modeling and design of power magnetics from many different angles.

36 MATERIALS SCIENCE↗

A case study in contrastive learning information combination: Application to technical forensics of additive manufacturing filament source identification

Combination of information from disparate data sources into a single decision is a core challenge in many fields, including the field of technical forensics. Technical forensics (TF) utilizes technical characterization of questioned samples to determine properties of that sample; these properties are then used to infer information of forensic interest, such as provenance, age, or attribution. TF is utilized in traditional forensic applications, such as the attribution of material fragments from an explosive, and in nuclear forensic applications, such as the attribution of actinides which have been interdicted out of regulatory control. The challenge of combining information from disparate sources, described alternately by many terms including “Data Fusion” and “Data Integration”, is exacerbated in the technical forensics domain due to at least two factors: the challenge of interpreting each information source singularly, and the relatively small data set sizes available. Extensive literature exists attempting to combine technical forensics information sources, both in manual and automated processes. These attempts are often bespoke to the specific information sources (such as the bi-, tri-, or quad-isotope chart (Moody, Grant, and Hutcheon 2005)), with some emerging examples of simple early- and late- fusion (, respectively). Simultaneous to the information combination efforts described in the previous paragraph, the field of natural language processing attempted (and largely succeeded) in combining information from multiple non-technical information sources. The ecosystem of “multi-modal” language models, which can take text and images as input, and generate text and images as output, became large and diverse by 2025 (Khan et al. 2025). In a generalized sense, many of these methods are trained by learning neural networks which can convert raw text or images into a vector of numbers describing the text or image, hereafter called “embeddings” and the neural networks performing the conversion are called “embedders”. By using a separate embedder for text and images, finding coincident text and images (such as images with their captions), and optimizing the parameters of the embedders such that the embeddings for the text and the image are similar, the field has found a bridge between text and images (Girdhar et al. 2023). It is the contention of the authors of this report that this insight is not limited to text and images but instead can be extended to any modality which can be found coincidently. The subject of the rest of this report is the application of this method to example multi-modal technical forensic data. Some details about the data used in this report are not appropriate for this report, and are included in a companion report (PNNL-38669).

36 MATERIALS SCIENCE↗

High rate information systems - Architectural trends in support of the interdisciplinary investigator

Data systems requirements in the Earth Observing System (EOS) Space Station Freedom (SSF) eras indicate increasing data volume, increased discipline interplay, higher complexity and broader data integration and interpretation. A response to the needs of the interdisciplinary investigator is proposed, considering the increasing complexity and rising costs of scientific investigation. The EOS Data Information System, conceived to be a widely distributed system with reliable communication links between central processing and the science user community, is described. Details are provided on information architecture, system models, intelligent data management of large complex databases, and standards for archiving ancillary data, using a research library, a laboratory and collaboration services.

Handley, Thomas H., Jr.↗

Study and prototype of data system interactions for the Earth Observing System Data and Information System

A crucial part of the Earth Observing System (EOS) is its Data and Information System (EOSDIS). The success of EOS depends not only on its instruments and science studies, but also on its ability to help scientists integrate data sets of geophysical and biological measurements taken by various instruments and investigators. NASA contractors have completed Phase B studies of EOSDIS, in particular its architecture, functionality, and user interfacing. At this point in time, it may seem impossible to exercise the EOSDIS or any of its components since they do not exist; i.e., if the EOSDIS is accepted as a totally new system, distinct from any existing DIS. However, if EOSDIS is seen as evolving from existing data systems, then some limited prototyping studies can be conducted by using currently functioning systems. In support of both the EOSDIS Science Advisory Panel and the EOSDIS Project, a prototyping activity was carried out by a cross section of interdisciplinary scientists. That prototyping activity is summarized and some conclusions are drawn that can be used by NASA-Goddard to evaluate and modify the specifications soon to be released in an RFP to build EOSDIS.

Emmitt, G. D.↗

Applications of LANDSAT data to the integrated economic development of Mindoro, Phillipines

LANDSAT data is seen as providing essential up-to-date resource information for the planning process. LANDSAT data of Mindoro Island in the Philippines was processed to provide thematic maps showing patterns of agriculture, forest cover, terrain, wetlands and water turbidity. A hybrid approach using both supervised and unsupervised classification techniques resulted in 30 different scene classes which were subsequently color-coded and mapped at a scale of 1:250,000. In addition, intensive image analysis is being carried out in evaluating the images. The images, maps, and aerial statistics are being used to provide data to seven technical departments in planning the economic development of Mindoro. Multispectral aircraft imagery was collected to compliment the application of LANDSAT data and validate the classification results.

Wagner, T. W.↗

Selection of Hyperspectral Narrowbands (HNBs) and Composition of Hyperspectral Twoband Vegetation Indices (HVIs) for Biophysical Characterization and Discrimination of Crop Types Using Field Reflectance and Hyperion-EO-1 Data

The overarching goal of this study was to establish optimal hyperspectral vegetation indices (HVIs) and hyperspectral narrowbands (HNBs) that best characterize, classify, model, and map the world's main agricultural crops. The primary objectives were: (1) crop biophysical modeling through HNBs and HVIs, (2) accuracy assessment of crop type discrimination using Wilks' Lambda through a discriminant model, and (3) meta-analysis to select optimal HNBs and HVIs for applications related to agriculture. The study was conducted using two Earth Observing One (EO-1) Hyperion scenes and other surface hyperspectral data for the eight leading worldwide crops (wheat, corn, rice, barley, soybeans, pulses, cotton, and alfalfa) that occupy approx. 70% of all cropland areas globally. This study integrated data collected from multiple study areas in various agroecosystems of Africa, the Middle East, Central Asia, and India. Data were collected for the eight crop types in six distinct growth stages. These included (a) field spectroradiometer measurements (350-2500 nm) sampled at 1-nm discrete bandwidths, and (b) field biophysical variables (e.g., biomass, leaf area index) acquired to correspond with spectroradiometer measurements. The eight crops were described and classified using approx. 20 HNBs. The accuracy of classifying these 8 crops using HNBs was around 95%, which was approx. 25% better than the multi-spectral results possible from Landsat-7's Enhanced Thematic Mapper+ or EO-1's Advanced Land Imager. Further, based on this research and meta-analysis involving over 100 papers, the study established 33 optimal HNBs and an equal number of specific two-band normalized difference HVIs to best model and study specific biophysical and biochemical quantities of major agricultural crops of the world. Redundant bands identified in this study will help overcome the Hughes Phenomenon (or "the curse of high dimensionality") in hyperspectral data for a particular application (e.g., biophysical characterization of crops). The findings of this study will make a significant contribution to future hyperspectral missions such as NASA's HyspIRI. Index Terms-Hyperion, field reflectance, imaging spectroscopy, HyspIRI, biophysical parameters, hyperspectral vegetation indices, hyperspectral narrowbands, broadbands.

Vegetation↗

NAND flash screening and qualification guideline for space application

All space missions have a need for nonvolatile memory (NVM), which maintains data integrity when unpowered. Types of NVM include PROM, EEPROM, NOR Flash, and NAND Flash. PROMs, NOR Flash, and EEPROMs are good choices for storing smaller file size data such as boot code or FPGA configurations. These products have excellent data retention characteristics and can reliably store data for years, unpowered, without any data corruption. They can also be read many times without disturbing the data. Another application of NVM is storage of science and engineering data, which requires large of amounts of memory. The highest density memories available today are SDRAM and NAND Flash. However, the power required to operate and store data in NAND Flash is far less than SDRAM. Wherever high density and low power is required, NAND Flash is very attractive.

Heidecker, Jason↗

JAXA-NASA Interoperability Demonstration for Application of DTN Under Simulated Rain Attenuation

As is well known, K-band or higher band communications in space link segment often experience intermittent disruptions caused by heavy rainfall. In view of keeping data integrity and establishing autonomous operations under such situation, it is important to consider introducing a tolerance mechanism such as Delay/Disruption Tolerant Networking (DTN). The Consultative Committee for Space Data Systems (CCSDS) is studying DTN as part of the standardization activities for space data systems. As a contribution to CCSDS and a feasibility study for future utilization of DTN, Japan Aerospace Exploration Agency (JAXA) and National Aeronautics and Space Administration (NASA) conducted an interoperability demonstration for confirming its tolerance mechanism and capability of automatic operation using Data Relay Test Satellite (DRTS) space link and its ground terminals. Both parties used the Interplanetary Overlay Network (ION) open source software, including the Bundle Protocol, the Licklider Transmission Protocol, and Contact Graph Routing. This paper introduces the contents of the interoperability demonstration and its results.

Suzuki, Kiyoshisa↗

Application of Low-Cost Fine Particulate Mass Monitors to Convert Satellite Aerosol Optical Depth Measurements to Surface Concentrations in North America and Africa

Low-cost particulate mass sensors provide opportunities to assess air quality at unprecedented spatial and temporal resolutions. Established traditional monitoring networks have limited spatial resolution and are simply absent in many major cities across sub-Saharan Africa (SSA). Satellites provide snapshots of regional air pollution but require ground-truthing. Low-cost monitors can supplement and extend data coverage from these sources worldwide, providing a better overall air quality picture. We investigate the utility of such a multi-source data integration approach using two case studies. First, in Pittsburgh, Pennsylvania, both traditional monitoring and dense low-cost sensor networks are compared with satellite aerosol optical depth (AOD) data from NASA's MODIS system, and a linear conversion factor is developed to convert AOD to surface fine particulate matter mass concentration (as PM2.5). With 10 or more ground monitors in Pittsburgh, there is a 2-fold reduction in surface PM2.5 estimation mean absolute error compared to using only a single ground monitor. Second, we assess the ability of combined regional-scale satellite retrievals and local-scale low-cost sensor measurements to improve surface PM2.5 estimation at several urban sites in SSA. In Rwanda, we find that combining local ground monitoring information with satellite data provides a 40 % improvement in surface PM2.5 estimation accuracy with respect to using low-cost ground monitoring data alone. A linear AOD-to-surface-PM2.5 conversion factor developed in Kigali, Rwanda, did not generalize well to other parts of SSA and varied seasonally for the same location, emphasizing the need for ongoing and localized ground-based monitoring, which can be facilitated by low-cost sensors. Overall, we find that combining ground-based low-cost sensor and satellite data, even without including additional meteorological or land use information, can improve and expand spatiotemporal air quality data coverage, especially in data-sparse regions.

AOD↗

Flight Testing of In-Time Safety Assurance Technologies for UAS Operations

Ongoing research at NASA is driven by a strategic plan defined by the Aeronautics Research Mission Directorate and a vision for future In-Time Aviation Safety Management Systems (IASMS) as described by the National Academies. In both visions, system safety awareness and provision are expanded through increased access to relevant data; integrated analysis and predictive capabilities; improved real-time detection and alerting of domain-specific hazards; decision support, and in some cases, automated risk mitigation strategies. One primary research focus is to develop means by which more timely (i.e., “in-time”) actions may be taken to mitigate precursors, anomalies, or trends that are observed during operations. In this paper, we describe such means as a collection of Services, Functions, and Capabilities (SFCs) that are supported by an underlying information system. For example, an integrated risk assessment capability is envisioned that continuously monitors safety-related metrics and margins and recommends timely operational changes. Assessment functions and/or services can be based on data analytics and predictive models derived from heterogeneous data sets that span relevant indicator metrics and their time histories. Likewise, on-board functions can identify and reduce susceptibility to precursor conditions that have led (and can lead) to aircraft loss-of-control or out-of-control accidents. This paper summarizes development and testing of such an information system tailored to hazards anticipated for future highly autonomous flight missions near and over densely populated areas. Testing is accomplished via simulation and by using small, unmanned aircraft operating over a test range at NASA’s Langley Research Center. Flight plans and test scenarios are defined to emulate several use-cases, including package delivery; reconnaissance; fire management; and urban air taxi vertiport operations. Two test phases are summarized with Phase 1 occurring in (2019-2020) and Phase 2 ongoing (2021-present). Results focus on SFC performance, technology readiness level assessment, and requirements discovery/validation. Companion papers are cited throughout for additional details on the recent testing.

safety management↗

Extraction and Analysis of Time Series Data from Building Automation Systems Using Large Language Models

Semantic schemas like Haystack 4, Brick and ASHRAE standard 223 enable the structured, standardized, and machine-readable representation of building data, facilitating interoperability, data integration, and advanced analytics. However, extracting information from these models requires specialized expertise in SPARQL and other programming languages, skills that are not commonly found among building professionals. Recent advancements in Large Language Models (LLMs), such as ChatGPT, enable the construction of queries using natural language, making it easier for individuals to interact with these systems in a manner that resembles everyday speech. However, these methods have not yet been tested on building semantic ontologies. This paper introduces a novel workflow and tool for enabling users to ask questions about a specific building's data, using natural language and receive answers automatically generated by GPT-4o. Our approach integrates semantic ontologies with advanced LLM capabilities to automate three critical steps: (1) generating SPARQL queries to retrieve time series references from ontological models, (2) extracting the corresponding time series data from the Building Automation System, and (3) performing computations and visualizations tailored to the user's query. The proposed method simplifies access to BAS data, allowing both domain experts and non-specialists to conduct sophisticated analyses without needing extensive technical knowledge of semantic web technologies. By demonstrating this pipeline, we facilitate more accessible and scalable data-driven decision-making in building operations and management.

Mulayim, Ozan Baris↗

Extraction and Analysis of Time Series Data from Building Automation Systems Using Large Language Models

Semantic schemas like Haystack 4, Brick and ASHRAE standard 223 enable the structured, standardized, and machine-readable representation of building data, facilitating interoperability, data integration, and advanced analytics. However, extracting information from these models requires specialized expertise in SPARQL and other programming languages, skills that are not commonly found among building professionals. Recent advancements in Large Language Models (LLMs), such as ChatGPT, enable the construction of queries using natural language, making it easier for individuals to interact with these systems in a manner that resembles everyday speech. However, these methods have not yet been tested on building semantic ontologies. This paper introduces a novel workflow and tool for enabling users to ask questions about a specific building's data, using natural language and receive answers automatically generated by GPT-4o. Our approach integrates semantic ontologies with advanced LLM capabilities to automate three critical steps: (1) generating SPARQL queries to retrieve time series references from ontological models, (2) extracting the corresponding time series data from the Building Automation System, and (3) performing computations and visualizations tailored to the user's query. The proposed method simplifies access to BAS data, allowing both domain experts and non-specialists to conduct sophisticated analyses without needing extensive technical knowledge of semantic web technologies. By demonstrating this pipeline, we facilitate more accessible and scalable data-driven decision-making in building operations and management.

Mulayim, Ozan Baris↗

LYNM-PE1 Seismic Parameters from Borehole Log, Laboratory, and Tabletop Measurements

The goal of this work is to provide a database of quality-checked seismic parameters that can be integrated with the Geologic Framework Model (GFM) for the LYNM-PE1 (Low Yield Nuclear Monitoring – Physical Experiment 1) testbed. We integrated data from geophysical borehole logs, tabletop measurements on collected core, and laboratory measurements. We reviewed for internal consistency among each measurement type, documented the caveats of measurement conditions, and integrated lithologic logs to check the validity of outlier values. The resulting consolidated parameter tables can be used as inputs for modeling and analysis codes and are designed to interface with the GFM, which is being actively developed.

58 GEOSCIENCES↗

Covariance of greenness and terrain variables over the Konza Prairie

An analysis is made of time-dependent covariance of the greenness vegetation index with mapped terrain variables over the Konza Prarie (Kansas) during the 1987 growing season. The analysis was part of an ongoing project to establish appopriate ground-sampling and data-integration strategies for satellite-based monitoring of land surface climate conditions. Greenness images for six dates between May and October were derived from atmospherically corrected thematic mapper (TM) data and coregistered with maps of woody vegetation, fire, and soils. Local variance in greenness peaked in mid-June, falling rapidly until mid-August, and declining gradually thereafter. Greenness images exhibited positive autocorrelation up to distances of 180-210 m, but the dominant scale of pattern occurred at a block size of 60 m by 60 m throughout the growing season. 40-44 percent of total scene variance in July and August was accounted for by the effects of woody vegetation (8.9 percent of the area), prairie burning, and soil type. The effect of these terrain variables was fairly consistent between June and late August and was manifested as additional high-frequency spatial variation in imagery from that period.

Davis, Frank W.↗

Reliable magnetic tape recording for real-time data acquisition systems

The design of a real-time magnetic tape I/O system using automatic error recovery and data spooling is discussed. Real time requirements (data acquired at one rate and recorded at another by queing) must be met and data integrity must be provided regardless of tape failures or I/O errors. Automatic blocking services should be used for multiple local records storage. Avoidance of operational delays and loss of data is essential, noting the problems of the tape drives of the NASA Langley cryogenic wind tunnel. A series of block diagrames and flow charts are included.

Richard W Ross↗