Search NASA⌕ Search

SEARCH · Search NASA

Results for “data sets”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 613 records · Page 34

Monitoring Global Precipitation Using Satellite Observations: Status and Future

The current status of monitoring global precipitation amounts and patterns is described using data sets from the Global Precipitation Climatology Project (GPCP) of the World Climate Research Program (WCRP) and from recent research satellites, especially the Tropical Rainfall Measuring Mission (TRMM). The GPCP monthly (and pentad) data set is a 23-year, globally complete precipitation analysis that is used to explore global and regional variations and trends. The data set is a blend of data mainly from low-orbit microwave satellites and geosynchronous infrared satellites, with additional input from satellite sounder data, Outgoing Longwave Radiation (OLR) data and raingauges. The monthly GPCP data set shows no significant global trend in precipitation over the twenty years, unlike the positive trend in global surface temperatures over the past century. Regional trends are also analyzed. A trend pattern that is a combination of both El Nino and La Nina precipitation features is evident in the 23-year data set. This pattern is related to an increase with time in the number of combined months of El Nino and La Nina during the 23-year period. This apparent trend may be a short-term variation, but also might be related to the increase with time of extreme precipitation events reported elsewhere. Patterns of precipitation variation related to ENSO and other phenomena are shown with clear signals extending from the Tropics into middle and high latitudes of both hemispheres. Also shown, as an example of higher time resolution data is the GPCP daily analysis, which is available for the last six years. A second focus of the talk is on TRMM precipitation data and how these newer data sets incorporating information from the first space-borne meteorological radar compare with the established GPCP data sets.

Adler, Robert F.↗

Alternate physical formats for storing data in HDF

Since its inception HDF has evolved to meet new demands by the scientific community to support new kinds of data and data structures, larger data sets, and larger numbers of data sets. The first generation of HDF supported simple objects and simple storage schemes. These objects were used to build more complex objects such as raster images and scientific data sets. The second generation of HDF provided alternate methods of storing data elements, making it possible to do such things as store extendible objects with in HDF, to store data externally from HDF files, and support data compression effectively. As we look to the next generation of HDF, we are considering fundamental changes to HDF, including a redefinition of the basic HDF object from a simple object to a more general, higher-level scientific data object that has certain characteristics, such as dimensionality, a more general atomic number type, and attributes. These changes suggest corresponding changes to the HDF file format itself.

Folk, Mike↗

Satellite Imagery of PV Site Storm Damage

"This repository contains multiple data sets focused on visible damage to photovoltaic (PV) installations following extreme weather events such as hailstorms and hurricanes. Data sets are split into two categories: the first category, the ‘manually labeled’ data, was compiled by researchers manually, and contains manually identified PV sites exposed to storms. The second data set, the ‘aggregated’ data, is a compilation of the manually labeled PV sites and deep learning-identified PV sites. The hail damage data set focuses on post-storm PV damage following a September 24, 2023 hailstorm in Austin, TX, which caused over $600 million in damages in the Austin metro area. The hurricane damage data set focuses on post-storm PV damage following Hurricanes Irma and Maria in Puerto Rico and the US Virgin Islands. Hurricanes Irma and Maria were back-to-back category 5 hurricanes, which pummeled the Caribbean and southeastern United States in September 2017, causing an estimated $115.2 billion in damages."

14 SOLAR ENERGY↗

Meteorological satellite products support for project COHMEX

The first year effort focussed on real-time support and satellite data collection during the field phase of COHMEX. Work efforts following the field phase of COHMEX concentrated on post-processing of the real-time data sets, and generation of enhanced, research-quality satellite data sets for selected COHMEX core days. These satellite-derived data sets will augment the special COHMEX conventional data base with high horizontal and temporal resolution information. The data sets will be examined for their usefulness in delineating important elements in the meteorological environment leading to convective activity. In addition, a limited research effort was conducted using the Cooperative Institute for Meteorological Satellite Studies (CIMSS) 4-d data assimilation system in conjunction with evaluating VISSR Atmospheric Sounder (VAS) and His-resolution Interferometer Sounder (HIS) data. The need to address the characteristics of the data types, and the problems they introduce into 4-d assimilation procedures is evident. The HIS instrument was flown aboard an ER-2 aircraft on several occasions during COHMEX. One of the flights was chosen for further study. Processed VAS soundings and COHMEX radiosonde data were also collected for this day. The case study included an evaluation of the HIS and VAS data and an impact study of the data on the assimilation system analysis.

Velden, Christopher S.↗

NSSDC data listing

The National Space Science Data Center (NSSDC) Data Listing is in an abbreviated form compared to the data catalogs normally published by NSSDC/WDC-A-R&S. It is organized by NSSDC spacecraft common name. The launch date and NSSDC ID are printed for each spacecraft. The experiments are listed alphabetically by the principal investigator's name and NSSDC ID are printed for each experiment. The data sets are listed by NSSDC ID following the experiment name. The data set name, data form code, quantity of data, and the time span of the data as verified by NSSDC are printed for each data set. Information on NSSDC facilities and ordering procedures are included.

Source record↗

NSSDC data listing

The first part of this listing, Satellite Data, is in an abbreviated form compared to the data catalogs published by NSSDC. It is organized by NSSDC spacecraft common name. The launch date and NSSDC ID are printed for each spacecraft. The experiments are listed alphabetically by the principal investigator's or team leader's last name following the spacecraft name. The experiment name and NSSDC ID are printed for each experiment. The data sets are listed by NSSDC ID following the experiment name. The data set name, data form code, quantity of data, and the time span of the data as verified by NSSDC are printed for each data set.

Source record↗

Accessing and Visualizing scientific spatiotemporal data

This paper discusses work done by JPL 's Parallel Applications Technologies Group in helping scientists access and visualize very large data sets through the use of multiple computing resources, such as parallel supercomputers, clusters, and grids These tools do one or more of the following tasks visualize local data sets for local users, visualize local data sets for remote users, and access and visualize remote data sets The tools are used for various types of data, including remotely sensed image data, digital elevation models, astronomical surveys, etc The paper attempts to pull some common elements out of these tools that may be useful for others who have to work with similarly large data sets.

data sets↗

Earth-Based Ring Occultation to the PDS Rings Node

The major task of this work has been to collect, document, and translate a wide variety of occultation data sets to a common format. This work has been completed. Table I lists the 1977-1987 Uranus occultation data sets that have been retrieved and which are ready for conversion to the PDS occultation standard format, once it has been standardized. The primary documentation and original data tapes for these occultations are part of the MIT data archive, and will be submitted to the PDS from MIT. I have been in contact with Jim Elliot and Steve McDonald, who are carrying out this work, and they have informed me that some of the original data tapes have bad blocks in them. As called upon, I will continue to provide the MIT group with whatever information they require in order that the most complete versions of the data sets be provided to the PDS. 2. All data sets from MIT NOVA disk platters were transcribed to Exabyte tapes, and software was written to enable the NOVA binary data formats to be converted to IEEE standard format. The number of these files which will be contributed to the PDS will depend on the success of the MIT group in reading the original data tapes for the events recorded on these disk platters. As needed, I will supply the MIT group with translated versions of these salvaged files for submission to the PDS. 3. The imaging observations from McDonald and Lick Observatories have been calibrated in both radius and optical depth, and are available in simple binary format. 4. Final submission of the Uranus ring data sets to the PDS is awaiting final agreement by the Rings Node Advisory Council on the PDS FITS format to be adopted for occultation data set.

French, Richard G.↗

A study of techniques for processing multispectral scanner data

A linear decision rule to reduce the time required for processing multispectral scanner data is developed. Test results are presented which justify the use of the new rule for digital processing whenever both accuracy and processing time are important. A method of evaluating the performance of the rule is also developed and applied to the problem of choosing a subset of channels. A technique used to find linear combinations of channels is described. The ability to extend signatures throughout a small area of approximately fifty square miles is tested. After preprocessing, signatures derived from the first of seven overlapping data sets are applied to all data sets. The test results show that the average probability of misclassification tends to increase with an increase in the number of data sets over which the signatures are extended.

Crane, R. B.↗

Geodynamics branch data base for main magnetic field analysis

The data sets used in geomagnetic field modeling at GSFC are described. Data are measured and obtained from a variety of information and sources. For clarity, data sets from different sources are categorized and processed separately. The data base is composed of magnetic observatory data, surface data, high quality aeromagnetic, high quality total intensity marine data, satellite data, and repeat data. These individual data categories are described in detail in a series of notebooks in the Geodynamics Branch, GSFC. This catalog reviews the original data sets, the processing history, and the final data sets available for each individual category of the data base and is to be used as a reference manual for the notebooks. Each data type used in geomagnetic field modeling has varying levels of complexity requiring specialized processing routines for satellite and observatory data and two general routines for processing aeromagnetic, marine, land survey, and repeat data.

Langel, Robert A.↗

From BERTopic to SysML: Informing Model-Based Failure Analysis With Natural Language Processing for Complex Aerospace Systems

The development of emerging complex aerospace systems will require new approaches for capturing safety incident scenarios as early as possible in the design phase. However, for novel systems, relevant data available is limited. In this work, we propose a framework informing model-based mission assurance activities with historical incident reports, lessons learned, or other relevant engineering documents using natural language processing. In doing so, we investigate whether there is useful information in data sets that are relevant, if not identical, to the system under design and whether, through rigorous systems engineering practice, this information can be effectively leveraged through model-based failure analysis. In a worked case study, we apply state-of-the-art topic modeling techniques to two data sets, a mission relevant data set and a system relevant data set. The sets of topics are merged and interpreted to form a preliminary list of failure topics that can be used to inform the identification of off-nominal modes in the model-based failure modes and effects analysis development. Once data from the system in operation is available, it can be used to update the topics identified. By extracting information about likely failures from relevant historical data sets and utilizing model-based mission assurance to ensure relevance and rigor, unanticipated failures can be reduced, and projects can more effectively learn from past missions.

Failure Analysis↗

From BERTopic to SysML: Informing Model-Based Failure Analysis With Natural Language Processing for Complex Aerospace Systems

The development of emerging complex aerospace systems will require new approaches for capturing safety incident scenarios as early as possible in the design phase. However, for novel systems, relevant data available is limited. In this work, we propose a framework informing model-based mission assurance activities with historical incident reports, lessons learned, or other relevant engineering documents using natural language processing. In doing so, we investigate whether there is useful information in data sets that are relevant, if not identical, to the system under design and whether, through rigorous systems engineering practice, this information can be effectively leveraged through model-based failure analysis. In a worked case study, we apply state-of-the-art topic modeling techniques to two data sets, a mission relevant data set and a system relevant data set. The sets of topics are merged and interpreted to form a preliminary list of failure topics that can be used to inform the identification of off-nominal modes in the model-based failure modes and effects analysis development. Once data from the system in operation is available, it can be used to update the topics identified. By extracting information about likely failures from relevant historical data sets and utilizing model-based mission assurance to ensure relevance and rigor, unanticipated failures can be reduced, and projects can more effectively learn from past missions.

Failure Analysis↗

A review of the LEC performance evaluation of UHMLE

Several candidate signature extensions algorithms were evaluated. Four simulated data sets and seven consecutive day data sets were used. Analysis and evaluation of the UHMLE test and recommendations on changes in the UHMLE algorithm motivated by the test are given for each data set. The criterion for evaluation of each algorithm was overall classification accuracy.

Decell, H. P., Jr.↗

Wildfire Emergency Response Hazard Extraction and Analysis of Trends (HEAT) through Natural Language Processing and Time Series

A methodology for Hazard Extraction and Analysis of Trends (HEAT) is proposed and conducted on a data set of wildfire incident response forms, known as ICS-209-PLUS.The HEAT processes: (1) extract a set of hazards from a data set, (2) calculate hazard-relevant metrics in a primary analysis, (3) analyze trends over time in metrics using timeseries, and (4) examine potential explanations for metric trends using a secondary analysis. Hazards are extracted from narrative data in the ICS-209-PLUS based on a framework previously developed by the authors, using natural language processing. Metrics examined for each hazard include operational time to occurrence, rate of occurrence, frequency, and severity. Primary results include a taxonomy of hazards present in the data set with relevant quantitative metrics. The most frequent hazards identified are environmental and include hazardous terrain. Most hazards occur on average between 35-55% containment. Incidents with hazards tend to have a higher average severity score when compared to the average score for all incidents. Time series of the metrics and relevant predictors, including fire characteristics, fire intensity, and operations, are created to facilitate further analysis. Secondary results used to determine which factors best predict hazard frequency include a correlation matrix and regression analysis. These findings are relevant to safety for current, as well as emerging wildfire operations, and are an exploratory first step in developing historical data-driven risk assessment models.

Sequoia R. Andrade↗

AVIRIS and TIMS data processing and distribution at the land processes distributed active archive center

The U.S. Government has initiated the Global Change Research program, a systematic study of the Earth as a complete system. NASA's contribution of the Global Change Research Program is the Earth Observing System (EOS), a series of orbital sensor platforms and an associated data processing and distribution system. The EOS Data and Information System (EOSDIS) is the archiving, production, and distribution system for data collected by the EOS space segment and uses a multilayer architecture for processing, archiving, and distributing EOS data. The first layer consists of the spacecraft ground stations and processing facilities that receive the raw data from the orbiting platforms and then separate the data by individual sensors. The second layer consists of Distributed Active Archive Centers (DAAC) that process, distribute, and archive the sensor data. The third layer consists of a user science processing network. The EOSDIS is being developed in a phased implementation. The initial phase, Version 0, is a prototype of the operational system. Version 0 activities are based upon existing systems and are designed to provide an EOSDIS-like capability for information management and distribution. An important science support task is the creation of simulated data sets for EOS instruments from precursor aircraft or satellite data. The Land Processes DAAC, at the EROS Data Center (EDC), is responsible for archiving and processing EOS precursor data from airborne instruments such as the Thermal Infrared Multispectral Scanner (TIMS), the Thematic Mapper Simulator (TMS), and Airborne Visible and Infrared Imaging Spectrometer (AVIRIS). AVIRIS, TIMS, and TMS are flown by the NASA-Ames Research Center ARC) on an ER-2. The ER-2 flies at 65000 feet and can carry up to three sensors simultaneously. Most jointly collected data sets are somewhat boresighted and roughly registered. The instrument data are being used to construct data sets that simulate the spectral and spatial characteristics of the Advanced Spaceborne Thermal Emission and Reflection Radiometer (ASTER) instrument scheduled to be flown on the first EOS-AM spacecraft. The ASTER is designed to acquire 14 channels of land science data in the visible and near-IR (VNIR), shortwave-IR (SWIR), and thermal-IR (TIR) regions from 0.52 micron to 11.65 micron at high spatial resolutions of 15 m to 90 m. Stereo data will also be acquired in the VNIR region in a single band. The AVIRIS and TMS cover the ASTER VNIR and SWIR bands, and the TIMS covers the TIR bands. Simulated ASTER data sets have been generated over Death Valley, California, Cuprite, Nevada, and the Drum Mountains, Utah using a combination of AVIRIS, TIMS, amd TMS data, and existing digital elevation models (DEM) for the topographic information.

Mah, G. R.↗

Mass and Ozone Fluxes from the Lowermost Stratosphere

Net mass flux from the stratosphere to the troposphere can be computed from the heating rate along the 380K isentropic surface and the time rate of change of the mass of the lowermost stratosphere (the region between the tropopause and the 380K isentrope). Given this net mass flux and the cross tropopause diabatic mass flux, the residual adiabatic mass flux across the tropopause can also be estimated. These fluxes have been computed using meteorological fields from a free-running general circulation model (FVGCM) and two assimilation data sets, FVDAS, and UKMO. The data sets tend to agree that the annual average net mass flux for the Northern Hemisphere is about 1P10 kg/s. There is less agreement on the southern Hemisphere flux that might be half as large. For all three data sets, the adiabatic mass flux is computed to be from the upper troposphere into the lowermost stratosphere. This flux will dilute air entering from higher stratospheric altitudes. The mass fluxes are convolved with ozone mixing ratios from the Goddard 3D CTM (which uses the FVGCM) to estimate the cross-tropopause transport of ozone. A relatively large adiabatic flux of tropospheric ozone from the tropical upper troposphere into the extratropical lowermost stratosphere dilutes the stratospheric air in the lowermost stratosphere. Thus, a significant fraction of any measured ozone STE may not be ozone produced in the higher Stratosphere. The results also illustrate that the annual cycle of ozone concentration in the lowermost stratosphere has as much of a role as the transport in the seasonal ozone flux cycle. This implies that a simplified calculation of ozone STE mass from air mass and a mean ozone mixing ratio may have a large uncertainty.

Schoeberl, Mark R.↗

Pilot climate data system user's guide

Instructions for using the Pilot Climate Data System (PCDS), an interactive, scientific data management system for locating, obtaining, manipulating, and displaying climate-research data are presented. The PCDS currently provides this supoort for approximately twenty data sets. Figures that illustrate the terminal displays which a user sees when he/she runs the PCDS and some examples of the output from this system are included. The capabilities which are described in detail allow a user to perform the following: (1) obtain comprehensive descriptions of a number of climate parameter data sets and the associated sensor measurements from which they were derived; (2) obtain detailed information about the temporal coverage and data volume of data sets which are readily accessible via the PCDS; (3) extract portions of a data set using criteria such as time range and geographic location, and output the data to tape, user terminal, system printer, or online disk files in a special data-set-independent format; (4) access and manipulate the data in these data-set-independent files, performing such functions as combining the data, subsetting the data, and averaging the data; and (5) create various graphical representations of the data stored in the data-set-independent files.

Reph, M. G.↗

Global Precipitation Variations and Long-term Changes Derived from the GPCP Monthly Product

Global and large regional rainfall variations and possible long-term changes are examined using the 25-year (1979-2004) monthly dataset from the Global Precipitation Climatology Project (GPCP). The emphasis is to discriminate among the variations due to ENSO, volcanic events and possible long-term changes. Although the global change of precipitation in the data set is near zero, the data set does indicate an upward trend (0.13 mm/day/25yr) and a downward trend (-0.06 mm/day/25yr) over tropical oceans and lands (25S-25N), respectively. This corresponds to a 4% increase (ocean) and 2% decrease (land) during this time period. Techniques are applied to attempt to eliminate variations due to ENSO and major volcanic eruptions. The impact of the two major volcanic eruptions over the past 25 years is estimated to be about a 5% reduction in tropical rainfall. The modified data set (with ENSO and volcano effect removed) retains the same approximate change slopes, but with reduced variance leading to significance tests with results in the 90-95% range. Inter-comparisons between the GPCP, SSWI (1988-2004), and TRMM (1998-2004) rainfall products are made to increase or decrease confidence in the changes seen in the GPCP analysis.

Adler, Robert F.↗