Search NASA⌕ Search

SEARCH · Search NASA

Results for “Archive”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 145 records · Page 8

The NASA Heliophysics Active Final Archive at the Space Physics Data Facility

The 2009 NASA Heliophysics Science Data Management Policy re-defined and extended the responsibilities of the Space Physics Data Facility (SPDF) project. Building on SPDF's established capabilities, the new policy assigned the role of active "Final Archive" for non-solar NASA Heliophysics data to SPDF. The policy also recognized and formalized the responsibilities of SPDF as a source for critical infrastructure services such as VSPO to the overall Heliophysics Data Environment (HpDE) and as a Center of Excellence for existing SPDF science-enabling services and software including CDAWeb, SSCWeb/4D Orbit Viewer, OMNIweb and CDF. We will focus this talk to the principles, strategies and planned SPDF architecture to effectively and efficiently perform these roles, with special emphasis on how SPDF will ensure the long-term preservation and ongoing online community access to all the data entrusted to SPDF. We will layout our archival philosophy and what we are advocating in our work with NASA missions both current and future, with potential providers of NASA and NASA-relevant archival data, and to make the data and metadata held by SPDF accessible to other systems and services within the overall HpOE. We will also briefly review our current services, their metrics and our current plans and priorities for their evolution.

McGuire, Robert E.↗

PDS Archive Release of Apollo 11, Apollo 12, and Apollo 17 Lunar Rock Sample Images

Scientists at the Johnson Space Center (JSC) Lunar Sample Laboratory, Information Resources Directorate, and Image Science & Analysis Laboratory have been working to digitize (scan) the original film negatives of Apollo Lunar Rock Sample photographs [1, 2]. The rock samples, and associated regolith and lunar core samples, were obtained during the Apollo 11, 12, 14, 15, 16 and 17 missions. The images allow scientists to view the individual rock samples in their original or subdivided state prior to requesting physical samples for their research. In cases where access to the actual physical samples is not practical, the images provide an alternate mechanism for study of the subject samples. As the negatives are being scanned, they have been formatted and documented for permanent archive in the NASA Planetary Data System (PDS). The Astromaterials Research and Exploration Science Directorate (which includes the Lunar Sample Laboratory and Image Science & Analysis Laboratory) at JSC is working collaboratively with the Imaging Node of the PDS on the archiving of these valuable data. The PDS Imaging Node is now pleased to announce the release of the image archives for Apollo missions 11, 12, and 17.

Garcia, P. A.↗

Lunar Data Node: Apollo Data Restoration and Archiving Update

The Lunar Data Node (LDN) of the Planetary Data System (PDS) is responsible for the restoration and archiving of Apollo data. The LDN is located at the National Space Science Data Center (NSSDC), which holds much of the extant Apollo data on microfilm, microfiche, hard-copy documents, and magnetic tapes in older formats. The goal of the restoration effort is to convert the data into user-accessible PDS formats, create a full set of explanatory supporting data (metadata), archive the full data sets through PDS, and post the data online at the PDS Geosciences Node. This will both enable easy use of the data by current researchers and ensure that the data and metadata are securely preserved for future use. We are also attempting to locate and preserve Apollo data which were never archived at NSSDC. We will give a progress report on the data sets we have been restoring and future work.

Williams, David R.↗

Stewardship of NASA's Earth Science Data and Ensuring Long-Term Active Archives

Program, NASA has followed an open data policy, with non-discriminatory access to data with no period of exclusive access. NASA has well-established processes for assigning and or accepting datasets into one of 12 Distributed Active Archive Centers (DAACs) that are parts of EOSDIS. EOSDIS has been evolving through several information technology cycles, adapting to hardware and software changes in the commercial sector. NASA is responsible for maintaining Earth science data as long as users are interested in using them for research and applications, which is well beyond the life of the data gathering missions. For science data to remain useful over long periods of time, steps must be taken to preserve: (1) Data bits with no corruption, (2) Discoverability and access, (3) Readability, (4) Understandability, (5) Usability' and (6). Reproducibility of results. NASAs Earth Science data and Information System (ESDIS) Project, along with the 12 EOSDIS Distributed Active Archive Centers (DAACs), has made significant progress in each of these areas over the last decade, and continues to evolve its active archive capabilities. Particular attention is being paid in recent years to ensure that the datasets are published in an easily accessible and citable manner through a unified metadata model, a common metadata repository (CMR), a coherent view through the earthdata.gov website, and assignment of Digital Object Identifiers (DOI) with well-designed landing product information pages.

Data Management↗

The NASA Ames Life Sciences Data Archive: Biobanking for the Final Frontier

The NASA Ames Institutional Scientific Collection involves the Ames Life Sciences Data Archive (ALSDA) and a biospecimen repository, which are responsible for archiving information and non-human biospecimens collected from spaceflight and matching ground control experiments. The ALSDA also manages a biospecimen sharing program, performs curation and long-term storage operations, and facilitates distribution of biospecimens for research purposes via a public website (https:lsda.jsc.nasa.gov). As part of our best practices, a tissue viability testing plan has been developed for the repository, which will assess the quality of samples subjected to long-term storage. We expect that the test results will confirm usability of the samples, enable broader science community interest, and verify operational efficiency of the archives. This work will also support NASA open science initiatives and guides development of NASA directives and policy for curation of biological collections.

Biobank↗

Use of Schema on Read in Earth Science Data Archives

Traditionally, NASA Earth Science data archives have file-based storage using proprietary data file formats, such as HDF and HDF-EOS, which are optimized to support fast and efficient storage of spaceborne and model data as they are generated. The use of file-based storage essentially imposes an indexing strategy based on data dimensions. In most cases, NASA Earth Science data uses time as the primary index, leading to poor performance in accessing data in spatial dimensions. For example, producing a time series for a single spatial grid cell involves accessing a large number of data files. With exponential growth in data volume due to the ever-increasing spatial and temporal resolution of the data, using file-based archives poses significant performance and cost barriers to data discovery and access. Storing and disseminating data in proprietary data formats imposes an additional access barrier for users outside the mainstream research community. At the NASA Goddard Earth Sciences Data Information Services Center (GES DISC), we have evaluated applying the schema-on-read principle to data access and distribution. We used Apache Parquet to store geospatial data, and have exposed data through Amazon Web Services (AWS) Athena, AWS Simple Storage Service (S3), and Apache Spark. Using the schema-on-read approach allows customization of indexing spatially or temporally to suit the data access pattern. The storage of data in open formats such as Apache Parquet has widespread support in popular programming languages. A wide range of solutions for handling big data lowers the access barrier for all users. This presentation will discuss formats used for data storage, frameworks with This presentation will discuss formats used for data storage, frameworks with support for schema-on-read used for data access, and common use cases covering data usage patterns seen in a geospatial data archive.

cloud applications↗

Publishing Variables Archived at GES DISC to Earth System Grid Federation (ESGF)

We present a straightforward and low-cost approach to publish variables archived at NASA Goddard Earth Sciences Data and Information Services Center (GES DISC) to the Earth System Grid Federation (ESGF). An ESGF publication requires a single standard-name variable aggregated over time to facilitate data inter-comparison. It also contains significant metadata to enable searching in ESGF. We look up standard names on high demand in ESGF search history, and using OPeNDAP and NcML technologies we aggregate the corresponding variables available in the GES DISC archive with augmented metadata required by CMIP6 and obs4MIPs Data Specification version 2.1. At this writing 10 variables from a standard product of the Atmospheric Infrared Sounder along with the Tech Notes are published in ESGF by NASA Center for Climate Simulation (NCCS). Users can view, analyze, and subset remotely, and download these aggregated variables via links in any ESGF node after searching. We plan to work on and publish more variables and data from different NASA missions and experiments in our archive.

Fan Fang↗

A Standard Reference Model for Data Archives

An implementable Data Archive Architecture is being developed for trusted digital repositories based on the Reference Model for an Open Archival Information System (OAIS) – ISO 14721. A set of interoperable protocols and interface specifications are planned that will offer capabilities for accessing, merging, and re-using data, both within and across the operational boundaries of trustworthy digital repositories. The model will also provide support for the fundamental scientific need to verify the reproducibility of results. This standards development task is being performed by the Data Archive Interoperability (DAI) working group within the Consultative Committee for Space Data Systems (CCSDS). The architecture integrates concepts from the OAIS Reference Model, the ISO/IEC 11179 Metadata Registry (MDR) standard, the CCSDS Reference Architecture for Space Information Management (RASIM), the proposed draft recommended practice document, Information Preparation to Enable Long Term Use (IPELTU), and three decades of digital repository development for science research.

Ambacher, Bruce↗

Validity of the Landsat Surface Reflectance Archive for Aquatic Science: Implications for Cloud-Based Analysis

Originally developed for terrestrial science and applications, the US Geological Survey Landsat surface reflectance (SR) archive spanning ~ 40 yr of observations has been increasingly utilized in large‐scale water‐quality studies. These products, however, have not been rigorously validated using in situ measured reflectance. This letter quantifies and demonstrates the quality of the SR products by harnessing a sizeable global dataset ( N = 1100). We found that the Landsat 8/9 SR in the green and red bands marginally meet the targeted accuracy requirements (30%), whereas the uncertainties in the blue and coastal‐aerosol bands ranged from 48% to 110%. We further observed > +25% biases in the visible bands of Landsat 5/7 SR, which can introduce an apparent downward trend when applied in time‐series analyses combined with Landsat 8/9. Users must exercise caution when using this archive for trend analyses, and progress in atmospheric correction is required to foster advanced applications of the Landsat archive for aquatic science.

Landsat, lakes, harmful algal blooms↗

High Performance Access to Archival Data Stored in HDF4 and HDF5 on Cloud Object Stores Without Reformatting the Files

Cloud computing offers numerous advantages for users of extensive Earth science data collections. These benefits encompass direct online access to data files and granules from any location, scalable access supporting parallel computing workflows, and flexible computing tools enabling innovative experimentation with processing techniques. However, older archival file formats designed for distinct computing systems hinder efficient access to decade-long time-series data when compared to data stored in modern cloud-optimized formats like Web Object Stores (WOS), exemplified by Amazon Web Services’ Simple Storage Service (S3). We describe DMR++ (Dataset Metadata Response plus plus), a technology facilitating efficient access to HDF5 (Hierarchical Data Format, version 5) and HDF4 files stored on WOS systems without requiring data reformatting. DMR++ achieves performance comparable to technologies like Zarr while preserving the original file structure, a substantial benefit considering the vast quantity of archival files held by organizations such as NASA. Moreover, DMR++ typically outperforms cloud-optimized versions of HDF5. Essentially an XML (Extensible Markup Language) document usually stored alongside the described data, DMR++ can also be generated on-the-fly but is generally created during data staging to the WOS. Archival files that use HDF4/5 often store large arrays of numerical data. The data in these files is often compressed, typically reducing their size by a factor of four or more. To achieve efficient access to portions of those arrays, they are 'chunked' into smaller sub-arrays, each individually compressed. The chunk size is a compromise, where spinning disks can efficiently access data in smaller chunks while S3 favors larger chunks. A simple optimization of aggregating smaller chunks that are stored adjacently, transferring them in a single access and then individually decompressing them will improve performance. NASA data pose an additional challenge: special Application Programmer Interface (API) libraries are often needed to compute some variables. These libraries are incompatible with WOS environments. Our solution involves storing computed values in the DMR++ document or a companion file, making them accessible like other variables and eliminating the need for specialized APIs. We outline specific optimizations for both satellite grid and swath data stored in HDF4-EOS2 (Earth Observing System).

James Gallagher↗

User’s Guide for the NASA High Efficiency Centrifugal Compressor Data Archive

The datasets contained in this archive are associated with the High Efficiency Centrifugal Compressor (HECC) in the Small Engine Components Compressor Test Facility, colloquially referred to as CE-18, at NASA Glenn Research Center. The archive is accessible at https://storage.googleapis.com/hecc-data/NASA-HECC-Data-Archive.zip. The documentation contained herein provides context for the data hosted on data.nasa.gov. The datasets and accompanying content in this document will be updated periodically as additional data is procured analyzed. The revision of the document is provided by date in the footer, and the revision updates are provided in the Revisions section. Please contact Trey Harrison (email: herbert.harrison@nasa.gov) for inquiries related to the dataset and documentation or to be added to an email list to be notified of updates and additions to the archive.

radial turbomachinery↗

Deep desert aquifers as an archive for Mid- to Late Pleistocene hydroclimate: An example from the southeastern Mediterranean

Many efforts have been made to illuminate the nature of past hydroclimates in semi-arid and arid regions, where current and future shifts in water availability have enormous consequences on human subsistence. Deep desert aquifers, where groundwater is stored for prolonged periods, might serve as a direct record of major paleo-recharge events. To date, groundwater-based paleoclimate reconstructions have mainly focused on a relatively narrow timescale (up to ∼40 kyr), limited by the relatively short half-life of the widely used radiocarbon (5.73 kyr). Here we demonstrate the usage of deep regional aquifers in the arid southeastern Mediterranean as a hydroclimate archive for earlier Mid-to-Late Pleistocene epochs. State-of-the-art dating tools, primarily the 81 Kr radioisotope (t 1/2 = 229 kyr), were combined with other atmosphere-derived tracers to illuminate the impact of four distinguishable wetter episodes over the past 400 kyr, with differences in climatic conditions and paleo-recharge locations. Variations in stable water isotope composition suggest moisture transport from more proximal (Mediterranean) and distal (Atlantic) sources to different parts of the region at distinct times. Large variability in the computed noble gas-based recharge temperature (NGT), ranging ~15–30 °C, cannot be explained by climate variations solely, and points to different recharge pathways, including geothermal heating in the deep unsaturated zone and recharge from high-elevation (colder) regions. The obtained groundwater record complements and enhances the interpretation of other terrestrial archives in the arid region, including a contribution of valuable information regarding the moisture source origin as reflected in the deuterium-excess values, which is unattainable from the common practice analysis of calcitic cave deposits. We conclude that similar applications in other deep (hundred-m-order) regional groundwater systems (e.g., the Sahara desert aquifers) can significantly advance our understanding of long-term (up to 1 Myr) paleo-hydroclimate in arid regions, including places where no terrestrial remnants, such as cave, lake, and spring sediments, are available.

Atom Trap Trace Analysis↗

Model Data Archive for Manuscript Titled "Evaluation of a Coupled Surface–Subsurface Hydrologic Model Using Dense Water‑Level Sensors in a Mixed Urban–Rural Watershed"

This archive provides scripts, input files, and datasets used for the implementation and evaluation of a fully coupled surface–subsurface hydrologic model in the Neches River Basin, southeast Texas. The study uses the Advanced Terrestrial Simulator (ATS) to simulate coupled surface–subsurface hydrologic processes over a mixed urban–rural watershed and evaluates model performance using a dense network of 136 in situ water-level sensors, nine U.S. Geological Survey (USGS) stream gauges, and SSEBop-derived evapotranspiration estimates during the period October 2014–June 2024. The workflow is implemented primarily in Python 3 using the Watershed Workflow package. The Jupyter notebooks can be executed using open-source software such as Anaconda JupyterLab or Visual Studio Code. Other data files include TXT, CSV, XML, SHP, TIF, NetCDF, HDF5, and ExodusII files, which can be processed using the provided Python scripts. ATS input files are provided in XML format and can be edited using any commonly used text editor. This archive contains: *Scripts and input files used to generate the ATS model setup, including watershed discretization, mesh generation, parameter mapping, and model configuration. *Jupyter notebooks used for preprocessing observational data, evaluating streamflow, water levels, and evapotranspiration, computing performance metrics, and generating the figures presented in the manuscript. *ATS simulation outputs and processed observational datasets, including OneRain and DD6 water-level sensors, USGS streamflow observations, GIS data, and supporting spatial datasets used throughout the study.

Dense water-level sensor network↗

Data Archival and Retrieval Enhancement (DARE) Metadata Modeling and Its User Interface

The Defense Nuclear Agency (DNA) has acquired terabytes of valuable data which need to be archived and effectively distributed to the entire nuclear weapons effects community and others...This paper describes the DARE (Data Archival and Retrieval Enhancement) metadata model and explains how it is used as a source for generating HyperText Markup Language (HTML)or Standard Generalized Markup Language (SGML) documents for access through web browsers such as Netscape.

The Defense Nuclear Agency DNA DARE Data Archival ↗

(abstract) Satellite Physical Oceanography Data Available From an EOSDIS Archive

The Physical Oceanography Distributed Active Archive Center (PO.DAAC) at the Jet Propulsion Laboratory archives and distributes data as part of the Earth Observing System Data and Information System (EOSDIS). Products available from JPL are largely satellite derived and include sea-surface height, surface-wind speed and vectors, integrated water vapor, atmospheric liquid water, sea-surface temperature, heat flux, and in-situ data as it pertains to satellite data. Much of the data is global and spans fourteen years.There is email access, a WWW site, product catalogs, and FTP capabilities. Data is free of charge.

oceanography Earth Observing System oceans data cl↗

GHRC: NASAs Hazardous Weather Distributed Active Archive Center

The Global Hydrology Resource Center (GHRC; ghrc.nsstc.nasa.gov) is one of NASA's twelve Distributed Active Archive Centers responsible for providing access to NASA's Earth science data to users worldwide. Each of NASA's twelve DAACs focuses on a specific science discipline within Earth science, provides data stewardship services and supports its research community's needs. Established in 1991 as the Marshall Space Flight Center DAAC and renamed GHRC in 1997, the data center's original mission focused on the global hydrologic cycle. However, over the years, data holdings, tools and expertise of GHRC have gradually shifted. In 2014, a User Working Group (UWG) was established to review GHRC capabilities and provide recommendations to make GHRC more responsive to the research community's evolving needs. The UWG recommended an update to the GHRC mission, as well as a strategic plan to move in the new direction. After a careful and detailed analysis of GHRC's capabilities, research community needs and the existing data landscape, a new mission statement for GHRC has been crafted: to provide a comprehensive active archive of both data and knowledge augmentation services with a focus on hazardous weather, its governing dynamical and physical processes, and associated applications. Within this broad mandate, GHRC will focus on lightning, tropical cyclones and storm-induced hazards through integrated collections of satellite, airborne, and in-situ data sets. The new mission was adopted at the recent 2015 UWG meeting. GHRC will retain its current name until such time as it has built substantial data holdings aligned with the new mission.

Data Archive↗

Kepler Archive Manual

A description of Kepler, its design, performance and operational constraints may be found in the Kepler Instrument Handbook (KIH, Van Cleve Caldwell 2016). A description of Kepler calibration and data processing is described in the Kepler Data Processing Handbook (KDPH, Jenkins et al. 2016; Fanelli et al. 2011). Science users should also consult the special ApJ Letters devoted to early Kepler results and mission design (April 2010, ApJL, Vol. 713 L79-L207). Additional technical details regarding the data processing and data qualities can be found in the Kepler Data Characteristics Handbook (KDCH, Christiansen et al. 2013) and the Data Release Notes (DRN). This archive manual specifically documents the file formats, as they exist for the last data release of Kepler, Data Release 25(KSCI-19065-002). The earlier versions of the archive manual and data release notes act as documentation for the earlier versions of the data files.

Kepler↗

How NASA is Building a Petabyte Scale Geospatial Archive in the Cloud

NASA's Earth Observing System Data and Information System (EOSDIS) is working towards a vision of a cloud-based, highly-flexible, ingest, archive, management, and distribution system for its ever-growing and evolving data holdings. This free and open source system, Cumulus, is emerging from its prototyping stages and is poised to make a huge impact on how NASA manages and disseminates its Earth science data. This talk outlines the motivation for this work, present the achievements and hurdles of the past 18 months and charts a course for the future expansion of Cumulus. We explore not just the technical, but also the socio-technical challenges that we face in evolving a system of this magnitude into the cloud. The NASA EOSDIS archive is currently at nearly 30 PBs and will grow to over 300PBs in the coming years. We've presented progress on this effort at AWS re:Invent and the American Geophysical Union (AGU) Fall Meeting in 2017 and hope to have the opportunity to share with FOSS4G attendees information on the availability of the open sourced software and how NASA intends on making its Earth Observing Geospatial data available for free to the public in the cloud.

NASA Archive↗