Search NASA⌕ Search

SEARCH · Search NASA

Results for “Metadata”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 361 records · Page 20

SPRUCE Air 13C and 14C Isotopes, Marcell Experimental Forest, Minnesota, 2016-2025

This data set reports 13C (carbon) and 14C signatures of air from Spruce and Peatland Responses Under Changing Environments (SPRUCE) experimental study plots located in the S1-Bog at the Marcell Experimental Forest in northern Minnesota from 2016-2025 (2016-04-13 to 2025-09-24). Air measurements were collected approximately five times throughout each active growing season beginning in 2016 and included the 10 SPRUCE experimental plots (Plots 4, 6, 8, 10, 11, 13, 16, 17, 19, and 20) two ambient co-located ambient plots (Plots 7 and 21) and a site at the Marcell Experiment Station’s S2-Bog meteorological station (MET) located 1.7 km northeast of the SPRUCE experimental site on the S1-Bog. Air samples were assessed for both 13C- and 14C-CO2 (carbon-carbon dioxide) signatures. 13C and 14C isotopic signatures can be used in end-member analysis and C-cycle models to track the movement of C within the experimental ecosystem, calculate turnover times within plant tissues, and to test mechanisms used in models. This dataset contains one data file in comma separate (*.csv) format. Additional metadata are provided: one data dictionary and a file-level metadata file in comma separate (*.csv) format and a user guide in PDF (*.pdf) format.

54 ENVIRONMENTAL SCIENCES↗

SPRUCE Root Tip and Ectomycorrhizal Fungi Colonization Measurements from Ingrowth Cores, 2017

This data set contains root tip and ectomycorrhizal fungi colonization measurements taken from ingrowth cores from the SPRUCE experiment (Hanson et al. 2017) that were deployed during the 2017 growing season (2017-06 to 2017-10-01). This study explored the relationship between warming treatments and fine-root growth. Increased fine-root growth may increase root exudates and accelerate turnover, representing an underlying mechanism for peat decomposition through priming, as exudates provide a labile carbon source to the microbial community. Roots of two tree species were studied: an evergreen conifer Picea mariana (black spruce) and a deciduous conifer Larix laricina (tamarack). Measurements include root tips counts and densities by tree species and the abundance of ectomycorrhizal colonization on root tips. This dataset contains one data file in comma separate (.csv) format. Additional metadata are provided: one data dictionary and a file-level metadata file in comma separate (.csv) format and a user guide in PDF (*.pdf) format.

black spruce [Picea mariana]↗

SPRUCE Photosynthesis and Respiration of Picea mariana and Larix laricina in SPRUCE Experimental Plots, 2019

This dataset contains physiological, morphological, and chemical measurements of the two dominant coniferous species, Picea mariana and Larix laricina, in August 2019 (2019-08-20 to 2019-08-22) at the SPRUCE (Spruce and Peatland Responses under Changing Environments) experiment site in the Marcell Experimental Forest in northern Minnesota, USA. These observations help to assess the effects of whole ecosystem scale warming and elevated carbon dioxide (CO2) concentrations on peatland ecosystems. Measurements include light-saturated photosynthesis and foliar dark respiration measurements under standard conditions and growth conditions involving varying temperatures and atmospheric CO2 concentrations, as well as leaf morphology measurements (leaf mass per unit leaf area) and nitrogen content based on mass and leaf area. Net photosynthesis and dark respiration measurements were taken using portable photosynthesis systems (LI6400XT, LI6800, LI-COR Biosciences, USA). This dataset contains one data file in comma-separate values (*.csv) format. Additional metadata are provided: a data dictionary and a file-level metadata file in comma-separate values (.csv) format and a user guide in PDF (*.pdf) format.

54 ENVIRONMENTAL SCIENCES↗

SPRUCE Surface N2O fluxes measured with LI-7820, 2024

This dataset contains N2O (nitrous oxide) efflux rates measurements from the Spruce and Peatland Responses Under Changing Environments (SPRUCE) experimental site within the Marcell Experimental Forest in northern Minnesota, USA. Measurements were made manually with a LiCor N2O/H2O analyzer (LI-7820) and paired SmartChamber (LI-8200-01S) in June, August, and October (2024-06-24 to 2024-10-22). During each measurement, the SmartChamber was placed on 8” PVC collars that were installed in May 2024. N2O flux was derived from 10-minute flux measurements processed using SoilFluxPro software (v5.3.1) and fit to a linear model. Model slope and R2 are reported along with soil water, soil temperature, and air temperature observations made with SmartChamber sensors. N2O is a gaseous N species formed during the microbial processes of denitrification and ammonia oxidation, and is a powerful greenhouse gas. This dataset contains one data file in comma-separate values (*.csv) format. Additional metadata are provided: one data dictionary and a file-level metadata file in comma-separate values (.csv) format and a user guide in PDF (*.pdf) format.

54 ENVIRONMENTAL SCIENCES↗

SPRUCE Root Production Assessed with Manual Minirhizotrons Resolved to Plant Functional Type, 2015-2021

This dataset contains raw root length and diameter for individual roots and estimated root population production measurements from the Spruce and Peatland Responses Under Changing Environments (SPRUCE) experimental site within the Marcell Experimental Forest in northern Minnesota, USA. Measurements started at the beginning of whole ecosystem warming manipulations in 2015 through 2021 (2015-05-26 to 2021-09-01). Root morphology and estimated production were quantified throughout the peat profile with manual minirhizotrons deployed within SPRUCE plots. Images were processed using commercial software to quantify the length and diameter of individual roots. Roots were visually assigned to a plant functional type (PFT) of either (ericaceous) shrub, herb (sedges and Maianthemum trifolium), or tree (Larix laricina, Picea mariana) based on expert opinion. The biomass of individual roots was estimated using PFT-specific allometric equations (Iversen et al., 2018). Production per day was estimated as the length of new roots produced between imaging sessions, divided by the number of days between imaging sessions. These values were placed on a m2 aboveground area basis and scaled to a standard depth of 1m (roots are not evenly distributed with depth, do not interpret value as being on a m3 basis). Maximum and average (weighted by production length) depth of each PFT were also estimated within each minirhizotron tube. Annual production was interpolated as the average of four methods to scale these data (see Weber et al, 2026). Standing crop of roots was estimated for each tube as the maximum visible amount (both length and mass) of roots of that PFT for that year. These data expand the ability of researchers to accurately estimate the belowground dynamics of peatland vegetation, as well as the role that fine roots may play in impacting the fluxes of carbon within peatlands. This dataset contains three data files in comma-separate values (*.csv) format. This dataset contains one data file in comma-separate values (.csv) format. Additional metadata are provided: three data dictionaries and a file-level metadata file in comma-separate values (.csv) format and a user guide in PDF (*.pdf) format.

54 ENVIRONMENTAL SCIENCES↗

SPRUCE Redox-Active Subsurface Organic Matter, Marcell Experimental Forest, Minnesota, 2023

This dataset contains measurements that report on the effects of the SPRUCE experimental treatments on redox-active organic matter (RAOM) reduction (Valenzuela and Cervantes, 2021). Measurements occurred at the SPRUCE Experiment site in the Marcell Experimental Forest in northern Minnesota, USA. This work is also a follow-up to Rush et al. (2021a) which investigated effects of temperature on RAOM reduction after two years of experimental warming (Rush et al. 2021b). This follow-up dataset addresses two main questions; (i) How does warming and elevated carbon dioxide (CO2) directly affect in situ RAOM reduction, and subsequent methane (CH4) and CO2 production, across the peat depth profile? and (ii) How has long-term warming and elevated CO2 changed the total RAOM pool, and subsequent CH4 and CO2 production, across the peat depth profile? This dataset reports electron shuttling capacity (a proxy for RAOM reduction; Keller, and Takagi, 2013) and carbon dioxide (CO2) and methane (CH4) concentrations both in one-week in situ incubations (2023-05-31 to 2023-08-01) and 42-day laboratory incubations from peat collected in 2023 (2023-05-31 to 2023-06-26). Laboratory incubations also measured acetate concentrations. The 2023 laboratory incubations were also compared with laboratory incubations conducted on peat collected in 2016 (Rush et al. 2021b). This dataset contains three data files in comma-separate (.csv) format. Additional metadata are provided: three data dictionaries and a file-level metadata file in comma separate (.csv) format and a user guide in PDF (*.pdf) format.

54 ENVIRONMENTAL SCIENCES↗

Thaw depth, soil moisture, and vegetation height, Teller and Kougarok sites, Seward Peninsula, Alaska, 2022

Thaw depth, soil moisture, and vegetation height sampled from locations on the Teller MM27, Kougarok MM80, and Kougarok MM82 NGEE-Arctic sites, Seward Peninsula, Alaska. These data were collected in support of the ongoing NGEE-Arctic and NASA ABoVE data synthesis work. Samples were collected in July 2022, including 6 transects covering the entire Teller MM27 watershed and 4 transects covering 4 thaw ponds at Kougarok MM80 and MM82. This data package includes sample information and thaw depth, soil moisture, vegetation height data (.csv). Metadata files include data descriptions (_dd.csv) for tabular data and file level metadata (.csv). The Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic), was a research effort to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research. The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska. Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

54 ENVIRONMENTAL SCIENCES↗

Leaf area index (LAI), Teller site, Seward Peninsula, Alaska, 2022

Leaf area index (LAI) sampled from locations on the Teller MM27 NGEE-Arctic site, Seward Peninsula, Alaska. These data were collected in support of the ongoing NGEE-Arctic and NASA ABoVE data synthesis work. Samples were collected in July 2022 at 100 locations that cover the down slope half of Teller MM27 hillslope. This data package includes sample information and LAI data (.csv), an ESRI shape file (.shp) and Keyhole Markup Language (.kml) that define the location and area of each LAI measurement. Metadata files include data descriptions (_dd.csv) for tabular data, file level metadata (.csv), and methods and the instrument manual (.pdf). The Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic), was a research effort to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research. The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska. Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

54 ENVIRONMENTAL SCIENCES↗

ISO 191* Revisions and Amendments

ISO standards have changed recently in response to new requirements from introduced by NASA, NOAA, and other ESIP members. These changes address important uses cases related to identification, parameter metadata, and platform/instrument/sensor metadata. Each change will be described along with discussion of how they may be used.

data quality↗

Using UMM-Var and E2E to Improve the User Experience for Accessing NASA EOSDIS Data Sets

The UMM-Variables (Var) Metadata Model has been evolved to support an End-to-End Services (E2E) capability, which enables variable level subsetting, data transformation, and data reformatting. This talk will discuss what is new with the model, how users can get their metadata ready for the E2E capability, and include a demo of how the model is being used to drive and improve the user experience in Earthdata Search Client when accessing EOSDIS data sets.

Metadata Modeling↗

New Ways of Facilitating Improved Data Discovery and Access for NASA's Suborbital Earth Science Observations

NASA conducts field research in various Earth Science disciplines utilizing airborne and other non-satellite platforms to acquire in situ and remotely sensed observations indicative of physical processes across a range of scales. Field efforts are key in the development and validation of instruments and satellite algorithm refinements. The heterogeneous data, with a range of file formats, scales, and acquisition methods, support research in several science areas. NASA’s archive process assigns data products to discipline-oriented Distributed Active Archive Centers (DAACs) for stewardship. Over time, individual DAACs have developed tools for data browsing and serving disparate user bases. As science becomes more interdisciplinary, researchers need to incorporate observations from multiple campaigns, and multiple DAACs, into their work. Motivated in part by this shifting paradigm of needs, the Catalog of Archived Suborbital Earth Science Investigations (CASEI) was created. CASEI provides a single starting point to browse, search, and discover airborne and field data. Contextual metadata are organized and inter-linked allowing intuitive, integrated exploration across all NASA DAACs. Campaign science objectives, platform and instrument configurations, geographical details, geophysical concepts, and more are tracked in CASEI’s database, facilitating multi-parameter search, browse, and discovery of relevant data products. Researchers are able to directly access associated data products, via DOI links, regardless of the DAAC where they reside. Significant events, key time periods of high science interest within the longer-duration campaign effort, are also indicated and allow for a more efficient identification of critical data subsets. This presentation describes CASEI’s development, intensive metadata curation process, and demonstrates the web interface experience. Initial content metrics and plans for continued maintenance will also be discussed.

metadata↗

Greenland and Canadian Arctic Ice Temperature Profiles Database

Here, we present a compilation of 95 ice temperature profiles from 85 boreholes from the Greenland ice sheet and peripheral ice caps, as well as local ice caps in the Canadian Arctic. Profiles from only 31 boreholes (36 %) were previously available in open-access data repositories. The remaining 54 borehole profiles (64 %) are being made digitally available here for the first time. These newly available profiles, which are associated with pre-2010 boreholes, have been submitted by community members or digitized from published graphics and/or data tables. All 95 profiles are now made available in both absolute (meters) and normalized (0 to 1 ice thickness) depth scales and are accompanied by extensive metadata. These metadata include a transparent description of data provenance. The ice temperature profiles span 70 years, with the earliest profile being from 1950 at Camp VI, West Greenland. To highlight the value of this database in evaluating ice flow simulations, we compare the ice temperature profiles from the Greenland ice sheet with an ice flow simulation by the Parallel Ice Sheet Model (PISM). We find a cold bias in modeled near-surface ice temperatures within the ablation area, a warm bias in modeled basal ice temperatures at inland cold-bedded sites, and an apparent underestimation of deformational heating in high-strain settings. These biases provide process level insight on simulated ice temperatures.

Greenland↗

Empowering Geothermal Research: The Geothermal Data Repository's New AI Research Assistant: Preprint

The Department of Energy's (DOE) Geothermal Data Repository (GDR) team has integrated a Large Language Model (LLM) with the metadata and supporting documents associated with GDR datasets to create an Artificially Intelligent (AI) research assistant. By leveraging work done to make GDR metadata machine-readable and an open-source LLM integration model called the Energy Language Model, developed by the National Renewable Energy Laboratory, AskGDR serves as a virtual research assistant to GDR users. It provides answers to a variety of user-provided questions using natural language processing and generative machine learning. Users can get answers to questions about specific datasets, including inquiries about the equipment, assumptions and methodologies used in the origination of the data; or more abstract questions, such as the applicability of data to specific research fields. AskGDR improves the discoverability of geothermal data by helping guide users to datasets beyond simple keyword searches. It enables users to find data based on properties of the data, discover information contained within supporting documents, and explore data from projects related to their research objectives.

access↗

Empowering Geothermal Research: The Geothermal Data Repository's New AI Research Assistant

The Department of Energy's (DOE) Geothermal Data Repository (GDR) team has integrated a Large Language Model (LLM) with the metadata and supporting documents associated with GDR datasets to create an Artificially Intelligent (AI) research assistant. By leveraging work done to make GDR metadata machine-readable and an open-source LLM integration model called the Energy Language Model, developed by the National Renewable Energy Laboratory, AskGDR serves as a virtual research assistant to GDR users. It provides answers to a variety of user-provided questions using natural language processing and generative machine learning. Users can get answers to questions about specific datasets, including inquiries about the equipment, assumptions and methodologies used in the origination of the data; or more abstract questions, such as the applicability of data to specific research fields. AskGDR improves the discoverability of geothermal data by helping guide users to datasets beyond simple keyword searches. It enables users to find data based on properties of the data, discover information contained within supporting documents, and explore data from projects related to their research objectives. This paper will outline the development, integration, output, and efficacy of the AskGDR LLM, including adherence to scientific rigor through improvements designed to increase the accuracy of generated answers, avoid speculation, and provide proper references for all resources used.

access↗

The Vertebrate Breed Ontology: Toward Effective Breed Data Standardization

Abstract Background Limited universally-adopted data standards in veterinary medicine hinder data interoperability and therefore integration and comparison; this ultimately impedes the application of existing information-based tools to support advancement in diagnostics, treatments, and precision medicine. Hypothesis/Objectives A single, coherent, logic-based standard for documenting breed names in health, production, and research-related records will improve data use capabilities in veterinary and comparative medicine. Animals No live animals were used. Methods The Vertebrate Breed Ontology (VBO) was created from breed names and related information compiled from the Food and Agriculture Organization of the United Nations, breed registries, communities, and experts, using manual and computational approaches. Each breed is represented by a VBO term that includes breed information and provenance as metadata. VBO terms are classified using description logic to allow computational applications and Artificial Intelligence–readiness. Results VBO is an open, community-driven ontology representing over 19 500 livestock and companion animal breed concepts covering 49 species. Breeds are classified based on community and expert conventions (e.g., cattle breed) and supported by relations to the breed's genus and species indicated by National Center for Biotechnology Information (NCBI) Taxonomy terms. Relationships between VBO terms (e.g., relating breeds to their foundation stock) provide additional context to support advanced data analytics. VBO term metadata includes synonyms, breed identifiers/codes, and attributed cross-references to other databases. Conclusion and Clinical Importance The adoption of VBO as a standard for breed names in databases and veterinary electronic health records enhances veterinary data interoperability and computability, supporting precision medicine.

Veterinary Sciences↗

A Data Processing Pipeline To Extract A Knowledge Graph From Sec Documents For Socio-technical Analysis Of Critical Infrastructure Influence

The code is written in Python and consists of the following pipeline that is implemented in Apache Airflow. This pipeline intends to understand the companies that are directly or indirectly involved with a type of critical infrastructure system at some point in that system's lifecycle. The pipeline takes a configuration file that specifies a list of initial companies to consider, a geographic region of interest (disk) expressed as a latitude/longitude point and distance, and a set of SEC form types from which to extract entities and relations. There are three main components to this pipeline as currently implemented: Social Network Extraction, Critical Infrastructure Network Extraction, and Inference and Fusion. First, Social Network Extraction, implemented as the `organizations_sec` component of the workflow graph queries the SEC EDGAR webservice using the list of initial companies from the configuration file. Given this, it extracts metadata that documents the number of each type of form for the given set of companies and their location. This forms metadata represents a catalog of data sources for the extracted social network knowledge graph. The pipeline then downloads these forms from the website and saves them in a build directory for further processing. These documents are then parsed for entities and relations. Second, the Critical Network Extraction component extracts entities and relations for a critical infrastructure sector. Currently, we focus on Electric Vehicle charging stations and this information is available via the Department of Energy (DOE) database on fueling stations maintained by NREL. Third, the Inference and Fusion component relates the social network graph to the critical infrastructure graph in order to understand the impact of a company within a geographic region. Relations include ownership of the EV Charging Station asset as well as maintenance/ownership of the EV payment networks. The fused network can be represented in many ways and currently we emit a knowledge graph.

Weaver, GabrielA.↗

ScholarGuard

The ScholarGuard framework aims to address the gap in archiving and preservation efforts for scholarly artifacts beyond traditional research papers, such as software source code, datasets, presentation slides, workflows, protocols, videos, and more. It introduces a prototype system designed to automatically track researchers' outputs across various scholarly productivity portals on the open web, including platforms like GitHub, Slideshare, Figshare, and Wikipedia. The system detects the availability of new scholarly artifacts and applies modern web archiving technology to create a durable archival record, including high-level metadata for each artifact. This metadata is displayed within the system, linking both to the live version and the archived version of the resource, ensuring long-term accessibility and preservation of diverse research outputs. The software serves as a critical tool for preserving the broader spectrum of scholarly contributions, facilitating visibility, searchability, and long-term access to research artifacts beyond the traditional scope of journal publications.

Balakireva, Lyudmila↗

bibcheck

SAND2026-16981O Bibcheck is designed to extract bibliographies from research papers and perform metadata searches to identify errors. It assists authors in checking their bibliographies for metadata errors during the writing process and helps reviewers identify errors in bibliographies of papers under review. The software uses large language models (LLMs) to extract bibliography entries from PDF documents, classifies the type of bibliography entry, and verifies referenced works. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Pearson, Carl [Sandia National Lab. (SNL-CA), Live↗