Search NASA⌕ Search

SEARCH · Search NASA

Results for “Metadata”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 253 records · Page 14

Enabling Cloud Services and Enhanced Data Discovery With Earthdata-Varinfo

NASA’s Earth Observing System Data and Information System (EOSDIS) contains thousands of Earth science datasets from satellites, models, and field campaigns. Each of these collections can contain hundreds of variables that describe each measurement within the dataset, therefore an automated method for generating UMM-Var records is necessary. The Unified Metadata Model for Variables (UMM-Var) provides a framework for variable metadata records in NASA’s Common Metadata Repository (CMR). The Python tool, earthdata-varinfo, was developed to solve this problem of automating the curation of UMM-Var records. Given either a collection DMR file or a netCDF-4 file, earthdata-varinfo can scrape variable metadata and return a CMR compliant UMM-Var record. Earthdata-varinfo can generate thousands of UMM-Var records in a matter of seconds, thus enabling subsetting capabilities and enhancing data discovery.

Eni Awowale↗

Extending CF Conventions to Enhance Data FAIRness for Atmospheric Composition Observations

The Hierarchical Data Format (HDF) and Network Common Data Form (NetCDF) are data file formats created to aid users in the creation or use of scientific data. These file formats are useful for handling large data volumes and hosting extensive metadata as global, group, or variable attributes and are popular with the modeling community. HDF and NetCDF files are widely used with atmospheric remote sensing data and have been used to support measurements from numerous field campaigns, from satellite to aircraft or ground and mobile based measurements. The files from airborne field studies, however, vary greatly in terms of the file structure and the amount and content of their metadata. Information relevant to the file that can be useful to the user such as the data producer, location where data was taken, variable descriptions, or information about the instrument might not be included in the file. Recently, the Measurements of Aerosols, Clouds, and their Interactions for Earth System Models (MACIE) group started a grassroots effort to develop a CF-based template for the HDF and NetCDF files for field studies, with the aim of making the data products more interoperable and usable. This template seeks to make the files more compliant to Climate and Forecast (CF) metadata conventions and to standardize the file structure and the global and variable attributes. The template would help to ensure that HDF and NetCDF files contain adequate metadata to better support their use for research, e.g., the modeling community, and to enhance the usability and interoperability of data for research communities at large. The draft template has been applied to recent field studies for various instruments and their merge files in support of the Atmosphere Observing System (AOS) project. The details of the revised template are to be presented, as well as examples of the implementation of these requirements for merge files and lidar observation data files and issues revealed during the implementation process.

Sean Leavor↗

Laboratory time series moisture manipulative experiment from sediment across the contiguous US: time series aerobic respiration and geochemistry (v2)

This dataset supports a broader study examining the effects of wetting and drying on hyporheic zone respiration across the contiguous United States (CONUS). The dataset provides data generated from a laboratory moisture manipulation experiment. The contents include time series aerobic respiration and moisture; dissolved oxygen; sediment geochemistry data; and field metadata (including qualitative information on instream and river corridor characteristics). Samples were collected as part of the WHONDRS CONUS-Scale Model-Sample Study (CM). This study was designed following ICON (integrated, coordinated, open, and networked) principles to facilitate a model-experiment (ModEx) iteration approach, leveraging crowdsourced sampling across the CONUS. The data package associated with the CM study is available at https://data.ess-dive.lbl.gov/view/doi:10.15485/1923689. CM sampling began in April 2022 and ended in October 2023. This study uses subsamples from a subset of CM samples collected between June 2022 and June 2023. The original field samples were labeled as CM_###. Subsequent subsamples for this study were labeled as EC_###. The labels from the field samples and the EC subsamples can be mapped directly based on the digits following the prefix and underscore (i.e., EC_001 is a subsample from CM_001). See the critical details section below for more details on sample naming. This data package was originally published in August 2024. It was updated in February 2026 (v2; new and modified files). See the change history section in the readme for more details. For details on how to navigate this data package, see this infographic from the River Corridor SFA https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. This dataset is comprised of one folder of raw Fourier transform ion cyclotron resonance mass spectrometry (FTICR-MS) data and one main data folder containing (1) file-level metadata; (2) data dictionary; (3) field metadata; (4) readme; (5) field protocol; and a (6) a subfolder with sediment sample data from the incubation experiment. The sample data subfolder contains (1) dissolved organic carbon (DOC, measured as non-purgeable organic carbon, NPOC); (2) total nitrogen (TN); (3) adenosine triphosphate (ATP); (4) percent carbon and nitrogen; (5) effect size; (6) iron (II); (7) gravimetric moisture; (8) respiration rates and raw dissolved oxygen values; (9) specific conductance; (10) pH; (11) temperature; (12) a summary containing median values of each data type for each treatment (wet and dry); (13) methods codes; (14) FTICR-MS methods; and (15) a subfolder of 9.4 Tesla FTICR-MS data. This folder contains three subfolders, one containing the sediment .xml data files, one containing the sediment CoreMS output files, the other containing instructions and scripts for processing the files in CoreMS (https://github.com/EMSL-Computing/CoreMS). All files are .csv, .pdf, .R, .ref, or .xml.

54 ENVIRONMENTAL SCIENCES↗

Metagenome-assembled genomes from Wind River Basin floodplain sediments Riverton, Wyoming site (June to October 2019)

Microorganisms play a key role in cycling nutrients and contaminants in the terrestrial environment depending on their genetic potential. Here we present metagenome-assembled genomes (MAGs) for the bacterial and archaeal community in floodplain sediment samples taken at three time points from June 12, 2019 to October 23,2019 at a location (PTT1) close to DOE Legacy Management well 855 at the Riverton, Wyoming floodplain site in the Wind River Basin (WRB). The groundwater at this site exhibits persistent U, Mo, and sulfate plumes and is one of the field sites in focus for the SLAC Groundwater Quality SFA program. Sediment samples were collected from 60 to 180 cm below surface every 30cm for microbial analyses through metagenomic sequencing. 15 metagenomes were sequenced through JGI and can be found under Gold sequencing project: Gs0131241. Metagenomes were assembled, binned, and refined using metawrap to generate MAGs (>50% complete and < 10% contamination based on checkM scores). This dataset includes a zip file of 780 MAG fasta files and a csv file with quality, taxonomic classification (GTDB RS220), and metagenome accessions for MAGs. This dataset also includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type. A sample metadata file (samples.csv) that contains site information has also been included.

54 ENVIRONMENTAL SCIENCES↗

Laboratory time series moisture manipulative experiment from sediment across San Antonio, Texas: time series aerobic respiration and geochemistry

This dataset supports a broader study examining the effects of wetting and drying on hyporheic zone respiration. The dataset provides data generated from a laboratory moisture manipulation experiment. The contents include time series aerobic respiration and moisture; dissolved oxygen; sediment geochemistry data; and field metadata (including qualitative information on instream and river corridor characteristics). Samples were collected as part of the WHONDRS Allison Veach collaboration (AV1). The data package associated with the AV1 study is available at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/2529428. AV1 sampling occurred across 7 perennial and 7 intermittent streams in San Antonio, Texas. Each stream/site was visited both in summer during base flow (July-September 2023) and winter during peak flow (January-February 2024). This study uses subsamples from a subset of AV1 samples. The original field samples were labeled as AV1_###. Subsequent subsamples for this study were labeled as EV_###. The labels from the field samples and the EV subsamples can be mapped directly based on the digits following the prefix and underscore (i.e., EV_001 is a subsample from AV1_001). See the critical details section below for more details on sample naming. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. This dataset is comprised of one main data folder containing (1) file-level metadata; (2) data dictionary; (3) field metadata; (4) readme; (5) field protocol; and a (6) a subfolder with sediment sample data from the incubation experiment. The sample data subfolder contains (1) effect size; (2) iron (II); (3) gravimetric moisture; (4) respiration rates; (5) raw dissolved oxygen values and plots; (6) specific conductance; (7) pH; (8) temperature; (9) a summary containing mean, median, and standard deviation values of each data type for each treatment (wet and dry); and (10) methods codes. All files are .csv or.pdf.

54 ENVIRONMENTAL SCIENCES↗

Groundwater and river water elevations and temperature from 2017 to 2022 across Meander Z in the East River Watershed, Colorado

This dataset includes groundwater and river water elevations and temperature data collected in the East River watershed located in the Upper Colorado River Basin. The data were collected in order to investigate the coupling between hydrology and biogeochemical processes in the floodplain. Data was collected at ten groundwater locations in Meander Z (MZ), located just upstream of the confluence with Brush Creek and two river locations directly adjacent to Meander Z from 2017-2019. From 2019-2022, data was collected at five groundwater locations in Meander Z. Note that location names, not location identifiers (IDs), are used in the related publication Dewey et al. (2022). Both location IDs and names are included in data files. Files in this dataset include the main data files for each location zipped into a single folder (waterlevel_data.zip), an installation methods file describing sensor installation (InstallationMethods.csv), a file containing field metadata including GPS (Global Positioning System) coordinates and ground surface elevations (transducers_locations.csv). This dataset also includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type. This dataset conforms to the ESS-DIVE hydrological reporting format. 2026-04-27 Update: The river water elevation data files (ER-MZR1.csv and ER-MZR2.csv) were corrected. The data for these two locations were inadvertently swapped in the original published data. This work was supported by the Watershed Function Science Focus Area at Lawrence Berkeley National Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-05CH11231.

54 ENVIRONMENTAL SCIENCES↗

Soil physical and chemical measurements for topsoils collected during NEON campaign in East River, CO (06/14/2018-06/28/2018)

The package is part of the DOE Watershed Function Science Focus Area (SFA) project and includes soil physical and chemical measurements from topsoils collected at the East River, Colorado, in conjunction with the National Ecological Observatory Network (NEON) Airborne Observation Platform (AOP) survey conducted in June 2018. The soil measurements include soil bulk density, soil volumetric water content, soil microbial biomass C (Carbon), N (Nitrogen) and C:N (C to N ratio), soil DNA yield, soil total extractable organic C, soil total extractable N, soil extractable nitrate, soil extractable ammonium, soil dissolved inorganic N, soil dissolved organic N, soil pH, soil TOC400 (total organic carbon at 400°C), soil ROC (residual oxidizable carbon), soil TIC (total inorganic carbon), soil TOC (total organic carbon), soil TC (total carbon), soil N, soil OM (organic matter) loss on ignition. Additional associated site metadata can be found in the ESS-DIVE package 10.15485/1618130. The dataset includes (1) 2018_NEON_soil_physical_chemical_measurements.csv: soil physical and chemical measurements indexed by soil sample IGSNs; (2) samples.csv: sample metadata file used to register International Generic Sample Numbers (IGSNs); (3) flmd.csv: file level metadata file; and (4) dd.csv: data dictionary file. This work was supported by the Watershed Function Science Focus Area at Lawrence Berkeley National Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-05CH11231.

2018 NEON and 2025 CHESS (Catchment Hydrology and ↗

Surface water nitrogen and sediment potential nitrate reduction rates, nutrient stocks, and stable isotopes from nine wetlands at the Tanglewood Biological Station, Alabama

This dataset supports a broader study investigating wetland hydrologic and biogeochemical responses to inundation disturbances. Bimonthly surface water and sediment sampling events were conducted at nine wetland sites situated within the Tanglewood Biological Station in Alabama from April 2023 to February 2024. The contents included in the data package include surface water nitrogen (nitrogen oxides and ammonium) and sediment potential nitrate reduction rates (measured as potential denitrification and dissimilatory nitrate reduction to ammonium processing), nutrient stocks (total carbon, total nitrogen, and organic matter), and stable isotopes (carbon and nitrogen). Water level data related to each wetland location can be found at https://data.ess-dive.lbl.gov/view/doi:10.15485/2530253 (Kirker et al., 2024) and related water geochemistry data can be found at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/3001967 (Forbes et al., 2025). In addition to a readme, this data package also includes a file-level metadata (FLMD) file that describes each file and a data dictionary (DD) that describes all column/row headers and variable definitions. This dataset is comprised of (1) file-level metadata; (2) data dictionary; (3) field metadata and international generic sample numbers (IGSNs); (4) readme; (5) the field protocol; and (6) a subfolder with sample data. The sample data subfolder contains (1) sediment potential denitrification rate, (2) sediment potential dissimilatory nitrate reduction to ammonium (DNRA) rate, (3) sediment total carbon and nitrogen content, (4) sediment stable isotopes (delta nitrogen-15 and delta carbon-13), (5) sediment percent organic matter, (6) surface water nitrous oxides, (7) surface water ammonium, and (8) methods codes. All files are .csv or .pdf.

Ammonium↗

Soil nitrogen mineralization rates, nutrient stocks, stable isotopes, and water volumetric measurements across terrestrial-aquatic interfaces from three wetlands at the Tanglewood Biological Station, Alabama

This dataset supports a broader study investigating wetland hydrologic and biogeochemical responses to inundation events. Soil samples were collected across four sampling events along terrestrial-aquatic gradients at three wetland sites located within the Tanglewood Biological Station in Alabama from April 2024 to June 2025. The contents in this data package include soil in-situ nitrogen mineralization rates (measurements of net nitrification, net ammonification, and net mineralization), nutrient stocks (total carbon, total nitrogen, and organic matter), stable isotopes (carbon and nitrogen), and water volumetric measurements (water-filled pore space). Water level data related to each wetland location can be found at https://data.ess-dive.lbl.gov/view/doi:10.15485/2530253 (Kirker et al., 2024), related water geochemistry data can be found at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/3001967 (Forbes et al., 2025), and related surface water sediment chemistry data can be found at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/3377325 (Molina Serpas et al., 2026). In addition to this readme, this data package also includes a file-level metadata (FLMD) file that describes each file and a data dictionary (DD) that describes all column/row headers and variable definitions. This dataset is comprised of (1) file-level metadata; (2) data dictionary; (3) field metadata and international generic sample numbers (IGSNs); (4) readme; (5) the field protocol; and (6) a subfolder with sample data. The sample data subfolder contains (1) net nitrification rate, (2) net ammonification rate, (3) areal net mineralization rate, (4) percent organic matter, (5) water-filled pore space, (6) total carbon content, (7) total nitrogen content, (8) stable carbon isotope (delta carbon-13), and (9) stable nitrogen isotope (delta nitrogen-15), and (10) methods codes. All files are .csv or .pdf.

13-C↗

Integrated Hourly Meteorological Database of 20 Meteorological Stations (1981-2022) for Watershed Function SFA Hydrological Modeling

This dataset contains (a) a script “R_met_integrated_for_modeling.R”, and (b) associated input CSV files: 3 CSV files per location to create a 5-variable integrated meteorological dataset file (air temperature, precipitation, wind speed, relative humidity, and solar radiation) for 19 meteorological stations and 1 location within Trail Creek from the modeling team within the East River Community Observatory as part of the Watershed Function Scientific Focus Area (SFA). As meteorological forcings varied across the watershed, a high-frequency database is needed to ensure consistency in the data analysis and modeling. We evaluated several data sources, including gridded meteorological products and field data from meteorological stations. We determined that our modeling efforts required multiple data sources to meet all their needs. As output, this dataset contains (c) a single CSV data file (*_1981-2022.csv) for each location (20 CSV output files total) containing hourly time series data for 1981 to 2022 and (d) five PNG files of time series and density plots for each variable per location (100 PNG files). Detailed location metadata is contained within the Integrated_Met_Database_Locations.csv file for each point location included within this dataset, obtained from Varadharajan et al., 2023 doi:10.15485/1660962. This dataset also includes (e) a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and (f) a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type. Review the (g) ReadMe_Integrated_Met_Database.pdf file for additional details on the script, methods, and structure of the dataset.The script integrates Northwest Alliance for Computational Science and Engineering’s PRISM gridded data product, National Oceanic and Atmospheric Administration’s NCEP-NCAR Reanalysis 1 gridded data product (through the `RCNEP` R package, Kemp et al., doi:10.32614/CRAN.package.RNCEP), and analytical-based calculations. Further, this script downscales the input data into hourly frequency, which is necessary for the modeling efforts.

54 ENVIRONMENTAL SCIENCES↗

Mineralogy of floodplain sediments from Meanders C, O, and Z in the East River Watershed, CO, USA

This dataset includes bulk X-ray diffraction data from floodplain sediments collected as a part of the Watershed Function Scientific Focus Area (SFA) located in the Upper Colorado River Basin. The data were collected in order to investigate the role of biogeochemical cycling and other river corridor processes on riverine export of solutes. Sediment cores were collected from Meander C, Meander O, and Meander Z in July 2016 to September 2017 to depths of approximately 40-95 cm. Sample metadata including locations, depths, and sample dates are included in a csv file ("sample_list_and_locations.csv"). The file "diffraction_data.csv" contains raw diffraction data, and mineral quantification is in the file "mineral_abundance.csv". This dataset also includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type.

54 ENVIRONMENTAL SCIENCES↗

Metagenome-assembled genomes from soil samples in control and warming plots in Blodgett Forest, CA (2014-2021)

The pathways of carbon transport and loss through and from soils—soil organic matter (SOM) depolymerization to dissolved organic carbon and mineralization to carbon dioxide (CO2)—are fundamentally driven by microbial activity, which is strongly regulated by environmental conditions. As part of LBNL (Lawrence Berkeley National Laboratory) TES (Terrestrial Ecosystem Science) Belowground Biogeochemistry Science Focus Area (SFA), we have established a novel whole-soil long-term warming experiment at the University of California (UC) Blodgett Forest Research Station (Sierra Nevada) in 2014, where we study the role of biogeochemical, microbial and geochemical process interactions in SOM decomposition and stabilization.Here, we present metagenome-assembled genomes (MAGs) for the bacterial and archaeal community from soil depth profiles collected from 2014 to 2021 from three paired control and warming plots. We collected soil samples across a range of depth profiles (spanning surface to 90 cm deep) from three paired control and warming plots from a temperate mixed forest in Northern California. Each paired plot had been subjected to experimental warming since June 2014 to simulate a predicted climate change scenario for northern California. 101 soil metagenomes were sequenced at JGI (Joint Genome Institute) and UCSF (University of California San Francisco) Center for Advanced Technology and can be found under the JGI (Joint Genome Institute) GOLD (Genomes Online Database) Sequencing project Gs0151586 and NCBI (National Center for Biotechnology Information) Projects PRJNA1225762 and PRJEB39497. Metagenomes were assembled using JGI (Joint Genome Institute) Metagenome Workflow (10.1128/mSystems.00804-20). For each metagenome, the assembled contigs were binned into genomes using 3 binning algorithms (cocacola, metabat, and maxbin) and the resulting bins were consolidated using dastool. The consolidated bins from all metagenomes were pooled, filtered by completeness (>50%) and contamination (<25%), and dereplicated at 99% ANI (average nucleotide identity) using dRep (https://github.com/MrOlm/drep).The dataset includes a zip file of 2321 MAG (Metagenome Assembled Genome) fasta files, the accession numbers for the underlying metagenomes, and a csv file with MAG (Metagenome Assembled Genome) quality metrics and taxonomic classification (GTDB -Genome Taxonomy Database-RS220). This dataset also includes a file-level metadata (flmd.csv) file that lists each file contained in the dataset with associated metadata and a data dictionary (dd.csv) file that contains column/row headers used throughout the files along with a definition, units, and data type. A sample metadata file (samples.csv) that contains site information has also been included.

54 ENVIRONMENTAL SCIENCES↗

Soil microbial ecology and microbiome-metabolite linkages improve understanding of ecosystem states along terrestrial-aquatic interfaces

These data are from Bandopadhyay et al., "Soil microbial ecology and microbiome-metabolite linkages improve understanding of ecosystem states along terrestrial-aquatic interfaces". This study aims to understand the soil microbial ecology along terrestrial-aquatic interfaces of a freshwater and estuarine region and how it relates to organic matter. We analyzed soil microbial (16S rRNA gene) and organic matter (Fourier-transform ion cyclotron resonance mass spectrometry, FTICR-MS) composition from upland (forested), transition (stressed forest), and wetland positions at three sites in each of the Lake Erie (freshwater) and Chesapeake Bay (estuarine) regions. This dataset includes 16S rRNA gene amplicon data (only processed file types included here) and organic matter composition from FTICR-MS data (raw and processed files included here) from upland (forested), transition (stressed forest), and wetland positions at three sites in each of the Lake Erie and Chesapeake Bay regions. These sites are part of the COMPASS-FME project (https://compass.pnnl.gov/FME/COMPASSFME). File formats and software needed to access files: 16S rRNA gene amplicon data: These files follow the format reported here https://ess-dive.gitbook.io/amplicon-sequencing-reporting-format#updates-in-v1.0.1. As per this format, there are four file types reported: 1. Taxon tables (also called sequence-by-sample or OTU (operational taxonomic unit)/ESV (exact sequence variant) tables) : available in a .txt file format and accessible using TextEdit or MS Excel. 2. Representative sequences (also called consensus sequences) : available in a .fasta format and accessible using TextEdit. 3. Sequencing metadata : available in a MS Excel workbook file format and CSV file format 4. Bioinformatic metadata : available in a MS Excel workbook file format and CSV file format FTICR-MS data: 1. Raw data converted to a processed file with intensities of the peaks in the given samples : available in a MS Excel CSV file format 2. Processed file used in analyses and visualizations (appended as icr_long_) : available in a MS Excel CSV file format 3. Metadata file for ICR features (appended as icr_meta) : available in a MS Excel CSV file format

54 ENVIRONMENTAL SCIENCES↗

Site and endmember spectra of terrestrial vegetation and soils for the Colorado Headwaters Ecological Spectroscopy Study, June-July 2025

This dataset provides site and endmember spectra collected during the 2025 Colorado Headwaters Ecological Spectroscopy Study (CHESS) campaign. The site spectra were collected to help validate airborne hyperspectral data acquired by the National Ecological Observatory Network's aerial observation platform (NEON AOP). Endmember spectra were collected to augment existing spectral libraries with additional samples of bare surfaces and non-photosynthetic vegetation. All measurements were acquired with an Analytical Spectral Devices (ASD) FieldSpec4 Hi-Res NG (Next Generation) spectroradiometer, which records radiance at 1nm (nanometer) intervals from the ultraviolet to the short-wave infrared (350-2500 nm). The dataset includes spectra measured at meadow sites where the CHESS team also collected vegetation samples for trait analyses. The site spectra were collected with the ASD FieldSpec4 palm grip attachment using an 8° field-of-view foreoptic. Site spectra are integrated measurements of the entire surface within the foreoptic’s field of view. For site-level spectra, the sun is the illumination source. A Spectralon panel mounted on a tripod was used for instrument optimization and white reference measurements for all site spectra. Site spectra were acquired within two hours of solar noon and within 48 hours of a NEON AOP overflight. Site spectra are labeled by date, sampling area, and site number according to the naming conventions of the CHESS campaign’s data management plan. The dataset also contains endmember spectra in the following categories: photosynthetic vegetation (PV), non-photosynthetic vegetation (NPV), bare (soil/rock), and flowers. Endmember measurements were acquired using either the contact probe or the leaf clip attachments of the ASD FieldSpec4. In these configurations, the bulb inside the spectrometer provides the light source for the measurements. The spectrometer was optimized and white reference measurements were recorded using the circular white pucks attached to the contact probe and leaf clip. Because they do not rely on solar illumination, contact probe and leaf clip measurements were collected during a broader time frame than the palm grip site spectra. Some endmembers were measured at CHESS meadow sites, while others were collected within the larger sampling area or in nearby locations (e.g. Gothic Townsite) with similar characteristics. Radiance, reflectance, and metadata files are split into three subfolders according to measurement type: proximal/palm grip (prx), contact probe (cp), and leaf clip (lc). Radiance spectra are provided in ASD file format (.asd file extension). All ASD files can be opened using the provided scripts. Metadata is provided in two formats: CSV file format (no geolocation) and GEOJSON file format (includes geolocation for each spectra). The dataset includes a set of pre-processed reflectance spectra as CSV files (yyyymmdd_rfl.csv). The python scripts and jupyter notebook used to calculate reflectance spectra from the ASD radiance data is included here and was previously published at: https://doi.org/10.3334/ORNLDAAC/2446. There is also a folder of JPEG photographs corresponding to selected spectra. We include a protocol document with detailed steps for ASD FieldSpec4 assembly and operations. This data additionally contains a file level metadata (flmd.csv) and data dictionary (dd.csv) file. Geospatial information: Geospatial data for mapping measurement site locations are in the files CHESS_polygons_lai_UTM.geojson, CHESS_polygons_shrub_UTM.geojson, and CHESS_polygons_meadow_UTM.geojson in the companion geospatial package for the 2025 CHESS campaign, ‘CHESS 2025: Location data for field observations and sampling’ (Henderson et al., 2026). CHESS Project Description: The Colorado Headwaters Ecological Spectroscopy Study (CHESS) comprised a multi-week airborne remote sensing and field observation campaign in the Upper Gunnison Basin, Colorado, conducted in June and July of 2025. Airborne remote sensing was conducted by the National Ecological Observatory Network Airborne Observation Platform (NEON AOP), concurrent with a field campaign run by the Rocky Mountain Biological Laboratory (RMBL), the Lawrence Berkeley National Laboratory (LBNL) and SLAC National Accelerator Laboratory Watershed Function Science Focus Area (SFA), and NASA-JPL (Jet Propulsion Laboratory) Earth Surface Mineral Dust Source Investigation (EMIT) program. Between June 10 and July 18, 2025, the NEON AOP flight team collected high-resolution aerial imaging spectroscopy and Light Detection and Ranging (LiDAR) data over three domains: the Upper East River (CRBU), Almont Triangle (ALMO), and the Upper Taylor Basin (UPTA). In coordination with the flights, a field campaign acquired ground-truth observations, including observations of vegetation composition, foliar traits, forest demography, and subsurface properties in 18 core sampling areas within the domains. Additional surface water observations were taken at over 380 point locations. All CHESS campaign datasets can be found within the CHESS ESS-DIVE data portal: https://data.ess-dive.lbl.gov/portals/chess. Funding Acknowledgment: This research was carried out at the Jet Propulsion Laboratory, California Institute of Technology, under a contract with the National Aeronautics and Space Administration (80NM0018D0004) and was funded by EMIT Extended Mission Phase E Science.

2018 NEON and 2025 CHESS Campaigns↗

SPRUCE Aboveground Vegetation Coverage in Root Ingrowth Core Plots, Marcell Experimental Forest, Minnesota, August 2022

This dataset contains vegetation survey measurements from root ingrowth core plots (Määttä et al. 2025) inside SPRUCE Experiment plots at the Marcell Experimental Forest in northern Minnesota. Vegetation surveys were conducted in 0.25 meter2 plots containing root ingrowth cores on August 8th and 9th, 2022 (2022-08-08 to 2022-08-09). The warming and elevated carbon dioxide (CO2) treatments in the dataset include the full treatment gradient: +0 degrees Celsius (C) (+0 and +500 parts per million (ppm) elevated CO2), +2.25 degrees C (+0 and +500 ppm), +4.5 degrees C (+0 and +500 ppm elevated CO2), +6.75 degrees C (+0 and +500 ppm) and +9 degrees C (+0 and +500 ppm elevated CO2) for both hummocks and hollows. This dataset includes measurements of the height and absolute coverage (%) for each vascular plant and moss species, as well as organic litter and dead overstory vascular plants, and the distance from the grid center to the nearest tree and the species of the nearest tree. These data were used as species-specific aboveground plant metadata for assessing the warming and elevated CO2 response of fine roots across different peatland microtopographical features (hummocks and hollows) and plant functional types (shrub, spruce and larch). This dataset contains one data file in comma separate (.csv) format. Additional metadata are provided: one data dictionary and a file-level metadata file in comma separate (.csv) format and a user guide in PDF (*.pdf) format.

ESS-DIVE CSV File Formatting Guidelines Reporting ↗

Data and scripts associated with “Moisture content modulates DOM thermodynamic regulation of oxygen consumption in drying streambed sediments”

This data package is associated with the publication “Moisture content modulates DOM thermodynamic regulation of oxygen consumption in drying streambed sediments” published in Scientific Reports (Garayburu-Caruso et al., 2026). The package contains processed data products and scripts used to quantify how drying and re-inundation of riverbed sediments influence dissolved organic matter (DOM) thermodynamic properties and their relationship with sediment oxygen (O₂) consumption across 33 stream sites in the contiguous United States. The data package contains DOM thermodynamic metrics (e.g., Gibbs free energy of carbon oxidation and thermodynamic efficiency), and O₂ consumption along with watershed-scale climate and land-cover metrics used as explanatory variables in the analyses. Underlying unprocessed and processed ultrahigh-resolution mass spectrometry data, oxygen consumption rates from laboratory moisture-manipulation experiments, within-sample environmental properties, sediment moisture content and contextual field measurements are archived separately at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/2428003 (Laan et al., 2024) and https://data.ess-dive.lbl.gov/datasets/doi:10.15485/1923689 (Forbes et al.,2023). A preliminary version of this data package was published in February 2026 at the time of manuscript submission. It was updated in June 2026, at the time of manuscript acceptance, to include the finalized data and additional metadata (readme, data dictionary, and file level metadata). For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. In addition to a readme, this data package also includes a file-level metadata (FLMD) file that describes each file and a data dictionary (DD) that describes all column/row headers and variable definitions. At the top level, the data package is organized into five main folders: (1) Data, (2)Figures, (3) Map, (4) GAM_Reulsts, and (5) src. The Data folder contains analysis-ready tabular files with oxygen consumption rates, DOM thermodynamic properties by site and treatment, site-level environmental variables, watershed-scale metrics, and other derived variables referenced in the manuscript. The Figures folder contains static image files associated with the main text and supplemental figures, while the Map folder includes spatial data and map-layer files used to create the sampling-location map. The GAM results folder contains the results for each of the general additive model (GAM).The src folder contains R scripts used to perform data processing, statistical analyses (including clustering, generalized additive models, and threshold analysis), and figure generation. This data package is associated with a GitHub repository found at https://github.com/WHONDRS-Hub/ECA_DOM_Thermodynamics.

Dissolved organic matter↗

Timeseries Photos of a Variably Inundated Stream: Umtanum Creek, Washington, United States

This dataset is associated with a broader study using game camera timeseries photos collected to evaluate stream variable inundation via changes in width (i.e. wet fraction). Four game cameras were deployed along Umtanum Creek (Washington, United States) to track changes in stream inundation over time. Drone imagery was collected at the same location on October 18, 2024 which was used to construct a digital elevation model (DEM) of the streambed topography. The associated paper and data can be found at https://doi.org/10.1016/j.envsoft.2025.106715 (Bao et al., 2025a)) and https://doi.org/10.15485/2589885 (Bao et al., 2025b), respectively. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. In addition to this readme, this data package also includes a file-level metadata (FLMD) files that describes each file and a data dictionaries (DD) that describe all column/row headers and variable definitions. This dataset is comprised of (1) file-level metadata; (2) data dictionary; (3) readme; (4) field metadata; (5) field protocol; and (5) folders containing game camera photos. Game camera photos are organized into folders for each camera (CDL, CUL, CDR, CUR; see readme for information on camera naming) by the month photos were collected. All files are .csv, .jpg, or .pdf.

AI image segmentation↗

Dated soil C–N–P profiles, water quality, and chamber fluxes across Ohio and Michigan wetlands (2024–2025)

This dataset includes dated soil core chemistry (bulk density, phosphorus, nitrogen and carbon concentrations), water quality, and chamber flux measurements collected from wetlands in the Midwest United States—12 sites in Ohio, one site in Indiana, one site in Michigan—collected in the spring or summer of 2024 or 2025, all in (.csv) format. These data were generated to examine how wetland restoration, management activities, and time since restoration affect biogeochemical processes, carbon sequestration, nutrient accumulation, water quality, and greenhouse gas emissions. Specifically, these data aim to investigate how restored wetlands differ from natural wetlands in terms of carbon, nitrogen, phosphorus dynamics, as well as carbon dioxide and methane fluxes. Also included are surface and porewater quality parameters and chamber flux measurements across these different wetlands. Sampling was conducted at various sites representing a range of restoration stages, from about 4 years post-restoration up to 105 years post-restoration, and also includes a natural wetland used as a reference in Michigan. These data can be used to determine carbon sequestration rates, nutrient cycling, and to enhance our understanding of biogeochemical responses to wetland restoration in temperate ecosystems. This data package contains (1) a csv file (Water_Quality.csv) containing water quality data (dissolved organic carbon, total dissolved nitrogen, and temperature) organized by location; (2) a csv file (Soil_C_N_P_Seq.csv) containing carbon, nitrogen, and phosphorus concentrations at each soil level and time of each soil level, as well as their sequestration rates; (3) a csv file (CH4_CO2_Flux.csv) including methane and carbon dioxide fluxes that were measured with a chamber; (4) a file-level metadata (FLMD.csv) file that lists each file contained in the dataset with associated metadata; (5) a data dictionary (DD.csv) file that contains terms/column headers used throughout the files along with a definition, units, and data type; and (6) a locations metadata file (Location_metadata.csv).

Earth Science > Atmosphere > Atmospheric Chemistry↗