Search NASA⌕ Search

SEARCH · Search NASA

Results for “Climate data records”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Downwelling Shortwave and Longwave Irradiance from CC-RIDER on Mount Soledad

The Clouds and Climate - Remote Integrated Deployment of Radiometers (CC-RIDER) is a suite of five Eppley Laboratory (Inc.) instruments that was deployed during EPCAPE at the secondary Mount Soledad site in La Jolla, CA. A primary and backup Precision Spectral Pyranometer (PSP Primary, PSP Backup) measured broadband downwelling shortwave irradiance in the spectral interval 280-2800 nm. A third Precision Spectral Pyranometer (PSP NIR) was fitted with a near-infrared long pass filter and measured downwelling shortwave irradiance in the spectral interval 780-2800 nm. A Total Ultraviolet Radiometer (TUVR) measured broadband downwelling broadband ultraviolet irradiance in the spectral interval 295-385 nm. A Precision Infrared Radiometer (PIR) Pyrgeometer measured downwelling broadband longwave irradiance in the spectral interval 3.5 - 50 microns. Data collection began on 18 April 2023 at 21:31 UTC and ended on 20 February 2024 at 22:34 UTC. Data were recorded by a Campbell Scientific (Inc.) CR1000X datalogger in one-minute intervals, for a total of 443584 data records. The datalogger was solar powered enabling data collection to proceed without interruption from start to finish.

Clouds and Climate – Remote Integrated DEployement↗

Evaluation of daily gridded climate products using in situ FLUXNET data and tree growth modeling

Gridded climate data products have facilitated research in climate and ecology by providing meteorological data continuously across large spatial scales. However, the sensitivity of scientific outcomes to dataset choice remains poorly understood, and evaluation using station-based records can favor datasets built heavily on weather stations. Here, we evaluate seven high-resolution daily gridded datasets covering the contiguous United States using independent meteorology from the FLUXNET2015 dataset, with a focus on the implications of dataset choice for process-based tree growth modeling. We find that gridded products tend to capture temperature accurately while consistently overestimating the magnitude and frequency of precipitation and its extremes. Moreover, datasets vary in how they define a ‘day,’ which significantly affects temporal alignment with FLUXNET2015 observations. Despite differences among the datasets, the interannual variability in tree ring simulations is insensitive to dataset choice, likely because daily-scale biases are averaged out through accumulated growth across several months. However, inaccuracies in temperature and precipitation can significantly bias modeled xylem cell production, with systematically higher annual precipitation in the gridded datasets leading to greater xylem production compared to simulations using in situ data. Our results suggest that model applications, especially those that integrate to time scales longer than one day, are likely insensitive to climate dataset choice, but applications that are sensitive to daily climate variations or to absolute climate values need to carefully consider biases in gridded climate products.

54 ENVIRONMENTAL SCIENCES↗

Attribution of the record-high 2023 SST using a deep-learning framework

Abstract The global-mean sea surface temperature (SST) reached a record high in 2023, exceeding the 2016 record by 0.14 °C. This unprecedented change in global-mean SST has major implications for our understanding of internal variability and the forced response in our changing climate. In this work, we use neural networks trained on simulated climate data to separate the contributions of internal variability and the forced response within observations. Performing attribution reveals that internal variability was responsible for +0.07 °C of the 2023 global mean SST, due to anomalously warm conditions in the Pacific, Atlantic, and Indian Ocean basins. Furthermore, these results provide a line of evidence for accelerated forced warming in recent years. Continued monitoring of the climate will be critical for understanding the drivers behind this unprecedented SST record.

Rader, Jamin K. (ORCID:0000000222045977)↗

Climate Conundrum: A Wet or Dry European and Northern African Climate During the Middle Miocene

Abstract End of 21st‐century hydroclimate projections suggest an expansion of subtropical dry zones, with Mediterranean and Sahel regions becoming much drier. However, paleobotanical assemblage evidence from the middle Miocene (17‐12 Ma), suggests both regions were instead humid environments. Here we show that by modifying regional sea surface temperatures (SST) in an Earth System Model (CESM1.2) simulation of the middle Miocene, the increased ocean evaporation and integrated water vapor flux overrides any drying effects associated with warming‐induced land‐surface evaporation driven by atmospheric CO 2 concentrations. These modifications markedly reduce the bias in the model‐data comparison for this period. A vegetation model (BIOME4) forced with simulated climatologies predicts both regions were dominated by mixed forest, which is largely consistent with the paleobotanical record. This study unveils the potential for wetter subtropical Mediterranean climates associated with warming, presenting an alternative scenario from future drying projections with localized SST warming governing regional climate change.

Acosta, R. P.↗

Stable Water Isotope Data for the East River Watershed, Colorado (2014-2025)

The stable water isotope data for the East River Watershed, Colorado, consists of delta2H (hydrogen) and delta18O (oxygen) values from samples collected at multiple, long-term monitoring sites including streams, groundwater wells, springs, and a precipitation collector used to establish a local meteoric water line (LMWL) for the watershed. These locations represent important and/or unique end-member locations for which stable isotope values can be diagnostic of the connection between precipitation inputs as snow and rain and riverine export. Such locations include drainages underline entirely or largely by shale bedrock, land covered dominated by conifers, aspens, or meadows, and drainages impacted by historic mining activity and the presence of naturally mineralized rock. Developing a long-term record of water isotope values from a diversity of environments is a critical component of quantifying the impacts of both climate change and discrete climate perturbations, such as drought, forest mortality, and wildfire, on water export. Such data may be combined with stream gaging stations co-located at each surface water monitoring site to relate seasonal variations in water export to their stable isotopic signature. Data for liquid water delta2H and delta18O values are reported in units of parts per thousand (per-mil; ‰). This data package contains (1) a zip file (isotope_data_2014-2025.zip) containing a total of 95 files: 96 data files of isotope data from across the Lawrence Berkeley National Laboratory (LBNL) Watershed Function Scientific Focus Area (SFA) which is reported in .csv files per location and a locations.csv (1 file) with latitude and longitude for each location; (2) a file-level metadata (v6_20260901_flmd.csv) file that lists each file contained in the dataset with associated metadata; and (3) a data dictionary (v6_20260901_dd.csv) file that contains terms/column_headers used throughout the files along with a definition, units, and data type. Missing values within the anion data files are noted as either "-9999" or "0.0" for not detectable (N.D.) data. There are a total of 43 locations containing isotope data. Update on 2022-06-10: versioned updates to this dataset was made along with these changes: (1) updated isotope data for all locations up to 2021-12-31 and (2) the addition of the file-level metadata (flmd.csv) and data dictionary (dd.csv) were added to comply with the File-Level Metadata Reporting Format. Update on 2022-09-09: Updates were made to reporting format specific files (file-level metadata and data dictionary) to correct swapped file names, add additional details on metadata descriptions on both files, add a header_row column to enable parsing, and add version number and date to file names (v2_20220909_flmd.csv and v2_20220909_dd.csv). Update on 2023-08-08: Updates were made to both the data files and reporting format specific files. New available anion data was added, up until 2023-03-13. The file level metadata and data dictionary files were updated to reflect the additional data added. Update on 2024-03-11: Updates were made to both the data files and reporting format specific files. New available anion data was added, up until 2024-02-19. Further, revisions to the data files were made to remove incorrect data points (from 1970 and 2001). The reporting format specific files were updated to reflect the additional data added. Update on 2025-05-15: Updates were made to both the data files and reporting format specific files. New available isotope data was added, up until the end of WY2024 (September 30, 2024). International Generic Sample Numbers (IGSNs), when registered, were added to the data files. The reporting format specific files were updated to reflect the additional data added. Update on 2026-09-01: Updates were made to both the data files and reporting format specific files. New available isotope data was added, up until the end of WY2025 (September 30, 2025).

54 ENVIRONMENTAL SCIENCES↗

Anion Data for the East River Watershed, Colorado (2014-2025)

The anion data for the East River Watershed, Colorado, consist of fluoride, chloride, sulfate, nitrate, and phosphate concentrations collected at multiple, long-term monitoring sites that include stream, groundwater, and spring sampling locations. These locations represent important and/or unique end-member locations for which solute concentrations can be diagnostic of the connection between terrestrial and aquatic systems. Such locations include drainages underlined entirely or largely by shale bedrock, land covered dominated by conifers, aspens, or meadows, and drainages impacted by historic mining activity and the presence of naturally mineralized rock. Developing a long-term record of solute concentrations from a diversity of environments is a critical component of quantifying the impacts of both climate change and discrete climate perturbations, such as drought, forest mortality, and wildfire, on the riverine export of multiple anionic species. Such data may be combined with stream gauging stations co-located at each monitoring site to directly quantify the seasonal and annual mass flux of these anionic species out of the watershed. This data package contains (1) a zip file (anion_data_2014_2025.zip) containing a total of 386 files: 387 data files of anion data from across the Lawrence Berkeley National Laboratory (LBNL) Watershed Function Scientific Focus Area (SFA) which is reported in .csv files per location and a locations.csv (1 file) with latitude and longitude for each location; (2) a file-level metadata (v7_20260901_flmd.csv) file that lists each file contained in the dataset with associated metadata; (3) a data dictionary (v7_20260901_dd.csv) file that contains terms/column_headers used throughout the files along with a definition, units, and data type; and (4) a anion MDL fact sheet (anion_MDLs_202608 in PDF and docx formats). Missing values within the anion data files are noted as either "-9999" or "0.0" for not detectable (N.D.) data. There are a total of 47 locations containing anion data. Update on 2022-06-10: versioned updates to this dataset was made along with these changes: (1) updated anion data for all locations up to 2021-12-31, (2) removal of units from column headers in datafiles, (3) added row underneath headers to contain units of variables, (4) restructure of units to comply with CSV reporting format requirements, and (5) the addition of the file-level metadata (flmd.csv) and data dictionary (dd.csv) were added to comply with the File-Level Metadata Reporting Format. Update on 2022-09-09: Updates were made to reporting format specific files (file-level metadata and data dictionary) to correct swapped file names, add additional details on metadata descriptions on both files, add a header_row column to enable parsing, and add version number and date to file names (v2_20220909_flmd.csv and v2_20220909_dd.csv). Update on 2022-12-20: Updates were made to both the data files and reporting format specific files. Conversion issues affecting ER-PLM locations for anion data was resolved for the data files. Additionally, the flmd and dd files were updated to reflect the updated versions of these files. Available data was added up until 2022-03-14. Update on 2023-08-08: Updates were made to both the data files and reporting format specific files. New available anion data was added, up until 2023-05-19. The file level metadata and data dictionary files were updated to reflect the additional data added. Update on 2024-03-11: Updates were made to both the data files and reporting format specific files. New available anion data was added, up until 2023-09-11. Further, revisions to the data files were made to remove incorrect data points (from 1970 and 2001). The reporting format specific files were updated to reflect the additional data added. Update on 2025-05-15: Updates were made to both the data files and reporting format specific files. New available anion data was added, up until the end of WY2024 (September 30, 2024). International Generic Sample Numbers (IGSNs), when registered, were added to the data files. The reporting format specific files were updated to reflect the additional data added. Update on 2026-09-01: Updates were made to both the data files and reporting format specific files. New available anion data was added, up until the end of WY2025 (September 30, 2025). An anion MDL document was included in this update.

54 ENVIRONMENTAL SCIENCES↗

New Particle Formation Event Dataset at the Southern Great Plains (SGP) Observatory from 2018 to 2023

This data set contains observations of new particle formation (NPF) events collected at the U.S. Department of Energy’s Atmospheric Radiation Measurement (ARM) Southern Great Plains (SGP) observatory from 2018 to 2023. Measurements include particle number size distributions, radiance measurements, and associated meteorological variables from onsite instrumentation relevant for identifying and characterizing NPF events. Events were identified using standardized criteria and documented to support investigations of aerosol nucleation, growth dynamics, and their interactions with local atmospheric conditions. The data set provides a multi-year record that enables evaluation of seasonal and interannual variability in NPF occurrence and intensity at a mid-continental site. These data are intended to support studies of aerosol-cloud-climate interactions and model evaluation within both ARM and the broader atmospheric science community. More information can be found within the README_NPF_SGP_2018_2023.docx file.

event↗

Utility-Scale Solar, 2024 Edition: Empirical Trends in Deployment, Technology, Cost, Performance, PPA Pricing, and Value in the United States [Slides]

Berkeley Lab’s “Utility-Scale Solar, 2024 Edition” presents analysis of empirical plant-level data from the U.S. fleet of ground-mounted photovoltaic (PV), PV+battery, and concentrating solar-thermal power (CSP) plants with capacities exceeding 5 MWAC (PV plants of 5 MWAC or less, including residential rooftop systems, are covered separately in Berkeley Lab’s companion annual report, Tracking the Sun). Key findings from this year’s report include: -18.5 GWAC of new utility-scale PV capacity came online in 2023, bringing cumulative installed capacity to more than 80.2 GWAC across 47 states. Installed costs continued to fall in 2023. Relative to 2022, capacity-weighted averages decreased by 8% to -$\$1.43$/WAC (or $\$1.08$/WDC). Costs, based on a 7.1 GWAC sample of 76 plants completed in 2023, have fallen by 75% (averaging 10% annually) since 2010. Plant-level capacity factors vary widely, from 6% to 36% (on an AC basis), with a sample median of 24%. -Levelized cost of energy (LCOE) of new 2023 projects increased slightly to $\$46$/MWh prior to the application of tax credits but continued to fall to $\$31$/MWh when accounting for federal incentives. PPA prices have largely followed the decline in solar’s LCOE over time, but newly signed longer-term PPA prices have increased since 2021, to an average of $\$35$/MWh (levelized, in 2023 dollars). -Solar’s average energy and capacity value (i.e., ability to offset costs of other power generation sources) across the U.S. was $\$45$/MWh in 2023. Solar’s average market value was lowest in CAISO ($\$27$/MWh), the market with the greatest solar generation share, and highest in ERCOT ($\$67$/MWh). -Newer solar projects had greater market value in 2023 than their generation costs, yielding $\$1.1$ billion in benefits. Projects built in 2022 delivered on average $\$15$/MWh more market value than their costs in 2023. -Solar’s combined value from wholesale electricity markets, public health and climate damage reduction were greater than generation costs and incentives, yielding $\$13.7$ billion in net benefits in 2023. We estimate U.S. health benefits of $\$24$/MWh and reduced global climate damages of $\$101$/MWh. -Adding battery storage is one way to increase the value of solar. Deployment of 52 new PV+battery hybrid plants set a record with 5.3 GW installed in 2023. Our public data file tracks metadata and PPA prices from more than 100 PV+battery hybrid projects that are already online or that have secured offtake arrangements. -Looking ahead, a massive pipeline of at least 1,085 GW of solar capacity dominates the nation’s interconnection queues at the end of 2023. Nearly 571 GW, or 53%, of that total was paired with a battery – in CAISO it was a staggering 98%. Historically only 10% of the requested solar capacity is built. -For more information, and to explore related interactive data visualizations, go to utilityscalesolar.lbl.gov.

14 SOLAR ENERGY↗

Climate change, malaria and neglected tropical diseases: a scoping review

Abstract To explore the effects of climate change on malaria and 20 neglected tropical diseases (NTDs), and potential effect amelioration through mitigation and adaptation, we searched for papers published from January 2010 to October 2023. We descriptively synthesised extracted data. We analysed numbers of papers meeting our inclusion criteria by country and national disease burden, healthcare access and quality index (HAQI), as well as by climate vulnerability score. From 42 693 retrieved records, 1543 full-text papers were assessed. Of 511 papers meeting the inclusion criteria, 185 studied malaria, 181 dengue and chikungunya and 53 leishmaniasis; other NTDs were relatively understudied. Mitigation was considered in 174 papers (34%) and adaption strategies in 24 (5%). Amplitude and direction of effects of climate change on malaria and NTDs are likely to vary by disease and location, be non-linear and evolve over time. Available analyses do not allow confident prediction of the overall global impact of climate change on these diseases. For dengue and chikungunya and the group of non-vector-borne NTDs, the literature privileged consideration of current low-burden countries with a high HAQI. No leishmaniasis papers considered outcomes in East Africa. Comprehensive, collaborative and standardised modelling efforts are needed to better understand how climate change will directly and indirectly affect malaria and NTDs.

60 APPLIED LIFE SCIENCES↗

Deep-Learning-derived Boundary Layer Height from Meteorological Data over the SGP, GOAMAZON, CACTI

The planetary boundary-layer (PBL) height (PBLH) is an important parameter for various meteorological and climate studies. This study presents a multi-structure deep neural network (DNN) model, designed to estimate PBLH by integrating morning temperature profiles with surface meteorological observations. The DNN model is developed by leveraging a rich data set of PBLH derived from long-standing radiosonde records and augmented with high-resolution micropulse lidar and Doppler lidar observations. We access the performance of the DNN with an ensemble of 10 members, each featuring distinct hidden layer structures, which collectively yield a robust 27-year PBLH data set over the Southern Great Plains from 1994 to 2020. The influence of various meteorological factors on PBLH is rigorously analyzed through the importance test. Moreover, the DNN model's accuracy is evaluated against radiosonde observations and juxtaposed with conventional remote-sensing methodologies, including Doppler lidar, ceilometer, Raman lidar, and micropulse lidar. The DNN model exhibits reliable performance across diverse conditions and demonstrates lower biases relative to remote-sensing methods. In addition, the DNN model, originally trained over a plain region, demonstrates remarkable adaptability when applied to the heterogeneous terrains and climates encountered during the GoAmazon (tropical rainforest) and CACTI (middle-latitude mountain) campaigns. These findings demonstrate the effectiveness of deep learning models in estimating PBLH, enhancing our understanding of boundary-layer dynamics with implications for enhancing the representation of PBL in weather forecasting and climate modeling.

54 ENVIRONMENTAL SCIENCES↗

Hot droughts in the Amazon provide a window to a future hypertropical climate

Tropical forests represent the warmest and wettest of Earth’s biomes, but with continued anthropogenic warming, they will be pushed to climate states with no current analogue. Droughts in the tropics are already becoming more intense as they occur at successively higher temperatures. Here, in this study, we synthesize multiple datasets to assess the effects of hot droughts on a central Amazon forest. First, a more than 30-year record of annually resolved forest demographic data from a selective logging experiment showed higher tree mortality during intense droughts, particularly among fast-growing pioneer species with low wood density. Second, analysis of ecophysiological field measurements from the 2015 and 2023 El Niño droughts identified a soil moisture threshold beyond which transpiration rates rapidly declined. As rainless days beyond this threshold continued, drought conditions intensified, increasing the potential for tree mortality from hydraulic failure and carbon starvation. Third, analyses from the Coupled Model Intercomparison Project Phase 6 demonstrated that under high-emission scenarios, a large area of tropical forest will shift to a hotter ‘hypertropical’ climate by 2100. Last, under a hypertropical climate, temperature and moisture conditions during typical dry season months will more frequently exceed identified drought mortality thresholds, elevating the risk of forest dieback. Present-day hot droughts are harbingers of this emerging climate, offering a window for studying tropical forests under expected extreme future conditions.

drought↗

Assessing Historical Extreme Weather Event Impacts

Understanding how past extreme weather events have affected a site is an integral part of site-level resilience planning. Energy and water resilience planning has been a key priority for the federal government for many years and agencies look to develop processes for identifying and addressing critical resilience gaps at their facilities and across their sites. The purpose of this information paper is to help inform how organizations could begin structuring a comprehensive process for recording the impacts of extreme weather events in order to facilitate climate vulnerability assessments, and thus, resilience planning. The paper highlights current limitations for developing event history assessments and suggests a framework for more consistently capturing key data points.

54 ENVIRONMENTAL SCIENCES↗

Estimating the Impacts of Increasing Temperatures and the Efficacy of Climate Adaptation Strategies in Urban Microclimates with Deep Learning

As urbanization and climate change progress, understanding and addressing urban heat becomes a priority for climate adaptation efforts. High temperatures concentrated in the urban core can drive increased risk of heat-related death and illness as well as increased energy demand for cooling. However, modeling the urban microclimate is an ongoing field of research typically burdened by an imprecise description of the built environment, incomplete observational records, significant computational cost, and a lack of high-resolution estimates of the impacts of increasing temperatures. Here, we present computationally efficient machine learning methods that can improve the accuracy of urban temperature estimates when compared to historical reanalysis data. These models are applied to a neighborhood in Los Angeles, and we compare the energy benefits of heat mitigation strategies to the impacts of climate change. We find that cooling demand is likely to increase substantially through midcentury, but engineered high-albedo surfaces could lessen this increase by more than 50 %. The corresponding increase in winter gas heating offsets the summer cooling benefit in the current climate, but total annual energy use from combined heating and cooling with electric heat pumps benefits from the engineered heat mitigation strategies under both current and future climates.

54 ENVIRONMENTAL SCIENCES↗

Pavement condition and climatic data in southeast Texas: A dataset for evaluating flood impacts on pavement performance

Effective pavement maintenance is essential for economic stability, optimal network performance, and roadway safety. Achieving this requires thorough evaluation of pavement conditions, including structural integrity, surface roughness, and distress characteristics. Pavement performance indicators play a critical role in influencing vehicle safety and ride quality. Recent advances have emphasized the use of data-driven modeling to anticipate pavement behavior, with the goal of optimizing resource allocation and refining Maintenance and Rehabilitation (M&R) strategies through accurate condition assessment. A foundational requirement for these modeling efforts is the availability of standardized, high-quality datasets that can support robust and reproducible infrastructure analysis. This data article presents a comprehensive dataset assembled to facilitate pavement performance prediction, with a geographic focus on Southeast Texas, particularly the flood-vulnerable area of Beaumont. The dataset encompasses pavement and traffic attributes, meteorological records, flood simulation outputs, ground deformation measurements, and topographic indices, enabling detailed examination of both load-associated and non-load-associated degradation mechanisms. Data preprocessing was performed using ArcGIS Pro, Microsoft Excel, and Python to ensure consistency and usability in data-driven modeling applications, including machine learning workflows. Key contributions of this dataset include its utility in analyzing the climatic and environmental factors affecting pavement conditions, identifying critical predictive features, and enabling in-depth correlation analysis across diverse variables. By filling existing gaps in input variable selection resources, this dataset supports the development of predictive tools for estimating future maintenance demand and enhancing the resilience of pavement networks in flood-impacted areas. The resource highlights the importance of standardized datasets for advancing pavement management practices and provides a robust foundation for ongoing infrastructure performance modeling.

42 ENGINEERING↗

East Antarctic Ice Sheet variability in the central Transantarctic Mountains since the mid Miocene

The response of the East Antarctic Ice Sheet to warmer-than-present climate conditions has direct implications for projections of future sea level, ocean circulation, and global radiative forcing. Nonetheless, it remains uncertain whether the ice sheet is likely to undergo net loss due to amplified melting coupled with dynamic instabilities or whether such losses will be balanced, or even offset, by enhanced accumulation under a higher-precipitation regime. The glacial depositional record from the central Transantarctic Mountains (TAM) provides a robust geologic means to reconstruct the past behaviour of the East Antarctic Ice Sheet, including during periods thought to have been warmer than today, such as the mid-Pliocene Warm Period (~3.3–3.0 Ma). This study describes a new surface-exposure-dated moraine record from Otway Massif in the central TAM spanning the last ~9 Myr and synthesises these data in the context of previously published moraine chronologies constrained with cosmogenic nuclides. The resulting record, although fragmentary, represents the majority of direct and unambiguous terrestrial evidence for the existence and size of the East Antarctic Ice Sheet during the last 14 Myr, and it thus provides new insight into the long-term relationship between the ice sheet and global climate. At face value, the existing TAM moraine record does not exhibit a clear signature of the mid-Pliocene Warm Period, thus precluding a definitive verdict on the East Antarctic Ice Sheet's response to this event. In contrast, an apparent hiatus in moraine deposition both at Otway Massif and the neighbouring Roberts Massif suggests that the ice sheet surface in the central TAM was potentially lower than present during the late Miocene and earliest Pliocene.

58 GEOSCIENCES↗

SetGo: Metadata Readiness for Scientific AI Datasets

Scientific datasets intended for AI use require both computational readiness for model training and metadata readiness for discovery, sharing, and reuse. The Readiness Engine for Data Integration (REDI) addresses computational readiness, but no corresponding tool evaluates whether a dataset’s metadata are sufficiently complete, governed, and standards-compliant for publication and agent-based consumption. Existing FAIR assessors operate only on published repository records, and no single system covers FAIR compliance, licensing, provenance, governance, reproducibility, and catalog readiness together. We present SetGo, an open-source Python toolkit that assesses and repairs metadata readiness across these six dimensions before a dataset is published or archived. Applied to four scientific corpora, SetGo surfaces deficiencies that general-purpose tools do not detect: ERA5 climate metadata scores 4% on ACDD 1.3 compliance; materials datasets fail OPTIMADE species-definition requirements; and PDB-derived proteomics data carries licensing terms incompatible with standard SPDX identifiers. Guided enrichment raises overall FAIR scores from 52–57% to 81–91%, and a single setgo publish command pushes to Hugging Face Hub, CKAN, or OpenMetadata with ML Commons Croissant 1.0 metadata sidecars. To support interactive and automated workflows, SetGo integrates with coding agents powered by large language models (LLMs) through a /setgo skill that enables natural-language execution of the full assess–enrich–publish loop, with user involvement limited to supplying missing metadata values.

Wilkinson, Sean [ORNL] (ORCID:0000000214437479)↗

NGEE Arctic Integrated Modeling (IM3): Improved snow-vegetation interaction

This data product represents the integration of new code capability for arctic tundra snow-vegetation-terrain interactions into the Energy Exascale Earth System Model (E3SM), through the E3SM Land Model (ELM) component. This code integration is the result of collaborative effort between the NGEE Arctic project and the E3SM project. The NGEE Arctic project developed a total of six Integrated Modeling (IM) modules informed by observations and experiments. New ELM capability represented by this data product (IM3) falls into three categories: 1) Downscaling from gridcell to topographic unit level when working through the existing coupler bypass code. 2) Four new parameters (taper, stocking, bendresist, and vegshape) have been added to ELM to allow for flexible definition of snow-vegetation interactions. 3) Vegshape and bendresist parameters are used to calculate the fraction of leaf area and/or stem area buried by snow for a given snow depth. This data record consists of a single document (pdf format) that describes the theoretical basis for the snow-vegetation-terrain interactions added to ELM, and describes the modifications made to the ELM code. The Methods section of this metadata record includes a link to the public E3SM code repository where the exact code modifications as integrated in E3SM can be accessed. The Next-Generation Ecosystem Experiments: Arctic (NGEE Arctic), was a research effort to reduce uncertainty in Earth System Models by developing a predictive understanding of carbon-rich Arctic ecosystems and feedbacks to climate. NGEE Arctic was supported by the Department of Energy's Office of Biological and Environmental Research. The NGEE Arctic project had two field research sites: 1) located within the Arctic polygonal tundra coastal region on the Barrow Environmental Observatory (BEO) and the North Slope near Utqiagvik (Barrow), Alaska and 2) multiple areas on the discontinuous permafrost region of the Seward Peninsula north of Nome, Alaska. Through observations, experiments, and synthesis with existing datasets, NGEE Arctic provided an enhanced knowledge base for multi-scale modeling and contributed to improved process representation at global pan-Arctic scales within the Department of Energy's Earth system Model (the Energy Exascale Earth System Model, or E3SM), and specifically within the E3SM Land Model component (ELM).

Thornton, Peter E [ORNL] (ORCID:0000000247595158)↗

Informing forest carbon inventories under the Paris Agreement using ground-based forest monitoring data

Human interactions with forests have shaped Earth's climate for millennia and will continue to do so as we target net-zero emission goals. Accurately characterizing these climate impacts requires making reliable forest carbon data available for forest monitoring and planning. Here, we develop a semi-automated process for submitting forest carbon measurements from the largest relevant scientific database to the International Panel on Climate Change's Emission Factor Database, which currently has sparse forest carbon data. Building this bridge from scientific research to international policy is an important step towards managing forests in a net-zero motivated future. Humans have been influencing Earth's climate via transformative impacts on forests for millennia, and forests are now recognized as critical to climate change mitigation under the Paris Agreement. The efficacy of climate change mitigation planning and reporting depends on quality data on forest carbon (C) stocks and changes. The Emission Factor Database (EFDB) of the International Panel on Climate Change (IPCC) is intended to be a definitive source for such data, but needs comprehensive and well-documented data to be so. To facilitate submission of forest C estimates from scientific studies to EFDB, we develop and document a process for semi-automated data submission from the Global Forest C database (ForC v4.0), which is the largest compilation of ground-based forest C estimates. We then assess the data currently available through ForC and provide recommendations for improving forest data collection, analysis, and reporting. As of September 2024, ForC contained ~19,286 records potentially relevant to EFDB, 1068 of which had been submitted and posted to EFDB. These represented 19% of the total EFDB records for forest land. Records were unevenly distributed across variables and geographic regions. ForC records (37%) reviewed could not be submitted because the original publication lacked required information. In the future, ground-based forest C estimates should target gaps in the record, and studies should ensure that they report all information necessary for inclusion in EFDB. Given that climate change is rapidly impacting the world's forests, timely reporting of recent estimates will be critical to accurate forest C inventories.

54 ENVIRONMENTAL SCIENCES↗