Search NASA⌕ Search

SEARCH · Search NASA

Results for “Python”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 739 records · Page 41

STM/S Grid LDOS Data and Analysis Code for Deciphering Majorana Zero Modes in Topological Superconductor

This dataset provides raw millikelvin scanning tunneling microscopy/spectroscopy (STM/S) grid spectroscopy data and Python analysis scripts supporting the manuscript “Deciphering Majorana Zero Modes in Topological Superconductor FeTe0.55Se0.45 with Machine-Learning-Assisted Spectral Deconvolution.” The dataset includes a raw grid spectroscopy file acquired on FeTe0.55Se0.45 at 40 mK under magnetic field, together with Python/Jupytext analysis scripts used for STM/S data processing, visualization, spectral deconvolution, Lorentzian peak fitting, feature extraction, machine-learning-assisted clustering, and figure generation. These files support the analysis of vortex-core local density of states and the identification of zero-bias-peak-related spectral components from complex in-gap states. The dataset is intended to provide a citable archival record of the data and analysis code associated with the published manuscript and to support transparency and reproducibility of the reported STM/S and machine-learning workflow.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND↗

Water4Energy Step-1 Band-M Ready-to-Train Samples for TVA Weeks-to-Years Prediction, Version 0

AI-ready Band-M (monthly) labelled training pack for the Water4Energy Genesis Task-1 project on weeks-to-years prediction of Tennessee Valley temperature and precipitation. The deposit includes leakage-aware issue-time samples (samples_M_v0.nc; N=486), train-only scalers, issue-time split table, supporting monthly panels, and Python generation scripts to recreate the pack from the companion Tier-1 raw observation collection (https://doi.org/10.13139/ORNLNCCS/3398576). Each sample pairs a 12-month lookback of teleconnection indices and SST box anomalies with TVA-mean ERA5 anomaly targets (t2m, tp, msl) at leads 1–3 months.

54 ENVIRONMENTAL SCIENCES↗

Microreactor Automated Control System Test Bed Digital Architecture for Real-Time, Hardware-in-the-Loop Simulation

This work describes progress made towards the development of a real-time hardware-in-the-loop (HIL) test bed for non-nuclear testing of microreactor control schemes and failure modes. Non-nuclear testing is a crucial step in developing robust control algorithms for managing microreactor dynamics. The creation of an HIL simulation harnesses the realistic dynamics of physical analogue systems while additionally considering the challenges of variable communication delay. This collaborative effort between Oak Ridge National Laboratory and Idaho National Laboratory has resulted in a LabVIEW-based gRPC communication protocol which couples a TRANSFORM Modelica simulation of nuclear components to the ViBRANT physical hardware for realistic feedback and visual representation of control action in real time. A modular python client structure is developed to manage FMU-based Modelica simulation and real-time gRPC communication. HIL testing suggests that the modeled reactor with natural convection molten salt loop coolant configuration responds well to PID control of drum positioning for modulation of reactor core power, however, future efforts will be made to explore the added thermal inertial delay of system level control and downstream demand changes. Development of this platform with a generalized methodology provides a foundation for exploring a variety of reactor configurations and failure modes in rapid order to provide insight into the most effective avenues of study for further research and development.

McConnell, Jono [ORNL] (ORCID:0000000238984741)↗

Replacing non-biomedical concepts improves embedding of biomedical concepts

Embeddings are semantically meaningful representations of words in a vector space, commonly used to enhance downstream machine learning applications. Traditional biomedical embedding techniques often replace all synonymous words representing biological or medical concepts with a unique token, ensuring consistent representation and improving embedding quality. However, the potential impact of replacing non-biomedical concept synonyms has received less attention. Embedding approaches often employ concept replacement to replace concepts that span multiple words, such as non-small-cell lung carcinoma, with a single concept identifier (e.g., D002289). Also, all synonyms of each concept are merged into the same identifier. Here, we additionally leveraged WordNet to identify and replace sets of non-biomedical synonyms with their most common representatives. This combined approach aimed to reduce embedding noise from non-biomedical terms while preserving the integrity of biomedical concept representations. We applied this method to 1,055 biomedical concept sets representing molecular signatures or medical categories and assessed the mean pairwise distance of embeddings with and without non-biomedical synonym replacement. A smaller mean pairwise distance was interpreted as greater intra-cluster coherence and higher embedding quality. Embeddings were generated using the Word2Vec algorithm applied to a corpus of 10 million PubMed abstracts. Our results demonstrate that the addition of non-biomedical synonym replacement reduced the mean intra-cluster distance by an average of 8%, suggesting that this complementary approach enhances embedding quality. Future work will assess its applicability to other embedding techniques and downstream tasks. Python code implementing this method is provided under an open-source license.

algorithms↗

SEGUID v2: Extending SEGUID checksums for circular, linear, single- and double-stranded biological sequences

Background Synthetic biology involves combining different DNA fragments, each containing functional biological parts, to address specific problems. Fundamental gene-function research often requires cloning and propagating DNA fragments, such as those from the iGEM Parts Registry or Addgene, typically distributed as circular plasmids. Addgene’s repository alone offers around 150,000 plasmids. To ensure data integrity, cryptographic checksums can be calculated for the sequences. Each sequence has a unique checksum, making checksums useful for validation and quick lookups of associated annotations. For example, the SEGUID checksum uniquely identifies protein sequences with a 27-character string. Objectives The original SEGUID, while effective for protein sequences and single-stranded DNA (ssDNA), is not suitable for circular DNA since there is no natural starting position nor for double-stranded DNA (dsDNA) since two separate sequences are present. Challenges include how to uniquely represent linear dsDNA, circular ssDNA, and circular dsDNA. To meet these needs, we propose SEGUID v2, which extends the original SEGUID to handle additional types of sequences. Conclusions SEGUID v2 produces orientation and rotation invariant checksums for single-stranded, double-stranded, possibly staggered, linear, and circular DNA and RNA sequences. Customizable alphabets allow for other types of sequences. In contrast to the original SEGUID, which uses Base64, SEGUID v2 uses Base64url to encode the SHA-1 hash. This ensures SEGUID v2 checksums can be used as-is in filenames, regardless of platform, and in URLs, with minimal friction. Availability SEGUID v2 is readily available for major programming languages, distributed under the MIT license. JavaScript package seguid is available on npm, Python package seguid on PyPi, R package seguid on CRAN, and a Tcl script on GitHub. These tools, along with documentation, examples, and an online SEGUID Calculator , can be found at https://www.seguid.org .

Pereira, Humberto↗

Spin-phonon coupling in AFM transition-metal mono-oxide

Time-of-flight INS measurements were performed on single crystal NiO with the Wide Angular Range Chopper Spectrometer (ARCS) at the Spallation Neutron Source. Experiments were performed on NiO single crystal mounted in an aluminum can and cooled using a closed-cycle helium refrigerator. Measurements were conducted at T = 100 K and 650 K, with the [HHL] scattering plane aligned horizontally. A Fermi chopper with slit spacing of 1.52mm, spinning at 300 Hz, was used to select an incident neutron energy of 100 meV. All datasets were normalized to a vanadium standard to correct for detector efficiency and solid angle coverage. The data sets include the .nxs files, the generated .hdf5 files (for use with Phonon Explorer), and Python scripts used to create them.

36 MATERIALS SCIENCE↗

EGS Collab Experiment 2: Microseismic Monitoring

This dataset contains continuous seismic waveform data recorded during stimulation and thermal circulation tests for the Enhanced Geothermal Systems (EGS) Collab Experiment #2, conducted from February to September 2022 at the Sanford Underground Research Facility in Lead, South Dakota. This experiment aimed to study and validate models of geothermal systems by injecting high-pressure fluids into rock formations 1200-1500 meters below the surface, inducing microseismic events. The seismic monitoring system included 16 three-component accelerometers and a 24-channel hydrophone array, installed in boreholes surrounding the test area. Data were recorded at high sampling rates using a continuous waveform recording system to monitor seismic activity in real time. The dataset contains the raw data stored in binary format, with files named based on timestamps, and includes calibration certificates for some sensors to facilitate corrections to real units. Users are strongly advised to consult the accompanying detailed report, which outlines the experimental setup, sensor specifications, installation procedures, and data processing methods. The report also describes important nuances, such as the hardware filters on hydrophones, sensor calibration details, and the naming conventions for the recorded data. Proper use of this dataset may require familiarity with seismic data analysis tools, such as the Obspy Python package, and an understanding of the SEED naming conventions used for channel identification.

15 GEOTHERMAL ENERGY↗

OCHRE

OCHRE™ uses a variety of input data sources to run time-series simulations. Building models can be taken from the ResStock™ database or generated using the Building Energy Optimization Tool (BEopt™) or other OpenStudio-HPXML workflows. EV charging profiles can be taken from datasets used in NLR's 2030 National Charging Network project. Weather data can be taken from the National Solar Radiation Database or EnergyPlus® weather files. There are no public datasets with OCHRE outputs at this time. However, a recent project dataset on water heater and EV demand flexibility can be requested. OCHRE is a Python-based energy modeling tool designed to model flexible loads in residential buildings. OCHRE includes detailed models and controls for flexible devices including HVAC equipment, water heaters, EVs, solar PV, and batteries. It is designed to run in co-simulation with custom controllers, aggregators, and grid models.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Upper-air soundings collected during the CROCUS Urban Canyons 2024 campaign in Chicago, Illinois USA

Funded by the Department of Energy’s Office of Science, Biological and Environmental Research program, Community Research on Climate and Urban Science (CROCUS) studies urban climate change and the impact it has on communities, with particular focus on disinvested, under-resourced communities. This information leads to new insights on urban climate challenges and informs future actions for mitigating and adapting to climate change at the street, neighborhood and regional levels.As part of the CROCUS effort, the Urban Canyons 2024 project was undertaken to study conditions at unprecedented detail over various neighborhoods in Chicago, Illinois. This dataset consists of upper air soundings that were collected as part of this effort. Soundings were launched during two intensive observing periods, IOP1 occurred on 22-23 July 2024, while IOP2 occurred on 27-28 July 2024. For IOP1, soundings were launched at coordinated times from three sites, Shedd Aquarium in Downtown Chicago, Abizu Campus High School in Humboldt Park, and Gary Comer Youth Center in West Woodlawn. For IOP2, the Gary Comer site was replaced by a neighborhood site in West Woodlawn, Chicago. The Abizu Campos site was operated by Valparaiso University and used iMET-4 rawinsondes, the other sites were operated by the University of Illinois Urbana-Champaign and used GRAW DFM-19 sondes.This dataset contains netCDF files containing quality-controlled temperature, dewpoint, geopotential height, pressure, and vector wind measurements at 1 second intervals following launch. These files are readable by the open-source netCDF software libraries available in many software packages (i.e., python, R, fortran, C++, etc.). The dataset also contains quicklook plots of each launch on a skew-T log-p thermodynamic diagram. These are in png format viewable by most web browsers.

54 ENVIRONMENTAL SCIENCES↗

Litter Production and Foliar Nutrient Resorption in Pioneer and Non-Pioneer Species in a Selective Logging Experiment in the Central Amazon, BIONTE, ZF-2, Manaus, 2022-23

This dataset was collected near the city of Manaus, Brazil, at the Experimental Station of Tropical Forestry (EEST, aka “ZF2”), inside the BIONTE (BIOmass and NuTrient Experiment). The experiment included three levels of increasing selective logging intensity, along with control, with 1-hectare permanent plots (12 total) located at the center of 4-hectare treatment plots. The vegetation has a high floristic diversity, the soils of the region are poor in nutrients, and the topography is characterized by plateaus (where BIONTE is located), and also valley bottoms and slopes. Three treatments of differing logging intensities were applied in the BIONTE experiment (T1, T2 and T3). The study was conducted in Treatment 3 (Block I – permanent plot), which represents the most intensive logging treatment, with 69% of the basal area (m²∙ha⁻¹) removed in 1988. The present dataset spans the period from May 1, 2022, to May 1, 2023. The data package includes leaf_nutrient_data, litterfall_total_data, leaf_litterfall_species_specific_data, and species_info, all provided in .csv format. These formats allow users to process and analyze the data in various software applications and programming languages, such as Python and R. This dataset was collected to advance knowledge on nutrient cycling in Amazonian forests, specifically distinguishing between species with two distinct functional traits: fast-growing and slow-growing. It also aims to improve Earth System Models, such as the E3SM Functionally Assembled Terrestrial Ecosystem Simulator (FATES). Additionally, it was used in a paper currently in preparation (Carvalho et al., in prep.), which aims to quantify seasonal litter production and foliar nutrient resorption in pioneer (fast-growing) and non-pioneer (slow-growing) tree species in the central Amazon. Specifically, it seeks to answer two key questions: 1) Is there a difference in leaf litter production, leaf nutrient flux and leaf nutrient concentration between pioneers and non-pioneers species? Is there a difference in the efficiency of foliar nutrient resorption between pioneers and non-pioneers species?

54 ENVIRONMENTAL SCIENCES↗

Model data for a watershed-scale study in the Portage River Basin (OH) examining the effects of subsurface drainage on the hydrologic response of an agricultural watershed.

This study builds on Rathore et al. (2024, WRR) and investigates the role of artificial tile-drainage on various aspects of watershed hydrological response, with a particular focus on peakflow. The model-data for the original modeling-focused paper (Rathore et al., 2024, WRR) is archived at Rathore et al. (2024, ESS-DIVE). Hence, this model-data archive provides scripts that are specific to this study that includes model updates, processing and analysis scripts. For details and models files of original model, readers are referred to Rathore et al. (2024, ESS-DIVE). The key difference between the model configuration in this study and Rathore et al. (2024, WRR) is that the tile drains are applied to the entire domain, to study the impact of tile-drains on different aspects of hydrological response. Additional scenario considering intensified precipitation after a dry period was also simulated. The Watershed Workflow package is implemented in Python3. The Jupyter notebooks can be executed through multiple open-source tools, for example, Anaconda Jupyter Lab, VS Studio Code, etc. Other data files include CSV and HDF5 files, which can be read through Python scripts.

54 ENVIRONMENTAL SCIENCES↗

CROCUS Air Quality Dataset from the University of Illinois Chicago (UIC), July 2024

This dataset was collected by the measurement system in the Atmosphere, Climate, and Ecosystems (ACE) Lab at the University of Illinois Chicago (UIC) from July 12 to July 31, 2024, as part of the Community Research on Climate and Urban Science (CROCUS) Urban Integrated Field Laboratory (UIFL) project, led by Argonne National Laboratory.To enhance understanding of urban air quality dynamics in Chicago, and as part of the CROCUS 2024 Urban Canyon Intensive Observation Period (IOP), several instruments were set up to provide continuous measurements of air quality parameters in Chicago during July 2024. These measurements cover both aerosols and gas-phase species. It focuses on particle size distribution (2.5–478 nm) measured by two Scanning Mobility Particle Sizers (SMPS) at a 4-min resolution, total particle number concentrations at a 1-s resolution, and chemical composition from a High-Resolution Time-of-Flight Aerosol Mass Spectrometer (AMS) at a 1-min resolution. Key gas-phase species, including NO, NO₂, SO₂, and O₃, are measured at a 1-min resolution, along with high-resolution NO and dimethyl sulfide (DMS) data from a Chemical Ionization Mass Spectrometer (CIMS). Volatile organic compound (VOC) data for toluene, isoprene, and benzene are provided by a GC-PID with a time resolution of 25 minutes.The data are formatted as NetCDF (.nc) files, making them easily accessible using common software such as MATLAB, R, and Python. Each parameter is stored in an individual dataset, which includes detailed instrument information in the header, as well as the corresponding sample start time and concentration/distribution data for each sample.

54 ENVIRONMENTAL SCIENCES↗

CROCUS 3-D wind data from Argonne Deployable Mast during Urban Canyon IOP July 2024

During the Community Research on Climate and Urban Science (CROCUS) Urban Integrated Field Laboratory (UIFL) project, led by Argonne National Laboratory, this dataset was collected by METEK uSonic-3 Class A MP sonic anemometer at 10 meter height of the Argonne Deployable Mast (ADM) from July 26 to July 28, 2024, .As part of the CROCUS 2024 Urban Canyon Intensive Observation Period (IOP), 3-D sonic anemometer on the ADM was set up to provide continuous measurements of wind and temperature in Chicago during July 2024. These measurements cover winds in X-, Y- and Z- direction and sonic temperature at 30-Hz resolution. The data are formatted as NetCDF (.nc) files, making them easily accessible using common software such as MATLAB, R, and Python.

54 ENVIRONMENTAL SCIENCES↗

Processed Soil Respiration at the TRACE experimental Warming project, Aug 2015 - Sep 2017, Sabana, Luquillo, Puerto Rico

This data package contains processed measurements of soil carbon dioxide (CO₂) efflux collected using LI-COR LI-8100 soil respiration chambers at the Tropical Responses to Altered Climate Experiment (TRACE) located at the Sabana Field Research Station near Luquillo, Puerto Rico. The TRACE site is a mature, closed-canopy tropical wet forest within the Luquillo Experimental Forest. These data quantify soil surface CO₂ fluxes from both ambient (control) and experimentally warmed plots to evaluate how long-term soil warming affects belowground carbon cycling in tropical ecosystems. The data files include time-series tables of CO₂ flux (µmol CO₂ m⁻² s⁻¹), soil temperature (°C), and ancillary environmental variables, stored in comma-separated values (CSV) format and viewable with any text editor, spreadsheet, or statistical software (e.g., R, Python, Excel). Associated metadata describe plot identifiers, measurement intervals, and processing steps. These data were generated to address the research question: How does sustained soil warming influence soil respiration and carbon flux dynamics in tropical wet forests?

54 ENVIRONMENTAL SCIENCES↗

Data and code for Daily and Multi-Day Extreme Rainfall Analysis Under Future Climates Using Stochastic Storm Transposition and NEX-GDDP-CMIP6 Over CONUS

This data package provides inputs, codes, and outputs for a comprehensive analysis of projected changes in extreme precipitation across 10 regions of the continental United States, using 34 downscaled Earth System Models (ESMs) from the NASA Earth Exchange Global Daily Downscaled Projections, Coupled Model Intercomparison Project Phase 6 (NEX-GDDP-CMIP6) dataset. These models are part of the Coupled Model Intercomparison Project Phase 6 (CMIP6), a coordinated climate modeling framework widely used to assess climate change impacts. The analysis applies a stochastic storm transposition method to quantify changes in extreme rainfall under two Shared Socioeconomic Pathway (SSP) climate scenarios—SSP2-4.5 (moderate emissions) and SSP5-8.5 (high emissions)—compared to historical conditions (1995–2014 vs. 2081–2100). The dataset includes rainfall depth estimates for extreme events with return periods from 2 to 500 years across multiple storm durations (1, 3, and 5 days) for each of the 10 U.S. regions. Weighted ensemble statistics are derived from individual ESM performance against historical precipitation patterns, enabling robust uncertainty quantification through both sign-based and permutation-test-based model agreement assessments. Key analyses address: (1) relative changes in extreme precipitation for each climate scenario, (2) differences between SSP scenarios (SSP5-8.5 vs. SSP2-4.5), (3) contrasts between rare and frequent events, and (4) variations between multi-day and daily storm durations. The workflow produces ensemble statistics—median, 5th, 25th, 75th, and 95th percentiles—along with model agreement metrics that identify regions and event types with robust climate change signals. The dataset includes: processed rainfall depth outputs (netCDF format) from the RainyDay Python package, ESM weights from historical performance evaluation using DayMet observations, ensemble statistics across all storm dimensions, and figures summarizing key findings.

54 ENVIRONMENTAL SCIENCES↗

Site and endmember spectra of terrestrial vegetation and soils for the Colorado Headwaters Ecological Spectroscopy Study, June-July 2025

This dataset provides site and endmember spectra collected during the 2025 Colorado Headwaters Ecological Spectroscopy Study (CHESS) campaign. The site spectra were collected to help validate airborne hyperspectral data acquired by the National Ecological Observatory Network's aerial observation platform (NEON AOP). Endmember spectra were collected to augment existing spectral libraries with additional samples of bare surfaces and non-photosynthetic vegetation. All measurements were acquired with an Analytical Spectral Devices (ASD) FieldSpec4 Hi-Res NG (Next Generation) spectroradiometer, which records radiance at 1nm (nanometer) intervals from the ultraviolet to the short-wave infrared (350-2500 nm). The dataset includes spectra measured at meadow sites where the CHESS team also collected vegetation samples for trait analyses. The site spectra were collected with the ASD FieldSpec4 palm grip attachment using an 8° field-of-view foreoptic. Site spectra are integrated measurements of the entire surface within the foreoptic’s field of view. For site-level spectra, the sun is the illumination source. A Spectralon panel mounted on a tripod was used for instrument optimization and white reference measurements for all site spectra. Site spectra were acquired within two hours of solar noon and within 48 hours of a NEON AOP overflight. Site spectra are labeled by date, sampling area, and site number according to the naming conventions of the CHESS campaign’s data management plan. The dataset also contains endmember spectra in the following categories: photosynthetic vegetation (PV), non-photosynthetic vegetation (NPV), bare (soil/rock), and flowers. Endmember measurements were acquired using either the contact probe or the leaf clip attachments of the ASD FieldSpec4. In these configurations, the bulb inside the spectrometer provides the light source for the measurements. The spectrometer was optimized and white reference measurements were recorded using the circular white pucks attached to the contact probe and leaf clip. Because they do not rely on solar illumination, contact probe and leaf clip measurements were collected during a broader time frame than the palm grip site spectra. Some endmembers were measured at CHESS meadow sites, while others were collected within the larger sampling area or in nearby locations (e.g. Gothic Townsite) with similar characteristics. Radiance, reflectance, and metadata files are split into three subfolders according to measurement type: proximal/palm grip (prx), contact probe (cp), and leaf clip (lc). Radiance spectra are provided in ASD file format (.asd file extension). All ASD files can be opened using the provided scripts. Metadata is provided in two formats: CSV file format (no geolocation) and GEOJSON file format (includes geolocation for each spectra). The dataset includes a set of pre-processed reflectance spectra as CSV files (yyyymmdd_rfl.csv). The python scripts and jupyter notebook used to calculate reflectance spectra from the ASD radiance data is included here and was previously published at: https://doi.org/10.3334/ORNLDAAC/2446. There is also a folder of JPEG photographs corresponding to selected spectra. We include a protocol document with detailed steps for ASD FieldSpec4 assembly and operations. This data additionally contains a file level metadata (flmd.csv) and data dictionary (dd.csv) file. Geospatial information: Geospatial data for mapping measurement site locations are in the files CHESS_polygons_lai_UTM.geojson, CHESS_polygons_shrub_UTM.geojson, and CHESS_polygons_meadow_UTM.geojson in the companion geospatial package for the 2025 CHESS campaign, ‘CHESS 2025: Location data for field observations and sampling’ (Henderson et al., 2026). CHESS Project Description: The Colorado Headwaters Ecological Spectroscopy Study (CHESS) comprised a multi-week airborne remote sensing and field observation campaign in the Upper Gunnison Basin, Colorado, conducted in June and July of 2025. Airborne remote sensing was conducted by the National Ecological Observatory Network Airborne Observation Platform (NEON AOP), concurrent with a field campaign run by the Rocky Mountain Biological Laboratory (RMBL), the Lawrence Berkeley National Laboratory (LBNL) and SLAC National Accelerator Laboratory Watershed Function Science Focus Area (SFA), and NASA-JPL (Jet Propulsion Laboratory) Earth Surface Mineral Dust Source Investigation (EMIT) program. Between June 10 and July 18, 2025, the NEON AOP flight team collected high-resolution aerial imaging spectroscopy and Light Detection and Ranging (LiDAR) data over three domains: the Upper East River (CRBU), Almont Triangle (ALMO), and the Upper Taylor Basin (UPTA). In coordination with the flights, a field campaign acquired ground-truth observations, including observations of vegetation composition, foliar traits, forest demography, and subsurface properties in 18 core sampling areas within the domains. Additional surface water observations were taken at over 380 point locations. All CHESS campaign datasets can be found within the CHESS ESS-DIVE data portal: https://data.ess-dive.lbl.gov/portals/chess. Funding Acknowledgment: This research was carried out at the Jet Propulsion Laboratory, California Institute of Technology, under a contract with the National Aeronautics and Space Administration (80NM0018D0004) and was funded by EMIT Extended Mission Phase E Science.

2018 NEON and 2025 CHESS Campaigns↗

Myco-CORPSE simulations assessing mycorrhizal carbon allocation across U.S. forests and global change scenarios

Plants allocate a substantial portion of their fixed carbon belowground to mycorrhizal fungi in exchange for nutrients and other benefits. However, most current ecosystem models omit mycorrhizal processes, limiting our ability to predict plant–soil carbon dynamics under environmental change. To address this gap, we used a mycorrhiza-explicit soil biogeochemical model, Myco-CORPSE (Mycorrhizal Carbon, Organisms, Rhizosphere, and Protection in the Soil Environment), to simulate tree carbon allocation to arbuscular mycorrhizal (AM) and ectomycorrhizal (ECM) fungi in temperate forests.The dataset includes outputs from two sets of model simulations:1. Perturbation experiments: Simulations across gradients of ECM dominance (0–100%), nitrogen deposition, soil temperature, and net primary productivity (NPP) to test how these factors affect mycorrhizal C allocation and nutrient cycling.2. FIA-based simulations: Model applications to over 1,800 U.S. forest sites using site-specific data from the U.S. Forest Inventory and Analysis (FIA) program, including vegetation composition, mycorrhizal type, climate, litter traits, soil properties, and N deposition.Model outputs include simulated mycorrhizal carbon allocation and related biogeochemical variables, such as soil and microbial carbon and nitrogen stocks. Data are provided in CSV format and organized by experiment type (in separate ZIP files). Python scripts for running simulations, plotting, and spatial mapping are also included and organized similarly. No proprietary software is required. These outputs support a peer-reviewed study and were used to generate figures and tables in the associated publication.

54 ENVIRONMENTAL SCIENCES↗

CROCUS Optical All Precipitation Gauge Data at Argonne National Laboratory Prairie Site

The APG (Optical Scientific Inc. All-Precipitation Gauge 815-DS) dataset contains one-minute measurements of precipitation rate, precipitation accumulation, air temperature, and present weather detection, both in 4680 format and decoded. Data were collected at the Argonne Testbed for Multiscale Observational Science (ATMOS), a 20-acre prairie site at Argonne National Laboratory in Lemont, Illinois. The data is presented as daily NetCDF (.nc) files, each containing approximately 24 hours of observations. Files follow the naming convention of: the project (CROCUS), location (atmos), instrument name (apg), data level (raw, a1), and date (year, month, day). The NetCDF format can be accessed using common scientific software such as Python using xarray, netCDF4 or act-doe.

54 ENVIRONMENTAL SCIENCES↗