Search NASA⌕ Search

SEARCH · Search NASA

Results for “value.”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

North America’s Potential for an Environmentally Sustainable Nickel, Manganese, and Cobalt Battery Value Chain

The Detroit Big Three General Motors (GMs), Ford, and Stellantis predict that electric vehicle (EV) sales will comprise 40–50% of the annual vehicle sales by 2030. Among the key components of LIBs, the LiNixMnyCo1−x−yO2 cathode, which comprises nickel, manganese, and cobalt (NMC) in various stoichiometric ratios, is widely used in EV batteries. This review reveals NMC cathodes from laboratory research. Furthermore, this study examines the environmental effect of NMC cathode production for EV batteries (including coating technologies), encompassing aspects such as energy consumption, water usage, and air emissions. Although gaps persist in NMC cathode environmental assessments (NMC111, NMC532, NMC622, and NMC811), limited life cycle assessments “(LCA)” have been conducted. Most available data originate from Asia (primarily China), accounting for 85% of the production of EV LIB cathode materials. The concept of battery passports for data collection on LIB components has been proposed to facilitate material traceability as a system for ensuring a sustainable supply chain for critical minerals. The automotive industry’s shift to electrification necessitates a sustainable supply chain from mine to vehicle end-of-life. As the critical mineral supply moves from Asia to North America, environmentally friendly industrial methods must be studied to provide this supply chain direction.

25 ENERGY STORAGE↗

Extraction of Value-Added Products from Food Processing Waste Using Dimethyl Ether

Poster for 2024 Intern Poster Session. Food production waste can be valorized to create a circular economy. Traditional extraction methods require pretreatment of the sample through heating or cell disruption, but this contributes to a majority of the process's energy usage for wet biomass. Using dimethyl ether extraction can combine the dewatering and extraction processes into one to skip the pretreatment step while still maintaining similar extraction rates.

09 BIOMASS FUELS↗

Extraction of Value-Added Products From Food Processing Waste Using Liquid Dimethyl Ether

Technical presentation for 2024 Intern Poster Session. Food production waste can be valorized to create a circular economy. Traditional extraction methods require pretreatment of the sample through heating or cell disruption, but this contributes to a majority of the process's energy usage for wet biomass. Using dimethyl ether extraction can combine the dewatering and extraction processes into one to skip the pretreatment step while still maintaining similar extraction rates.

09 BIOMASS FUELS↗

SCA Tools - SCRM Value Add or Lossy Noise Machines

Software supply chain risk management (SCRM) depends upon accurate information regarding the software components that comprise any given software system. The collection of components included in a software package can be organized within a software bill of materials, or SBOM. SBOMs are ideally generated when the software components are put together, such as at compile time, but for many reasons that has not and is not always possible. For example, legacy or proprietary software packages often do not have SBOMs available to downstream consumers of that software. It’s not just end users that are affected, manufacturers themselves also must deal with this problem. To answer these questions, the market has seen the rise of several commercial software composition analysis (SCA) tools. These tools aim to peer into completed software systems, automatically identifying hidden software dependencies and looking up known vulnerabilities associated with those dependencies to enable end-users to enhance their cyber supply chain risk management processes. These tools are potentially a huge boon to end users of legacy and proprietary software – and a potential bane, depending on how accurate they are. This research asks that question – how accurate are currently available binary SCA tools – and provides answers to several other questions: What does it mean to be “accurate”? What limitations do the tools have in identifying common edge cases that take place in modern software development? Can they help you avoid a devastating supply chain attack, or is it all just noise? After researching SCA tools on the market, we identified three vendors that fit our use case and would provide analysis on compiled binaries. Using these tools, we submitted firmware for critical infrastructure devices for analysis and SBOM generation. The SBOM outputs were then cross referenced with SBOMs generated through manual analysis for comparison. In addition to the firmware samples, we also submitted edge case samples based off a popular open-source library that were specifically crafted to evaluate each tools’ ability to accurately identify components. These samples were customized to be consistent with modifications we have seen in modern software development as well as a couple that are representative of supply chain attacks.

97 MATHEMATICS AND COMPUTING↗

TRAILS Output Files

Overview This data repository contains ZIP files that store compressed versions of the output of running the WaterPaths utility planning and management tool in the DU Re-Evaluation mode (to download the tool, please see this GitHub repository). The tool was used to simulate the six-utility North Carolina Research Triangle problem. Details on the contents of each ZIP file can be seen below. Data details Temporal range: Weekly data for 2,344 weeks from 2015 to 2060 (45 years). Spatial range: Six water utilities in the North Carolina Research Triangle region (0: Chapel Hil/OWASA, 1: Durham, 2: Cary, 3: Raleigh, 4: Pittsboro, and 5: Chatham) File types: CSV and OUT Different solutions available The solution numbers correspond to the different pathway strategies (henceforth referred to as "solutions") discussed in paper's main and supporting text (abstract and link to the paper here). They are as follows: Sol92: The Durham-focused pathway strategy Sol132: The Raleigh-focused pathway strategy Sol140: The regionally-robust pathway strategy Objectives files These files can be accessed by unzipping solXX_objectives_pathways.zip that contains 1,000 Objectives_RDMXX_solsXX_to_XX.csv files. Each CSV file will consist of a row representing all the objective values for that specific solution, while every six columns represents the reliability, restriction frequency, infrastructure net present value ($ mil), peak financial cost, worst-case cost, and unit cost ($ per MG; in that order) for each of the six utilities. There will be 1,000 such files, denoting the performance of the six utilities across the 1,000 deeply uncertain states of the world (DU SOWs). Pathway files These files can be accessed by unzipping solXX_objectives_pathways.zip that contains 1,000 Pathways_sXX_RDMXX.out file. Each OUT corresponds to the set of infrastructure being triggered in a specific DU SOW, and each file will have the name file will consist of four tab-delimited columns that are described as follows: Realization: The realization in which an infrastructure options being triggered utility: The utility currently triggering infrastructure week: The week in which a specific infrastructure option is being triggered infra.: The infrastructure option being triggered If the OUT file contains only the header line, no infrastructure was triggered for that specific DU SOW. Policies files These files can be obtained by unzipping Policies.zip. Each of the 1,000 CSV files within the unzipped folder will contain weekly water use restriction policies for all 1,000 hydroclimatic realizations within a specific DU SOW. The column structure is as follows: 0rest_m: restriction multiplier for utility 0 (values between 0 and 1) 1rest_m: restriction multiplier for utility 1 (values between 0 and 1) 2rest_m: restriction multiplier for utility 2 (values between 0 and 1) 3rest_m: restriction multiplier for utility 3 (values between 0 and 1) 4rest_m: restriction multiplier for utility 4 (values between 0 and 1) 5rest_m: restriction multiplier for utility 5 (values between 0 and 1) 0transf: transfer volume for utility 0 (in MGD) 1transf: transfer volume for utility 1 (in MGD) 2transf: transfer volume for utility 2 (in MGD) 3transf: transfer volume for utility 3 (in MGD) 4transf: transfer volume for utility 4 (in MGD) 5transf: transfer volume for utility 5 (in MGD) Water Sources files These files can be obtained by unzipping WaterSources_subset.zip. Each of the 100 CSV files within the unzipped folder will contain weekly state variables at each water source for all 1,000 hydroclimatic realizations within a specific DU SOW. The column structure is as follows: Xvolume: available water volume from source X (in MGD) Xs_area: surface area of source X (in ACF) Xdemand: demand drawn from a water source from source X (in MGD) Xup_spill: upstream spillage from source X (in MGD) Xww_inflow: wastewater inflow from source X (in MGD) Xcatch_inflow: upstream catchment inflow to source X (in MGD) Xevap: evaporation multiplier for source X (values between 0 and 1) Xds_spill: downstream spillage from source X (in MGD) X_Y_alloc_cap: the allocated capacity from source X to utility Y (values between 0 and 1) X_Y_alloc_dem: the allocated demand from source X to utility Y (values between 0 and 1) Xtrmt_alloc_Y: the allocated treatment capacity from source X to utility Y (values between 0 and 1) Utilities files These files can be obtained by unzipping Utilities_subset.zip. Each of the 100 CSV files within the unzipped folder will contain weekly state variables at each utility for all 1,000 hydroclimatic realizations within a specific DU SOW. The column structure is as follows: Xst_vol: total available storage volume of utility X (in MG) Xcapacity: total storage capacity of utility X (in MG) Xnet_inf: : net inflow for all storage infrastructure for utility X (in MGD) Xst_rof: short term ROF for utility X (values between 0 and 1) Xst_stor_rof: short-term storage ROF for utility X (values between 0 and 1) Xst_trmt_rof: short-term treatment ROF for utility X (values between 0 and 1) Xlt_rof: long-term ROF for utility X (values between 0 and 1) Xlt_stor_rof: long-term storage ROF for utility X (values between 0 and 1) Xlt_trmt_rof: long-term treatment ROF for utility X (values between 0 and 1) Xrest_demand: restricted demand for utility X (in MGD) Xunrest_demand: unrestricted demand for utility X (in MGD) Xunfulf_demand: unfulfilled demand for utility X (in MGD) Xwastewater: wastewater return for utility X (in MGD) Xtreat_capacity: total treatment capacity for utility X (in MG) Xcont_fund: reserve (contingency) fund balance for utility X Xins_pout: insurance payout for utility X (% annual volumetric revenue) Xins_price: insurance price for utility X (% annual volumetric revenue) Xinfra_npv: infrastructure net present value for utility ($mil) Xst_vol: total available storage volume of utility X (in MG) Xdebt_serv: debt service for utility X (usually once per year if the infrastructure is triggered; % annual volumetric revenue) Xstor_vol: total stored volume (in MGD) Xobs_ann_dem: observed annual demand for utility X (in MGD) Xproj_dem: projected annual demand for utility X (in MGD) Xpv_debt_serv: present value of debt service payments for utility X (% annual volumetric revenue) Xgross_rev: gross revenue for utility X ($mil) Acknowledgment IM3 is a multi-institutional effort led by Pacific Northwest National Laboratory and supported by the U.S. Department of Energy's Office of Science as part of research in MultiSector Dynamics, Earth and Environmental Systems Modeling Program.

Artificial Intelligence↗

HDG-1 Experiment Irradiation Monitoring Data Qualification Final Report

SUMMARY The U.S. Department of Energy (DOE) Advanced Reactor Technologies (ART) Graphite Research and Development (GRD) Program is conducting a series of six experiments to quantify the effects of irradiation on nuclear-grade graphite. This report documents the qualification of irradiation monitoring data for the fifth experiment, High Dose Graphite-1 (HDG-1). Qualified monitoring data are required by the ART program to support the design and licensing of the first high-temperature reactor (HTR) nuclear plant. Data are classified as Qualified if they meet the usage requirements described in the experiment planning and quality assurance (QA) documents, Failed if they do not meet those requirements and provide no usable information, or Trend if they do not fully meet all requirements but still provide useful information subject to an assessment of how any deficiencies may affect a particular use of the data. HDG-1 irradiation began with Advanced Test Reactor (ATR) Cycle 168B on August 24, 2020, and concluded after Cycle 173C on January 27, 2025. The HDG-1 capsule was removed from the reactor core twice—during core internal change (CIC) Cycle 170A and powered axial locator mechanism (PALM) Cycle 172A—to prevent overheating of the graphite specimens during high-power PALM cycles. The capsule was therefore irradiated during a total of seven normal ATR cycles: 168B, 169A, 171A, 171B, 173A, 173B, and 173C. Irradiation monitoring data evaluated in this report include thermocouple (TC) temperature, gas flow rate, gas moisture, gas pressure, specimen load, and graphite stack displacement. Temperature. A total of 14,508,065 TC temperature records were captured. Of these, 13,901,785 (95.8%) are Qualified and 606,280 (4.2%) are Failed. The principal source of failed temperature data was the instrument failure of TC-9 (Zone 2) on June 24, 2024, and TC-10 (Zone 1) on July 5, 2024, near the end of Cycle 173A, which resulted in 595,554 Failed readings. An additional 379 missing values and 10,347 slightly negative values from TC-13 during ATR outages are also Failed. Neither TC-9 nor TC-10 was used as a temperature-control TC, and their failures did not compromise capsule condition monitoring. Correlation analysis of all 13 TCs found no evidence of virtual junction formation. Control chart analysis revealed clear downward drift of approximately 80°C for TC-6 (Zone 3) relative to other stable TCs, and possible downward drift of approximately 60°C for TC-13 relative to the Zone 5 control TC (TC-1), though TC-13 remained consistent with the Zone 2 control TC (TC-12). Gas flow. A total of 20,088,090 gas flow rate records were captured. Of these, 19,941,463 (99.3%) are Qualified and 146,627 (0.7%) are Failed due to missing values. All argon, helium, and total gas flow data were within expected ranges throughout the irradiation. Gas moisture. A total of 1,116,005 outlet gas moisture values were captured. Of these, 1,101,421 (98.7%) are Qualified and 14,584 (1.3%) are Failed, comprising 14,556 out-of-range values and 28 missing values. The out-of-range moisture values exceeded 22,000 ppmv for approximately 1 week at the beginning of Cycle 173A, when accumulated moisture evaporated after the capsule was retrieved from water storage during PALM Cycle 172A and reinserted into the east flux trap. Moisture levels returned to below 25 ppmv for the remaining three cycles, and the transient high-moisture event did not affect the integrity of specimen irradiation. Gas pressure. A total of 7,812,035 gas pressure values were captured. Of these, 6,642,048 (85.0%) are Qualified and 1,169,987 (15.0%) outlet pressure values are Failed, comprising 718,537 zero outlet pressure values due to sensor failure from Cycle 168B through Cycle 171B, 54,550 missing values, and 396,900 too-low outlet pressure values, ranging from 1.1 to 1.6 psia after sensor replacement during Cycle 173A. Load. A total of 6,696,030 load values were captured. Of these, 6,694,580 (99.98%) are Qualified and 1,450 (0.02%) are Failed due to missing values. Applied loads to the six specimen stacks were stable throughout the irradiation. Stack displacement. A total of 6,696,030 displacement values were captured. Of these, 5,713,297 (85.32%) are Qualified and 3,781 (0.06%) are Failed due to missing values. Stack displacement increased consistently throughout the irradiation, reaching approximately 3.08 in. for Channels 5 and 6 by the end of irradiation. 978,952 (14.62%) substantially elevated displacements observed for Channel 6 beginning in Cycle 171A and for Channel 5 beginning in Cycle 173A are assigned Trend status. Raising pressure. A total of 1,115,999 raising pressure values were captured. Of these, 1,115,430 (99.95%) are Qualified and 569 (0.05%) are Failed due to missing values. Ram pressure. A total of 6,696,030 ram pressure values were captured. Of these, 6,692,249 (99.95%) are Qualified and 3,484 (0.05%) are Failed due to missing values. Stack raising was perf

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Energy Infrastructure Futures: A Multiscale Evaluation of Projected Power Plant Siting Across the Western Interconnection

Energy Infrastructure Futures: A Multiscale Evaluation of Projected Power Plant Siting Across the Western Interconnection Description This dataset contains input and output data for the manuscript Mongird, K. et al. (under review) titled "Energy Infrastructure Futures: A Multiscale Evaluation of Projected Power Plant Siting Across the Western Interconnection". Input data corresponds to gridded spatial siting attributes that are necessary to conduct a random forest machine learning analysis of siting feature importance. Output data includes SHAP feature analysis outputs, and classification report values. For data on power plant siting results referred to in the manuscript, please refer to the CERF: IM3 Projected Western US Power Plant Locations data download page. The downloadable data includes values for eight different future scenarios for the Western US. The scenarios include combinations of two Shared Socioeconomic Pathways (SSP3 and SSP5) with four high-resolution climate projections specific to the United States (see, https://tgw-data.msdlive.org/). These climate projections include "hotter" and "cooler" variants for two Representative Concentration Pathways (RCP4.5 and RCP8.5). The resulting eight simulations are: rcp45cooler_ssp3 rcp45cooler_ssp5 rcp45hotter_ssp3 rcp45hotter_ssp5 rcp85cooler_ssp3 rcp85cooler_ssp5 rcp85hotter_ssp3 rcp85hotter_ssp5 Technical Information The dataset includes two sets of data files: (1) CERF gridded siting parameters and (2) Feature analysis outputs and classification reports. All downloadable data is in csv file format. Files with x/y coordinate information use the Albers Equal Area Conic projection (ESRI:102003). 1. CERF Gridded Siting Parameters This directory provides a balanced sample of gridded CERF siting parameters data for eight different scenarios for the Western US through 2055, seven different technologies, and eight timesteps. This data serves as input to the feature analysis. It contains the following parameters. region_name - name of region (i.e., state) sited - binary value representing whether the grid cell received a siting of that technology type (1=True) rcp - binary value representing scenario resource concentration pathway (0 = RCP4.5, 1 = RCP8.5) ssp - binary value representing scenario shared socioeconomic pathway (0 = SSP3, 1 = SSP5) climate - binary value representing cooler (0) or hotter (1) GCM forcing tech_name - generation technology name sited_year - year that values correspond to transmission_cost - cost of transmission interconnection pipeline_cost - cost of natural gas pipeline interconnection interconnection_cost - total interconnection cost (sum of transmission cost and gas pipeline cost) lmp - associated locational marginal value ($/MWh) associated with the grid cell, timestep, scenario, and technology xcoord - x-coordinate of location ycoord - y-coordinate of location 2a. Feature Analysis Output The dataset includes the feature analysis shap output for locational marginal price and interconnection cost. It contains the following parameters. technology - generator technology name scenario - name of scenario feature - name of feature, either locational_marginal_price or interconnection_cost value - the mean of absolute value of SHAP values for given feature 2b. Feature Analysis Classification Report This download includes the classification report associated with each random forest model. The dataset contains the following parameters. technology - generation technology name scenario - name of scenario test - one of precision (the proportion of predicted positives that are actually correct), recall (the proportion of actual positives that were correctly identified), f1-score (the harmonic mean of precision and recall) 0.0 - value of test for classification of 0 (grid cell not chosen for siting) 1.0 - value of test for classification of 1 (grid cell chosen for siting) accuracy - accuracy of model (i.e., fraction of all predictions that were right) macro avg - Simple average of test values for all classes weighted avg - Weighted average of test values for all classes, weighted based on Acknowledgment IM3 is a multi-institutional effort led by Pacific Northwest National Laboratory and supported by the U.S. Department of Energy's Office of Science as part of research in MultiSector Dynamics, Earth and Environmental Systems Modeling Program. License This data is made available under a CCBY4 License Disclaimer This material was prepared as an account of work sponsored by an agency of the United States Government. Neither the United States Government nor the United States Department of Energy, nor the Contractor, nor any or their employees, nor any jurisdiction or organization that has cooperated in the development of these materials, makes any warranty, express or implied, or assumes any legal liability or responsibility for the accuracy, completeness, or usefulness or any information, apparatus, product, software, or process disclosed, or represents that its use would not infringe privately owned rights. Reference herein to any specific commercial product, process, or service by trade name, trademark, manufacturer, or otherwise does not necessarily constitute or imply its endorsement, recommendation, or favoring by the United States Government or any agency thereof, or Battelle Memorial Institute. The views and opinions of authors expressed herein do not necessarily state or reflect those of the United States Government or any agency thereof. PACIFIC NORTHWEST NATIONAL LABORATORYoperated byBATTELLEfor theUNITED STATES DEPARTMENT OF ENERGYunder Contract DE-AC05-76RL01830

Mongird, Kendall [Pacific Northwest National Labor↗

Marginal Soils Index Analysis & Geospatial Data

This data package contains output files associated with Mongird et al. (in prep) organized into four dataset directories. Each dataset is described in more detail below. 1. Marginal Soils Index Analysis Description: This folder contains a csv file with land needs and availability by state, power generating technology type, and scenario in 2050 when suitable siting areas are additionally constrained to areas with increasing levels of soil marginality. Files: msi_constrained_siting_availability_2050.csv Variables: Scenario - Projected 2050 scenario name State - US state abbreviation Technology - Generating technology type solar = solar photovoltaic gas_cc_re = natural gas combined cycle (recirculating cooling) wind = onshore wind gas_cc_ccs_re = natural gas combined cycle with carbon capture sequestration (recirculating cooling) gas_cc_dry = natural gas combined cycle with (dry cooling) gas_cc_pond = natural gas combined cycle with (pond cooling) coal_conv_ccs_re = conventional coal with carbon capture sequestration (recirculating cooling) Req_Capacity_MW - The amount of rated capacity required in 2050 of the given technology type in the given state and under the given scenario from the capacity expansion plan Req_Capacity_Factor - The assumed capacity factor (fraction between 0 and 1) for the given technology type in the given state and under the given scenario by the capacity expansion plan Req_Land_km2 - The amount of land required (in km-squared) to host the required generating capacity that is capable of meeting the specified capacity factor for the given technology type in the given state and under the given scenario Req_Energy_TWh - Product of Req_Capacity_MW, Req_Capacity_Factor, and 8760/1e6 for the given technology type in the given state and under the given scenario MSI_Case - The level of MSI that siting the given technology is additionally constrained to, where >0 means siting is additionally constrained to suitable land areas that have an MSI value greater than 0 >=1 means siting is additionally constrained to suitable land areas that have an MSI value greater than or equal to 1 >=2 means siting is additionally constrained to suitable land areas that have an MSI value greater than or equal to 2 >=3 means siting is additionally constrained to suitable land areas that have an MSI value greater than or equal to 3 Soil Attribute Rasters Description: This folder contains geospatial raster files for individual soil parameters upscaled to the listed grid resolution (30m or 1 km). 1 km resolution files are a spatial average of non-missing 30m resolution values. All raster files use the USA Contiguous Albers Equal Area Conic (ESRI:102003) projection. Files: avg_cond_raster_ .tif - Average conductivity of the saturation extract across all soil horizons within a depth of 40 inches, measured in mmhos/cm max_cond_raster_ .tif - Maximum conductivity of the saturation extract across all soil horizons within a depth of 40 inches, measured in mmhos/cm min_ph_raster_ .tif- Min pH values across all soil horizons within a depth of 40 inches. avg_ph_raster_ .tif - Average pH value across all soil horizons within a depth of 40 inches. max_ph_raster_ .tif- Max pH value across all soil horizons within a depth of 40 inches. erosion_factor_raster_ .tif - Product of k-factor and percent slope flood_freq_raster_ .tif - Number of months of the year during which the area is commonly, frequently, or very frequently flooded. max_sar_raster_ .tif - Maximum sodium adsorption ratio across all horizons within a depth of 40 inches rock_frac_raster_ .tif - Fraction of the upper 6 inches of soil composed of rock fragments larger than 3 inches. temp_regime_raster_ .tif - Soil temperature regime with the following key: 0 = pergelic 1 = gelic 2 = cryic 3 = frigid 4 = isofrigid 5 = mesic 6 = isomesic 7 = thermic 8 = isothermic 9 = hyperthermic 10 =isohyperthermic Marginal Soils Index Rasters Description: This folder contains geospatial raster files of the Marginal Soils Index at the listed grid resolution (30m or 1 km). 1 km resolution files are a spatial average of 30m resolution. Both raster files use the USA Contiguous Albers Equal Area Conic (ESRI:102003) projection. A value of 0 indicates that there were no soil attributes present that indicate marginal soil. NA values indicate that data was unavailable or bodies of water. Files: marginal_soils_index_30m_raster.tif marginal_soils_index_1km_raster.tif Marginal Soils Index Resource Potential Rasters Description: This folder contains geospatial raster files of the Marginal Soils Index + Resource Potential (MSI+RP) score at 1km resolution for geothermal, solar, and wind technologies. Raster files use the USA Contiguous Albers Equal Area Conic (ESRI:102003) projection. NA values indicate that the location is not suitable for siting the given technology due to policy, environmental, socioeconomic, topological, and other constraints regardless of soil marginality level. Areas with values greater than or equal to zero represent the product of the normalized MSI value and the normalized resource potential value. Files: geothermal_msi_ep_score_raster.tif solar_msi_ep_score_raster.tif wind_msi_ep_score_raster.tif Acknowledgment IM3 is a multi-institutional effort led by Pacific Northwest National Laboratory and supported by the U.S. Department of Energy's Office of Science as part of research in MultiSector Dynamics, Earth and Environmental Systems Modeling Program. Disclaimer This material was prepared as an account of work sponsored by an agency of the United States Government. Neither the United States Government nor the United States Department of Energy, nor the Contractor, nor any or their employees, nor any jurisdiction or organization that has cooperated in the development of these materials, makes any warranty, express or implied, or assumes any legal liability or responsibility for the accuracy, completeness, or usefulness or any information, apparatus, product, software, or process disclosed, or represents that its use would not infringe privately owned rights. Reference herein to any specific commercial product, process, or service by trade name, trademark, manufacturer, or otherwise does not necessarily constitute or imply its endorsement, recommendation, or favoring by the United States Government or any agency thereof, or Battelle Memorial Institute. The views and opinions of authors expressed herein do not necessarily state or reflect those of the United States Government or any agency thereof. PACIFIC NORTHWEST NATIONAL LABORATORYoperated byBATTELLEfor theUNITED STATES DEPARTMENT OF ENERGYunder Contract DE-AC05-76RL01830

Agriculture↗

Marginal Soils Index Analysis & Geospatial Data

This data package contains output files associated with the Mongird et al. paper entitled "Can US power grid expansion avoid prime agricultural lands?" and is organized into four dataset directories. Each dataset is described in more detail below. 1. Marginal Soils Index Analysis Description: This folder contains a csv file with land needs and availability by state, power generating technology type, and scenario in 2050 when suitable siting areas are additionally constrained to areas with increasing levels of soil marginality. Files: msi_constrained_siting_availability_2050.csv Variables: Scenario - Projected 2050 scenario name State - US state abbreviation Technology - Generating technology type solar = solar photovoltaic gas_cc_re = natural gas combined cycle (recirculating cooling) wind = onshore wind gas_cc_ccs_re = natural gas combined cycle with carbon capture sequestration (recirculating cooling) gas_cc_dry = natural gas combined cycle with (dry cooling) gas_cc_pond = natural gas combined cycle with (pond cooling) coal_conv_ccs_re = conventional coal with carbon capture sequestration (recirculating cooling) Req_Capacity_MW - The amount of rated capacity required in 2050 of the given technology type in the given state and under the given scenario from the capacity expansion plan Req_Capacity_Factor - The assumed capacity factor (fraction between 0 and 1) for the given technology type in the given state and under the given scenario by the capacity expansion plan Req_Land_km2 - The amount of land required (in km-squared) to host the required generating capacity that is capable of meeting the specified capacity factor for the given technology type in the given state and under the given scenario Req_Energy_TWh - Product of Req_Capacity_MW, Req_Capacity_Factor, and 8760/1e6 for the given technology type in the given state and under the given scenario MSI_Case - The level of MSI that siting the given technology is additionally constrained to, where >0 means siting is additionally constrained to suitable land areas that have an MSI value greater than 0 >=1 means siting is additionally constrained to suitable land areas that have an MSI value greater than or equal to 1 >=2 means siting is additionally constrained to suitable land areas that have an MSI value greater than or equal to 2 >=3 means siting is additionally constrained to suitable land areas that have an MSI value greater than or equal to 3 Soil Attribute Rasters Description: This folder contains geospatial raster files for individual soil parameters upscaled to the listed grid resolution (30m or 1 km). 1 km resolution files are a spatial average of non-missing 30m resolution values. All raster files use the USA Contiguous Albers Equal Area Conic (ESRI:102003) projection. Files: avg_cond_raster_ .tif - Average conductivity of the saturation extract across all soil horizons within a depth of 40 inches, measured in mmhos/cm max_cond_raster_ .tif - Maximum conductivity of the saturation extract across all soil horizons within a depth of 40 inches, measured in mmhos/cm min_ph_raster_ .tif- Min pH values across all soil horizons within a depth of 40 inches. avg_ph_raster_ .tif - Average pH value across all soil horizons within a depth of 40 inches. max_ph_raster_ .tif- Max pH value across all soil horizons within a depth of 40 inches. erosion_factor_raster_ .tif - Product of k-factor and percent slope flood_freq_raster_ .tif - Number of months of the year during which the area is commonly, frequently, or very frequently flooded. max_sar_raster_ .tif - Maximum sodium adsorption ratio across all horizons within a depth of 40 inches rock_frac_raster_ .tif - Fraction of the upper 6 inches of soil composed of rock fragments larger than 3 inches. temp_regime_raster_ .tif - Soil temperature regime with the following key: 0 = pergelic 1 = gelic 2 = cryic 3 = frigid 4 = isofrigid 5 = mesic 6 = isomesic 7 = thermic 8 = isothermic 9 = hyperthermic 10 =isohyperthermic Marginal Soils Index Rasters Description: This folder contains geospatial raster files of the Marginal Soils Index at the listed grid resolution (30m or 1 km). 1 km resolution files are a spatial average of 30m resolution. Both raster files use the USA Contiguous Albers Equal Area Conic (ESRI:102003) projection. A value of 0 indicates that there were no soil attributes present that indicate marginal soil. NA values indicate that data was unavailable or bodies of water. Files: marginal_soils_index_30m_raster.tif marginal_soils_index_1km_raster.tif Marginal Soils Index Resource Potential Rasters Description: This folder contains geospatial raster files of the Marginal Soils Index + Resource Potential (MSIxRP) score at 1km resolution for geothermal, solar, and wind technologies. Raster files use the USA Contiguous Albers Equal Area Conic (ESRI:102003) projection. NA values indicate that the location is not suitable for siting the given technology due to policy, environmental, socioeconomic, topological, and other constraints regardless of soil marginality level. Areas with values greater than or equal to zero represent the product of the normalized MSI value and the normalized resource potential value. Files: geothermal_msi_rp_score_raster.tif solar_msi_rp_score_raster.tif wind_msi_rp_score_raster.tif Acknowledgment IM3 is a multi-institutional effort led by Pacific Northwest National Laboratory and supported by the U.S. Department of Energy's Office of Science as part of research in MultiSector Dynamics, Earth and Environmental Systems Modeling Program. Disclaimer This material was prepared as an account of work sponsored by an agency of the United States Government. Neither the United States Government nor the United States Department of Energy, nor the Contractor, nor any or their employees, nor any jurisdiction or organization that has cooperated in the development of these materials, makes any warranty, express or implied, or assumes any legal liability or responsibility for the accuracy, completeness, or usefulness or any information, apparatus, product, software, or process disclosed, or represents that its use would not infringe privately owned rights. Reference herein to any specific commercial product, process, or service by trade name, trademark, manufacturer, or otherwise does not necessarily constitute or imply its endorsement, recommendation, or favoring by the United States Government or any agency thereof, or Battelle Memorial Institute. The views and opinions of authors expressed herein do not necessarily state or reflect those of the United States Government or any agency thereof. PACIFIC NORTHWEST NATIONAL LABORATORYoperated byBATTELLEfor theUNITED STATES DEPARTMENT OF ENERGYunder Contract DE-AC05-76RL01830

Agriculture↗

Integrated GW Farm ABM

This Data Repository includes data used for the integrated groundwater- farm ABM model, raw model output from scenario ensemble, and processed outputs that isolate the groundwater storage depletion outcomes for the 35,000 farm cells. Model Inputs: Farm ABM Inputs: This folder contains the input data used by the integrated groundwater - farm ABM modelling script (Python file) used for the high performance computing (HPC) experiments. The sub-folder "data inputs" contains all of the farm attribute data, while the three files in the folder have the hydrogeological data lookup table (NLDAS Cost Curve Attributes.csv), a lookup table (Theis well function table.csv) for the groundwater cost curve function, and the farm indexes and corresponding NLDAS ids for all of the cells run in this experiment (nldas farms subset final.csv). NLDAS Cost curve hydrogeological data: Hydrogeological data aggregated to 1/8 degree resolution and aligned with the NLDAS grid. Parameters include: water depth below ground surface [meters], subsurface porosity [unitless], aquifer depth from ground surface to aquifer bottom [meters], annual average recharge (USGS: mm, Doll: meters), and three different hydraulic conductivity (K) values (meters/day). The three K values represent the mean value from Gleeson et al. (2018), one standard deviation above the mean from Gleeson et al. (2018), and the de Graaf et al. 2020 modifications to certain lithologies. Additional information about these datasets and their processing are documented in the supplement to Yoon et al. 2025 (in review). Output: Raw outputs: This folder contains a .zip file that has model outputs for the entire scenario ensemble. There is one csv for each farm id, using the format "farm farmid cases.csv". The relationship between the farm id and NLDAS id is defined by the "nldas farms subset final.csv" located in the Farm ABM Inputs folder. Each csv has 625 rows, corresponding to 625 combinations of different scenario parameter values. Each row (scenario) represents the outcome of a 100 year simulation. Columns define scenario settings and summary statistics for each scenario. The first four columns define the scenario settings: "hydro ratio," "econ ratio," "K scenario," and "gamma scenario." The hydro and econ ratios are values passed to the modeling script that influence multipliers for other model parameters, as documented in the supplement to Yoon et al. 2025 (in review). The gamma multiplier is a coefficient multiplier applied to the baseline gamma values (values below 1 represent lower unobserved costs compared to baseline, values above 1 represent higher costs). The K scenario names represent K values of: "low": 0.5 m/d, "int 1": 2.5 m/d, "int 2": 10 m/d, "high": 50 m/d, and "gleeson": mean Gleeson K value. "Perc vol depleted" is the fraction of groundwater depleted at the end of the 100 simulation. Processed Output: Derived depletion outcomes from raw outputs: All of the individual csv files from the Raw outputs were aggregated into a single file that has the scenario settings and fraction depletion "Perc vol depleted" for every farm cell, for every scenario. The other two files define relationships between the farm id, NLDAS id, and local and major aquifer units, used for aquifer-level depletion analysis.

Agent based modeling↗

Package Data for CERF-Data Centers

This dataset contains sample input 100m resolution raster files for running the CERF-DC python package (see https://github.com/IMMM-SFA/cerf_data_centers) at the state level across the CONUS. Due to data availability constraints, some of the items included in this dataset are proxies or assumptions for siting factors used in the model. These are individually noted in the item descriptions and can be exchanged with more detailed information upon availability. Data Descriptions The following raster files are included in the data download: state_siting_region.tif — State areas identified by state FIPS code composite_siting_suitability.tif — Value of 1 indicates suitable siting location, 0 otherwise. The following areas are excluded from siting: Areas within 300m of a federal airport runway Waterbodies Areas with slope >16% Areas susceptible to sinkholes High coastal or inland flood risk areas Local, state, and federal parks, leisure areas, and cemeteries Areas >2 km away from electric substations Areas >5 km away from a municipal water supplier service area Areas >2 km away from high-speed fiber provider service territory Protected Areas Database of the United States (PAD-US) areas Railroads, major roadways, and minor roadways Military areas and training grounds Developed lands Areas >0.8 km (0.5 miles) from developed lands land_value_dollar_per_sqft.tif — USD per square foot (sqft) derived from USDA $/acre land cost personal_property_tax_rate.tif — Personal property tax rate by state. Uses an assumed 0.0125 personal property tax rate for states with personal property tax, 0 for states without personal property tax. real_property_tax_rate.tif — Real property tax rate. Based on county level residential real estate property tax rates. sales_tax_rate.tif — Sales tax rate by state. mechanical_cooling_fraction.tif — Fraction of year (values between 0 and 1, inclusive) that the data center would be cooled through mechanical processes based on local water stress and humidity levels. water_cooling_fraction.tif — Fraction of year (values between 0 and 1, inclusive) that the data center would be cooled through evaporative (water cooled) processes based on local water stress and humidity levels. distance_to_substation.tif — Distance to nearest substation in hundreds of meters (i.e., value of 1 equals a distance of 100m). Offshore areas have a value of 0. industrial_electricity_rates_dollar_per_kwh.tif — USD/kWh industrial electricity rates. Represents the average industrial rate across all utilities that operate within a given county. Values are derived from the US Utility Rate Database. commercial_electricity_rates_dollar_per_kwh.tif — USD/kWh commercial electricity rates. Represents the average commercial rate across all utilities that operate within a given county. Values are derived from the US Utility Rate Database. data_center_market_locations.tif — Grid cells with positive values represent the centroid of existing data center market clusters. The value of non-zero grid cells represents the number of data centers in the market cluster. All other grid cells have a value of 0. Geospatial Metadata CRS: Albers Equal Area Conic (ESRI:102003) Extent: -2415585.0000000023283064,-1441981.2605773280374706 : 2384414.9999999976716936,1708018.7394226719625294 Dimensions: X: 48000 Y: 31500 Bands: 1 Origin: -2415585.0000000023283064,1708018.7394226719625294 Pixel Size: 100,-100 Acknowledgment IM3 is a multi-institutional effort led by Pacific Northwest National Laboratory and supported by the U.S. Department of Energy's Office of Science as part of research in MultiSector Dynamics, Earth and Environmental Systems Modeling Program. License This data is made available under a CCBY4.0 License Disclaimer This material was prepared as an account of work sponsored by an agency of the United States Government. Neither the United States Government nor the United States Department of Energy, nor the Contractor, nor any or their employees, nor any jurisdiction or organization that has cooperated in the development of these materials, makes any warranty, express or implied, or assumes any legal liability or responsibility for the accuracy, completeness, or usefulness or any information, apparatus, product, software, or process disclosed, or represents that its use would not infringe privately owned rights. Reference herein to any specific commercial product, process, or service by trade name, trademark, manufacturer, or otherwise does not necessarily constitute or imply its endorsement, recommendation, or favoring by the United States Government or any agency thereof, or Battelle Memorial Institute. The views and opinions of authors expressed herein do not necessarily state or reflect those of the United States Government or any agency thereof. PACIFIC NORTHWEST NATIONAL LABORATORYoperated byBATTELLEfor theUNITED STATES DEPARTMENT OF ENERGYunder Contract DE-AC05-76RL01830

Mongird, Kendall↗

Package Data for CERF-Data Centers

This dataset contains sample input 100m resolution raster files for running the CERF-DC python package (see https://github.com/IMMM-SFA/cerf_data_centers) at the state level across the CONUS. Due to data availability constraints, some of the items included in this dataset are proxies or assumptions for siting factors used in the model. These are individually noted in the item descriptions and can be exchanged with more detailed information upon availability. Data Descriptions The following raster files are included in the data download: state_siting_region.tif — State areas identified by state FIPS code composite_siting_suitability.tif — Value of 1 indicates suitable siting location, 0 otherwise. The following areas are excluded from siting: Areas within 300 m of a federal airport runway or within an airport area boundary Waterbodies Areas with slope >16% Areas susceptible to sinkholes High coastal or inland flood risk areas Local, state, and federal parks, leisure areas, and cemeteries Areas >2 km away from electric substations Areas >5 km away from a municipal water supplier service area Areas >2 km away from high-speed fiber provider service territory USGS Protected Areas Database of the United States (PAD-US) GAP status 1, 2, or 3 areas US National Parks Wetlands USFWS critical habitats BIA land areas Railroads, major roadways, and minor roadways Military areas and training grounds NLCD developed lands Areas >0.8 km (0.5 miles) from NLCD developed lands land_value_dollar_per_sqft.tif — USD per square foot (sqft) derived from USDA $/acre land cost personal_property_tax_rate.tif — Personal property tax rate by state. Uses an assumed 0.0125 personal property tax rate for states with personal property tax, 0 for states without personal property tax. real_property_tax_rate.tif — Real property tax rate. Based on county level residential real estate property tax rates. sales_tax_rate.tif — Sales tax rate by state. mechanical_cooling_fraction.tif — Fraction of year (values between 0 and 1, inclusive) that the data center would be cooled through mechanical processes based on local water stress and humidity levels. water_cooling_fraction.tif — Fraction of year (values between 0 and 1, inclusive) that the data center would be cooled through evaporative (water cooled) processes based on local water stress and humidity levels. distance_to_substation.tif — Distance to nearest substation in hundreds of meters (i.e., value of 1 equals a distance of 100m). Offshore areas have a value of 0. industrial_electricity_rates_dollar_per_kwh.tif — USD/kWh industrial electricity rates. Represents the average industrial rate across all utilities that operate within a given county. Values are derived from the US Utility Rate Database. commercial_electricity_rates_dollar_per_kwh.tif — USD/kWh commercial electricity rates. Represents the average commercial rate across all utilities that operate within a given county. Values are derived from the US Utility Rate Database. data_center_market_locations.tif — Grid cells with positive values represent the centroid of existing data center market clusters. The value of non-zero grid cells represents the number of data centers in the market cluster. All other grid cells have a value of 0. Geospatial Metadata CRS: Albers Equal Area Conic (ESRI:102003) Extent: -2415585.0000000023283064,-1441981.2605773280374706 : 2384414.9999999976716936,1708018.7394226719625294 Dimensions: X: 48000 Y: 31500 Bands: 1 Origin: -2415585.0000000023283064,1708018.7394226719625294 Pixel Size: 100,-100 Acknowledgment IM3 is a multi-institutional effort led by Pacific Northwest National Laboratory and supported by the U.S. Department of Energy's Office of Science as part of research in MultiSector Dynamics, Earth and Environmental Systems Modeling Program. License This data is made available under a CCBY4.0 License Disclaimer This material was prepared as an account of work sponsored by an agency of the United States Government. Neither the United States Government nor the United States Department of Energy, nor the Contractor, nor any or their employees, nor any jurisdiction or organization that has cooperated in the development of these materials, makes any warranty, express or implied, or assumes any legal liability or responsibility for the accuracy, completeness, or usefulness or any information, apparatus, product, software, or process disclosed, or represents that its use would not infringe privately owned rights. Reference herein to any specific commercial product, process, or service by trade name, trademark, manufacturer, or otherwise does not necessarily constitute or imply its endorsement, recommendation, or favoring by the United States Government or any agency thereof, or Battelle Memorial Institute. The views and opinions of authors expressed herein do not necessarily state or reflect those of the United States Government or any agency thereof. PACIFIC NORTHWEST NATIONAL LABORATORYoperated byBATTELLEfor theUNITED STATES DEPARTMENT OF ENERGYunder Contract DE-AC05-76RL01830

Mongird, Kendall↗

129 I, 99 Tc, and U Distribution Coefficients of Subsurface Sediments Collected from the Proposed Site of the Environmental Management Disposal Facility

Performance Assessment calculations were completed in 2020 to evaluate the Environmental Management Disposal Facility (EMDF), a proposed new low-level radioactive waste (LLW) disposal facility on the U.S. Department of Energy’s Oak Ridge Reservation (ORR). Among the large number of input parameters needed for such calculations, are distribution coefficients (K d values; radionuclide concentration solid: liquid ratio) that provide a measure of the tendency of radionuclides to bind to sediments. The objective of this study was to measure K d values of three radionuclides that may pose a disproportionately large amount of risk, U, iodine-129 ( 129 I) and technetium-99 ( 99 Tc). The average 129 I K d value for the 14 geological materials recovered from the proposed EMDF site was 37.8 mL/g and ranged from 0.45 to 140.9 mL/g. These values were consistent, but somewhat larger than previous measurements made with ORR sediments and were about an order of magnitude greater than those used in previous EMDF PA calculations. The median 99 Tc K d value was 365.7 mL/g, much greater than previously reported using ORR geological materials. Five of the 14 tested geological materials sorbed large quantities of 99 Tc, suggesting that the weakly sorbing 99 Tc(VII) species had been reduced to the sparingly soluble 99 Tc(IV) species. The five strongly sorbing sediments had apparent 99 Tc solubility values of approximately <10 -8 mol/L. The median U K d value was 5,726 mL/g. All of the tested geological materials had large K d values, ranging from 625 to >10,208 mL/g. Among the sediment samples that exhibited strong U binding, the apparent solubility value was approximately <10 -9 mol/L. Based on sediment properties and general ORR geological considerations, it was proposed that much of the 129 I and 99 Tc retention could be attributed to the site materials exhibiting low pH (average pH = 4.94), and/or the elevated levels of iron oxides, manganese oxides, and natural organic matter. Similarly, the extremely high U binding measured in these sediments may also be attributed to the low conditions of carbonates, which can complex and therefore solubilize uranyl in these tests due to the low pH, and also the relatively high concentrations of iron and organic coatings on these samples. An implication of this study is that the areas of the EMDF subsurface environment may have natural properties for attenuating 129 I, 99 Tc, and U movement, and potentially other radionuclides, thereby possibly reducing risk posed by burial of LLW at this site.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Soil Water Retention and Hydraulic Conductivity Data and Model at Pump House in East River Watershed, Colorado 2019-2024

This data package includes soil water retention and hydraulic conductivity data and model fitting results from measurements of ex-situ soil samples and in-situ soil sensors near Pump House at Mount Crested Butte in the East River Watershed. Soil water retention curves (SWRC) characterize soil water content as a function of soil water potential. SWRC depends on soil texture and pore structure and can be used to describe the constraints on biogeochemical processes in terms of soil water availability. In this data package, the sample identification follows the format ER-X-Y, where ER refers to East River, X is the location identifier, and Y is the depth identifier at the same X (shallow Y=1). Specifically, ER-PHS, ER-LMC, ER-LMF, and ER-SMN are associated with ecohydrology sites under the East-Taylor Watershed Community Observatory Sites directory, and ER-RBTn (upslope n=1) are sampling transects during the 2019 Rootball Campaign. The sample and location information can be found in metadata.csv. Sampling and Measurements Each sample falls into one of the three sampling methods – (1) intact cores, (2) repacked samples, or (3) soil sensors – and one of the two measurement methods – (a) laboratory or (b) in-situ. Both intact cores and repacked samples were measured using the laboratory methods, which include measurements of soil water potential (HYPROP & WP4C, METER), saturated (KSAT, METER) and unsaturated hydraulic conductivity (HYPROP). The in-situ method uses a pair of co-located soil sensors to measure volumetric water content (TEROS12, METER) and soil water potential (TEROS21, METER), and the hydraulic conductivity was not measured. In comparison, the laboratory methods progress from full saturation to dry conditions, and the in-situ method includes both dry-to-wet and wet-to-dry cycles. The sampling and measurement methods for each sample can be found in metadata.csv, and more information about the measurements is detailed in the Methods section below. Models Retention and hydraulic conductivity data were fitted with four van-Genuchten-type models (specified by “model_name” column in the files): (1) traditional constrained van Genuchten model (“vG_constrained”), (2) traditional unconstrained van Genuchten model (“vG_unconstrained”), (3) PDI-variant of the constrained van Genuchten model (“vG_constrained_PDI”), and (4) PDI-variant of the unconstrained van Genuchten model (“vG_unconstrained_PDI”). The difference between the constrained (1: n) and the unconstrained (2: n, m) van Genuchten models is the number of pore-size distribution parameters in the model equations, giving the unconstrained model more degrees of freedom when fitting the data. Between the traditional and the PDI-variant models, model fitting differs the most at the dry end of the measurements. The traditional models allow infinite suction at the residual water content (water content does not drop below residual water content), and the PDI-variant models enforce a soil water potential value of pF=6.8 (~ -630 MPa) at oven-dryness (water content reaches 0). The inclusion of the van-Genuchten-type models is due to their common application. If other retention models are required, users can access the data in data.csv for further data fitting. More information about the models can be found in the Methods section below. Fitting Tasks The model fitting can be categorized into three levels of tasks (specified by “fitting_task” column in the files). Level 1 (“fit_retention”) only includes retention data fitting (the only level available for the in-situ method). Level 2 (“fit_retention_conductivity”) includes both retention and hydraulic conductivity data fitting, and the saturated hydraulic conductivity (Ks, a parameter of the hydraulic conductivity functions) is fixed by the measurements from KSAT. Level 3 (“fit_retention_conductivity_Ks”) also includes both retention and hydraulic conductivity data fitting, but Ks is a fitted parameter without the constraints from KSAT measurements. Among the same retention models (e.g. vG_constrained models of the same sample), level 1 should produce the best retention data fitting. Level 2 should have the highest misfit of the retention and hydraulic conductivity data, because the retention and hydraulic conductivity functions share common model parameters, and the unsaturated hydraulic conductivity (HYPROP) data fitting is subject to Ks measured independently by KSAT. Level 3 should have mid-level misfits of the retention and hydraulic conductivity data. While level 3 fits the hydraulic conductivity data better than level 2, the fitted Ks value might be unreasonable due to the lack of constraints at the wet end of the measurements. General recommendation when using this data package: (1) Choice of sampling methods: Intact cores and in-situ soil sensors could be prioritized because these sampling methods are less destructive. While the repacked samples were packed to the target bulk density (estimated post-sampling, when sample volume was known), these samples had altered pore structures. Nevertheless, intact cores might suffer from sample gaps that would lead to overestimation of Ks (sample gaps can be inferred from the “soil_sample_volume” column in metadata.csv when the value is < 249). In-situ method also has higher uncertainty in characterizing the wet end of the SWRC because of sensor limitations and the difficulty in reaching full saturation under natural conditions. (2) Choice of fitting tasks: When only retention data is needed, level 1 (“fit_retention”) should be prioritized. When both retention and hydraulic conductivity data are needed, level 2 (“fit_retention_conductivity”) could be prioritized. (3) Choice of models: This could depend on what the downstream models call for. If no specific model is required, model misfit could be used as a ranking criterion. Model misfit values in terms of RMSE can be found in model_parameters.csv. The following files are included in this data package: (1) metadata.csv – This file includes the general information of each sample, including location (description, geocoordinates, elevation), sampling and measurements details (method, depth, time or period, volume, instruments), and soil physical properties (bulk density, saturated hydraulic conductivity, only applicable to physical soil samples). (2) data.csv – This file includes soil water potential, volumetric water content, and unsaturated hydraulic conductivity data of each sample. Column “instrument” specifies the instrument (HYPROP, WP4C, or TEROS) used to perform the measurements. (3) model_fit.csv – This file includes soil water potential, volumetric water content, and unsaturated hydraulic conductivity fitted from the four models and three fitting tasks. Column “model_name” specifies the retention model used, and “fitting_task” specifies the level of data fitting. Missing values indicate that the variable does not apply to that fitting task. (4) model_parameters.csv – This file includes the fitted model parameters, model misfits, and conventional water content thresholds (field capacity and wilting point) from the four models and three fitting tasks. Column “model_name” specifies the retention model used, and “fitting_task” specifies the level of data fitting. Missing values indicate that the parameter does not apply to that model and/or that fitting task. (5) data_Ks.csv – This file includes the saturated hydraulic conductivity measurements from KSAT. (6) /figure/*.png – This folder includes three quick visualizations of the data, retention model fitting results and misfits, and hydraulic conductivity model fitting results, misfits, and parameters. The model fitting results are separated by samples and fitting tasks and colored by models. Zoom-in required. (7) /hyprop/*.bdhx – This folder includes proprietary hyprop files that require the free Labros SoilView-Analysis (METER) to open. Users can explore data fitting using other retention models (i.e. Brooks-Corey, Fredlund-Xing, Kosugi, bimodal models). Be aware that Ks value is pre-entered under “Fitting tab, Conductivity functions parameters” for level 2 fitting. If the value is lost, please refer to metadata.csv under “Ks” column. (8) Six file-level metadata that summarize file, header, column, and variable information of all files. This work was supported by the Watershed Function Science Focus Area at Lawrence Berkeley National Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-05CH11231.

EARTH SCIENCE > LAND SURFACE > SOILS↗