Search NASA⌕ Search

SEARCH · Search NASA

Results for “River corridor model”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

PFLOTRAN modeling data and scripts associated with “Refining the Hydrogeologic Framework of a Large River Corridor Model Using Waterborne Transient Electromagnetics”

NOTE: The manuscript associated with this data package is currently in review. The data may be revised based on reviewer feedback. Upon manuscript acceptance, this data package will be updated with the final dataset and additional metadata. This data package is associated with the publication “Refining the Hydrogeologic Framework of a Large River Corridor Model Using Waterborne Transient Electromagnetics” submitted to Water Resources Research (Terry et al. 2025). The data package contains the groundwater modeling dataset from PFLOTRAN software. It includes the python script for mesh generation, boundary condition setting, PFLOTRAN input deck formation and postprocessing. It couples groundwater flow and species transport for Hanford Reach river corridor and pipelines the model generation and processing. This model can be used to easily generate the model and analysis for Hanford site. It can also be adjusted to other hydrologic area with ease. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. The data package consists of 6 folders: (1) “data” contains all necessary data as input and intermediate data for processing; (2) “mesh” contains all mesh related files to generate mesh in Hanford Reach river corridor; (3) “model_run” contains the generated script for PFLOTRAN modeling; (4) “notebooks” contains all the Python script to generate the model; (5) “output” contains all the output from the computation; (6) “postprocessing” contains the Python script to generate scientific figure for manuscript. All files are .csv (comma-separated values), .h5 (HDF5 format), .in (input files), .ipynb (Jupyter notebooks), .p (Python pickle), .png (images), .PNG (images), .py (Python scripts), .pyc (Python bytecode), .r (R scripts), .sh (shell scripts), .txt (text files), .vtu (3D mesh/visualization format), .xz (compressed archive), or .zip (compressed archive).

54 ENVIRONMENTAL SCIENCES↗

Model Inputs, Outputs, and Scripts associated with: “Combined effects of stream hydrology and land use on basin-scale hyporheic zone denitrification in the Columbia River Basin”

This data package is associated with the publication “Combined effects of stream hydrology and land use on basin‐scale hyporheic zone denitrification in the Columbia River Basin”, published in Water Resource Research (Son et al.2022) available at https://doi.org/10.1029/2021WR031131. This data package includes the key model inputs/outputs of the river corridor model for the Columbia River Basin (CRB) and the model source codes used in the manuscript. The model is a carbon-nitrogen-coupled river corridor model (RCM), and the model is used to quantify hyporheic zone (HZ) denitrification at the NHDPLUS stream reach scales. The RCM used in this study combines empirical substrate models derived from observations and three microbially driven reactions, including two-step denitrification and aerobic respiration, are considered within the HZ. The key input data of the model are exchange flux, residence time, and stream solute (dissolved organic carbon (DOC), dissolved oxygen (DO), and nitrate concentrations). These inputs are constant over time and represent long-term averaged values. This study uses the RCM to explore the spatial patterns of HZ denitrification across reaches with different sizes and land use in the CRB. Our main objective is to use the RCM as a virtual reality model, and the machine-learning models as surrogates that encapsulate the complexities of the physics-based model while identifying the importance of different variables that are not evident in the model conceptualization. We do not include a direct comparison of the modeled HZ denitrification and measurements; however, the RCM can capture the overall spatial patterns of the HZ denitrification because the model inputs and its reaction networks are based on well-established theory and a physical-based model. The combination of the model-based predictions and a machine-learning approach (e.g., random forest) is used to improve our understanding of what variables of the model are associated with spatial patterns of the modeled denitrification across reaches with different sizes and land uses, and to develop a proxy model using measurable variables to reproduce the simulated patterns.This dataset contains five folders: (1) model_inputs, (2) model_outputs, (3) Rscripts, (4) figures, and (5) model_codes. It also contains a readme, file level metadata (FLMD), and data dictionary (dd). Please see the FLMD for a list of all the files contained in this data package and descriptions for each. The model_inputs folder contains the model inputs used to drive the model simulations. The model_outputs folder contains key model output files from the river corridor model. The Rscripts folder contains the Rscripts for pre- and post- processing model results. The figures folder contains the raw figures associated with the manuscript. The model_codes folder includes key model source codes/input files. All files are .jpg, .jpeg, .out, .e, .od, .dat, .sub, .F90, .0, .R, .sbx, .cpg, .sbn, .shx, .shp, .dbf, .prj, .tfw, .tif, .xml, .pdf, or .csv.

54 ENVIRONMENTAL SCIENCES↗

Model Inputs, Outputs, and Scripts associated with: “Spatial microbial respiration variations in the hyporheic zones within the Columbia River Basin”

This data package is associated with the publication “Spatial microbial respiration variations in the hyporheic zones within the Columbia River Basin” published in the Journal of Geophysical Research: Biogeosciences (Son et al. 2022) available at doi: 10.1029/2021JG006654. This data package includes the key model inputs/outputs of the river corridor model for the Columbia River Basin (CRB) and the model source codes, which were used in the manuscript. The model is a carbon-nitrogen-coupled river corridor model (RCM), and the model is used to quantify hyporheic zone (HZ) aerobic and anaerobic respiration at the NHDPLUS stream reach scales. The RCM used in this study combines empirical substrate models derived from observations and three microbially driven reactions to compute respiration of the HZ for each National Hydrography Dataset (NHD) reach within the CRB. The reactions in HZs of each NHD reach include anaerobic respiration and two-step anaerobic respiration via denitrification. Our HZ respiration estimates are limited to the lotic (or flowing) stream/river systems, and do not account for the respiration process in water column. Note that the RCM only simulates the HZ’s contribution to the dissolved carbon dioxide (CO2) concentrations in the streams, and the CO2 emissions to the atmosphere are not modelled. The model computes at hourly timesteps because of the fast reaction rates. The key input data of the model are exchange flux, residence time, and stream solute (dissolved organic carbon (DOC), dissolved oxygen (DO), and nitrate concentrations). These inputs are constant over time and represent long-term averaged values.This modeling framework successfully quantified HZ respiration components over multiple scales. It revealed key mechanisms driving the spatial variation of HZ aerobic and anaerobic respiration in reaches with varying hydrologic and substrate conditions. Thus, this modeling study offers a testing hypothesis in different river system (e.g., climate and biomes) for the HZ respiration processes, and can be used as a sampling design tool for large-scale HZ experimental studies.This dataset contains five folders: (1) model_inputs, (2) model_outputs, (3) Rscripts, (4) figures, and (5) model_codes. It also contains a readme, file level metadata (FLMD), and data dictionary (dd). Please see the FLMD for a list of all the files contained in this data package and descriptions for each. The model_inputs folder contains the model inputs used to drive the model simulations. The model_outputs folder contains key model output files from the river corridor model. The Rscripts folder contains the Rscripts for pre- and post- processing model results. The figures folder contains the raw figures associated with the manuscript. The model_codes folder includes key model source codes/input files. All files are .jpg, .jpeg, .out, .e, .od, .dat, .sub, .F90, .0, .R, .sbx, .cpg, .sbn, .shx, .shp, .dbf, .prj, .tfw, .tif, .xml, .pdf, or .csv.

54 ENVIRONMENTAL SCIENCES↗

Combined Effects of Stream Hydrology and Land Use on Basin‐Scale Hyporheic Zone Denitrification in the Columbia River Basin

Abstract Denitrification in the hyporheic zone (HZ) of river corridors is crucial to removing excess nitrogen in rivers from anthropogenic activities. However, previous modeling studies of the effectiveness of river corridors in removing excess nitrogen via denitrification were often limited to the reach‐scale and low‐order stream watersheds. We developed a basin‐scale river corridor model for the Columbia River Basin with random forest models to identify the dominant factors associated with the spatial variation of HZ denitrification. Our modeling results suggest that the combined effects of hydrologic variability in reaches and substrate availability influenced by land use are associated with the spatial variability of modeled HZ denitrification at the basin scale. Hyporheic exchange flux can explain most of spatial variation of denitrification amounts in reaches of different sizes, while among the reaches affected by different land uses, the combination of hyporheic exchange flux and stream dissolved organic carbon (DOC) concentration can explain the denitrification differences. Also, we can generalize that the most influential watershed and channel variables controlling denitrification variation are channel morphology parameters (median grain size (D50), stream slope), climate (annual precipitation and evapotranspiration), and stream DOC‐related parameters (percent of shrub area). The modeling framework in our study can serve as a valuable tool to identify the limiting factors in removing excess nitrogen pollution in large river basins where direct measurement is often infeasible.

54 ENVIRONMENTAL SCIENCES↗

Spatial Study 2022: Water Column, Sediment, and Total Ecosystem Respiration Rates across the Yakima River Basin, Washington, USA (v2)

This dataset supports a broader study examining the drivers of spatial variability in sediment respiration rates in the Yakima River Basin and is associated with the manuscript “Sediment-associated processes account for most of the spatial variation in ecosystem respiration in the Yakima River basin” submitted to Nature Communications Earth & Environment (Garayburu-Caruso et al., in review). The dataset provides ecosystem metabolism estimates generated from streamMetabolizer (Appling et al.; 2018) using data collected during the same five-week period at 48 sites within multiple rivers throughout the Yakima River Basin in Washington, USA. Additionally, it includes the scripts used for the analysis and producing the figures in the manuscript. The contents include streamMetabolizer inputs and outputs and additional relevant data needed to generate the main manuscript results. The data included are: total ecosystem respiration, water respiration, calculated sediment-associated respiration, gross primary production outputs from the river corridor model for the Yakima River Basin, median grain size (d50), depth, dissolved oxygen, water temperature, pressure, and annual oxygen consumption. The associated GitHub repository can be found at https://github.com/river-corridors-sfa/SSS_metabolism. Samples collected during this study were labeled as “Second Spatial Study” or “SSS.” Raw time series sensor data, total suspended solids, and depth data from SSS were published at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/1969566. A subset of data from the SSS samples were published in the contiguous United States (CONUS)-Scale Model-Sample (CM) study data package available at https://data.ess-dive.lbl.gov/view/doi:10.15485/1923689 that presents data from across the CONUS. They include dissolved organic carbon (DOC, measured as non-purgeable organic carbon, NPOC), total nitrogen (TN), grain size, aerobic sediment respiration, dissolved oxygen (DO), and temperature. Parent IDs and Site IDs are consistent between the SSS and CM data packages, and they can be mapped directly so data across packages can be used together. Field metadata for the samples in this da This dataset is comprised of one main data folder with four subfolders. The main data folder contains of (1) file-level metadata; (2) data dictionary; (3) total/water column/sediment respiration; (4) gross primary production (GPP); (5) median grain size (d50); and (6) annual oxygen consumption. The “Figures” subfolder contains the figures used in the paper and all intermediate files (including geospatial files). The “Published_Data” contains a readme directing the user to download the public data to reproduce analyses and figures. The “Scripts” folder contains all scripts used in the analyses that were not part of running StreamMetabolizer. Lastly, the “Stream_Metabolizer” folder contains all files associated with running StreamMetabolizer including (1) model input files, (2) model output files, (3) processing scripts, (4) histogram plots of the outputs, and (5) an R project. All files are .csv, .pdf, .R, .Rmd, .Rproj, .html, .png, .txt, .qgz, .cpg, .dbf, .prj, .shp, .shp.ea.iso.xml, .shp.iso.xml, .shx, .sbn. ta package can be found at either link. We acknowledge the Yakama Nation as owners and caretakers of the lands where we collected these data. We thank the Confederated Tribes and Bands of the Yakama Nation Tribal Council and Yakama Nation Fisheries for working with us to facilitate sample collection and optimization of data usage according to their values and worldview.

54 ENVIRONMENTAL SCIENCES↗

Spatial Microbial Respiration Variations in the Hyporheic Zones Within the Columbia River Basin

Abstract While the hyporheic zone (HZ) accounts for a significant portion of whole stream CO 2 concentrations, HZ respiration modeling studies are lacking in quantifying their contributions to the total CO 2 at large watershed/basin scales. Quantifying the contribution of anaerobic respiration is also underappreciated. This study used a carbon‐nitrogen‐coupled river corridor model to quantify HZ aerobic and anaerobic respiration and determined the key factors controlling their spatial variability within the Columbia River Basin (CRB). The modeled respiration patterns showed high spatial variability. Among the nine sub‐basins composing the CRB, the Lower Columbia and the Willamette, which receive higher precipitation, had higher respiration. Medium‐sized rivers (fourth to sixth orders) produced the highest aerobic and anaerobic respiration among reaches of different sizes. At the basin scale, aerobic respiration is dominant, representing approximately 98.7% of the total respiration across the CRB. While most of the reaches were dominant with aerobic respiration, reaches in agricultural land showed a relatively higher anaerobic respiration (18%) ratio. A variable importance analysis showed that hyporheic exchange flux controlled most of the spatial variability of HZ respiration, dominating over other physical variables such as residence time, stream dissolved organic carbon (DOC), nitrate, and dissolved oxygen (DO). The influence of substrate concentration (DOC and DO) is larger in modeling anaerobic respiration than aerobic respiration. Future efforts will focus on improving the estimation of the HZ exchange flux and the implementation of spatially explicit parameterizations for the reactions of interest to reduce model uncertainty.

54 ENVIRONMENTAL SCIENCES↗

Data and scripts associated with “Allometric scaling of hyporheic respiration across basins in the Pacific Northwest USA"

This data package is associated with the publication “Allometric scaling of hyporheic respiration across basins in the Pacific Northwest USA” submitted to JGR-Biogeosciences (Regier et al. 2025).This study used reach-scale modeled estimates of hyporheic aerobic respiration made by the River Corridor Model (Fang et al. 2020) and watershed characteristics across the Willamette and Yakima River basins to explore potential allometric scaling (i.e., power-law relationships between size and function) of cumulative hyporheic respiration across catchment-to-basin scales. Scaling was explored quantitatively via the R2, slope, and y-intercept of relationships between cumulative hyporheic respiration and watershed area, divided into hyporheic exchange flux (HEF) quantiles. We also explored relationships between allometric scaling and other watershed characteristics through linear regression, spatial patterns, and mutual information analyses. Our results also suggest variability of hyporheic respiration allometry for middle exchange flux quantiles, and in relation to land-cover. Our findings provide initial evidence that allometric scaling may be useful for predicting hyporheic biogeochemical dynamics across watersheds from reach to basin scales. This data package is associated with the GitHub repository found at https://github.com/peterregier/rc_wrb_yrb_scaling. The data package is organized into several key directories. The “data” folder contains multiple CSV files, including landscape heterogeneity, scaling analysis, and watershed boundary data. The “figures” folder has all figure files in both PDF and PNG formats. Core analysis scripts and figure generation scripts are in the “scripts” directory, systematically numbered for sequential execution. The root directory includes essential project files; please see the file ending in “flmd.csv” for a list and description of all files contained in this data package and the file ending in “dd.csv” for data dictionaries used to describe tabular column headers.

54 ENVIRONMENTAL SCIENCES↗

Allometric Scaling of Hyporheic Respiration Across Basins in the Pacific Northwest United States

Abstract Hyporheic zones regulate biogeochemical processes in streams and rivers, but high spatiotemporal heterogeneity makes it difficult to predict how these processes scale from individual reaches to river basins. Recent work applying allometric scaling (i.e., power‐law relationships between size and function) to river networks provides a new paradigm for understanding cumulative hyporheic biogeochemical processes. We used previously published model predictions of reach‐scale hyporheic aerobic respiration to explore patterns in allometric scaling across two climatically divergent basins with differing characteristics in the Pacific Northwest, United States. In the model, hydrologic exchange fluxes (HEFs) regulate hyporheic respiration, so we examined how HEFs might influence allometric scaling of respiration. We found consistent scaling behaviors where HEFs were either very low or very high, but differences between basins when HEFs were moderate. Our findings provide initial model‐generated hypotheses for factors influencing allometric scaling of hyporheic respiration. These hypotheses can be used to optimize new data generation efforts aimed at developing predictive understanding of allometries that can, in turn, be used to scale biogeochemical dynamics across watersheds.

59 BASIC BIOLOGICAL SCIENCES↗

Numerical evaluation of photosensitive tracers as a strategy for separating surface and subsurface transient storage in streams

For this work, we numerically evaluated photosensitive tracers as a potential strategy for separating the effects of surface and hyporheic storage zones (SSZs and HSZs, respectively) on stream corridor transport. Correctly separating HSZ and SSZ effects is critical to estimating the hydro-biogeochemical function of a stream because HSZs and SSZs expose solutes to significantly different biogeochemical conditions, like sunlight exposure, microbial processes, and oxygen availability. Our numerical experiments used a multiscale river-corridor transport model implemented in the ATS code, which accommodates multiple storage zones with distinct travel time distributions and biogeochemical reactions. For parameter inferences, we used Bayesian inverse modeling. We found that breakthrough curves for photo-decaying tracers from day and night injection can delineate surface and hyporheic transient storage contributions, but only when interpreted jointly through a two-storage zone model. Numerical experiments that used only daytime injection or interpreted breakthrough curves with a single storage zone model yielded good fit to breakthrough curves, but parameter estimates were biased and controlling processes misattributed, examples of good model fits for the wrong reasons. Using those biased parameter estimates in reactive transport simulations resulted in significantly different projections of denitrification, which underscores the potential for stream function to be mischaracterized if tracer tests are interpreted through an inappropriately simplified model for transient storage. More generally, this study highlights the role of modeling in evaluating the experimental design and identifying the potential of system mischaracterization and its implications.

54 ENVIRONMENTAL SCIENCES↗

WHONDRS Surface Water and Sediment Geochemistry and Organic Matter Characterization Data from Streams across HJ Andrews Experimental Forest, Oregon (v2)

This dataset supports a broader study developing conceptual models for river corridor critical zone processes across spatial scales and was generated in collaboration with the HJ Andrews River Corridor Critical Zone Workshop in 2025. The dataset provides surface water geochemistry (dissolved organic carbon, total dissolved nitrogen) from 48 sites across the HJ Andrews Experimental Forest, Oregon (https://andrewsforest.oregonstate.edu). Some of the sites have been impacted by the Holiday Farm Fire and the Lookout Fire in 2020 and 2023, respectively. Related data were collected as part of the workshop and will be published separately in collaboration with other workshop attendees and available at http://www.hydroshare.org/resource/b274c4a234bf4b12b7cb8a54a696c629. Related genomic data can be found on the National Center for Biotechnology Information (NCBI) under BioProject PRJNA1503030 (see critical details section below for more information). Additional related data collected in 2016 from a similar effort can be found at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/3377027 and http://www.hydroshare.org/resource/ea6c0832885a46c3939e7bb22e48e754 and are described within https://doi.org/10.5194/essd-11-1-2019 (Ward et al., 2019). This data package was originally published in March 2026. It was updated in August 2026 (v2; new and modified files). See the change history section in the readme for more details. For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. In addition to a readme, this data package also includes a file-level metadata (FLMD) file that describes each file and a data dictionary (DD) that describes all column/row headers and variable definitions. This dataset is comprised of (1) a folder of field photos, (2) a folder of raw Fourier transform ion cyclotron resonance mass spectrometry (FTICR-MS) data, (3) a data checks report, (4) a folder of sample data, (5) file-level metadata, (6) data dictionary, (7) field metadata, (8) readme, (9) international generic sample number (IGSN) mapping file; and (10) field protocol. The sample data subfolder contains surface water and sediment (1) dissolved organic carbon (DOC, measured as non-purgeable organic carbon, NPOC) data and averages, (2) total dissolved nitrogen data and averages, (3) methods codes, (4) FTICR-MS methods; and (5) a subfolder of 9.4 Tesla (9.4T) FTICR-MS data. This folder contains the CoreMS processed data and seven subfolders, thee containing .xml files for each sample type (sediment, surface water and blank samples), three containing the sediment CoreMS output files for each sample type (sediment, surface water and blank samples), and the other containing instructions and scripts for processing the files in CoreMS (https://github.com/EMSL-Computing/CoreMS). All files are .csv, .pdf, .R, .xml, .Rmd, .py, .cal, .json, .jpg, or .jpeg.

Biogeochemistry↗

On the transferability of residence time distributions in two 10-km long river sections with similar hydromorphic units

Quantifying hydrologic exchange fluxes (HEFs) at the stream-groundwater interface and their residence time distributions (RTDs) in the subsurface are important for managing the water quality and ecosystem health in dynamic river corridors. However, direct simulating high-spatial resolution HEFs and RTDs can be time-consuming, especially for watershed-scale modeling. Efficient surrogate models linking RTDs to hydromorphic units (HUs) can be alternatives for simulating RTDs in large-scale models. A common concern of these surrogate models, though, is the transferability of the relationship between the RTDs and HUs from one river corridor to another. To address this issue, this work evaluates the HEFs and resulting RTD-HU relationships for two 10-km long river corridors along the Columbia River leveraging a one-way coupled three-dimensional transient surface-subsurface water transport modeling framework we previously developed. Applying such a framework at the two river corridors with similar HUs allows for quantitative comparisons of HEFs and RTDs using both statistical tests and machine learning classification models. Finally, our comparison shows that the similarity and transferability of the RTD-HU relationship is very low for the two investigated river sections, which suggests that devising a general algorithm to estimate RTDs based solely on surface water hydrodynamics and short-distance river channel topography data, as well as HU classification, might be nearly impossible.

54 ENVIRONMENTAL SCIENCES↗

Quantifying Groundwater Response and Uncertainty in Beaver‐Influenced Mountainous Floodplains Using Machine Learning‐Based Model Calibration

Abstract Beavers ( Castor canadensis ) alter river corridor hydrology by creating ponds and inundating floodplains, and thereby improving surface water storage. However, the impact of inundation on groundwater, particularly in mountainous alluvial floodplains with permeable gravel/cobble layers overlain by a soil layer, remains uncertain. Numerical modeling across various floodplain structures considers topographic and sediment complexity and multidirectional flow, linking inundation to groundwater response. This study develops a model‐data integration workflow to address uncertainty in groundwater response to beaver‐induced inundations in a mountainous alluvial floodplain in the Upper Colorado River Basin. Uncertain factors include seasonal hydrologic dynamics, hydraulic conductivities, floodplain structures, and meteorological forcings. We employed an ensemble of groundwater models, based on geophysical and hydrologic data, with machine learning‐based calibration using a neural density estimator. This allowed us to quantify the vertical flux from the soil layer to the permeable gravel bed, the down‐valley underflow within the gravel bed, and their ratios. Results show a significant increase in the vertical flux relative to down‐valley underflow, from 2 during dry pond periods to 20 during wet periods, serving as an analogy for conditions without and with beaver ponds. The study highlights the influence of floodplain structure on groundwater storage, water balance, and water quality impacted by beaver ponds. A thick gravel bed layer, with a large down‐valley underflow, minimizes the effect of beaver‐induced inundation on water quality. We emphasize the need for field‐scale measurements of floodplain structure and improved characterization of evapotranspiration changes to reduce uncertainty in groundwater response. Plain Language Summary Beavers change the flow of water in river corridors by creating ponds, expanding wetlands, and flooding floodplains. This increases surface water area, promotes plant growth, and enhances biodiversity. However, the impact of this flooding on groundwater flow is not well understood, especially in mountainous areas with gravel layers where water moves easily beneath soil. In this study, we used numerical modeling to investigate how beaver ponds influence groundwater in a mountainous floodplain of the Upper Colorado River Basin. We adapted a machine learning method to validate our numerical models using multiple field data sets. Our findings show that beaver ponds significantly increase vertical water flow from the soil to the gravel during wet periods, compared to when the ponds are fully drained. The study also highlights the importance of floodplain structure in controlling both water flow in gravel layers along the river direction and vertical flow from the soil to the gravel with the presence of beavers. To reduce uncertainty in groundwater response, we emphasize the need for more field‐scale measurements of floodplain structure, hydraulic properties, and evapotranspiration changes. Key Points Floodplain structures and hydraulic conductivities are important for groundwater response with beaver ponds in mountainous floodplains Large down‐valley underflow in permeability‐stratified floodplains reduces beaver‐induced impacts on groundwater storage and water quality Machine learning‐based model calibration methods are effective for estimating posterior distributions of groundwater model parameters

Wang, Lijing↗

Data and scripts associated with a manuscript on residence time distribution simulation in two 10-kilometer long river sections

This data package is associated with the publication “On the Transferability of Residence Time Distributions in Two 10-km Long River Sections with Similar Hydromorphic Units” submitted to the Journal of Hydrology (Bao et al. 2024).Quantifying hydrologic exchange fluxes (HEFs) at the stream-groundwater interface, along with their residence time distributions (RTDs) in the subsurface, is crucial for managing water quality and ecosystem health in dynamic river corridors. However, directly simulating high-spatial resolution HEFs and RTDs can be a time-consuming process, particularly for watershed-scale modeling. Efficient surrogate models that link RTDs to hydromorphic units (HUs) may serve as alternatives for simulating RTDs in large-scale models. One common concern with these surrogate models, however, is the transferability of the relationship between the RTDs and HUs from one river corridor to another. To address this, we evaluated the HEFs and the resulting RTD-HU relationships for two 10-kilometer-long river corridors along the Columbia River, using a one-way coupled three-dimensional transient surface-subsurface water transport modeling framework that we previously developed. Applying this framework to the two river corridors with similar HUs allows for quantitative comparisons of HEFs and RTDs using both statistical tests and machine learning classification models. This data package includes the model inputs files and the simulation results data. This data package contains 10 folders. The modeling simulation results data are in the folders 100H_pt_data and 300area_pt_data, for the study domain Hanford 100H and 300 area respectively. The remaining eight folders contain the scripts and data to generate the manuscript figures. The file-level metadata file (Bao_2024_Residence_Time_Distribution _flmd.csv) includes a list of all files contained in this data package and descriptions for each. The data dictionary file (Bao_2024_Residence_Time_Distribution _dd.csv) includes column header definitions and units of all tabular files.

54 ENVIRONMENTAL SCIENCES↗

Exploring the determinants of organic matter bioavailability through substrate-explicit thermodynamic modeling

Microbial decomposition of organic matter (OM) in river corridors is a major driver of nutrient and energy cycles in natural ecosystems. Recent advances in omics technologies enabled high-throughput generation of molecular data that could be used to inform biogeochemical models. With ultrahigh-resolution OM data becoming more readily available, in particular, the substrate-explicit thermodynamic modeling (SXTM) has emerged as a promising approach due to its ability to predict OM degradation and respiration rates from chemical formulae of compounds. This model implicitly assumes that all detected organic compounds are bioavailable, and that aerobic respiration is driven solely by thermodynamics. Despite promising demonstrations in previous studies, these assumptions may not be universally valid because OM degradation is a complex process governed by multiple factors. To identify key drivers of OM respiration, we performed a comprehensive analysis of diverse river systems using Fourier-transform ion cyclotron resonance mass spectrometry OM data and associated respiration measurements collected by the Worldwide Hydrobiogeochemistry Observation Network for Dynamic River Systems (WHONDRS) consortium. In support of our argument, we found that the incorporation of all compounds detected in the samples into the SXTM resulted in a poor correlation between the predicted and measured respiration rates. The data-model consistency was significantly improved by the selective use of a small subset (i.e., only about 5%) of organic compounds identified using an optimization method. Through a subsequent comparative analysis of the subset of compounds (which we presume as bioavailable) against the full set of compounds, we identified three major traits that potentially determine OM bioavailability, including: (1) thermodynamic favorability of aerobic respiration, (2) the number of C atoms contained in compounds, and (2) carbon/nitrogen (C/N) ratio. We found that all three factors serve as “filters” in that the compounds with undesirable properties in any of these traits are strictly excluded from the bioavailable fraction. This work highlights the importance of accounting for the complex interplay among multiple key traits to increase the predictive power of biogeochemical and ecosystem models.

59 BASIC BIOLOGICAL SCIENCES↗

The importance of explicitly representing the streambed in watershed models

Abstract The streambed is the critical interface between the aquatic and terrestrial systems and hosts important biogeochemical hot spots within river corridors. Although the streambed characteristics are significantly different from those of its surrounding soil, the streambed itself has not been explicitly represented in watershed models. Here, we explicitly incorporated a streambed layer into an integrated hydrologic model through model parameterization and discretization. We examined the hydrological effects of streambed characteristics, including hydraulic conductivity ( K ), layer thickness, and resolution, on the exchange fluxes across the streambed as well as the streamflow at the watershed outlet. The numerical experiments were performed in the American River Watershed, a headwater, mountainous watershed within the Yakima River Basin in central Washington. Despite having a negligible effect on the watershed streamflow, an explicit representation of the streambed with distinctive properties dramatically changed the magnitude and variability of the exchange flux. In general, a larger streambed K along with a thicker streambed layer induced larger exchange fluxes. The exchange flux was most sensitive to the streambed resolution. A finer streambed resolution increased exchange fluxes per unit area while reducing the overall exchange volumes across the entire streambed. The amount of baseflow decreased by 6% as the streambed resolution increased from 250 to 50 m. This finding is important because these hydrological changes may, in turn, affect the exchange of nutrients and contaminants between surface water and groundwater and the associated biogeochemical processes. Our work demonstrated the importance of representing streambeds in fully distributed, process‐based watershed models to better capture the exchange flow dynamics in river corridors.

54 ENVIRONMENTAL SCIENCES↗

Data, model inputs, and analysis scripts associated with a manuscript on stream intermittency controls across spatial scales in Pacific Northwest watersheds

NOTE: The manuscript associated with this data package is currently in review. The data may be revised based on reviewer feedback. Upon manuscript acceptance, this data package will be updated with the final dataset and additional metadata. This data package is associated with the manuscript "Hydroclimatic Memory and Watershed Template Shape Stream Intermittency: Multi-scale Attribution Using Process-based Simulation and Explainable ML" by Niroula et al. (2026), submitted to Water Resources Research (WRR). The study investigates the dominant controls on stream intermittency across local, reach, and watershed scales using a coupled process-based simulation and explainable machine-learning framework. Long-term daily simulations from the Advanced Terrestrial Simulator (ATS) were used to generate wetness states and ponded-depth responses over river-corridor cells. These ATS outputs were then aggregated across scales and used to train XGBoost (eXtreme Gradient Boosting) models. SHAP (SHapley Additive exPlanations) was applied to quantify the relative importance of hydroclimatic forcings, watershed template attributes, and antecedent-memory effects in shaping intermittency behavior. The analysis is carried out for three contrasting Pacific Northwest watersheds: Oak Creek (OCW), American River Watershed (ARW), and H.J. Andrews (HJA). Across these testbeds, the package contains ATS-ready watershed inputs, ATS run configuration and selected output files, model-evaluation data products, intermittency-analysis datasets, machine-learning target-feature tables, SHAP outputs, and notebooks used to organize, analyze, and visualize results. At a high level, the package documents a workflow in which ATS provides the physically based simulation backbone and explainable machine learning is used as a post-processing attribution tool. The contents are intended to support interpretation of the manuscript figures and results, provide context for how intermittency metrics were generated at multiple scales, and preserve the key artifacts needed to understand and reuse the analysis workflow. The package contains a high-level directory summary file (`summary.txt`) and four main content folders (1) `evaluation_plots` contains evaluation figures and supporting evaluation datasets; (2) `intermittency_plots` contains intermittency-focused analysis notebook and prepared datasets; (3) `ml-training-and-shap_values_plots` contains ML training inputs, SHAP outputs, and figure-generation notebooks; and (4) `watershed_mesh_and_ats_input` contains ATS model setup materials, forcing inputs, geometry, and selected run files. More specifically, the `evaluation_plots` folder contains the notebook used for ATS evaluation plotting and site-specific evaluation datasets. These include evapotranspiration and water-balance products for three watersheds, as well as an Oak Creek field-measurement discharge file. The `intermittency_plots` folder contains the notebook used for intermittency analysis and the prepared datasets used to analyze intermittent and non-intermittent wetness behavior across the study watersheds. The `ml-training-and-shap_values_plots` folder contains notebooks and outputs for the machine-learning and explainability workflow. This includes the main XGBoost and SHAP notebook(s), a beeswarm plotting notebook, target-feature tables for machine-learning training, SHAP summary tables, and per-sample SHAP value archives. The `watershed_mesh_and_ats_input` folder contains ATS-related watershed inputs and supporting materials. This includes mesh and shape products, ATS-readable LAI and meteorological forcing inputs, selected ATS spinup and transient-run files, and a watershed workflow example notebook. Subdirectories are organized by watershed where applicable.All files are .cpg (codepage files), .csv (comma-separated values), .dbf (database files), .exo (Exodus mesh format), .h5 (HDF5 format), .ipynb (Jupyter notebooks), .pkl (Python pickle), .prj (projection files), .sh (shell scripts), .shp (shapefile geometry), .shx (shapefile index), .txt (text files), or .xml (markup data).

Advanced Terrestrial Simulator↗

Data and scripts associated with the manuscript "Organic Molecules are Deterministically Assembled in River Sediments"

This data package is associated with the publication "Organic Molecules are Deterministically Assembled in River Sediments" submitted to Scientific Reports (Stegen et al., 2024). The study applies community ecology methods to dissolved organic matter (DOM) chemistry from variably inundated riverbed sediments to uncover principles governing DOM composition at a reach-scale. This data package documents the workflow used to process and generate the main findings in the manuscript. The R scripts reference the raw, unprocessed Fourier transform ion cyclotron resonance mass spectrometry (FTICR-MS) data from another data package, available on ESS-DIVE at https://data.ess-dive.lbl.gov/view/doi:10.15485/1834208. The scripts then process the raw FTICR-MS data and generate the findings and figures presented in the associated manuscript. In brief, this study demonstrates that DOM assemblages in variably inundated sediments are primarily governed by deterministic variable selection, including sediment moisture effecting the degree of deterministic assembly. See the manuscript for more details pertaining to interpretation and implications of the findings. This data package is associated with the GitHub repository found at https://github.com/WHONDRS-Hub/ECA_2020_Sed.This data package is comprised of 6 scripts and 7 folders. The file-level metadata file (file ending in "flmd.csv") lists all files contained in this data package and descriptions for each. The data dictionary (file ending in "dd.csv) describes all tabular data columns and their respective definitions and units. The FTICR_Processing_Scripts produce the outputs found in the "Processed_Data" folder. The remaining scripts (located in the parent directory) produce the outputs found in the following four folders: (1) "MCD_Dendrograms", "MCD_Randomizations", "MCD_bNTI_Outcomes", and "OM_Null_Modeling". The fifth script additionally takes the three comma-separated values (CSV) files found in the parent directory as input ("VGC_texture.csv", "merged_weights.csv", and "ECA2_FTICR_BetaDisp.csv"). The outputs of each of the five scripts serve as the input to the following script, with the final outputs stored in the folder "OM_Null_Modeling".

54 ENVIRONMENTAL SCIENCES↗

Integrated Modeling Driven Evaluation of Opportunities for Climate‐Resilient Perennial Biomass Crop Plantings in Flood‐Prone Agricultural Landscapes

Adapting to future climate change in flood-prone landscapes will require climate-resilient agricultural systems. Planting perennial crops, like switchgrass and willow, along river corridors can mitigate future flooding while supporting bioenergy markets. We developed an integrated assessment linking climate, hydrologic, and inundation model results to assess future flood risk to river-adjacent agricultural lands in the Mid-Atlantic Region (MAR) and explore this opportunity. We produced ensemble streamflow projections for every MAR stream using a hydrologic model driven by a suite of downscaled and bias-corrected Coupled Model Intercomparison Project Phase 6 climate projections. We then conducted high-resolution inundation mapping based on projected flood frequencies for baseline and future periods. Results show that in the near-term future, at least two-thirds of the streams will experience 100-year floods more severe than the baseline 200-year floods. Riparian zones are projected to face a median rise of inundation by 9.5%–24.1%. Results show that there is an opportunity to mitigate flooding in over half of MAR's counties with the quantities of switchgrass and willow plantings anticipated for mature bioenergy markets, even under the most extreme (200-year) flood events. Our integrated modeling framework can guide similar regions to evaluate opportunities for flood-resilient agricultural systems under climate change.

60 APPLIED LIFE SCIENCES↗