Search NASA⌕ Search

SEARCH · Search NASA

Results for “USGS”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

A regional comparison of sub-daily flow variability in regulated and unregulated rivers in the United States

Regulating rivers for hydropower or other purposes can dramatically alter river flow patterns, including creating substantial changes in flow over short, minutes-to-hours-long timespans known as sub-daily flow variability (SDFV). The impacts of flexible hydropower production on flow and aquatic organisms are increasingly documented in research. However, the degree to which flow alteration relates to different hydropower operational modes in distinct geographical regions and seasons is not well understood. This study offers a methodology for regional- and species-appropriate evaluations of potential impacts of flow on fish based on sub-daily flow characteristics of hydropower operational modes. We analyzed 15-min discharge data between 2018 and 2021 from 69 USGS stream gages to compare SDFV in hydropeaking, run-of-river, and unregulated systems in the US Southeast and Pacific Northwest. Regulated systems exhibited significant SDFV downstream from hydropower facilities relative to unregulated systems, but specific impacts differed between regions. Regulated systems in the Southeast were characterized by high flow coefficients of variation and ratios (hydropeaking only) and extended durations of daily upramping flow phases. Regulated systems in the Pacific Northwest were characterized by many short flow phases per day and large portions of the day spent upramping. Pacific Northwest unregulated systems displayed the strongest seasonal flow patterns while Southeastern hydropeaking systems displayed the greatest SDFV. Given that SDFV impacts multiple dimensions of fish ecology, region-specific sub-daily flow signatures have important implications for understanding and mitigating potential community-, species-, and age-specific effects on fish in different parts of the country.

Fish↗

Knowledge-guided graph machine learning for spatially distributed prediction of daily discharge and nitrogen export dynamics

Spatially distributed prediction of streamflow and nitrogen export dynamics is essential for precision management of agricultural watersheds. While temporal deep learning models such as Long Short-Term Memory (LSTM) have shown strong performance at basin scales, their ability to generalize spatially is limited by insufficient representation of spatial dependencies and flow paths, particularly under data-scarce conditions. To address this gap, we propose HydroGraphNet, a knowledge-guided graph machine learning framework that integrates process-based knowledge and explicit spatial learning into temporal modeling. This framework incorporates directed graph topology to encode watershed connectivity and upstream inflows, with mass balance constraints to improve physical consistency. To enhance generalization in sparsely monitored regions, HydroGraphNet is pretrained on synthetic data generated by the SWAT+ (Soil and Water Assessment Tool Plus) model. We evaluated HydroGraphNet in the Upper Sangamon River Basin (44 HUC-12 subwatersheds, 2001–2020) against two LSTM baselines: a lumped basin-level model and a distributed variant. When benchmarked on SWAT+ simulations in pretraining, HydroGraphNet improved test NSEs by 8.9% (discharge) and 13.7% (NO₃–N load) in temporal extrapolation, and by 27.1% and 34.7% in spatial extrapolation, relative to the Lumped LSTM baseline. After fine-tuning with USGS monitoring data, the model achieved mean test NSE (KGE) scores of 0.768 (0.861) for discharge and 0.626 (0.664) for NO₃–N load, substantially outperforming baselines. Attribution analysis further highlighted the importance of upstream inflow representation and graph-based spatial learning in capturing cross-subwatershed dependencies. The model also reproduced seasonal hydrological and biogeochemical patterns consistent with known processes, demonstrating its robustness and process fidelity for spatially distributed prediction. Altogether, HydroGraphNet advances the integration of physical knowledge and spatially explicit learning in hydrological modeling, offering a generalizable framework for distributed modeling to support spatially targeted water quality management in data-scarce watersheds.

54 ENVIRONMENTAL SCIENCES↗

Consolidation and Permeability of the B1 and D1 Gas Hydrate Bearing Sands and Associated Seal Sediments of the Extended-Duration Gas Production Test Site on the Alaska North Slope

Gas hydrate, a solid combination of gas (mostly methane in nature) and water molecules stable at low temperatures and elevated pressures, occurs naturally in marine and permafrost-associated environments. Gas hydrate reservoirs, such as those in the Alaska North Slope, have been considered potential energy resources for gas production. To understand the petrophysical and geo-mechanical characteristics of the reservoir, core samples retrieved from the site of the JOGMEC-DOE-USGS collaborative gas hydrate R&D project have been analyzed in the laboratory for their hydraulic and mechanical properties. This paper focuses on both seal and reservoir samples associated with the B1 and D1 sands, which are evaluated for index properties (including porosity, grain size distribution, liquid and plastic limits, specific surface area, and specific gravity), consolidation, permeability, and water retention. Furthermore, the reservoir core samples were tested with pore-filling, laboratory-grown tetrahydrofuran hydrate, in order to assess reservoir behavior during gas production from hydrates. Under simulated in situ stress conditions, the seal and hydrate-free reservoir cores had a permeability anisotropy ratio of k h /k v = 3.0−5.0, and k h /k v = 2.4−3.0 for the reservoir tetrahydrofuran hydrate-bearing cores. The data suggest that depressurizing the reservoir to induce hydrate dissociation alters the reservoir effective permeability in three ways: permeabilities decrease due to porosity lost (e.g., the initial reservoir thickness can decrease by up to 5% upon 7 MPa depressurization), permeability increases due to the loss of solid hydrate in the pore space, and permeability anisotropy k h /k v decreases in response to the evolving pore-space geometry. We show that given the simulated in situ gas hydrate saturations (i.e., S h = 32% in core 7P-2E and S h = 21% in core 20P-4), gas production from the dissociation of tetrahydrofuran hydrate in the two tested cores results in a net increase in effective permeability and a decrease in k h /k v . This study highlights the importance of investigating seal and reservoir sediments and the impacts of depressurization on the porosity and permeability responses during production.

Geological materials↗

Modeling the Effects of Artificial Drainage on Agriculture-Dominated Watersheds Using a Fully Distributed Integrated Hydrology Model

In agriculture-dominated watersheds where natural drainage is poor, agricultural ditches (narrow engineered channels) and tile drains (perforated pipes) are widely employed to enhance surface and subsurface drainage, respectively. Despite their relatively small scale, these features exert substantial control over the hydro-biogeochemical function of watersheds and their effects need to be represented in the models. We introduce a novel strategy to incorporate the effects of artificial agricultural drainage into a fully distributed basin-scale integrated surface-subsurface hydrology models. In our approach, narrow agriculture ditches for surface drainage are resolved efficiently using ditch-aligned computational meshes that are hydrologically conditioned to ensure connectivity in the stream/ditch network. For tile drainage in the subsurface, we use the physically based Hooghoudt's drainage equation as a subgrid model and route the water drained through tiles to the nearest ditch. Without site-specific calibration, this model reproduced observed streamflow in the Portage River Watershed (>1,000 km 2 ) as recorded by a USGS gauge with good accuracy (normalized KGE = 0.81) and outperformed a calibrated SWAT model (normalized KGE = 0.68). Numerical experiments confirm that artificial drainage reduces surface inundations and effectively controls the water table. At the watershed scale, artificial drainage increases baseflow but has little effect on watershed discharges above the 90th percentile. The strong physical underpinnings and reduced need for calibration allow us to study the impacts of artificial drainage on distributed hydrological response in terms of fluxes and states and provide a platform for investigating watershed-scale nutrient transport.

54 ENVIRONMENTAL SCIENCES↗

An Integrated Modeling Framework for Sediment Dynamics During Urban Flooding: Application to Hurricane Harvey in Houston

Floodwater can mobilize and redistribute large volumes of sediment from upland to downstream urban areas, threatening infrastructure, water quality, and ecosystem health. However, existing modeling approaches often fail to capture sediment dynamics in urban floodplains due to the lack of integration between upland hydrological processes and riverine sediment transport. This study presents the first integrated modeling framework that couples the Energy Exascale Earth System Model (E3SM) land component, which simulates runoff and hillslope erosion, with TELEMAC-GAIA, a two-dimensional hydrodynamic and sediment transport model. This framework enables the fully distributed, process-based simulation of high-resolution (as fine as 30 m) sediment dynamics from hillslopes to floodplains. Applied to a highly urbanized watershed in Houston during Hurricane Harvey, this framework reproduced observed water levels at 16 USGS gauges (median R 2 = 0.83 and KGE = 0.78), key sediment dynamics such as sediment transport and deposition processes, and reproduced spatial deposition patterns consistent with LiDAR-derived data. Based on the simulation, we estimate 8.0 million m 3 of event-scale sediment deposition, including 5.7 million m 3 trapped in the flood-control reservoirs and 2.3 million m 3 deposited along major channels and floodplains. Using a representative unit removal cost, this corresponds to an estimated dredging cost of $581 million for total deposition. These results provide a first-order, physically based quantification of Harvey-scale sediment impacts. This study provides a valuable tool for the holistic analysis of sediment dynamics triggered by extreme urban flooding, supporting flood-resilience planning. More broadly, it highlights the importance of integrating physically based hydrological processes for urban flooding and sediment research.

Hurricane Harvey↗

FFTSF: Revisiting Sub-Seasonal Streamflow Forecasting with Simple Feedforward Network

Accurate short-to-subseasonal streamflow forecasts are vital for water management, including flood preparedness, drought mitigation, hydropower scheduling, and ecosystem protection. However, extending a forecast beyond a few days remains challenging due to complexity of hydrological processes. While recent self-attention based transformer architectures such as iTransformer have gained traction in time-series forecasting, these models suffer from several critical limitations: (1) significant computational overhead that scales quadratically with sequence length, (2) vulnerability to overfitting on limited hydrological datasets, (3) degraded performance on long-horizon forecasts due to attention decay, and (4) excessive architectural complexity that hampers interpretability and operational deployment. In this study, we propose a simple Feedforward Time Series Forecasting (FFTSF) network that directly addresses these limitations through its lightweight architecture and long-range forecasting capabilities. We evaluate FFTSF across 178 USGS stream gauges spanning diverse climate regimes by forecasting lead times of 1-, 7-, 14-, and 30-days. Our results demonstrate that FFTSF achieves competitive performance at short lead times (NSE of 0.778 for 1-day forecasts) while substantially outperforming complex baselines at longer forecast period, achieving the highest NSE (0.271) at 30-day forecasts with greater robustness and stability. For 30-day forecasts, FFTSF achieves a 71% improvement over NLinear, 57% improvement over DLinear and 12% improvement over the computationally intensive iTransformer while requiring fewer computational resources. Our findings reveal that architectural complexity is not necessary for hydrological forecasting, demonstrating that well-designed simple models can outperform attention mechanisms for subseasonal streamflow forecasting. The computational efficiency and consistent long-range performance of FFTSF make it suitable for water management applications where reliable extended forecasts are essential.

Krishnan Kutty Ambika, Anukesh [ORNL] (ORCID:00000↗

Turbidity and suspended sediment data for Gwynns Falls, Baisman Run, and Pond Branch, Baltimore County and Baltimore City, MD, USA

This resource includes turbidity and suspended sediment data collected at two sampling stations located on Gwynns Falls in Baltimore County, MD, USA. In addition, two forested reference sites, Baisman Run and Pond Branch at Oregon Ridge, and two urban sites, Dead Run and Maiden's Choice Run (tributaries to Gwynns Falls), were sampled in Baltimore County and Baltimore City, MD, USA. Turbidity sensor data were collected at a 5-minute frequency using YSI EXO2 sondes. Suspended sediment was collected using ISCO samplers for the purpose of establishing correlations between turbidity and suspended sediment concentration. The six sites are co-located with USGS stream gages. This resource is part of the Baltimore Social-Environmental Collaborative Urban Integrated Field Laboratory supported by Department of Energy as well as the Critical Zone Collaborative Network supported by National Science Foundation. This resource includes a technical report summarizing the findings.

58 GEOSCIENCES↗

Seismic Contingency Auto Generator

This code takes in premade earthquake scenario XML files from USGS, power grid data, and converts them into a contingency file (.con file) that can be used by power grid solvers. Within the .con file are a number (Specified by the user) of contingencies that have randomly failed power transformers based on their likelihood of failure and peak ground acceleration (PGA) value around the transformer. The transformers' likelihood of failure was calculated based on a variety of finite element modeling on various transformer designed for specific transformer voltage classes. Parameters from these FEM were used to create generic fragility curves for transformers within a specific voltage class, which correspond with earthquake PGA values to produced a probability of failure for a given earthquake scenario. More refined versions of this process, such as specifying specific transformer design categories within a voltage class, could also be applied in future iterations of the software.

Vaagensmith, Bjorn [Idaho National Laboratory (INL↗

Location Identifiers, Metadata, and Map for Field Measurements at the East-Taylor Watershed Community Observatory, Colorado, USA (Version 3.3)

This dataset contains identifiers, metadata, and a map of the locations where field measurements have been conducted at the East-Taylor Watershed Community Observatory located in the Upper Colorado River Basin, United States. This is version 3.3 of the dataset and replaces the prior version 3.2 (see below for details on changes between the versions). Dataset description: The East River-Taylor Watershed is the primary field site of the Watershed Function Scientific Focus Area (WFSFA) and the Rocky Mountain Biological Laboratory. Researchers from several institutions generate highly diverse hydrological, biogeochemical, climate, vegetation, geological, remote sensing, and model data at the East-Taylor Watershed in collaboration with the WFSFA. Thus, the purpose of this dataset is to maintain an inventory of the field locations and instrumentation to provide information on the field activities in the East-Taylor Watershed and coordinate data collected across different locations, researchers, and institutions. The dataset contains (1) a README file with information on the various files, (2) three csv files describing the metadata collected for each surface point location, plot and region registered with the WFSFA, (3) csv files with metadata and contact information for each surface point location registered with the WFSFA, (4) a csv file with with metadata and contact information for plots, (5) a csv file with metadata for geographic regions and sub-regions within the watershed, (6) a compiled xlsx file with all the data and metadata which can be opened in Microsoft Excel, (7) a kml map of the locations plotted in the watershed which can be opened in Google Earth, (8) a jpg image of the kml map which can be viewed in any photo viewer, and (9) a zipped file with the registration templates used by the SFA team to collect location metadata. The zipped template file contains two csv files with the blank templates (point and plot), two csv files with instructions for filling out the location templates, and one compiled xlsx file with the instructions and blank templates together. Additionally, the templates in the xlsx include drop down validation for any controlled metadata fields. Persistent location identifiers (Location_ID) are determined by the WFSFA data management team and are used to track data and samples across locations. Dataset uses: This location metadata is used to update the Watershed SFA’s publicly accessible Field Information Portal (an interactive field sampling metadata exploration tool; https://wfsfa-data.lbl.gov/watershed/), the kml map file included in this dataset, and other data management tools internal to the Watershed SFA team. Version Information: The latest version of this dataset publication is version 3.3. This version contains 167 new point locations, 1 new plot, and 2 new geographic regions. Overall, there are a total of 1439 point locations, 75 plots, and 54 geographic regions. Additionally, the kml map of locations and image now includes two boundaries (Upper Ohio Creek (UO) and Carbon Creek (CA)) outside of the East River watershed (USGS HUC-10) and accompanying stream network that represents areas of focus. Refer to methods for further details on the version history. This dataset will be updated on a periodic basis with new measurement location information. Researchers interested in having their East-Taylor Watershed measurement locations added to this list should reach out to the WFSFA data management team at wfsfa-data@googlegroups.com. Acknowledgments: Please cite this dataset if using any of the location metadata in other publications or derived products. If using the location metadata for the 2018 NEON hyperspectral campaign, additionally cite Chadwick et al. (2020). doi:10.15485/1618130. This work was supported by the Watershed Function Science Focus Area at Lawrence Berkeley National Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-05CH11231. Part of this work was performed at SLAC Accelerator Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-76SF00515.

2018 NEON and 2025 CHESS Campaigns↗

Groundwater and Surface Water Flow (GSFLOW) model files to explore bedrock circulation depth and porosity in Copper Creek, Colorado

This data package contains integrated hydrological model input and output files for Copper Creek, Colorado (24 km2), a tributary of the East River located in the headwaters of the Upper Colorado River Basin. The model code is the U.S. Geological Survey (USGS) Groundwater and Surface Water Flow (GSFLOW) model. The model contains a 100-m grid resolution and a daily timestep. The land surface model is dynamically linked to a three-dimensional groundwater flow model that allows for streamflow gaining and losing conditions. The groundwater model contains 12 model layers and extends 400 m below land surface. The original Copper Creek model was modified to contain geologic layers representing saprolite, shallow bedrock, and deep bedrock. Endmember depth versus hydraulic conductivity relationships and porosity values for fractured crystalline rock are simulated. For the shallow case, median flow depths occur in the shallow saprolite at depths <8 m, while the deep case promotes a median groundwater flow depth of 100 m. With this modeling framework we compare streamflow response to a plausible worst-case drought lasting up to five years. Streamflow metrics of analysis include average streamflow, fraction of stream network that is dry, no-flow duration, average groundwater flow to streams and time to recovery following the drought. Results and implications are presented in a paper submitted to Geophysical Research Letters titled, "The role of bedrock circulation depth and porosity in mountain streamflow response to prolonged drought" by Rosemary WH. Carroll, Andrew H. Manning and Kenneth H Williams. A Readme.txt file provides instructions on how to download all model files and execute each model scenario. In addition to the GSFLOW output/prms/copper_drought.csv file containing daily basin water stores and fluxes (refer to GSFLOW manual) and the output/prms/copper_drought_statvar.dat file with output defined in the gsflow3.control file (refer to GSFLOW Manual), output files also include spatially distributed daily values of total evapotranspiration, canopy evaporation, precipitation, snowfall, infiltration, snow water equivalent, potential evapotranspiration, recharge, sublimation, soil moisture, contributing interflow, water table elevations, changes in groundwater storage, groundwater evapotranspiration, interbasin groundwater flow (limited to the alluvium below the stream outlet), and surface-groundwater exchanges within the river system.

54 ENVIRONMENTAL SCIENCES↗

Dataset: "Widespread Drought-driven Declines in Streamflows and Water quality in the Upper Colorado River Basin (1998-2022)"

This data package contains the associated data and scripts for Nagamoto, E., Ombadi, M., Ciulla, F. et al. Widespread drought-driven declines in streamflows and water quality in the Upper Colorado River Basin during 1998-2022. Commun Earth Environ 7, 734 (2026). https://doi.org/10.1038/s43247-026-03890-5. This purpose of this study was to investigate the impact of the 21st century drought on water quantity and quality at catchments throughout the Upper Colorado River Basin (UCRB). We used stream flow, water temperature, specific conductance, air temperature, precipitation, and catchment attribute data for over 200 sites in the UCRB, collected from the National Water Information System using Basin3D (Varadharajan, 2023), GAGESII (Falcone, 2010), and the Google Earth Engine. We identified years of severe drought between 1998 and 2022 using the Standardized Precipitation Evaporation Index (SPEI), then calculated the relative change percentage of the stream flow, water temperature, and specific conductance from drought versus non-drought years. We used the attribute information from GAGESII to investigate what physical traits of catchments are associated streamflow vulnerability (greater relative change) or resilience to drought. We used land cover data from the National Land Cover Database (USGS, 2024) to assess any changes to physical attributes that may not be represented in the static attributes information in GAGESII. To increase data availability, we modeled stream temperature using methods from Willard, 2023. While the study period is water years 1998 to 2022, the raw water quantity and quality data extends to 1950 and the meteorological data extends to 1980. The data and code can be downloaded via the UCRB_drought.zip. Within the zip, the files are organized as follows: - INPUTS: Contains all input data used in UCRB_Drought_Workflow.ipynb - OUTPUTS: Contains all intermediate data created from UCRB_Drought_Workflow.ipynb as well as final products including the calculated Standardized Evapotranspiration Index (SPEI) - climatic_variables: The code used to collect meteorologic data from Google Earth Engine - feature_importance: The code used for the catchment attributes analysis - preprocessing: Code used in UCRB_Drought_Workflow_Preprocessing.ipynb - pyeto: Code used in UCRB_Drought_Workflow_Preprocessing.ipynb - calculations: Code used in UCRB_Drought_Workflow_Impacts.ipynb - plotting: Code used in UCRB_Drought_Workflow_Impacts.ipynb - README.md - UCRB_Drought_Workflow_Preprocessing.ipynb: The code used to prep raw data for the analysis - UCRB_Drought_Workflow_Impact.ipynb: The code which uses the prepped raw data for analysis, and plots all figures - requirements_ucrb-drought_v2.yml: The requirements file to create a virtual environment and Jupyter Lab kernel to run the code The INPUTS folder is organized into the following major directories and sub-directories. The "RDC_WT_SC_RAW" folder contains raw data for streamflow, water temperature, and specific conductance in a ".h5" file. The "NLCD_RAW" folder contains ".csv" files with annual land cover percentages for counties within the UCRB. The "MET_RAW" folder contains a ".csv" file with monthly meteorological data (air temperature and precipitation) for the sites in the UCRB which was obtained from code in the climatic_variables folder. The "GAGESII" folder contains ".csv" files with physical catchment attribute variables for catchments across the country. The "WT_LSTM_data" folder contains ".csv" files with calculated WT (Willard, 2023) and the associated RMSEs. The "Upper_Colorado_River_Basin_Boundary" folder contains geographic data including a shapefile for plotting in the UCRB_Drought_Workflow.ipynb. The "RESERVOIRS_RAW" folder contains ".csv" files for each reservoir in the UCRB with daily reservoir storage. There are also two files in the INPUTS folder that have combined reservoir storage data and reservoir metadata. The OUTPUTS folder is organized into the following major directories and sub-directories. The "RDC_WT_SC_data" folder contains a folder "Water_year" with the associated cleaned data, metadata, and data availability information in ".csv" files, a folder "Median_Relchange" with the relative change comparing drought to non-drought years in ".csv" files, and a folder "Peak95_Min5_Relchange" that has ".csv" files for the relative change in peak (95th %) and minimum (5th %) variables. The "NLCD_data" folder contains the difference in land cover from the beginning to end of the study period and the percentage of the county that is within UCRB bounds can be found in Nagamoto et al (2025)). The "MET_data" folder contains separated monthly air temperature and precipitation data and the calculated PET in ".csv" files. The "SPEI_data" folder contains ".csv" files with calculated SPEI values (one restricted to the study period and the other with information from the entire MET data period). The "Paper_Tables" folder contains two ".csv" files containing site information and data availability and information about the GAGESII trait aggregated categories. The base directory includes the file “flmd.csv” for a list and description of all files and the file “dd.csv” for data dictionaries. Scripts for preprocessing, analysis, and figure generation are located in the associated GitHub repository found at [https://github.com/iNAIADS/drought-impacts/tree/develop/UCRB-drought]. UPDATE 1: Title and code file updated to match submitted manuscript 10-15-2025. UPDATE 2: Code and data files updated to match revised manuscript 3-4-2026. UPDATE 3: Code and data files updated to match revised manuscript 6-7-2026. ** NOTE: DD and FLMD have not been updated yet. UPDATE 4: Added associated Manuscript information and DD and FLMD have been updated. To cite this code, please use the following BibTeX: @misc{nagamoto2025drought, author = {Emily Nagamoto and Fabio Ciulla and Mohammad Ombadi and Jared Willard and Rosemary Carroll and Charuleka Varadharajan}, title = {Dataset: "Widespread Drought-driven Declines in Streamflows and Water quality in the Upper Colorado River Basin (1998-2022)"}, year = {2025}, doi = {10.15485/2551894}, publisher = {ESS-DIVE Repository}, url = {https://data.ess-dive.lbl.gov/datasets/doi:10.15485/2551894} }

54 ENVIRONMENTAL SCIENCES↗

15-minute Parker River gap-filled tide height and salinity data, PIE LTER, Plum Island Sound, MA (2014–2023), for ELM PFLOTRAN modeling

This dataset contains 15-minute tide height and salinity data from the Typha site along the Parker River, part of the Plum Island Ecosystems Long Term Ecological Research (PIE LTER) site in Plum Island Sound, Massachusetts (MA) 2014-2023. Tide height (in NAVD88) was compiled from measurements conducted at the mouth of Plum Island Sound and corrected for time lags. Gap-filling of missing periods were done by fitting tidal constituents to the time series. Salinity was measured (and is stored on ESS DIVE ) in 2022 and 2023 using HOBO U24-002 conductivity loggers. River discharge is the most important control on tidal river water salinity at the location (Vallino & Hopkinson, 1998). An artificial neural network was trained to predict river water salinity at the location using Parker River discharge (USGS station 01101000, Parker River at Byfield, MA) and gap-filled salinity observations from a long-term monitoring station ca. 3km downstream from the Typha site (LTER station ‘Middle Road’) as input variables to create continuous time series information. The data set was used in the spin up and simulations of a land surface model coupled to a biogeochemical reaction network (ELM PFLOTRAN) assessing impacts of hydrology and salinity input on methane fluxes in 2022 and 2023 (Sulman et al., 2024). Metadata files ELMPFLOTRAN_tide_salinity_dd.csv and ELMPFLOTRAN_tide_salinity_flmd.csv provide details on site location, data variables, and QA/QC methods .

54 ENVIRONMENTAL SCIENCES↗

Data From: "Warming and snow loss increase reliance on old groundwater in a Colorado River headwater"

This repository contains the data and code associated with the paper titled "Warming and snow loss increase reliance on old groundwater in a Colorado River headwater," published in Nature Geoscience, 2026. This study seeks to answer how various ages of groundwater interact with mountainous streamflow in mountainous headwaters such as the East River. It includes various model-data processing scripts, primarily for ParFlow-CLM analysis of simulated water years 2015-2021, and two numerical warming experiments (+2.5 and +4.0 degrees C), including run scripts, forcing scripts, and post-processing, as well as comparison to observation datasets, detailed below. This data requires the use of R (.r, .rmd), Python (.py), Jupyter Notebook or Jupyter Lab (.ipynb), ParFLOW-CLM, EcoSLIM. Further information on the use of all file formats mentioned below (e.g. .tff. .nc) are provided within the associated scripts and directory where the files are located. Contents & Usage ASO/: ​​Contains the bash and python scripts used to convert airborne snow observatory (ASO) data (ASO, 2023) in various data formats (georeferenced tiff file, NetCDF, UTM, and to latitude/longitude) then regrided to the ParFlow equivalent grid. Output data are in regrid_regll_data.zip and subsequently visualized and analyzed in plot_and_compare.py for Supplementary Figures A14 and A15. The wksht_ASO_comparison.xlsx spreadsheet is used to calculate the data for Supplementary Figure A16. EcoSLIM/: Contains the scripts and input files to run the EcoSLIM particle tracking simulations (/run_scripts) and the post-processing python script (/plot_scripts/eco_agedist_plots.ipynb). Jasechko et al./: Contains the jupyter notebook (Extract_Elevation.ipynb) to determine the outlet elevations of the 260 watersheds used in Jasechko et al. (2016), and the corresponding table, Table_S1_Watersheds_alt.csv. Used to create Supplementary Information Figure A2. PLM_Wells/: Contains the QA/QC-ed groundwater level time series of the PLM-1 and PLM-6 Monitoring Wells from Faybishenko et al. (2023), reformatted to water years used for Supplementary Figures A19 and and A20. ParFlow/: Contains the input files and run scripts to run ParFlow-CLM (/run_scripts), the python and tool command language (Tcl) scripts to create and distribute the ParFlow forcing simulation files (/forcing), and various scripts and intermediary files to analyze the model outputs (/post_process). SQUIRE/: Contains the processing scripts and intermediary files for the Surface QUantitatIve pRecipitation Estimation (SQUIRE) data (Grover, 2023) used to generate Supplementary Figure A18. USGS_Streamflow/: Contains the raw and gap-filled United States Geological Survey streamflow data (U.S. Geological Survey, 2026) used at the Almont station (site number 09112500). Gap-filling is performed in the R script with data from the Taylor station (site number 09110000). (/USGS_09112500_EAST_RIVER_AT_ALMONT_GAP_FILLED/code_almont_streamflow_gap_fill.Rmd). discharge/: Contains the gap-filled discharge data at the Watershed Function SFA East River pumphouse site (Newcomer et al., 2022) used to generate Supplementary Figure A13 and to compute hourly Nash-Sutcliffe model efficiency coefficients (NSE) in Table A4. snotel_and_flux_tower/: Contains the snow telemetry data (U.S. Department of Agriculture, 2024) from the Butte (site ID 380) and Schofield (site ID 737) stations, reformatted by water year, accessed with the snotelr R package. Used to create Supplementary Figure A17. Also contains the flux tower observational data (FluxTower_Pumphouse_ESS-DIVE.ET_only.h.txt) from Ryken et al. (2022) and sap flux transpiration data (MaxB_Transpiration_5Sites.daily_sums.h.txt) from Ryken (2021), used to create Supplementary Figures A22 and A23, respectively. Raw EcoSLIM model outputs are in excess of 24TB, and are stored on National Energy Research Scientific Computing Center (NERSC) and publicly available via the external link provided in the paper.

atmospheric warming↗

A bespoke model of Arctic river basins based on hillslope delineation: Model Archive

This dataset is a model archive of the paper A bespoke model of Arctic river basins based on hillslope delineation (in prep), which introduces a watershed decomposition and parameterization method for large scale permafrost hydrology simulation. With this dataset, this study aims to address the research question: whether a computationally efficient hillslope-based modeling framework can reliably simulate discharge at Arctic river-basin scales. This dataset contains model input and output data for five modeling scenarios at a study site located in the Sagavanirktok River basin. The five modeling scenarios include three modeling cases under temperate conditions using full 3D, decomposed 3D, and decomposed 2D modeling strategies; and two modeling cases under actual Arctic conditions with permafrost using full 3D and decomposed 2D modeling strategies. Simulations were performed using the Advanced Terrestrial Simulator (ATS, v1.6 for three temperate scenarios and v1.5 for two Arctic scenarios), a physics-rich integrated surface–subsurface hydrologic model with cryo-hydrology features. For the three temperate models, simulations were conducted for the period of 10/01/1993 - 09/30/2002; and for the two Arctic models, simulations were conducted for the period of 01/01/1994 - 12/31/2002. To facilitate reproducibility of simulations, all datasets are organized hierarchically. The dataset contains: (1) Mesh files (.exo) for full 3D model, decomposed 3D models, and decomposed 2D models, located in huc/190604020802_gauge15906000/mesh/. Mesh files can be visualized through Paraview or read by Python. (2) Climate forcings (.h5) for full 3D model and decomposed 3D/2D models are located in huc/190604020802_gauge15906000/daymet_onePiece/, and huc/190604020802_gauge15906000/vp_pr_revised_daymet_1980_2006_with_wind/ separately. Accessible by Python. (3) Raw measured gage discharge (.csv) from USGS, located in huc/190604020802_gauge15906000/gaged_basin15906000_discharge_usgs/. Accessible by Python. (4) Delineated subdomain raster (.tif) and shape files (.shp), and the final parameterized results (.npy) for decomposed models, located in huc/190604020802_gauge15906000/data_preprocessed-meshing. Accessible by Python. (5) Temperate models are located in nonpermaf_huc190604020802_gauge15906000/, which includes three cases: decomposed 2D models (inside model_0*-hillslope_*), decomposed 3D models (inside model_1*-subcatchment_*), and full 3D model (inside model_2*-onepiece_*). Two step spin-up results (checkpoint_final.h5) are located in model_*1-*_spinup_steadystate and model_*2-*_spinup_cycle, separately, which are used to initialize real transient models. The input files (.xml) and output results (.dat) of the real transient models are located in model_*3-*_transient/. Especially, for two example hillslope models (ID=-11 and 11), additional h5py files are included in model_03-hillslope_transient/hillslope-11/, model_03-hillslope_transient/hillslope11, model_13-subcatchment_transient/subcatchment-11/, model_13-subcatchment_transient/subcatchment/11, respectively, which are used to plot the saturation figure (Figure 5) in the manuscript. Accessible by Python. (6) Arctic models are located in huc190604020802_gauge15906000/, which includes two cases: decomposed 2D models (inside model_04-hillslope_transient), and full 3D model (inside model_05-onepiece_transient_mannp1_ra). Three step spin-up results (checkpoint_final.h5) are located in model_01-column_freezeup/, model_02-column_spinup/, model_03-hillslope_spinup/, respectively, which are used to initialize real 2D transient hillslope models. The input files (.xml) and output results (.dat) of transient 2D hillslope models are located in model_04-hillslope_transient/. The input files (.xml) and output results (.dat) of the full 3D transient model is located in model_05-onepiece_transient_mannp1_ra/. The full 3D transient model is initialized by model_02-column_spinup/. Accessible by Python. (7) The MOSART routed discharge results (.csv) under Arctic conditions is located in huc190604020802_gauge15906000/MOSART/. Accessible by Python. (8) All Python codes (.py) used to parameterize full 3D model to decomposed 2D models are located in script/. These codes fit with watershed workflow (a watershed delineation tool) v1.4 under the branch gaob/v1.4 from https://github.com/gaobhub/watershed-workflow.git.

EARTH SCIENCE > CRYOSPHERE↗

Constituent Data Replacement Tool

The purpose of this tool is to estimate key parameters that may be missing in public wastewater composition datasets. The tool can be applied to develop complete treatment and critical mineral extraction profiles for leachate, produced water and other aqueous waste streams. The tool applies machine learning algorithms to replace missing data in a user’s water data set that are adjusted based on user preferences for options including algorithm type, number of features, and classification variables. The tool can use the user’s data alone or combine user data with the NEWTS USGS Produced Water Database for more robust training. This research was funded by the U.S. Department of Energy’s Office Fossil Energy and Carbon Management (FECM) through National Energy Technology Laboratory’s ongoing research under the Water Management for Power System Field Work Proposal, DE-FECM 1022428 and Critical Minerals Field Work Proposal, DE-FECM 1022420.

Aqueous Chemistry↗

Datasets and U-Net Model for "A Deep Learning Based Framework to Identify Undocumented Orphaned Oil and Gas Wells from Historical Maps: a Case Study for California and Oklahoma"

This dataset has results and the model associated with the publication Ciulla et al., (2024). It contains a U-Net semantic segmentation model (unet_model.h5) and associated code implemented in tensorflow 2.0 for the model training and identification of oil and gas well symbols in USGS historical topographic maps (HTMC). Given a quadrangle map (7.5 minutes), downloadable at this url: https://ngmdb.usgs.gov/topoview/, and a list of coordinates of the documented wells present in the area, the model returns the coordinates of oil and gas symbols in the HTMC maps. For reproducibility of our workflow, we provide a sample map in California and the documented well locations for the entire State of California (CalGEM_AllWells_20231128.csv) downloaded from https://www.conservation.ca.gov/calgem/maps/Pages/GISMapping2.aspx. Additionally, the locations of 1,301 potential undocumented orphaned wells identified using our deep learning framework or the counties of Los Angeles and Kern in California, and Osage and Oklahoma in Oklahoma are provided in the file found_potential_UOWs.zip. The results of the visual inspection of satellite imagery in Osage County is in the file visible_potential_UOWs.zip. The dataset also includes a custom tool to validate the detected symbols in the HTMC maps (vetting_tool.py). More details about the methodology can be found in the associated paper: Ciulla, F., Santos, A., Jordan, P., Kneafsey, T., Biraud, S.C., and Varadharajan, C. (2024) A Deep Learning Based Framework to Identify Undocumented Orphaned Oil and Gas Wells from Historical Maps: a Case Study for California and Oklahoma. Accepted for publication in Environmental Science and Technology. The geographical coordinates provided correspond to the locations of potential undocumented orphaned oil and gas wells (UOWs) extracted from historical maps. The actual presence of wells need to be confirmed with on-the-ground investigations. For your safety, do not attempt to visit or investigate these sites without appropriate safety training, proper equipment, and authorization from local authorities. Approaching these well sites without proper personal protective equipment (PPE) may pose significant health and safety risks. Oil and gas wells can emit hazardous gasses including methane, which is flammable, odorless and colorless, as well as hydrogen sulfide, which can be fatal even at low concentrations. Additionally, there may be unstable ground near the wellhead that may collapse around the wellbore. This dataset was prepared as an account of work sponsored by the United States Government. While this document is believed to contain correct information, neither the United States Government nor any agency thereof, nor the Regents of the University of California, nor any of their employees, makes any warranty, express or implied, or assumes any legal responsibility for the accuracy, completeness, or usefulness of any information, apparatus, product, or process disclosed, or represents that its use would not infringe privately owned rights. Reference herein to any specific commercial product, process, or service by its trade name, trademark, manufacturer, or otherwise, does not necessarily constitute or imply its endorsement, recommendation, or favoring by the United States Government or any agency thereof, or the Regents of the University of California. The views and opinions of authors expressed herein do not necessarily state or reflect those of the United States Government or any agency thereof or the Regents of the University of California.

Artificial Intelligence↗

A Regional Phase Amplitude Model of 2-D Attenuation for North America

We analyzed seismic attenuation patterns across the North America using over 70,000,000 Lg wave amplitudes recorded by the various networks at frequencies of 0.1-32 Hz. Our inversion solved for laterally varying attenuation, site terms, moments, and apparent stress following Phillips et al., (2016). The inversion was anchored by independently constrained: corner frequencies (via coda spectral ratios) to control the attenuation-stress tradeoff and moment measurements of teleseismic (GCMT, USGS) and regional (St. Louis University and UC Berkeley) earthquakes to provided absolute scaling. The quality factor (Q) shows clear regional patterns: low values in coastal, volcanic, and tectonically active regions, and high values in stable areas like the Great Plains and major plateaus throughout North America. These 2-D Q models enable improved regional source characterization, magnitude estimation, and yield determination.

58 GEOSCIENCES↗

Targeted Rare Earth Element Extraction from Mine Drainage Treatment Solids Informed by Advanced Characterization

In support of a clean energy transition in the U.S., National Energy Technology Laboratory (NETL) has collaborated with staff at Hedin Environmental and students at the University of Pittsburgh to characterize critical mineral content and recovery potential from acid mine drainage treatment solids (AMD solids). AMD solids in Appalachia are an unconventional feedstock of rare earth elements (REEs), with potential of suppling 1,102 tons REE/year. To inform recovery efforts, select AMD solids were examined using synchrotron microprobe analysis in conjunction with USGS-developed geochemical modeling to indicate likely phases hosting critical minerals (REE, Co, Ni, etc.) and associated metals . More than 100 AMD solids were collected from 94 passive AMD treatment systems in Pennsylvania, where limestone aggregates are used for acidity neutralization. As pH increases, dissolved metals and critical minerals in AMD are attenuated as surface coatings on limestone. The collected AMD solids contained up to 2000 mg/kg REE, up to 13,000 mg/kg transition metals (Co, Ni, Zn) and up to 440 mg/kg Li. Regardless of the diverse chemical compositions from AMD solids (Al-rich, Mn-rich, or Al,Fe,Mn-rich), REEs were mostly associated with Al and Mn (hydr)oxides, while select heavy REEs (e.g., Gd, Dy) were co-localized with Fe (hydr)oxides. Co and Ni have different distribution zones, while both co-localized with Mn (hydr)oxides. Based on this characterization, NETL developed a patent-pending innovative step-leaching protocol, “Targeted Rare Earth Extraction (TREE)” to effectively recover up to 90% REE and 60% Co in separate steps. In addition, select post-TREE solid residuals (purified Al oxides, or Mn oxides) can be further developed into functional materials (e.g., lithium and CO2 sorbents) needed for green energy transition and carbon management. This characterization-informed approach as well as TREE processing from AMD solids can be used for other legacy wastes (e.g., coal ash, oil and gas drill cutting, mine tailings), and offers an opportunity to transform waste streams into environmental and economic assets that meet U.S. Department of Energy and U.S. Environmental Protection Agency goals.

characterization and extraction of rare earth elem↗