Search NASASearch

SEARCH · Search NASA

Results for “Geologic data”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

International Offshore Geologic Carbon Storage Inventory and Data Collection

We present an interactive data collection to aggregate, understand, and disseminate the data that are publicly available to support offshore GCS which can be leveraged by stakeholders to understand where GCS may be viable offshore, create GCS project analogs, and address challenges to GCS in offshore environments. The Offshore Geologic Carbon Storage Data Collection is an Experience Builder web application of multiple web mapping applications, aggregated into a single tool for each data type for access, visualization, and exploration. We also present a spatial inventory of global offshore GCS efforts to visualize the scale and locations of actualized and potential offshore GCS. It includes project location, project type and stage, CO2 storage resource potential, injection rate, reservoir and seal geology, and key literature references. Quantitative and qualitative comparisons of the distribution and magnitude of projects by their attributes lends spatial insight into the status of global GCS operations and storage resource potential, thereby enabling comparative assessments and cross-cutting knowledge transfer for projects in development. These datasets illuminate trends in ongoing offshore projects and can be leveraged by stakeholders to estimate storage resources, identify subsurface analogs, review regulations, and address challenges to offshore GCS. Additionally, opportunities for concurrent decarbonization strategies can be identified.

Mulhern, Julia

MRCI Subtask 2.3: Developing Industrial Partnerships and Regional Technical Collaboration Final Technical Summary Report

Under the objective of regional data collection and helping accelerate deployment, MRCI collaborated with industrial stakeholders in their project planning, characterization, and analysis. Some examples of these collaborations are given below. The data and information shared by the industrial collaborations added to the regional CCS framework development and were incorporated into the overall datasets, while addressing any proprietary data requirements. Three examples of collaborative partnerships with industry that have provided geologic characterization data relevant and beneficial to the MRCI program are discussed below, including: the UIC Class II Injection Facility in Eastern Ohio, the Core Energy CO2-EOR (enhanced oil recovery) operation in Otsego County Michigan, and the Marquis ethanol plant in Hennepin Illinois.

CCS,CCUS,MRCI,Midwest USA,Technical Challenges,inj

International Offshore Geologic Carbon Storage Project Inventory and Data Collection

We present an interactive data collection to aggregate, understand, and disseminate the data that are publicly available to support offshore GCS which can be leveraged by stakeholders to understand where GCS may be viable offshore, create GCS project analogs, and address challenges to GCS in offshore environments. The Offshore Geologic Carbon Storage Data Collection is an Experience Builder web application of multiple web mapping applications, aggregated into a single tool for each data type for access, visualization, and exploration. We also present a spatial inventory of global offshore GCS efforts to visualize the scale and locations of actualized and potential offshore GCS. It includes project location, project type and stage, CO2 storage resource potential, injection rate, reservoir and seal geology, and key literature references. Quantitative and qualitative comparisons of the distribution and magnitude of projects by their attributes lends spatial insight into the status of global GCS operations and storage resource potential, thereby enabling comparative assessments and cross-cutting knowledge transfer for projects in development. These datasets illuminate trends in ongoing offshore projects and can be leveraged by stakeholders to estimate storage resources, identify subsurface analogs, review regulations, and address challenges to offshore GCS. Additionally, opportunities for concurrent decarbonization strategies can be identified.

Mulhern, Julia

Data From: "Warming and snow loss increase reliance on old groundwater in a Colorado River headwater"

This repository contains the data and code associated with the paper titled "Warming and snow loss increase reliance on old groundwater in a Colorado River headwater," published in Nature Geoscience, 2026. This study seeks to answer how various ages of groundwater interact with mountainous streamflow in mountainous headwaters such as the East River. It includes various model-data processing scripts, primarily for ParFlow-CLM analysis of simulated water years 2015-2021, and two numerical warming experiments (+2.5 and +4.0 degrees C), including run scripts, forcing scripts, and post-processing, as well as comparison to observation datasets, detailed below. This data requires the use of R (.r, .rmd), Python (.py), Jupyter Notebook or Jupyter Lab (.ipynb), ParFLOW-CLM, EcoSLIM. Further information on the use of all file formats mentioned below (e.g. .tff. .nc) are provided within the associated scripts and directory where the files are located. Contents & Usage ASO/: ​​Contains the bash and python scripts used to convert airborne snow observatory (ASO) data (ASO, 2023) in various data formats (georeferenced tiff file, NetCDF, UTM, and to latitude/longitude) then regrided to the ParFlow equivalent grid. Output data are in regrid_regll_data.zip and subsequently visualized and analyzed in plot_and_compare.py for Supplementary Figures A14 and A15. The wksht_ASO_comparison.xlsx spreadsheet is used to calculate the data for Supplementary Figure A16. EcoSLIM/: Contains the scripts and input files to run the EcoSLIM particle tracking simulations (/run_scripts) and the post-processing python script (/plot_scripts/eco_agedist_plots.ipynb). Jasechko et al./: Contains the jupyter notebook (Extract_Elevation.ipynb) to determine the outlet elevations of the 260 watersheds used in Jasechko et al. (2016), and the corresponding table, Table_S1_Watersheds_alt.csv. Used to create Supplementary Information Figure A2. PLM_Wells/: Contains the QA/QC-ed groundwater level time series of the PLM-1 and PLM-6 Monitoring Wells from Faybishenko et al. (2023), reformatted to water years used for Supplementary Figures A19 and and A20. ParFlow/: Contains the input files and run scripts to run ParFlow-CLM (/run_scripts), the python and tool command language (Tcl) scripts to create and distribute the ParFlow forcing simulation files (/forcing), and various scripts and intermediary files to analyze the model outputs (/post_process). SQUIRE/: Contains the processing scripts and intermediary files for the Surface QUantitatIve pRecipitation Estimation (SQUIRE) data (Grover, 2023) used to generate Supplementary Figure A18. USGS_Streamflow/: Contains the raw and gap-filled United States Geological Survey streamflow data (U.S. Geological Survey, 2026) used at the Almont station (site number 09112500). Gap-filling is performed in the R script with data from the Taylor station (site number 09110000). (/USGS_09112500_EAST_RIVER_AT_ALMONT_GAP_FILLED/code_almont_streamflow_gap_fill.Rmd). discharge/: Contains the gap-filled discharge data at the Watershed Function SFA East River pumphouse site (Newcomer et al., 2022) used to generate Supplementary Figure A13 and to compute hourly Nash-Sutcliffe model efficiency coefficients (NSE) in Table A4. snotel_and_flux_tower/: Contains the snow telemetry data (U.S. Department of Agriculture, 2024) from the Butte (site ID 380) and Schofield (site ID 737) stations, reformatted by water year, accessed with the snotelr R package. Used to create Supplementary Figure A17. Also contains the flux tower observational data (FluxTower_Pumphouse_ESS-DIVE.ET_only.h.txt) from Ryken et al. (2022) and sap flux transpiration data (MaxB_Transpiration_5Sites.daily_sums.h.txt) from Ryken (2021), used to create Supplementary Figures A22 and A23, respectively. Raw EcoSLIM model outputs are in excess of 24TB, and are stored on National Energy Research Scientific Computing Center (NERSC) and publicly available via the external link provided in the paper.

atmospheric warming

Data and code from: Multivariate bayesian regression model for predicting disposed ash composition at U.S. coal fired power stations

This dataset contains the code and data files needed for implementation of a Multivariate Bayesian Regression model, described in Jin et al. (2025), for the historical prediction of the chemical composition of disposed coal ash at U.S. coal fired power plants as a function of annualized coal purchase data. The integrated coal supply data file (CoalSupplyDataset.csv) represents a compilation of monthly fuel purchase records for the period 1973-2022 at major U.S. power stations. These records were obtained from the U.S. Energy Information Administration. The CSV file also contains, for each coal purchase record, the coal region of the mine as defined by the U.S. Geological Survey. Data entry errors and data gaps in the EIA records were corrected as described in Jin et al. This CSV file represents the integrated coal supply data after corrections were made. The model structure and fitting parameters are encoded in pickle file format (Bayesian.pkl). The model was developed with the coal supply data and coal ash composition data, apportioned according to the Stratified Shuffle Split for training and testing subsets. The model was built using Python and the PyMC library. Reference Publication: Jin, Z.; Huang, J.; Hower, J.C.; Hsu-Kim, H.(2025). Predictive Assessment of the Chemical Composition of Coal Ash in Reserve at U.S. Disposal Sites. Environmental Science & Technology.

Coal ash composition

Curating Carbon Storage Data for Reuse: Enabling Research and Modeling from Earth’s Surface to Subsurface

The volume of public geologic carbon storage (GCS) data resources has continued to increase in recent years as the result of an increase in funding from government, industry, and academia towards national, basin, regional and field scale studies to ensure carbon capture and storage becomes a commercially viable operation. Despite the increasing volume of data, GCS data applied towards analyses such as geologic, cost, and risk modeling continues to be multi-sourced and often disparate in nature, published across government agencies, websites, data repositories and buried in derivative reports and documents. Much of the time preparing for an analysis and derivative product development is spent collecting, aggregating, transforming and preparing input data. There have been significant efforts within the DOE National Energy Technology Laboratory’s Carbon Storage Program to optimize multi-source, multi-scale subsurface geologic data curation and aggregation to support data discovery, interoperability, and reuse. Methods include the use of artificial intelligence, machine learning, and data science techniques. This talk will discuss the workflows, best practices, and processes developed to support the aggregation and curation of data through the whole system – surface to subsurface data - that support multi-scale, multi-purpose analysis for carbon storage research.

Morkner, Paige

WELLBASE - An Interactive Platform for Wellbore Material Assessment

This project seeks to build an open-source wellbore material data repository with adequate material performance and contextual data to support Geological Carbon Storage (GCS). By appropriately evaluating the data types as mentioned earlier made available by the WELLBASE tool, stakeholders can make more informed decisions regarding well selections, risk assessment, and economic analysis for geologic carbon storage projects. Advanced Natural Language Processing models and other custom python scripts will be deployed in an automated process to extract unstructured data from documents, reports, and web applications and subsequently parse to more usable formats. The processed data will then be integrated into a robust and comprehensive database architecture, optimizing data accessibility, and usability for analytical purposes. The final data products will be accessible through a user-friendly visualization platform that will allow users to query and visualize the data, as well as download data in usable formats.

Tetteh, Daniel A.

Geoanalytical Evaluation of Saline Storage (GEESS) Geodatabase v2.0

The Geoanalytical Economic Evaluation of Saline Storage (GEESS) geodatabase was developed to support the United States Department of Energy (DOE) and National Energy Technology Laboratory (NETL) in their geologic carbon storage efforts by characterizing saline geologic formations present in the FECM/NETL CO2 Saline Storage Cost Model (CO2_S_COM) [1]. Using publicly available literature and data, the GEESS geodatabase characterizes 57 geologic formations across the lower-48 U.S. states in what are called Fully Integrated Geodatabases (FIGs). The FIG is a vector polygon feature containing thousands or tens of thousands of individual polygons, which each contain discrete geologic parameter values. A list of the critical geologic parameters that are characterized in the GEESS geodatabase are described in the “Processing Steps and Workflow” part of the ReadMe file, as well as the Data Catalog accompanying the GEESS geodatabase. The FIG is the basis of the GEESS geodatabase and is the direct representation of the collected geologic data. In addition to the FIG, the GEESS system contains grid files. Due to the complexity of the FIGs, grids are used to sample the geologic data so they can be exercised within CO2_S_COM. The grid files contain the geologic data sampled from the FIG, as well as estimates of “Plume Uncertainty Diameter” and “First-year Break-even Price of CO2” derived from CO2_S_COM based on the GEESS grid data.

carbon

Alabama Carbon Storage: Data Sharing and Engagement (Final Report)

This report is the final technical report on Alabama Carbon Storage: Data Sharing Engagement (ACS:DSE) project activities. The goals of the ACS:DSE project are to compile geologic, geophysical, infrastructure, and other relevant CCUS datasets for the study area and develop a geologic model of the study area; develop an online platform to serve data to stakeholders; engage with the public, students, and industry to educate them about CCUS and the data platform; and ensure energy and environmental justice is central to all aspects of the project. Datasets compiled and expanded include formation depths and elevations, digital geophysical well logs, reservoir properties, geologic structures, and geologic models. The geologic data were used to create a three-dimensional geologic model, structure grids, structure contour maps, and fault trace maps. In addition to downloadable datasets, links to CCUS relevant regulatory agencies (e.g., OGB, U.S. Environmental Protection Agency) and sources for infrastructure and educational information were included on the website Educational materials on CCUS for use by K-12 teachers were produced as part of the ACS:DSE project.

01 COAL, LIGNITE, AND PEAT

Stochastic Ensemble Generation for Improved Characterization of Representing Geologic Variability in a Reservoir: IBDP Case Study for SMART Initiative

This document is a poster covering the findings from activities on training data generation, specifically geologic ensemble generation. The generated geologic realizations captured the range of possible permeability distributions of the subsurface at the Illinois Basin - Decatur Project (IBDP) site, based on available well log variabilities. The percentages of reservoirs and baffles in the injection zone and a truncation of baffle permeability led to more variance in the simulations. This will be used to build forward modeling, history matching, and optimization workflows. The geologic realizations were also ranked according to dynamic measures of hydraulic diffusivity, and simulations confirm a greater contrast between the reservoir and the baffles during injection.

stochastic ensemble generation

HydroBio: Hydropower Capacity and Freshwater Biodiversity in Conterminous United States Sub-basins

This dataset summarizes existing and potential hydropower capacity and freshwater biodiversity at the sub-basin level throughout the conterminous United States (CONUS). It contains descriptive information regarding each sub-basin (e.g., 8-digit hydrologic unit code identifier, name, states, and size) along with sub-basin-level summaries of: 1) existing hydropower capacity (MW), 2) potential nominal non-powered dam (NPD) capacity (MW), 3) potential capacity of new stream reach development (NSD) (MW), and 4) freshwater biodiversity, including the total richness of fish, crayfish, and mussels and metrics that account for how rare and threatened those species tend to be. Hydropower data were obtained from Oak Ridge National Laboratory data resources (Existing Hydropower Assets, Non-Powered Dam Technical Potential, and New Stream Reach Development). Freshwater biodiversity data were obtained from NatureServe. Sub-basin characteristic information was obtained from the United States Geological Survey. Additionally, long data that provide lists of unique elements within each sub-basin for each constituent data resource (e.g., NatureServe, Existing Hydropower Assets) are provided to enhance dataset utility for users. The dataset provides, for the first time, a national-level assessment of existing and potential hydropower capacity in the context of freshwater biodiversity and is a valuable resource for stakeholders tasked with providing affordable, reliable energy to the American public while maintaining or enhancing invaluable freshwater resources. The dataset contains six data files in comma separated (*.csv) format that are within a zipped file.

Bozeman, Bryan [Oak Ridge National Laboratory (ORN

Carbon Storage Site Mapping Inquiry Tool (MapIT)

To date, 48 projects, consisting of 139 wells, are currently under review with the Environmental Protection Agency’s (EPA) Underground Injection Control (UIC) Program for Class VI – wells used for geologic sequestration of carbon dioxide. The number of applications submitted is expected to increase in coming years with the increase of the 45Q tax credit available to projects that initiate construction prior to 2033. The amount of data collected to submit a Class VI permit is vast, and often disparate, coming from state, federal, and commercial entities, as well as field-specific data collected within an area of interest. When preparing for site selection and permitting, the initial aggregation of relevant public data can be time intensive. The Carbon Storage Site Mapping Inquiry tool (MapIT) was created to support and accelerate the discovery and accessibility of open-source data and information available across the USA. Data was aggregated and organized based on data types described within the EPA UIC Class VI permit documentation. The online tool enables users to explore hundreds of geospatial data layers and connect to additional external resources, leveraging API and REST services where possible to ensure updates to data in real time. MapIT enables users to explore state and federal data related to geologic, geophysical, structural, hydrologic, and contextual information. In addition to displaying spatial data and linking to external resources, MapIT leverages custom widgets to ensure that internal data and external data are discoverable and accessible. The widgets connect users to resources such as the USGS publications and the USGS Earthquake Catalog based on a user-defined location. This talk will describe data aggregation workflows, data types, data preparation, and tool development for MapIT. The Carbon Storage Site Mapping Inquiry Tool and underlying database are valuable, intuitive resources that empower government, academic, commercial and industry stakeholders to explore, analyze, and acquire carbon storage related data.

Morkner, Paige

WHOLESCALE - Water & Hole Observations Leverage Effective Stress Calculations And Lessen Expenses (Final Technical Report 2020 - 2024)

The WHOLESCALE acronym stands for Water & Hole Observations Leverage Effective Stress Calculations and Lessen Expenses. The goal of the WHOLESCALE project is to simulate the spatial distribution and temporal evolution of stress in the geothermal system at San Emidio in Nevada, United States. To reach this goal, the WHOLESCALE team has developed a methodology to incorporate and interpret data from four methods of measurement into a multi-physics model that couples thermal, hydrological, and mechanical (T H-M) processes. The WHOLESCALE team has applied this methodology at the San Emidio geothermal field, located ~100 km north of Reno, Nevada in the northwestern Basin and Range province. The WHOLESCALE team includes 30 individuals working at two universities, two national laboratories, and one industry partner. Two master-degree students and five post-doctoral researchers have gained professional experience and earned partial financial support via the WHOLESCALE project. The WHOLESCALE team has taken advantage of the perturbations created by changes in pumping operations during planned shutdowns in 2016, 2021, and 2022 to infer temporal changes in the state of stress in the geothermal system at San Emidio, Nevada, U.S. The WHOLESCALE results support the working hypothesis that increasing pore-fluid pressure reduces the effective normal stress acting across fault zones. During normal operations, pumping in deep production wells decreases fluid pressures and thus increases the effective normal stresses on faults, reducing microseismicity. During planned shutdowns, the cessation of production increases pore-fluid pressure and reduces effective normal stress. The WHOLESCALE products generated during the 4-year period between 2020 and 2024 include: three articles published in the open-access, peer-reviewed scientific literature, two master’s theses, 20 presentations or papers at scientific conferences, and 17 data sets available on public repositories. The WHOLESCALE project has been completed in two phases that included three performance periods separated by two Go/No-go Stage Gate Reviews. Tasks were classified by data type (i.e., Geologic Structure, Borehole, Geodesy, Hydrology, Seismology, and Modeling). The first phase of the project started July 31, 2020 and included ongoing project coordination (Task 1), a project kickoff (Task 2), analysis of existing data (Task 3), development of the initial stress model & deployment design (Task 4), and Go/No-go Decision Point #1 (Task 5). Phase II began with implementing the 2022 deployment (Task 6), followed by Go/No-go Decision Point #2 (Task 7) The remainder of Phase II consisted of analyzing data collected during deployment (Task 8), calibration of the stress model on all observations (Task 9), and the Final Review (August 23, 2024) & Reporting (Task 10).

15 GEOTHERMAL ENERGY

A novel conditional generative model for efficient ensemble forecasts of state variables in large-scale geological carbon storage

Integrating monitoring data to efficiently update reservoir pressure and CO 2 plume distribution forecasts presents a significant challenge in geological carbon storage (GCS) applications. Inverse modeling techniques are commonly used to fuse observational data and refine reservoir model parameters, thereby improving state variable forecasts. However, these techniques often rely on linear or Gaussian assumptions, which can limit their effectiveness in accurately predicting state variables. Moreover, simulating large-scale three-dimensional (3D) GCS problems is computationally expensive, making iterative runs in inverse problems prohibitive. To address these challenges, we propose a conditional generative model utilizing the score-based diffusion method for real-time 3D pressure and saturation field distribution predictions. Our approach involves solving the score function with a mini-batch-based Monte Carlo estimator to generate labeled data. This data is subsequently employed to train a fully connected neural network, enabling it to learn the conditional sample generator within a supervised learning framework. This method enables the rapid generation of a large ensemble of predictions, facilitating comprehensive uncertainty quantification of state variables. Here we applied our method to forecast the dynamic 3D distributions of pressure and saturation fields over a 30-year injection period. The statistical assessment with low root mean square error (RMSE) values demonstrates that our method can accurately predict the spatiotemporal distributions of both pressure and saturation fields. Moreover, the developed conditional generative model shows high computational efficiency by generating 100 ensemble forecasts of 3D state variables in less than 10 min. The consistency between ensemble averages and ground truth values further illustrates the model’s capability to capture state variable dynamics during the CO 2 plume injection process. Notably, the ground truth values fall within the ensemble forecasts, indicating that our uncertainty quantification effectively captures variability and potential noise in the observations. Thus, the developed conditional generative model proves to be a more efficient, accurate, and practical tool for GCS applications, facilitating timely risk analysis and informed decision-making.

58 GEOSCIENCES

Monitored Natural Attenuation (MNA) Assessment for the Chemicals, Metals, and Pesticides (CMP) Pits Operable Unit (OU) and the Pen Branch Wetland

In October 2024, Savannah River National Laboratory (SRNL) was tasked to conduct an independent review of groundwater data and assess the monitored natural attenuation (MNA) performance related to the Chemicals, Metals, and Pesticides (CMP) Pits Operable Unit (OU) located in the central portion of the Savannah River Site (SRS). The SRNL assessment of MNA entailed an independent analysis of groundwater concentration data, groundwater elevation data, available surface water concentration data, soil concentration data, and relevant historical geological characterization logs for the CMP Pits OU. Historical data tables and records were obtained from both the Savannah River Nuclear Solutions - Area Completion Projects (SRNSACP) team and South Carolina State University (SCSU) and condensed into new data sets by the SRNL project team for more targeted analysis of MNA performance characteristics. All data pertaining to groundwater concentration, surface water concentration, and groundwater elevation were restricted to collection dates after any known active remediation for the CMP OU.

54 ENVIRONMENTAL SCIENCES

Characterization of Pliocene and Miocene Formations in the Wilmington Graben, Offshore Los Angeles, for Large-Scale Geologic Storage of CO2

The project Characterization of Pliocene and Miocene Formations in the Wilmington Graben, Offshore Los Angeles, for Large-Scale Geologic Storage of CO2 is one of 9 site characterization projects that were implemented as part of ARRA (American Recovery and Reinvestment Act). Data from this project was used to improve resolution of data in NATCARB in the area of study. Data related to this study has already been incorporated in NATCARB Atlas. The Los Angeles Basin presents an opportunity for large-scale geologic CO2 storage. Due to its large population and historical and geologic setting as one of the most prolific oil and gas producing basins in the United States, the region is home to more than 12 major power plants and oil refineries that produce more than 5 million metric tons of fossil fuel-related CO2 emissions each year. GeoMechanics Technologies worked to characterize the Pliocene and Miocene sediments in the Wilmington Graben, offshore of Los Angeles, California, for high-volume CO2 storage. The Graben is located offshore of the Los Angeles and Long Beach Harbor area, making it accessible yet geologically isolated from the nearby Wilmington oilfield and onshore areas. These sediments span more than 5,000 feet of vertical interval with an estimated storage resource of more than 100 million metric tons of CO2. The project team analyzed and interpreted existing geologic data within the region, including detailed exploration well log data and 2-D and 3-D seismic data. New seismic lines were acquired to fill in current data gap areas and two new characterization wells were drilled and logged. This information was integrated with existing geologic interpretations for adjacent onshore areas to help characterize optimal areas for CO2 storage and seals to safely store CO2. Integrated 3-D geologic and geomechanical models for the Wilmington Graben were developed to simulate the fate and transport of injected CO2 in the subsurface and to assess risks. This project contributed to the understanding of injectivity, containment mechanisms, rate of dissolution and mineralization, and storage capacity of the Wilmington Graben and associated analogous basins. This effort also provided greater insight into the potential for offshore geologic formations to safely and permanently store CO2.

.las

Machine Learning Applications in Analyzing the Role of Shale Barriers and Baffles for CO2 Storage

This study uses machine learning to analyze microseismic data from the Illinois Basin Decatur Project (IBDP) and quantify CO₂ plume extents. By leveraging well logs, microseismic records, and CO₂ injection metrics, the research predicts subsurface CO₂ plume dynamics. Findings show vertical clustering of microseismic events near the injection well, with CO₂ periodically breaching barriers due to buoyancy. K-Means clustering performed best, achieving the highest Silhouette Score and lowest Davies-Bouldin Index. This capability is crucial for real-time monitoring and management of CO₂ sequestration sites, validated against physical models and IBDP data, reinforcing CO₂ geological sequestration's viability and enhancing management tools.

Carr, Timothy