Search NASASearch

SEARCH · Search NASA

Results for “geospatial analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 91 records · Page 5

Machine-Learning-Based Mapping and Modeling of Solar Energy with Ultra-High Spatiotemporal Granularity

Despite the rapid growth of solar energy, we still lack a dynamic, high-fidelity database that tracks the spatiotemporal variations of solar PVs and their associated infrastructures across different places at a spatially resolved scale. The absence of such data presents a barrier to various applications such as solar PV growth projection, solar energy integration, solar incentive design, and climate risk assessment. In this project, we aim to bridge this gap by developing AI-based algorithms to extract granular information about solar PV installations and their associated infrastructures (i.e., distribution grids) from widely available unstructured data like remote sensing images and street views. As a result, we have built the Solar Energy Atlas, a fine-grained, large-scale geospatial overlay of distributed solar PVs and distribution grids. On top of it, we have advanced the understanding of solar adoption and distribution grid vulnerability to climate-induced extremes. Our major contributions can be summarized as follow: (1) By developing new AI algorithms, we have built the most comprehensive solar PV spatiotemporal database covering the entire US. This is the first time we obtained the exact GPS locations, size, subtype, and installation year information for rooftop solar PVs across the US. This database can be used for solar PV growth projection, solar energy integration, solar energy policy analysis and design, and spatially-resolved climate risk assessment. (2) Leveraging this database, we have uncovered the socioeconomic driving factors that are correlated with earlier onset of solar adoption and higher saturated adoption levels. We have identified the heterogeneity in the effects of different types of financial incentives on solar adoption and provided implications for tailoring incentive design based on local income levels to promote equitable solar adoption. (3) We have developed a distribution grid GIS mapping algorithm which can obtain granular geospatial and topology information about distribution grids using multi-modal open data, reducing the dependency on hard-to-obtain smart meter data of conventional approaches. It shows effectiveness in both the U.S. and Sub-Saharan Africa. Using this algorithm, we have uncovered the non-uniform vulnerability of distribution grids to wildfires in California in the aspects of undergrounding protection and Distributed Energy Resources (DER) preparedness. This has provided important implications for improving the affordability and equity of grid adaptation approaches. (3) We have made our produced database publicly available and provided user-friendly interface to enable various stakeholders and the general public to interact with the data. We have also integrated the produced data into the Data Commons platform to enable the public to access the data and correlate it with other location-specific characteristics simply using natural language as queries. The impact of our project is three-fold: (1) New algorithms for mapping solar PVs and distribution grids across space and time, which are open source to facilitate researchers and industry; (2) New databases of solar PVs and distribution grids that have been made publicly available for engineering, social, and policy applications; (3) New understandings and actionable insights on the potential approaches to promoting solar adoption and reducing energy infrastructure vulnerabilities. In this report, we start by discussing the project background and motivation (section 5), followed by the overview of project objectives (section 6). Results and discussion for each task are presented in section 7. Significant accomplishments are summarized in section 8. This report will be concluded by discussing the paths forwards (section 9), products (section 10), and team roles (section 11).

14 SOLAR ENERGY

Data for "Land-based Resources for Engineered Carbon Dioxide Removal in the United States Exceed the Expected Needs"

Gigatonne-scale atmospheric carbon dioxide removal (CDR), alongside deep emission cuts, is critical to stabilizing the climate. However, some of the most scalable CDR technologies are also the most land intensive. Here, we examine whether adequate land resources exist in the contiguous United States to meet CDR targets when prioritizing grid emissions reduction, food production, and the protection of sensitive ecosystems. We focus on biomass carbon removal and storage (BiCRS) and direct air capture and storage (DACS) and show that suitable lands exceed the expected needs: 37.6 million hectares of land are available for BiCRS, resulting in 0.26 GtCO2 of CDR/year, and 34 million hectares are suitable for wind- and solar-powered DACS, resulting in 4.8 GtCO2 of CDR/year if facilities are co-located with geologic CO2 storage. We identify biomass and energy supply hotspots to meet CDR targets while ensuring land protection and minimizing land competition.

carbon

Locating Undocumented Wells Using Historical Oil and Gas Exploration Maps: A Case Study in Osage County, Oklahoma

Undocumented oil and gas wells lack reliable information about their locations and characteristics, making them difficult to identify. These wells can result in unanticipated delays and costs in the development of nearby surface and subsurface resources, and, if improperly plugged, can cause contamination. This study leverages historical petroleum exploration maps to locate such wells, focusing on Osage County, Oklahoma. Two sets of early 20th century oil and gas exploration maps by the United States Geological Survey were georeferenced and analyzed using a computer vision model to detect well symbols. The locations of detected wells were compared to the location of known wells in the database from the Bureau of Indian Affairs Osage Agency to identify potential undocumented wells. The analysis yielded over 500 potential undocumented wells, with dry holes constituting the largest fraction. Field verification confirmed the presence of some undocumented wells. Comparison with prior work revealed limited overlap, underscoring the complementary value of historical oil and gas maps for locating undocumented wells. This approach demonstrates the utility of integrating historical cartographic resources with modern geospatial and machine learning techniques to improve the identification and management of undocumented wells.

Energy - Petroleum

Floating Photovoltaic Technical Potential: A Novel Geospatial Approach on Federally Controlled Reservoirs in the United States

Floating photovoltaic generation is a rapidly expanding sector of the solar energy industry, and understanding the quantity that can feasibly be installed is a crucial step to understand its role in future energy systems. This paper presents a novel spatially explicit methology of FPV potential for federally owned and managed reservoirs in the United States that uses site-specific attributes of reservoirs to estimate available area and potential generation capacity. The analysis finds that the proportion reservoir area that is found to be available for FPV development is similar to assumed values used in previous research on average, however there is a wide variability in this proportion on a site by site basis. Potential FPV generation capacity on these reservoirs is estimated to be in the range of 861 to 1,042 GWdc depending on input assumptions, likely representing a significant portion of future US solar generation needs. This work represents an advancement in methods used to estimate FPV potential that presents many natural extensions for further research.

floating solar

Meteoric 10Be Flux Calibration Data for the East River Watershed, Colorado, USA

This data package contains tabular and geospatial data used to quantify and model meteoric beryllium-10 fluxes in the East River watershed, Colorado, USA. The tabular component includes calibration-site data from five glacial moraine sites and includes environmental variables used to evaluate spatial controls on meteoric 10Be delivery, including elevation, mean annual precipitation (MAP), mean snow depth, and mean snow water equivalent (SWE). These site-level data were used to compare observed fluxes with environmental gradients across the watershed and to evaluate the effects of erosion correction on flux estimates. The package also includes supporting slope and curvature values used to assess topographic inputs to the erosion analysis. A second component of the data package contains updated manuscript tables and regression outputs used to summarize the relationships between meteoric 10Be flux and environmental predictors. These tables include meteoric 10Be sample information and AMS results, site-level environmental values, site-level meteoric 10Be inventory and flux values, watershed-averaged predicted fluxes, soil bulk density measurements, fine-fraction values, soil pH measurements, and regression statistics including slope, intercept, coefficient of determination, and p-value. The regression products include both standard linear regressions and regressions in which the intercept is constrained to pass through zero, and they support the analyses presented in the companion manuscript. Together, these tabular files provide the numerical basis for the manuscript tables and the regression-based interpretation of meteoric 10Be flux variability in a snow-dominated mountain watershed. The geospatial component of the package consists of GeoTIFF raster files used to generate the map products presented in Figures 2 and 6 of the companion manuscript. These rasters represent watershed-scale spatial layers for environmental variables and regression-based predictions of meteoric 10Be flux. This dataset contains comma-separated values files (.csv), Microsoft Excel files (.xlsx), GeoTIFF raster files (.tif), and upporting metadata files, including CSV data dictionaries and readme text files (.csv, .txt). The tabular files can be opened with standard spreadsheet software, and the raster files can be viewed and analyzed in GIS software such as ArcGIS Pro or QGIS. Together, these files document the numerical and spatial datasets used to calibrate and predict meteoric 10Be delivery in the East River watershed.

East River

Development of a Geothermal Module in reV: Quantifying the Geothermal Potential While Accounting for the Geospatial Intersection of the Grid Infrastructure and Land Use Characteristics: Preprint

The Renewable Energy Potential (reV) model is a geospatial platform for estimating technical potential and developing renewable energy supply curves, initially developed for wind and solar technologies. The model evaluates deployment constraints, considering land use, environmental, and cultural factors, and estimates the distance to existing grid features to connect future plants (Maclaurin et al., 2021). A pressing deficiency in the reV model, however, is representation of geothermal electricity generation technologies. To address this gap, we developed a novel geothermal generation module for reV that allows for representation and analysis at the same level of detail as other renewable technologies. This paper describes our process for evaluating data sources for the modeling, and presents five preliminary reV geothermal results. More specifically, we present two sets of resource data that represent upper and lower bounds for geothermal potential. We then present several sensitivity runs using the upper bound resource data; the results are encouraging that levelized cost of electricity (LCOE) can be reduced by optimizing the location and estimated capacity of the spatially diverse geothermal resource while considering the distance to existing grid infrastructure. Our preliminary supply curves and levelized cost of electricity (LCOE) results should be considered with care due to the highly uncertainty in geothermal resource potential data. We present median LCOE values for the conterminous U.S. for five scenarios: four hydrothermal (3.5km depth) and one EGS (4.5km depth). The capital and operating costs for each respective technology are modeled. We also compare results using two different resource data sources.

exclusions

A machine learning pipeline for identifying infiltration managed aquifer recharge locations from satellite imagery in the San Joaquin Valley, California

This study focuses on an agricultural region in California’s Central Valley, USA, where Managed Aquifer Recharge (MAR) is widely implemented to mitigate groundwater depletion under increasing water demand and climate variability. A deep learning and machine learning framework was developed to identify infiltration-MAR locations using satellite imagery and environmental data. The framework integrates surface water detection from Sentinel-2 imagery, geospatial delineation of water bodies, spatiotemporal tracking of water body dynamics, and supervised classification using meteorological, environmental, and topographic variables. The framework was applied to a 2379 km² study area southwest of Fresno, where 765 water bodies were detected, including 139 identified MAR sites based on publicly available datasets and expert knowledge. The classification model achieved an accuracy of 0.94 and an F1 score of 0.85. Feature importance analysis indicates that cropland, normalized difference vegetation index (NDVI), and evaporation are among the most influential predictors for infiltration-MAR. Notably, the framework suggests that engineered water management in infiltration-MAR systems can disrupt or even reverse the expected positive correlation between surface water extent and precipitation. These findings provide physically interpretable insights into the characteristics of existing infiltration-MAR facilities and demonstrate the potential of the proposed framework as a reproducible, interpretable, and potentially transferable tool for data-driven infiltration-MAR identification and inventory development under growing climatic and hydrological uncertainty.

Classification

Data from: "Towards CONUS-Wide ML-Augmented Conceptually-Interpretable Modeling of Catchment-Scale Precipitation-Storage-Runoff Dynamics"

This data package was generated to support the manuscript “Towards CONUS-Wide Machine Learning-Augmented Conceptually Interpretable Modeling of Catchment-Scale Precipitation-Storage-Runoff Dynamics.” It provides input files, model outputs, plotting data, scripts, notebooks, and documentation used to develop, evaluate, and reproduce Mass-Conserving Perceptron (MCP)-based hydrologic modeling experiments across 513 selected Catchment Attributes and Meteorology for Large-sample Studies in the United States (CAMELS-US) basins. The files are organized by modeling component and analysis purpose, including rainfall–runoff experiments, snow module experiments, coupled hydrologic-snow experiments, Long Short-Term Memory (LSTM) benchmark results, model skill metrics, initialization and epoch records, cell-state normalization files, Akaike Information Criterion (AIC)-based model comparison files, and data used to generate manuscript figures. Tabular files can be opened using standard spreadsheet software or Python/R data-analysis tools. Python scripts, Jupyter notebooks, and selected MATLAB scripts are included for model execution, postprocessing, plotting, and statistical analysis. Quality assurance and quality control were conducted through the source-data selection and modeling workflow. Meteorological forcing, streamflow, and static catchment attributes were derived from the CAMELS-US dataset, and snow water equivalent data were derived from the University of Arizona (UA) Snow Water Equivalent dataset. Selected basins and time periods were screened during the associated research workflow to avoid missing observations or poor-quality cases. Static geospatial features were processed primarily using Quantum Geographic Information System (QGIS) and Geospatial Data Abstraction Library (GDAL) workflows. Additional details are provided in the associated manuscript and documentation.

ESS-DIVE CSV File Formatting Guidelines Reporting

Qualitative Risk Assessment of Legacy Wells within the Estimated Prairie State Generating Company Area of Review

This report details the digitization of a legacy wellbore database, including data processing assumptions, parameter estimation, and risk assessment methodology. The database, comprising 6,454 documents, was provided by ISGS. It includes valuable data from the Prairie State Generating Company (PSGC) and One Earth Energy (OEE) sites of the CarbonSAFE Phase III – Illinois Storage Corridor project. The report focuses on wells within a 15-mile radius from the Lively Grove #1 (LG#1) well at PSGC site, evaluating subsurface conditions and potential risks. A total of 4,386 wellbores within 15 miles of the LG#1 well were filtered based on depth and formation codes. LG#1 is the stratigraphic well at the PSGC site drilled in 2021. Ninety-four (94) wells penetrating the Maquoketa Shale Group (the primary confining unit) within the estimated area-of-review (AoR) for the PSGC site were evaluated using a qualitative risk assessment (QRA) methodology. The QRA developed by Arbad et al. 2022 focuses on legacy wells within the AoR and categorizes them based on well construction details. The QRA identifies wells that need immediate attention by categorizing them based on penetration depth and protection. Wells within the AoR were categorized into nine groups based on penetrations and protections. These categories range from Type 1 wells, with no documentation, to Type 9 wells, which do not penetrate the primary confining unit or storage reservoir (unit). Well accessibility within the AoR varies based on well status, including Dry & Abandoned (DA), Plugged & Abandoned (PA), Injection (INJ), Oil/Gas Producing (PROD), and Observation (Obs) wells. Accessibility levels were determined by well construction, with DA wells being the least accessible and Observation wells the most accessible, impacting gas leakage detection possibilities. Remedial action priority of wells decreases from Type 1 to Type 9 wells. Type 1 to Type 6 wells with status DA and PA require immediate attention, while Type 7 and Type 8 wells are low priority. A risk matrix used to prioritize corrective actions for legacy wells is proposed to categorize wells within an AoR based on penetrations, protections, and accessibility. The methodology involves data acquisition, well categorization into nine types, and determining CO 2 leakage pathways using well schematics and geospatial mapping. This approach is particularly useful for managing the integrity of legacy wells throughout the lifecycle of a Carbon Capture and Storage (CCS) project. A qualitative risk assessment of 94 wells within the AoR of the PSGC site identified 54 wells with high priority for corrective action due to penetration of the primary containment seal. The assessment utilizes color-coded maps to categorize well types and prioritize corrective actions, providing a comprehensive analysis. Schematics of wells penetrating the primary confining unit were drawn, and leakage pathways were identified. Details of all wells penetrating the confining zone are provided in the appendix, including information on well types, plugging, and casing status.

01 COAL, LIGNITE, AND PEAT

Mauka Energy FEVER Tool Dataset

Mauka Energy’s dataset, developed under the Forestry Electric Vehicle Energy Routing (FEVER) project and funded by the U.S. Department of Energy’s Small Business Innovation Research program, is a high-resolution geospatial resource designed to support energy modeling for electric log trucks in complex forestry environments. The dataset integrates detailed spatial and road network data to enable accurate simulation of vehicle performance across varied terrain. At its core, the dataset incorporates lidar-derived elevation models, road alignments, and surface classifications from Oregon State University’s McDonald-Dunn Research Forest. These data capture fine-scale variations in slope, curvature, and surface conditions across forest road systems, allowing for vehicle-level analysis of energy consumption and recovery. The dataset also includes data collected on the surrounding public and private road networks in Benton County, Oregon, used in real-world haul routes. These connecting segments provide critical context for modeling transitions between forest operations and regional transportation infrastructure, incorporating attributes such as grade profiles, elevation change, and speed constraints. This combined dataset underpins the development of Mauka Energy’s rolldown tool, which quantifies energy use and regenerative braking potential on downhill and variable-grade segments. By leveraging high-resolution terrain and road data, the FEVER project enables more accurate assessment of electric vehicle feasibility and performance in forestry applications.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

Artificial Intelligence and Machine Learning Applications in Modern Power Systems

Machine learning (ML) and artificial intelligence (AI) algorithms offer valuable tools for the analysis and interpretation of large datasets. These tools have the capability to uncover insights that may not be readily apparent within these datasets. In recent years, the integration of ML and AI has become increasingly prevalent in various applications within the power system domain. One of the earliest instances of machine learning in power systems can be traced back to demand forecasting, where artificial neural networks were employed for short-term load forecasting. In contemporary power systems, an abundance of high-resolution geospatial and temporal data is generated at various time intervals, ranging from sub-seconds (Phasor Measurement Units or PMUs) to seconds (Supervisory Control and Data Acquisition or SCADA), minutes (Process Information or PI), and extending to days, months, and years. These datasets contain valuable information concerning system reliability and performance. This information holds the potential to offer critical insights into system operations, as well as solutions for predicting and mitigating contingencies to prevent cascading outages. Despite the immense power of machine learning tools, system operators, planners, and utilities often exhibit hesitancy in fully embracing AI-enabled system operations and planning. This cautious approach persists, even as numerous diverse applications of machine learning continue to emerge in the realm of power systems. In this chapter, our focus will delve deep into ML and AI applications tailored for power systems. These applications aim to furnish system operators with enhanced situational awareness and augment their decision-making capabilities, especially during challenging operating conditions. Specific areas of interest encompass root cause analyses of electricity market datasets and the strategic selection of representative samples from vast power system databases for training ML/AI models. Finally, the chapter will conclude with a short discussion on the future of ML/AI in power systems and possible directions that the industry is moving towards.

power system applications, machine learning (ML),

Geothermal Power Systems Analysis: Outcome of Industry Stakeholders Workshop: Preprint

Geothermal cost and performance evaluation implemented via technoeconomic assessment (TEA) modeling is critical for the Department of Energy (DOE) and other geothermal industry stakeholders in assessing the current state of geothermal technologies and to identify existing hurdles to commercially viable geothermal development. The Geothermal Electricity Technology Evaluation Model (GETEM) is a major TEA tool used in estimating the economic feasibility and levelized cost of energy (LCOE) of conventional hydrothermal systems and enhanced geothermal systems (EGS). Since 2021, GETEM has been transitioning from an intricate spreadsheet model to a user-friendly tool within the System Advisor Model (SAM) developed by the National Renewable Energy Laboratory (NREL). Apart from enabling an expanded visibility of the geothermal model among other renewable resources, having GETEM in SAM has the advantage of simulation automation, better usability, updates tracking, active user inputs/feedback, and extended financial modeling. GETEM is used in developing supply curves for the Annual Technology Baseline (ATB). The ATB data are inputs to the Renewable Energy Potential (reV) and the Regional Energy Deployment System (ReEDS) models. The geothermal module in NREL’s reV model assesses the geothermal energy potential in the conterminous United States by defining the geospatial intersection of geothermal resources with existing grid infrastructure within the constraint of land use characteristics. The ReEDS model is a capacity expansion model used for simulating the long-term build-out and operation of the US generation and transmission system based on current energy costs and policies. To ensure enhanced representation of current industry trends in our model transitions and development, we organized a two-day virtual workshop to elicit geothermal industry stakeholder input and recommendations on our current approaches and assumptions on technoeconomic, resource assessment, and deployment scenarios modeling of geothermal technologies. Participants included developers, operators, investors, regulatory agencies, system modelers, national laboratory researchers, consultants, and other stakeholders. In this workshop, we gained stakeholder insights on current geothermal plant performance (i.e., capacity factors), updated drilling costs and learning curves, and next generation technologies such as closed loop and superhot rock geothermal. Other outcomes from this workshop and its impact on future geothermal development feasibility, resource availability, and capacity expansion studies are compiled and discussed.

Annual Technology Baseline

Historic climate, cosmogenic 10Be, denudation-rate, and geospatial datasets from the Pikes Peak region, Colorado, USA

This data package contains geographic information system (GIS) layers and tabular datasets associated with the study of elevation-dependent denudation rates on Pikes Peak in the Front Range of the Rocky Mountains, Colorado, USA. The package includes GIS layers used to produce the study-area map, including sample locations, sample watershed boundaries, the Pikes Peak batholith, Pleistocene glacier extent, weather station locations, and elevation and hillshade rasters, together with comma-separated value (CSV) tables and matching CSV data dictionaries. These mapped layers provide the geographic framework for interpreting denudation patterns across the Pikes Peak region and for relating sample locations to watershed geometry, bedrock setting, glacial history, and nearby climate stations. The first group of tables reports climate and geospatial context for the study area. These files include station-based temperature and precipitation data used to characterize elevational gradients in mean annual climate and monthly climate seasonality, sample locations, denudation-rate and topographic metrics, fixed frost-cracking model parameters, frost-cracking intensity and precipitation-frequency metrics, and stream-power inversion results. Together, these data provide the basis for evaluating how denudation varies with elevation, climate, and landscape form across sampled catchments on Pikes Peak. The second group of tables reports cosmogenic nuclide and erosion-model results used in the denudation analysis. Included files contain accelerator mass spectrometry (AMS) measurements for in situ-produced cosmogenic beryllium-10 (10Be), including sample identifiers, measured 10Be:9Be ratios, analytical uncertainties, carrier mass, quartz mass, blank corrections, blank-group statistics, and calculated 10Be concentrations and uncertainties. Additional tables summarize stream-power-law inversion results for sampled catchments, including optimized model parameters, predicted erosion rates, residual metrics, channel-pixel counts, and convergence status, as well as regression equations and summary statistics used to evaluate relationships among elevation, climate, frost cracking, precipitation forcing, and denudation rate. The package contains GIS files, comma-separated value files (.csv), Microsoft Excel files (.xlsx), CSV data dictionaries, a file-level metadata table, and a readme text file.

10Be cosmogenic nuclides

Geothermal Power Systems Analysis: Outcome of Industry Stakeholders Workshop

Geothermal cost and performance evaluation implemented via techno-economic assessment (TEA) modeling is critical for the U.S. Department of Energy (DOE) and other geothermal industry stakeholders in assessing the current state of geothermal technologies and to identify existing hurdles to commercially viable geothermal development. The Geothermal Electricity Technology Evaluation Model (GETEM) is a major TEA tool used in estimating the economic feasibility and levelized cost of energy (LCOE) of conventional hydrothermal systems and enhanced geothermal systems (EGS). Since 2021, GETEM has been transitioning from an intricate spreadsheet model to a user-friendly tool within the System Advisor Model (SAM) developed by the National Renewable Energy Laboratory (NREL). Apart from enabling an expanded visibility of the geothermal model among other renewable resources, having GETEM in SAM has the advantage of simulation automation, better usability, updates tracking, active user inputs/feedback, and extended financial modeling. GETEM is used in developing supply curves for NREL's Annual Technology Baseline (ATB), which provides inputs to the Renewable Energy Potential (reV) and the Regional Energy Deployment System (ReEDS) models. The geothermal module in NREL's reV model assesses the geothermal energy potential in the conterminous United States by defining the geospatial intersection of geothermal resources with existing grid infrastructure within the constraint of land use characteristics. The ReEDS model is a capacity expansion model used for simulating the long-term build-out and operation of the U.S. generation and transmission system based on current energy costs and policies. To ensure enhanced representation of current industry trends in our model transitions and development, we organized a two-day virtual workshop to elicit geothermal industry stakeholder input and recommendations on our current approaches and assumptions on techno-economic, resource assessment, and deployment scenarios modeling of geothermal technologies. Participants included developers, operators, investors, regulatory agencies, system modelers, national laboratory researchers, consultants, and other stakeholders. In this workshop, we gained stakeholder insights on current geothermal plant performance (i.e., capacity factors), updated drilling costs and learning curves, and next-generation technologies such as closed-loop and superhot rock geothermal. Other outcomes from this workshop and its impact on future geothermal development feasibility, resource availability, and capacity expansion studies are compiled and discussed.

annual technology baseline

PSH Assessment and Site Identification [Slides]

NLR's Pumped Storage Hydropower (PSH) geospatial and cost model algorithms are applied to the Chemehuevi Reservation to assess the potential for PSH within the reservation. The algorithm identifies both "open-loop" PSH opportunities formed by constructing a new reservoir within the reservation paired with bordering Lake Havasu and "closed-loop" systems formed by two new reservoirs within the reservation. Options range from 28 to 538 MW of electrical generation power at maximum generation and 10 hours of storage. CAPEX is estimates as 4422 2022 $\$$/kW of generating power, which is meaningfully higher than the lowest cost systems identified in NLR's national scale assessments. Further economic analysis is needed to fully evaluate whether this would be an attractive option to meet the Chemehuevi Reservation's goals.

25 ENERGY STORAGE

A Data Processing Pipeline To Extract A Knowledge Graph From Heterogeneous Data For Socio-technical Analysis Of Critical Infrastructure Influence

The code is written in Python and consists of the following pipeline that is implemented in Apache Airflow. This pipeline intends to understand the companies that are directly or indirectly involved with a type of critical infrastructure system at some point in that system's lifecycle. The pipeline takes a configuration file that specifies a list of initial companies to consider, a geographic region of interest, and a set of SEC form types as well as other data sources (e.g. CrunchBase) from which to extract entities and relations. There are four main components to this pipeline as currently implemented: Entity Extraction, Network Construction, Analysis, and Visualization. First, Entity Extraction, is implemented as the `topear-extract_organizations` Apache Airflow workflow. Given an initial query that specifies a geographic region of interest and a time interval, the software will extract CI facilities of interest and organizations that have a direct influence relationship to those facilities (e.g. ownership). During the course of the LDRD, we focused on Electric Vehicle charging stations and this information is available via the Department of Energy (DOE) database on fueling stations maintained by NREL. Within the context of the DOE CESER project, we have focused on Battery Energy Storage Systems (BESS). Second, the Network Extraction component will iteratively construct a social network graph given the set of organizations and people extracted in the previous step. Organizations (and eventually People if desired) are then fed as a query to the `topgear-construct_social_network` Apache Airflow workflow which given a set of initial companies and data sets (e.g. SEC EDGAR form types, OpenCorporates, Crunchbase). This Airflow workflow will iteratively query such data sources to discover relationships with new organizations and people. For example, this module can iteratively query SEC EDGAR for metadata that documents the number of each type of form for the given set of companies and their location. This forms metadata represents a catalog of data sources from SEC EDGAR for the extracted social network knowledge graph. The pipeline then downloads these forms from the website and saves them in a build directory for further processing. These documents are then parsed for entities and relations. Again, we note that in additional to SEC data sources, this step can also pull in information on organizations via API services such as CrunchBase and OpenCorporates or bulk data sources. At the end of this step, the resultant social network, the Critical Infrastructure network, and the edges that encode relationships between organizations and CI facilities, form the Adversarial Socio-Technical Network (ASTN) that informs the analysis. Third, the Analysis component processes these generated ASTN. Previously, that has included the ability to compare prevalence of different vendors for a given infrastructure component type across different regions as well as identify common public and private investors across those vendors. This was demonstrated for EV Charging Stations across several different metropolitan areas within an IEEE PES GridEdge publication. More recently, we have looked at ways to identify infrastructure owners and operators of BESS with the most nameplate capacity across different states as well as other indictors of risk resulting from changes in ownership over time. Finally, the Visualization component consists of an HTML/CSS/JS framework by which users can interact geospatial, operational, and organizational relationships across a given portfolio of Critical Infrastructure facilities. The objective is to provide a library of UI/UX modules that can be repurposed for stakeholder-specific dashboards. All of the modules are related via a common event model that enables UI actions in one view to percolate across the other views.

Weaver, Gabriel [Idaho National Laboratory (INL),

AquaPV: Foundational Analysis and Industry Guidance on Floating PV Results

This dataset contains results of the technical potential analysis of floating photovoltaics (FPV) on federally owned or permitted reservoirs in the United States. Estimates of the area of reservoirs that is technically feasible for FPV development are provided for reservoirs that are owned by the US Army Corps of Engineers or the US Bureau of Reclamation, or are associated with hydropower dams licensed by the Federal Energy Regulatory Commission. Other associated data are included, such as estimates of the potential evaporative losses of FPV development, and whether the waterbody is associated with an energy community or disadvantaged community that could qualify it for investment or production tax credits. Data is provided both as a spreadsheet and in a geospatial format including the waterbody geometries. Please read the included readme for more detailed field definitions.

area

BAMCensus (The Behavior and Advanced Mobility Census Dataset Aggregator) [SWR-25-120]

This software is a high-performance tool developed in Rust for downloading and processing large-scale geospatial datasets, specifically focusing on US Census data. It is designed to address scaling limitations found in existing tools, such as R's [tidycensus](https://walker-data.com/tidycensus/), by providing performant streaming dataset JOIN operations between various US Census datasets (like ACS and LEHD) and their corresponding geometries stored on the TIGER/Lines web server. The tool automates the process of joining these data sources, returning aggregated data to the user based on a specified census GEOID type. The tool automates the process of joining these data sources, returning aggregated data to the user based on a specified census GEOID type. Its primary motivation stems from the need for a high-performance solution to combine spatial datasets with graph traversals within the context of mobility analysis tooling being developed at NREL's Behavior and Advanced Mobility (BAM) group.

Fitzgerald, Robert [National Renewable Energy Labo