Search NASASearch

SEARCH · Search NASA

Results for “geospatial analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Biochemical Conversion of Herbaceous Biomass to Renewable Diesel: Biorefinery Marginal Air Quality Impacts and Comparison to Feedstock Production

This study assesses the air quality impacts of an advanced biorefinery that produces renewable diesel blendstock (RDB) from lignocellulosic biomass via aerobic respiration (Davis et al. 2022) by estimating fine particulate matter (PM2.5) impacts from biorefinery emissions. It continues a prior analysis that used a geospatial assessment to identify source regions for biomass feedstocks and studied the impact of feedstock production emissions on air quality (Thind et al. 2022). Thind et al. (2022) identified RDB biorefineries that can use corn stover feedstocks of 2,000; 5,200; and 9,100 dry metric tons per day (DMT/day), based in Iowa, and suggested 7 unique counties can serve as hosts for a biorefinery that draws biomass feedstock from neighboring counties. Given 13 unique county-biorefinery size combinations and two waste lignin end uses at the biorefinery (lignin as a fuel for electricity generation and lignin for pellet production), the air-quality-related sustainability aspects of each of these 26 scenarios are assessed by estimating the annual average impacts of biorefinery emissions on the dispersion and formation of secondary PM2.5 in the atmosphere using a novel reduced-complexity air quality model called the Intervention Model for Air Pollution (InMAP). The 26 biorefinery design combinations help capture how a biorefinery's emissions of air pollutants and their resulting impact on local and regional air quality are influenced by the magnitude of production scale, lignin utilization strategy, and location of a proposed biorefinery. Methods developed in Thind et al. (2022) are applied to estimate the constraints on primary PM2.5 and secondary PM2.5 precursor emissions based on compliance with U.S. Environmental Protection Agency's (EPA's) annual primary National Ambient Air Quality Standard (NAAQS) of PM2.5 (i.e.12.0 micrograms per cubic meter (microgram/m3)) at downwind receptors of a biorefinery. Incremental PM2.5 concentrations caused by the emission of biorefining corn stover into RDB are assessed and compared to those of corn stover production. To illuminate which upstream supply chain stage of renewable diesel production contributes most to air quality impacts, marginal PM2.5 concentrations are compared between both stages at multiple downwind air quality monitor locations. In addition, through a hotspot analysis, we identify the primary contributing factors of emissions within the feedstock production and biorefinery stage operations. In doing so, we provide insights for improving the air pollutant emission-related sustainability of advanced lignocellulosic biofuel production.

09 BIOMASS FUELS

Renewable hydrogen horizon: Geospatial techno-economic feasibility and life cycle greenhouse gas analysis in the Middle East and North Africa

Renewable hydrogen is receiving increasing attention for its potential as a flexible energy carrier in sectors such as transportation and industry. Specific cost and carbon intensity (CI) of renewable hydrogen production vary largely based on the location, owing to differences in renewable energy resources, as well as the supply chain dynamics. This study maps the techno-economic and life cycle greenhouse gas emissions of renewable hydrogen production in the Middle East and North Africa region, leveraging abundant solar and wind resources. The work investigates the variability in hydrogen costs and CI, optimally sizing proton-exchange membrane (PEM) electrolyzers to account for partial and cyclic loading, and explores standalone versus grid-connected systems. PEM capacity ratios of 52 %–63 % for photovoltaic (PV) systems and 28 %–82 % for wind systems were identified as optimal, with hydrogen production costs ranging from $\$3.8$-$\$4.8$/kg for PV and $2.0-$7.0/kg for wind. CIs span from 1.9 to 3.7 kg CO 2 ,eq /kg H 2 for PV and 0.4–7.7 kg CO 2,eq /kg H 2 for wind systems. The study highlights significant cost and CI reductions achievable with technological advancements and co-product revenue from oxygen and excess electricity sales.

Carbon Intensity

More land is needed for solar and wind infrastructure under a high renewables scenario in the Western US by 2050

Expanding United States electricity infrastructure to meet growing demand could require extensive power plant development footprints and land use conversion, depending on the mix of generation types chosen. Understanding where future power plant sitings are likely to take place and identifying potential conflicts and land-use tradeoffs will be key to identifying feasible and affordable investments and evaluating regional planning coordination needs. Here we use an integrated modeling framework that combines capacity expansion planning, hourly grid operations, and geospatial techno-economic analysis to develop projections (2025-2050) of power plant sitings in the Western United States (US) at a 1 km 2 resolution for a business-as-usual scenario and a high renewables penetration scenario. We find that 30% more land will be needed in the high renewables scenario as compared to business-as-usual, and that 75% of that development is projected to be located within 10 km of natural areas.

Mongird, Kendall [Pacific Northwest National Labor

Vegetation classification map and covariates associated with NEON AOP survey, East River, CO 2018

This package includes geospatial data layers developed to investigate how environmental gradients—specifically topography and near-surface soil properties—drive the spatial arrangement of dominant plant communities in mountainous watersheds. The geospatial products, which support the analysis of these ecological relationships, are derived from airborne hyperspectral and LiDAR datasets acquired by the National Ecological Observatory Network (NEON) Airborne Observation Platform (AOP), in conjunction with an extensive ground field campaign conducted in summer 2018. This work is part of the DOE Watershed Function Science Focus Area (SFA) and features geospatial datasets developed based on observations and ground data collected at East River, Colorado, in collaboration with the National Ecological Observatory Network (NEON) Airborne Observation Platform (AOP) survey in June 2018. Classification Map: - Classification Map (PNG, GeoTIFF): Derived from hyperspectral and LiDAR airborne data using a machine learning approach. - Class Code Mapper (CSV): Associates pixel values with corresponding vegetation/non-vegetation classes. - Classification Reference Data (CSV): Reference data used in the machine learning procedure. LiDAR-Derived Products: - Topographical Metrics (GeoTIFFs): Elevation, slope, curvature, TWI, TPI, solar insolation, and canopy height model (CHM), smoothed with a 5x5 pixel window. Vegetation Indices: - GeoTIFFs of NDVI, NDNI, NDWI: Vegetation indices derived from hyperspectral data. Urban Masks: - Urban Mask (GeoTIFF): Applied to the mapping to convert bare soil classes to urban classes. Software Compatibility: GeoTIFFs: Can be visualized with GIS software or libraries that support GeoTIFF images. CSV Files: Can be opened with any software that handles comma-separated values. The FLMD file provides details and links to the source datasets used to derive the products. The manuscript (in the Method session) provides details on how each product was derived. This work was supported by the Watershed Function Science Focus Area at Lawrence Berkeley National Laboratory funded by the US Department of Energy, Office of Science, Biological and Environmental Research under Contract No. DE-AC02-05CH11231. Update on 2026-03-25: Since the original dataset publication date of 02/28/2020, this package has a new classification map derived by an improved methodology. This update also includes additional ground data that improved the representation of some of the communities. See the methods for further details on what has changed between versions.

2018 NEON and 2025 CHESS Campaigns

Developing Data-Driven Synthetic Infrastructure Models for Resilience Analysis

Research on infrastructure resilience has produced promising methods to simulate and optimize complex networks to improve performance. However, restrictions on sharing infrastructure models and the steep cost of developing and maintaining infrastructure models presents a roadblock to adoption. To overcome this limitation, this research focuses on methods to create data-driven infrastructure models that will help improve infrastructure resilience and security. The analysis couples incomplete utility data, geospatial data, machine learning, and synthetic network generation methods to rapidly develop and update infrastructure models. The methods are validated using realistic utility models and site-specific data, with a focus on Puerto Rico due to its unique infrastructure challenges and available data. This research highlights promising opportunities for the use of synthetic network generation and machine learning to create infrastructure models when very little data is available. Results demonstrate that hybrid methods, which combine sparse utility data with synthetic models, can enhance model accuracy, and machine learning can predict model attributes using training data from other models. However, the complexity of infrastructure systems means that even minor changes in network connectivity can significantly impact simulation results. Resilience analysis using synthetic infrastructure models shows that while some system behaviors are preserved, the magnitude of disruptions may not be accurately represented, indicating the need for more research and validation before using synthetic models for critical infrastructure investment decisions. The framework outlined in this report represents a significant advance to infrastructure model development and could be applied to additional domains and sites. Future research will continue to streamline and validate methods to help reduce roadblocks to resilience analysis.

24 POWER TRANSMISSION AND DISTRIBUTION

Data from: 'Abiotic influences on continuous conifer forest structure across a subalpine watershed'

This package archives the core data used for analysis and inference in 'Abiotic influences on continuous conifer forest structure across a subalpine watershed' (Worsham et al., 2025). All data were collected in the East River, Washington Gulch, Slate River, and Coal Creek watersheds of Colorado. In the paper, we quantified the relative influence of climate, topographic, edaphic, and geologic factors on conifer stand structure and composition, and their functional relationships, at the watershed scale. We used waveform LiDAR data to derive spatially continuous stand structure metrics. We fused these with a species-level classification map to estimate tree species abundance. We applied generalized additive and generalized boosted models to evaluate the covariability of structural and compositional metrics with abiotic variables. The package contains the essential products required for reproducing our analysis and the tables and figures reported in the publication. The products comprise four classes: (1) geospatial data, (2) tabular data used for inferential analysis, (3) tabular data describing analytical results and performance statistics, and (4) a data user guide. (1) includes discretized waveform LiDAR data, locations and attributes of individual tree crowns, sampling locations and domain boundaries, a canopy height model, and raster files of estimated forest structural and compositional metrics at 100 m grid scale. (2) includes all response and explanatory variable values applied in inferential models. Response variables include conifer forest stand density, basal area, 95th percentile height, quadratic mean diameter, and others. Explanatory variables include climatic water deficit, actual evapotranspiration, elevation, heat load, soil available water content, and others. (3) includes results of training and testing several individual tree detection (ITD) algorithms, as well as inferential modeling results. (4) is a PDF user guide for this data package, including detailed descriptions and data dictionaries for all files. The data package root contains 17 assets: 8 compressed tape archive (.tar.gz) files, 5 comma-separated values (.csv) files, 3 Geographic Tagged Image File Format (GeoTIFF) (.tif) files, and 1 Portable Document Format (.pdf) file. The compressed .tar.gz archives contain ESRI shapefiles (.shp) .tif, compressed LASer (.laz), and .csv files. The archives must first be decompressed using the widely distributed command-line software utility TAR. All other files, including constituent files within the .tar.gz archives, can be opened in the open-source R statistical computing environment. Alternatively, .csv files may also be read in any simple text editor software or Microsoft Excel. Geospatial files including .shp and .tif files can also be opened in GIS software, such as QGIS (open-source) or ESRI ArcGIS (proprietary). The .pdf Data User Guide can be read with Adobe Acrobat Reader or other compatible readers.

2018 NEON and 2025 CHESS Campaigns

Evaluating opportunity for distributed wind energy in rural and agricultural areas

Wind energy is among the most mature renewable energy technologies, accounting for 11% of the current US electricity generation in 2024, with the lowest average levelized cost. While it is known that substantial opportunity exists for further development, a key question has been where wind energy is best suited compared to other technologies. This study leverages an immense dataset of parcel-resolved technoeconomic potential for the contiguous United States, focusing on distributed wind (DW) energy—a configuration where one or more turbines, typically 30–60 m in height are used to satisfy nearby energy needs. The analysis is conducted at multiple spatial scales and considers land use, crop land, census, and incentive program data to determine the most opportune areas for market development. The results show that rural, agricultural and residential areas are most suited to DW. Connection type (in front of, or behind the meter) and regulations determine the best application, while siting constraints, economics, demand and the wind resource determines the optimal size of turbine.

17 WIND ENERGY

Where to cool off: a geospatial framework for placement of cooling centers

Indoor cooling is essential to reduce heat stress and increase passive survivability during heatwaves. Although air conditioning (AC) is recommended for maintaining indoor thermal comfort, low- and medium-income households in the U.S. often do not own an AC and/or limit AC usage to reduce energy consumption and associated costs, thereby risking their health and safety. With the frequency and intensity of heatwaves increasing, cooling centers are considered an appropriate alternative to indoor cooling and a possible mitigation strategy to prevent adverse health impacts of heat exposure. However, these centers are limited in numbers and not always accessible. This requires (i) developing a geospatial framework using physical and social factors for optimal siting of cooling centers to meet future needs and (ii) ranking of existing and potential cooling centers (schools, libraries, religious institutions) based on their accessibility among vulnerable populations and proximity to healthcare facilities. We developed and deployed a geospatial framework based on the Multi-criteria Decision Analysis approach in five U.S. cities (Los Angeles (LA), Phoenix, Austin, Atlanta, Miami) to evaluate the effectiveness of the framework in ranking cooling centers based on accessibility and population coverage. The results revealed that (i) access to cooling centers varies across cities and 32.2–50.7% of centers are within walking distance of the most vulnerable populations, (ii) vulnerable populations exposed to Urban Heat Island (UHI) effects are more likely to experience energy burden, and (iii) about 21.2–49.4% of population with high energy burden have access to these centers. Considering that more cooling centers are needed to assist energy burdened households alleviate heat exposure impacts, the framework developed herein could be adapted to incorporate other factors (e.g. health impacts, policies) to assess site suitability of existing shelters, identify potential sites for new cooling centers, and geo-target communities where energy efficient emerging technologies could be deployed to reduce heat stress.

58 GEOSCIENCES

Towards Prospective LCA Using Life-Cycle Assessment Integration into Scalable Open-Source Numerical Models (LiAISON) Framework for Analyzing Emerging Low-Carbon Technologies

Decarbonizing the industrial sector is a significant challenge in achieving a net-zero greenhouse gas (GHG) emissions economy by 2050 and the Paris Agreement, i.e., a global climate change mitigation target of achieving a maximum average temperature change potential of 1.5 Degrees Celsius or less by 2100 with respect to pre-industrial levels. In the United States (US), the industrial sector accounts for 23% of total GHG emissions and is home to a number of hard-to-electrify activities. The chemicals subsector has the single largest subsector emissions profile after direct emissions from fossil fuel combustion and leakage from fossil fuel distribution systems. Within the chemicals subsector, many processes depend on hydrogen or ammonia precursors. Decarbonizing these two commodities would contribute significantly to decarbonizing the industrial sector as hydrogen could also be used for low carbon steel production (e.g., hydrogen-based direct reduction of iron) and other industrial applications. Emerging technologies require the application of prospective life cycle assessment (LCA), which can account for technology (foreground) scaling and process improvements via learning-by-doing, among others. In many cases, the future system context (background) in which the technologies are assumed to operate in is equally relevant. Background scenarios generated by integrated assessment models (IAM) can coherently incorporate potential future dynamics of the energy-climate-human-land system. Further, IAM scenarios are harmonized across socioeconomic and climate change mitigation pathways, which facilitates the comparability of prospective LCAs using different IAMs. We introduce an open source prospective LCA framework, the Life-cycle Assessment Integration into Scalable Open-source Numerical models (LiAISON), to analyze the non-linear relationships between technology foreground and the future energy system background across a series of midpoint and resource use metrics The integration of LCA and IAM data is achieved using prospective environmental Impact assessment (PREMISE). We showcase it by assessing two Power-to-Hydrogen (PtH2) processes, namely Solid Oxide Electrolysis (SOE) and Polymer Electrolyte Membrane Electrolysis (PEME). We compare the technologies to a baseline of hydrogen production via natural gas-based Steam Methane Reforming (SMR) in a US context of multiple energy system and climate change mitigation futures. Besides providing an analysis that specifies the LCA results ranges with temporal and geospatial explicitness across the two technologies, metrics, and impact assessment methods, this research also aims to establish a base framework that can be expanded to use other IAM generated scenarios and US open-source life cycle inventory (LCI) databases. We find that the temporal environmental performance of either technology or their difference to SMR is directly influenced by the underlying background dynamics. Additionally we compare our results by linking two other prospective models with LiAISON - GCAM(Global Change Assessment Model) and ReEDS (Regional Energy Deployment System) to analyze the effect of changing background scenarios using varying predictions in life cycle analysis.

emissions

An Analysis of Future Wind Energy Resources and Cost Uncertainties Across the United States

Wind power is a growing source of energy generation that relies on complex global atmospheric and earth system processes. There has been evidence of reductions in average wind speeds over land in North America since the 1980s, and several models project that average wind speeds will continue to decrease. Concurrently, the cost of wind energy systems in the United States has been decreasing since around 2010, a trend also projected to continue. There is considerable uncertainty in these future projections, with quantitative estimates of future wind resource and system costs varying widely. To study this, we run wind energy models with possible future system costs, turbine designs, and meteorological inputs from multiple downscaled earth system models over the contiguous United States. Changes in mean annual energy production from 2000-2019 to 2040-2059 can be as high as +10% in South Texas or as low as -20% in Iowa. Larger turbines and moderate reductions in system costs can offset the largest projected decreases in wind resource, but much uncertainty remains in the extent to which wind resources will change and to what extent system costs can be reduced. An analysis of variance shows, in several states in the Midwest, uncertainty in future wind resources can influence the cost of wind energy nearly as much as uncertainty in future system costs.

17 WIND ENERGY

Life-Cycle Assessment Integration into Scalable Open-Source Numerical Models (LiAISON) for Analyzing Emerging Low-Carbon Technologies

Decarbonizing the industrial sector is a significant challenge in achieving a net-zero greenhouse gas (GHG) emissions economy by 2050 and the Paris Agreement, i.e., a global climate change mitigation target of achieving a maximum average temperature change potential of 1.5 degrees C or less by 2100 with respect to pre-industrial levels. In the United States (US), the industrial sector accounts for 23% of total GHG emissions and is home to a number of hard-to-electrify activities. The chemicals subsector has the single largest subsector emissions profile after direct emissions from fossil fuel combustion and leakage from fossil fuel distribution systems. Within the chemicals subsector, many processes depend on hydrogen or ammonia precursors. Decarbonizing these two commodities would contribute significantly to decarbonizing the industrial sector as hydrogen could also be used for low carbon steel production (e.g., hydrogen-based direct reduction of iron) and other industrial applications. Emerging technologies require the application of prospective life cycle assessment (LCA), which can account for technology (foreground) scaling and process improvements via learning-by-doing, among others. In many cases, the future system context (background) in which the technologies are assumed to operate in is equally relevant. Background scenarios generated by integrated assessment models (IAM) can coherently incorporate potential future dynamics of the energy-climate-human-land system. Further, IAM scenarios are harmonized across socioeconomic and climate change mitigation pathways, which facilitates the comparability of prospective LCAs using different IAMs. We introduce an open source prospective LCA framework, the Life-cycle Assessment Integration into Scalable Open-source Numerical models (LiAISON), to analyze the non-linear relationships between technology foreground and the future energy system background across a series of midpoint and resource use metrics The integration of LCA and IAM data is achieved using prospective environmental Impact assessment (PREMISE). We showcase it by assessing two Power-to-Hydrogen (PtH2) processes, namely Solid Oxide Electrolysis (SOE) and Polymer Electrolyte Membrane Electrolysis (PEME). We compare the technologies to a baseline of hydrogen production via natural gas-based Steam Methane Reforming (SMR) in a US context of multiple energy system and climate change mitigation futures. Besides providing an analysis that specifies the LCA results ranges with temporal and geospatial explicitness across the two technologies, metrics, and impact assessment methods, this research also aims to establish a base framework that can be expanded to use other IAM generated scenarios and US open-source life cycle inventory (LCI) databases. We find that the temporal environmental performance of either technology or their difference to SMR is directly influenced by the underlying background dynamics. Under baseline projections (i.e., no decarbonization goals), neither process reaches parity with the incumbent technology across several environmental metrics. Under the decarbonization scenarios, the underlying sectoral shifts result in declining impacts over time, compared to 2020 levels, except for metal depletion levels, which increase. The background shifts postulate a heavily decarbonized economy and energy system, which help technologies reach parity with SMR between 2040-2050 (RCP2.6) and 2030-2040 (RCP1.9) for global warming. Despite declines across several other metrics over time, neither PtH2 technology break even with SMR by 2100 besides for global warming.

decarbonizing

Materials data science using CRADLE: A distributed, data-centric approach

Abstract There is a paradigm shift towards data-centric AI, where model efficacy relies on quality, unified data. The common research analytics and data lifecycle environment (CRADLE™) is an infrastructure and framework that supports a data-centric paradigm and materials data science at scale through heterogeneous data management, elastic scaling, and accessible interfaces. We demonstrate CRADLE’s capabilities through five materials science studies: phase identification in X-ray diffraction, defect segmentation in X-ray computed tomography, polymer crystallization analysis in atomic force microscopy, feature extraction from additive manufacturing, and geospatial data fusion. CRADLE catalyzes scalable, reproducible insights to transform how data is captured, stored, and analyzed. Graphical abstract

97 MATHEMATICS AND COMPUTING

Applying a Multisector Scenario Framework to Evaluate Past and Future Public Surface Water Supply Infrastructure Strategies in Texas

Datasets supporting the index model and scenario analysis used in evaluating surface water supply strategies across different water system types in Texas. These data underpin the scenario development and application of five key indicators: Water Availability Index (WAI), Water Quality Index (WQI), Energy Requirement Index (ERI), Water Treatment Cost (WTC), and Water Infrastructure Cost (WIC). The datasets are organized by system type—stream reaches (flowlines), waterbodies, and reservoirs—and include both raw and standardized index values. The integrated datasets also provide scenario classifications (original and adjusted) based on infrastructure and planning priorities, enabling comparison across Shared Socioeconomic Pathways (SSPs). Additional strategy-level data are included to support evaluation of state-level new reservoir projects in relation to cost and availability tradeoffs. Please refer to the README file provided in Files for more details. Descriptions of the datasets are provided below. Dataset(s) Descriptions Folder: Index_model_database.zip Subfolder: Stream_reach.zip Fl_wf.csv, Fl_wq.csv, Fl_er.csv, Fl_wf_wtcUV.csv, Fl_wf_wtcnoUV.csv, Fl_allfac_wic1.csv, Fl_allfac_wic2.csvDatasets for computing WAI, WQI, ERI, WTC, and WIC for surface water systems classified as stream reaches (flowlines). Subfolder: Waterbody.zip Wb_wf.csv, Wb_wq.csv, Wb_er.csv, Wb_wf_wtcUV.csv, Wb_wf_wtcnoUV.csv, Wb_allfac_wic1.csv, Wb_allfac_wic2.csvEquivalent index model datasets for waterbodies, reflecting hydrologic and infrastructure attributes specific to impounded natural systems. Subfolder: Reservoir.zip Rs_wf.csv, Rs_wq.csv, Rs_er.csv, Rs_wf_wtcUV.csv, Rs_wf_wtcnoUV.csv, Rs_allfac_wic1.csv, Rs_allfac_wic2.csvIndex model datasets specific to regulated reservoir systems, incorporating both resource indicators and cost parameters. Folder: Integrated data.zip combined_merged_data.csv, combined_merged_data_scenario.csvDatasets integrating index model indicators (both raw and scaled) with scenario classifications, including adjustments reflecting SSP-aligned transitions and planning shifts. Folder: Additional data.zip wai_supplystrat_wic_merged.csvCurated dataset capturing proposed major reservoir-based municipal water supply strategies in Texas. Integrates site-level planning data with estimated capital infrastructure costs and water availability scores for comparative assessment.

geospatial

OpenPATH - Leveraging Technology to Measure Travel Behavior

Shifting transportation to more sustainable modes is a key piece of the decarbonization puzzle. However, mobility behavior and travel patterns are difficult to influence because they are difficult to measure. OpenPATH provides a tool to capture longitudinal behaviors through a smartphone application. Agencies interested in gathering data about a population's travel behavior can set up a deployment of the app customized to the needs of their community. Partners can choose between simple mode and purpose labels or surveys for each trip to balance the level of user engagement with the associated burden. The labels, trip surveys, and an initial demographic survey can all be tailored to the specific context of the deployment. The OpenPATH tool is unique in its open-source nature, ability to gather detailed longitudinal travel data, and design allowing direct engagement with travelers. A valuable technological advancement, this tool enables partners to measure the way changes in the transportation landscape impact their community. The suite of tools includes both public and administrator dashboards. The public dashboard supports continuous data analysis through charts presenting trip information updated daily. The administrator dashboard displays geospatial data and supports data export. Example applications have included e-bike programs; gathering valuable metrics on increased access to opportunities and reduction in VMT, and studies aimed at understanding existing mobility behavior to see where advancements such as electric vehicles could fit into these habits. OpenPATH collects travel data in association with an initial demographic survey, enabling detailed insight into the behavior patterns or impact of a certain program on different populations.

ADVANCED PROPULSION SYSTEMS

Geospatial Data Workflow Orchestration and Architecture

In an era characterized by explosive growth in geospatial data, the selection of appropriate technologies for data storage, processing, and orchestration is critical for organizations aiming to maintain competitive advantages. This white paper provides a comprehensive analysis of how Oak Ridge National Laboratory (ORNL) has effectively employed various cloud technologies, including containerized applications, container orchestrators, and workflow orchestrators, to develop robust geospatial data processing solutions. We explore the fundamental concepts behind these technologies and compare multiple deployment models tailored to diverse use cases. Our findings conclude that while Kubernetes has emerged as the preferred platform for truly scalable and fault-tolerant production workflows, the choice of workflow orchestration tool requires careful consideration of team needs, pipeline complexity, and deployment environments. This paper aims to serve as a strategic guide for organizations leveraging geospatial data, articulating the balance between technology choices and practical implementation to enhance workflow efficacy and scalability.

97 MATHEMATICS AND COMPUTING

Circularity Futures Workshop Series: Summary Report

The aim of this report is to synthesize key feedback received from the three-part Circularity Futures workshop series held in Spring 2024. The workshop series was conducted by the National Renewable Energy Laboratory (NREL) on behalf of U.S. Department of Energy, Office Energy Efficiency and Renewable Energy (EERE), and was broken into three workshops: Workshop 1 - Circularity Analysis Needs and Priorities; Workshop 2 - Circularity Metrics and Indicators; and Workshop 3 - Circularity Data. Together, the workshops focused on identifying the existing priorities and gaps in the circularity modeling space, understanding different stakeholders' use and interpretation of circularity metrics and indicators, identifying common data gaps and data quality challenges, and assessing the robustness of available solutions. The workshop series brought a diverse group of stakeholders - including representatives from U.S. government offices, national labs, nonprofit organizations, industry, and academia - to collect first-hand feedback on needs, priorities, challenges and opportunities in the circularity modeling and analysis space. The workshop discussions highlighted numerous common needs, priorities and challenges among the interviewed groups. Several topics were frequently discussed, including: 1) Circularity as a pathway for sustainable economic growth: While circularity is generally defined in terms of resource conservation and reducing wasteful disposal of materials, participants agreed that circular strategies should serve broader economic, environmental, and social goals. It is therefore crucial for circularity analysis to look beyond waste reduction and instead evaluate a variety of impact metrics such as cost savings, job creation, air quality, and pollutant emissions. Mutli-criteria decision-making frameworks may be useful for making sense of disparate metrics and evaluating tradeoffs between impact categories.; 2) Economic and social factors are not well understood: Underdevelopment of existing end-of-life (EOL) management infrastructure, inconsistent standardization codes and policy space in reusing recycled content, and suboptimal collection and sorting strategies collectively contribute to uncertainty about the economic potential of circular pathways. The latter observation is consistent among all technologies but more emphasized for renewable energy systems. Social impacts of circularity practices are less understood and less researched than other sustainability aspects.; 3) Inconsistent methods for assessing emerging technologies: LCA and TEA results vary widely depending on the assumptions made with regards to market adoption of new technologies. Emerging technologies suffer limited availability of data needed to conduct a robust circularity analysis. Yet, understanding projected impacts of proposed nascent technology is a key need for different stakeholder groups.; and 4) Lack of temporally and geospatially explicit data: There is a need for open data that represents variations in circularity technologies over time and location. The lack thereof leads to aggregated and potentially misrepresented results in circularity analysis. Sensitivity analyses should be included to verify whether options perceived as more sustainable align with real-world practices.

29 ENERGY PLANNING, POLICY, AND ECONOMY

LandScan Global 2023: Silver Edition

For a quarter of a century, the LandScan Global (LSG) project has annually released a global, high-resolution gridded population dataset representing the ambient or unwarned population at a 30 arcsecond resolution. LSG supports a range of applications such as emergency management, disaster response, and human health and security for understanding populations at risk. The 2023 release of LSG, the LandScan Silver Edition, represents a major methodological leap forward while also leveraging previous knowledge—the previous year was the baseline for the current annual update carrying forward valuable knowledge of the built environment for the past quarter century—to train the machine learning models. Compared with annual releases over the past 24years, multiple advancements were made to different aspects of the methodology to achieve reproducibility, transparency, and consistent global propagation of solutions to modeling or population distribution issues identified during the review process. These novel changes include incorporation of the latest available geospatial inputs across the globe, machine learning models instead of manual modifications, population feature importance analysis, open-source solutions vs. proprietary software, generation of multiple global versions, analytic validations, and human-in-the-loop revisions to produce the final version. Additionally, algorithms—such as anomaly detection—were introduced to quickly identify areas of focus to develop a new and robust systematic review. Significant changes in modeled population distributions were observed between the 2022 and 2023 releases, largely attributable to improvements in data and methods and discussed thoroughly within this report. In summation, the LandScan Silver Edition leverages the best of the past quarter century of LSG legacy knowledge and continues a tradition of applying cutting-edge enhancements to serve as a new benchmark for accurate, actionable gridded population data

Lebakula, Viswadeep

CERF: IM3 Projected Western US Power Plant Locations

Overview The Capacity Expansion Regional Feasibility (CERF) model is an open-source geospatial python package that provides new power plant locations at a 1km resolution. The model ingests U.S. state or regional-scale electricity system capacity expansion plans, such as those produced by the Global Change Analysis Model (GCAM-USA), and identifies feasible, site-specific locations for individual new power plants (renewable and non-renewable). CERF combines high-resolution geospatial suitability analyses with an economic algorithm that selects individual plant siting locations based on grid interconnection costs and the locational marginal value of new generation. The model incorporates a wide range of dynamic constraints and opportunities, such as protected lands, population density, existing infrastructure, and water availability. This dataset provides CERF power plant siting results for IM3 Phase 2 simulations across eight different scenarios for the Western US through 2055. The scenarios include combinations of two Shared Socioeconomic Pathways (SSP3 and SSP5) with four high-resolution climate projections specific to the United States (see, https://tgw-data.msdlive.org/). These climate projections include "hotter" and "cooler" variants for two Representative Concentration Pathways (RCP4.5 and RCP8.5). The resulting eight simulations are: rcp45cooler_ssp3 rcp45cooler_ssp5 rcp45hotter_ssp3 rcp45hotter_ssp5 rcp85cooler_ssp3 rcp85cooler_ssp5 rcp85hotter_ssp3 rcp85hotter_ssp5 CERF siting results in this dataset correspond to capacity expansion plans in the GCAM-USA IM3 Phase 2 simulation data and are available for each of the above scenarios. Data Details Temporal Range: 2015-2055 in 5-year timesteps. Note that 2015 is the experiment base year and 2020 and beyond represent model simulation years. Spatial Range: Plant locations are provided for the eleven states in the Western US including Arizona, California, Colorado, Idaho, Montana, New Mexico, Nevada, Oregon, Utah, Washington, and Wyoming. Spatial Resolution: 1 km-squared, provided in x and y coordinates Geospatial Projection: Albers Equal Area Conic (ESRI:102003) File Type: csv The dataset contains subdirectories for each of the eight scenarios described in the overview. Each scenario folder contains two subfolders with the following information: 1. Power Plant Data This directory contains a single .csv file of power plant locations for both pre-existing (non-CERF sited plants in operation in 2015) and new (CERF-sited) power plants across the temporal range along with additional CERF model output parameters for CERF-sited plants. Plant with a siting year earlier than 2020 correspond to facilities that are operational leading into the first timestep CERF simulation. For a more detailed description of CERF model output parameters, see the CERF model documentation. Note that the cerf_plant_id parameter is unique within each scenario file but not across scenario files. Parameter Descriptions scenario - Name of scenario cerf_plant_id - Unique siting identifier cerf_sited - If True, indicates that plant was sited by CERF model. If False, indicates pre-existing facility region_name - Name of region (state) tech_id - Technology ID tech_name - Full generation technology name inclusive of cooling type (if applicable) and additional characteristics tech_simple - Simplified generation technology type unit_size_mw - Power plant unit size (MW) xcoord - X coordinate in the default CRS (meters) ycoord - Y coordinate in the default CRS (meters) index - Index position in the flattend 2D array buffer_in_km - Exclusion buffer around site (km) sited_year - Year of siting retirement_year - Year of retirement lmp_zone - Locational marginal price (LMP) zone ID locational_marginal_price_usd_per_mwh - Locational marginal price ($/MWh) generation_mwh_per_year - Generation output (MWh/yr) operating_cost_usd_per_year - Cost of plant operations ($/yr) net_operational_value - Net operational value based on LMP and and operating costs ($/yr) interconnection_cost - Cost of interconnection for transmission & gas pipeline (if applicable) net_locational_cost -- Difference of interconnection cost and operating value ($/yr) capacity_factor_fraction - Capacity factor (fraction) carbon_capture_rate_fraction - Carbon capture rate (fraction) fuel_co2_content_tons_per_btu - Fuel CO2 content (tons/Btu) fuel_price_usd_per_mmbtu - Fuel price ($/MMBtu) fuel_price_esc_rate_fraction - Fuel price escalation rate (fraction) heat_rate_btu_per_kWh - Heat rate (Btu/kWh) lifetime_yrs - Technology lifetime for annuity (years) operational_life_yrs - Operational lifetime for retirement (years) variable_om_usd_per_mwh - Variable operation and maintenance costs of yearly capacity use ($/MWh) variable_om_esc_rate_fraction - Variable operation and maintenance costs escalation rate (fraction) carbon_tax_usd_per_ton - Carbon tax ($/ton) carbon_tax_esc_rate_fraction - Carbon tax escalation rate (fraction) 2. Storage Data This directory contains information on new and pre-existing energy storage facilities operational in each timestep along with various storage operational parameters. The 2015 timestep provides pre-existing energy storage data and corresponds with facilities that are operational leading into the first model simulation timestep. Note that coordinates in the storage files correspond to the interconnection point on the grid (substation location), not individual energy storage locations. Energy storage is added in a cumulative process at each given interconnection point. That is, each individual file provides the total operational storage capacity interconnected to the specified substation for the given timestep, inclusive of previously installed storage at that location and new storage installed in that timestep at that location. Parameters scenario - Name of scenario timestep - Simulation timestep name - Unique storage identifier s_typ - Type of energy storage technology (battery or pumped storage hydro) s_node - Node ID of interconnecting substation xcoord - X coordinate in the default CRS (meters) ycoord - Y coordinate in the default CRS (meters) charge_rate - Maximum charge rate (power capacity) of storage system (MW) discharge_rate - Maximum discharge rate (power capacity) of storage system (MW) duration - Duration of storage system (hours) max_SoC - Allowed maximum state of charge (energy capacity) of storage system (MWh) min_SoC -Allowed minimum state of charge (energy capacity) of storage system (MWh) charge_eff - Efficiency of charge (fraction between 0 and 1) discharge_eff - Efficiency of discharge (fraction between 0 and 1) Acknowledgment IM3 is a multi-institutional effort led by Pacific Northwest National Laboratory and supported by the U.S. Department of Energy's Office of Science as part of research in MultiSector Dynamics, Earth and Environmental Systems Modeling Program.

CERF