Search NASA⌕ Search

SEARCH · Search NASA

Results for “Modelling”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 559 records · Page 31

Challenges and Opportunities for Electric Utility Modeling and Asset Valuation Frameworks: Case Study on Valuing New Pumped Storage Hydropower

Asset valuation by electric utilities is becoming increasingly difficult in the rapidly changing electric sector. Rapid deployment of variable generation and inverter-based storage systems along with uncertain demand growth, climate, policies, and other factors create a challenging environment for understanding the value proposition of a new potential asset. This report describes an effort between the Tennessee Valley Authority (TVA) and three U.S. Department of Energy laboratories to perform a detailed review of utility modeling and analysis practices for asset valuation and identify challenges and opportunities for advancing its methods into the future. It focuses on a case study of new potential pumped storage hydropower (PSH) because of growing interest in new PSH capacity to provide energy balancing, firm capacity, and a range of ancillary services. Staff from the DOE labs conducted systematic interviews about current practices in capacity expansion modeling, production-cost modeling, hydrological modeling, and transmission stability modeling while also discussing how scenario analysis is conducted and how models and data are integrated. The effort resulted in a set of model, integration, and scenario recommendations that could be valuable to TVA, other utilities, system operators, and other stakeholders conducting integrated grid analysis. Individual model recommendations suggest exploring computational tradeoffs with detail and resolution across spatiotemporal structure, supply- and demand-side details, transmission overlays, market interactions, and ancillary services. Automated processes to pass data between models and conduct larger scenario suites could also enhance valuation practices by enabling a more consistent study of asset value across a broader range of uncertain future grid conditions where PSH could be particularly valuable. TVA and other industry stakeholders can learn from and adapt applied research-grade methods developed by DOE laboratories and other research institutions to improve decision making and accelerate progress towards a reliable, economic, sustainable energy system.

13 HYDRO ENERGY↗

Biophysical model of eelgrass and water quality in Coos Bay, OR shows greater mitigation potential for ocean acidification than hypoxia

Seagrass beds provide important ecosystem services and are valued, in part, for their potential to mediate stressors such as ocean acidification and hypoxia (OAH) for sensitive species. However, the susceptibility of seagrasses to anthropogenic impacts and recent declines motivate the need to better understand the drivers of seagrass and the water quality consequences that occur with variation in seagrass abundance. To meet this need, we leveraged existing monitoring data (water quality and seagrass), hydrodynamic circulation model, and biogeochemical model framework with seagrass submodel, to produce a biophysical model of Coos Bay estuary, Oregon, U.S. The model includes biogeochemical processes involving water quality, plankton, seagrass, and sediment-water interactions. Ecosystem models like this are useful for evaluating complex estuarine systems because they allow us to extend our understanding of system dynamics beyond existing observations and perform experiments to identify the processes driving observed patterns. We used the biophysical model of Coos Bay to evaluate the dynamics of water quality and native eelgrass (Zostera marina) under three eelgrass abundance scenarios (zero eelgrass, current extent, and maximum observed extent) to elucidate the relationship between eelgrass and OAH. Including eelgrass in the Coos Bay model produced results that more closely resembled water quality observations - dissolved oxygen (DO) and pH were more dynamic in simulations with eelgrass, often having both higher highs and lower lows. While there were some areas of the estuary where DO improved with the addition of eelgrass to the model there was overall a small net increase in harmful DO conditions (based on a salmon physiological threshold). In contrast, ocean acidification conditions, pH and calcium carbonate saturation state for aragonite (Ω), were improved (based on oyster requirements) with the addition of eelgrass - although the magnitude of improvement differed seasonally and spatially. Our new model represents a useful tool - one which accounts for and controls the relevant physical and biogeochemical processes - to evaluate conditions that confer resilience or enhance vulnerability to OAH in an important Pacific Northwest coastal estuary and results can inform the OAH-related dynamics occurring in other eastern boundary current estuaries.

FVCOM-ICM↗

Comparative Performance of Gaussian Plume and Backward Lagrangian Stochastic Models for Near-Field Methane Emission Estimation Using a Single Controlled Release Experiment

Methane (CH 4 ) is a major component of natural gas and a potent greenhouse gas. Increasing atmospheric methane concentrations are attributed to emissive anthropogenic activities by an average of 13 ppb per yr since 2020 and are linked to a changing global climate. Mitigating CH 4 emissions from oil and gas production sites has recently become a target to reduce overall greenhouse gas emissions; however, monitoring the efficacy of mitigation strategies depends on accurate quantification of CH 4 emissions at the facility-level. Near-field quantification of methane (CH 4 ) emissions from oil and gas (O&G) facilities remains challenging due to the effects of atmospheric variability and sensor configuration on atmospheric dispersion models. This study evaluates the performance of two atmospheric dispersion models, the Gaussian plume (GP) and backward Lagrangian stochastic (bLS), by comparing calculated CH 4 emissions to controlled single-point emissions between 0.4 and 5.2 kg CH 4 h −1 . Emissions were calculated by both models using 121 individual sets of measurements comprising five-minute averaged downwind methane mixing ratios and matching meteorological data. The comparison shows that the bLS approach achieved a higher proportion of emission estimates within a factor of two (FAC2) of the known emission rates compared to the GP approach. The emissions calculated by the bLS model also had a lower multiplicative error and reduced bias relative to GP. Other error-based metrics further confirmed the bLS model performed better, as it yielded lower RMSE and MAE than GP. Statistical analysis of the emission data shows that the lateral and vertical alignment of the source and the sensor plays a critical role in emission estimations, as measurements made closer to the plume centerline and at a distance between 40 and 80 m downwind yielded the best FAC2 agreement. High wind meander degraded the ability of both approaches to generate representative emissions, particularly with the GP approach, as it violates the modeling approach’s assumption of steady-state emissions. Data suggest emissions calculated by the bLS model are comprehensively in better agreement, but the computational demands of the modeling approach and integration into fenceline systems limit real-time applicability. While these results provide insight into model performance under controlled near-field conditions, their applicability to more complex or heterogeneous oil and gas production environments (e.g., the regions Marcellus or Unita Basins) remains limited and uncertain.

gaussian plume↗

Observationally constrained analysis on the distribution of fine- and coarse-mode nitrate in global models

Nitrate plays an important role in the Earth system and air quality. A key challenge in simulating the life cycle of nitrate aerosol in global models is to accurately represent mass size distribution of nitrate aerosol. In this study, we evaluate the performance of the Energy Exascale Earth System Model version 2 (E3SMv2) and the Community Earth System Model version 2 (CESM2), along with Aerosol Comparisons between Observations and Models (AeroCom) phase III models, in simulating spatial distribution of fine-mode nitrate, the mass size distribution of fine- and coarse-mode nitrate, and the gas–aerosol partitioning between nitric acid gas and nitrate, using long-term ground-based observations and measurements from multiple aircraft campaigns. We find that most models underestimate the annual mean PM 2.5 (particulate matter with diameter less than 2.5 µm) nitrate surface concentration averaged over all sites. The observed nitrate PM 2.5 / PM 10 and PM 1 / PM 4 ratios are influenced by the relative contribution of fine sulfate or organic particles and coarse dust or sea salt particles. Overall, the ground-based observations give an annual mean surface nitrate PM 2.5 / PM 10 ratio of 0.7. Most models underestimate the annual mean PM 2.5 / PM 10 ratio in all regions. There are large spreads in the modeled nitrate PM 1 / PM 4 ratios, which span the full range from 0 to 1. Most models underestimate the surface molar ratio of nitrate to total inorganic nitrate averaged across all sites. Our study indicates the importance of gas–aerosol partition parameterization and the simulation of dust and sea salt in correctly simulating the mass size distribution of nitrate.

Nitrate↗

Benchmarking soil moisture and its relationship to ecohydrologic variables in Earth System Models

Soil moisture (SM) is a key regulator of ecosystem biogeophysics, influencing plant water relations and land-atmosphere energy exchanges. We evaluate the representation of SM in 16 Earth System Models from the Coupled Model Intercomparison Project Phase 6 (CMIP6) using the International Land Model Benchmarking (ILAMB) framework, focusing on surface (0–5, 0–10 cm) and rootzone (0–100 cm) depths, as well as key ecohydrological variables like gross primary productivity (GPP), leaf area index (LAI), and evapotranspiration (ET), and their coupling. Models are benchmarked against multiple observational and assimilated datasets to assess both state variables and cross-variable relationships. Surface SM is generally well represented (r > 0.87), while rootzone SM variability is systematically overestimated (normalized standard deviation > 1). ET shows strong agreement with observations (r > 0.9), whereas GPP and LAI exhibit larger inter-model spread. Skill in individual variables does not guarantee realistic SM–ecohydrology coupling, which varies strongly across models and depends on the reference dataset. Köppen-based regional analyses reveal strong regime dependence, with several models performing well in Tropical and Temperate regions but degrading in Continental (high-latitude) zones. Across both global and regional benchmarks, models cluster by land surface framework, indicating that structural choices in soil hydrology and soil–plant coupling exert a first-order control on performance. These results provide process-relevant benchmarks and suggest that improving the representation of vertical soil structure, rooting depth distributions, and soil–plant hydraulic coupling will be central to advancing soil moisture realism in next-generation Earth system models.

CMIP6↗

Leveraging Large Language Models for Real-World Data Evidence: A Framework for Automated Treatment Extraction and Data Harmonization

Background: The ability to comprehensively collect treatment information from cancer patient medical records would enable studies to evaluate real-world benefits and risks tied to specific treatments. Currently, it is difficult to system- atically collect high-quality treatment information because it is often stored in unstructured text. Manually extracting and standardizing drug and regimen data is time-intensive. Recent advances in large language models (LLMs) offer a potential solution for automated extraction of structured treatment information from clinical text. Objective: This study systematically evaluates the utility of four LLMs from the Llama family for automated extraction of oncology treatment information from clinical text. This information can guide researchers using cancer registry data to provide insights into cancer care and outcomes beyond clinical trials. Methods: Four instruction-tuned Llama models with varying parameter counts (1B, 3B, 8B, and 70B) were evaluated for their ability to extract treatment information from clinical documents. A unified oncology knowledge base integrating seven major public data sources was developed to standardize and normalize extracted entities—a critical step for harmonizing data from diverse sources. Extracted treatment data were compared against expert-annotated ground truth. Model performance was assessed using accuracy metrics (Precision, Recall, F1-Score) and opera- tional feasibility metrics, including processing speed and structural compliance of the output. Results: A strong positive correlation was observed between model size and extraction accuracy. F1-score improved from 0.609 for the 1B model to 0.710 (3B), 0.807 (8B), and 0.828 (70B). While larger models demonstrated superior accuracy and compliance, they incurred higher computational costs. The modest performance difference between 8B and 70B suggests diminishing returns with increasing model size. Conclusions: LLMs represent a viable technology for automating oncology treatment extraction. The 8B-parameter model emerged as a highly effective option, balancing high accuracy and computational efficiency. Selecting an appropriate LLM for deployment in cancer registries involves a trade-off between desired accuracy and available operational resources. Harmonizing extracted entities with the oncology knowledge base facilitates standardized integration into common data models, enhancing data quality for real-world evidence analyses.

artificial intelligence↗

Mapping Incidence and Prevalence Peak Data for SIR Modeling Applications

Infectious disease modeling and forecasting have played a key role in helping assess and respond to epidemics and pandemics. Recent work has leveraged data on disease peak infection and peak hospital incidence to fit compartmental models for the purpose of forecasting and describing the dynamics of a disease outbreak. Incorporating these data can greatly stabilize a compartmental model fit on early observations, where slight perturbations in the data may lead to model fits that forecast wildly unrealistic peak infection. We introduce a new method for incorporating historic data on the value and time of peak incidence of hospitalization into the fit for a Susceptible-Infectious-Recovered (SIR) model by formulating the relationship between an SIR model’s starting parameters and peak incidence as a system of two equations that can be solved computationally. We demonstrate how to calculate SIR parameter estimates – which describe disease dynamics such as transmission and recovery rates – using this method, and determine that there is a noticeable loss in accuracy whenever prevalence data is misspecified as incidence data. To exhibit the modeling potential, we update the Dirichlet-Beta State Space modeling framework to use hospital incidence data, as this framework was previously formulated to incorporate only data on total infections. This approach is assessed for practicality in terms of accuracy and speed of computation via simulation.

97 MATHEMATICS AND COMPUTING↗

Modeling Framework for Data Center

This chapter highlights the critical need for advanced modeling of data centers due to their rapidly increasing energy consumption and impact on grid reliability. Driven by the demand for AI applications, data centers are projected to consume a significant portion of US energy by 2028, putting stress on an already challenged power grid. The chapter emphasizes the importance of "fast" time-scale models to understand the dynamic interactions between data centers and the grid, especially given the rapid power fluctuations of AI workloads. It outlines a modeling framework that includes both offline and real-time EMT domain simulations, detailing the necessary representations for various components like utility interfaces, transformers, IT loads, UPS, cooling loads, Battery Energy Storage Systems (BESS), generators, protection systems, and higher-level control systems. While standard simulation tools like PSCAD offer basic models, custom development is often required to accurately capture the unique and fast-changing behaviors of modern data centers. The chapter also discusses key metrics and test cases for validating these models, focusing on transient load responses, protection relay coordination, and demand flexibility. Finally, it addresses the challenges of modeling large-scale data centers, such as computational complexity and the trade-off between model fidelity and practicality, suggesting hybrid modeling approaches as a solution. The overarching goal is to create a robust framework that helps assess data center impacts on grid stability, identify vulnerabilities, and inform the development of standards for reliable integration of these large loads into the bulk power system.

25 ENERGY STORAGE↗

Characteristics of Fluid‐Solid Interaction Constitutive Models Within Poroelastodynamics at Higher Strain‐Rates and Large Deformations Implemented in 1D

The large deformation, mixed formulation, finite element (FE) modeling approach presented in Irwin et al. 2024 is extended herein to include improved constitutive models for representing dynamic solid-fluid interactions at higher strain rates (𝒪⁢(1⁢0 2 −1⁢0 3 )⁢s −1 ) and larger overpressure magnitudes (𝒪⁡(1⁢0 2 )⁢kPa) within a biphasic soft porous material using Theory of Porous Media (TPM) at finite strain. Specifically, these constitutive modeling improvements are the following: (i) a more physically robust constitutive model for pore fluid seepage velocity via inclusion of pore fluid viscous stress, and (ii) a modified deformation-dependent-permeability model and updated hyperelastic constitutive model better suited for handling larger volumetric compressions and extensions. The novelty of the present work is mainly the contribution (i): inclusion of pore fluid viscous stress at higher strain-rate and large deformations, which requires 𝐶 1 continuity in the weak formulation, accomplished by employing Hermite cubic interpolation functions within a mixed nonlinear poromechanical finite element formulation. In (ii), the model is updated to weakly enforce solid phase incompressibility, such that this assumption is not violated numerically, which provides improved numerical stability for achieving larger overpressure magnitudes on 𝒪⁡(1⁢0 2 ) kPa, which were not achievable with the previous Kozeny–Carman model in Irwin et al. 2024. Also in (ii), the volumetric part of the solid skeleton free energy function is modified to ensure proper bounds on the solid skeleton Jacobian of deformation 𝐽 s related to incompressibility of the solid phase. Uniaxial strain, unidirectional flow examples at higher strain rates (𝒪⁢(1⁢0 2 −1⁢0 3 )⁢s −1 ) and larger deformations (up to 0.2 (or 20%) nominal axial strain) demonstrate the improved physical representation—and numerical stability—of these constitutive model improvements.

42 ENGINEERING↗

Deriving Stable Peak Models to Fit Complex XPS Data From Cu Contaminated Pt Electrocatalysts

X-ray Photoelectron Spectroscopy spectra peak models, designed to partition photoemission signals emanating from different elements or chemical states within an atom, are fitted to data limited to an energy interval over which inelastically scattered photoemission signal can be estimated. While the choice of background approximation and line shapes of components to the peak model requires careful consideration, the energy interval used to define the data to which the peak model is optimized has a significant impact on the final peak model. The relationship between the background intensity and data intensity at the start and end of the energy interval dictates the line shapes used in the peak model. In this work, we devise a method to peak fit a complex overlapping Cu 3p and Pt 4f XPS peak structure to perform the elemental quantification. We first use an Al 2s peak to illustrate how background curves approach data at the limits of the energy interval over which the background is defined, influencing the analysis of XPS spectra. Next, we demonstrate the nature of interactions between specific line shapes (Voigt and pseudo-Voigt profiles) suitable for photoemission peaks and a specific background curve (Shirley) and a peak model is presented that includes components to the peak model that accommodates background intensity during fitting of the peak model to data. The peak model allowed for quantification of the contributions of Pt 4f peaks emanating from the substrate that exhibits strong asymmetry in the presence of the inhomogeneously distributed Cu species, mostly of Lorentzian character.

XPS↗

Towards robust surrogate models: Benchmarking machine learning approaches to expediting phase field simulations of brittle fracture

Data-driven approaches have the potential to make modeling complex, nonlinear physical phenomena significantly more computationally tractable. For example, computational modeling of fracture is a core challenge where machine learning techniques have the potential to provide a much needed speedup that would enable progress in areas such as multi-scale modeling and uncertainty quantification. Currently, phase field modeling (PFM) of fracture is one such approach that offers a convenient variational formulation to model crack nucleation, branching and propagation. To date, machine learning techniques have shown promise in approximating PFM simulations. While standard fracture benchmarks represent realistic scenarios frequently observed in practice, they typically do not provide sufficiently challenging tests for data-driven methods. Here, to address this gap, we introduce a challenging dataset based on PFM simulations designed to benchmark and advance ML methods for fracture modeling. This dataset includes three energy decomposition methods, two boundary conditions, and 1000 random initial crack configurations for a total of 6000 simulations. Each sample contains 100 time steps capturing the temporal evolution of the crack field. Alongside this dataset, we also implement and evaluate Physics Informed Neural Networks (PINN), Fourier Neural Operators (FNO), and UNet models as baselines, and explore the impact of ensembling strategies on prediction accuracy. With this combination of our dataset and baseline models drawn from the literature we aim to provide a standardized and challenging benchmark for evaluating machine learning approaches to solid mechanics. Our results highlight both the promise and limitations of popular current models, and demonstrate the utility of this dataset as a testbed for advancing machine learning in fracture mechanics research.

Benchmark dataset↗

Integration and validation of some modules for modelling of high-speed chemically reactive flows in two-phase gas-droplet mixtures

Three modules are integrated into the built-in OpenFOAM rhoCentralFoam solver towards accurate and efficient modelling of high-speed chemically reactive flows in two-phase gas-droplet mixtures within the OpenFOAM 10.0 framework. The first module is the mixture-averaged diffusion model. The second module is the built-in OpenFOAM Lagrangian solver coupled with optimised droplet drag coefficient and convective heat transfer coefficient sub-models. The last module is a sparse stiff chemistry solver based on dynamic adaptive hybrid integration (AHI-S). The optimised droplet sub-models are first verified in correct implementation for subsequent simulations in this work. Further, they show good accuracy against experimental and analytical data in the modelling of ammonia droplet acceleration and cooling in the flowing and/or low-temperature air. The accuracy and efficiency gains related to the mixture-averaged diffusion model and the AHI-S chemistry solver are examined by simulating 1-D detonation propagation in ammonia droplet-free/laden ammoniaoxygen mixtures. Numerical results of detonation propagation speed, gaseous temperature, density, and species distributions around the induction zone show good agreement with experimental data and analytical solutions. Compared to the built-in OpenFOAM diffusion model, the mixture-averaged diffusion model provides different numerical predictions of pulsating instabilities in detonation propagation. It shows better accuracy in depicting the detonation structure within the droplet-free section attributed to improved multi-component diffusion modelling. Compared to the built-in OpenFOAM solver EulerImplicit (backward Euler), the AHI-S chemistry solver reduces the computational cost by around 50%. It achieves satisfactory accuracy in calculating detonation propagation speed within the droplet-free section with the optimal efficiency when the safety factor, β, equals 0.5.

42 ENGINEERING↗

Inference of phase field fracture models

The phase field approach to modeling fracture uses a diffuse damage field to represent cracks. This representation mollifies singularities that arise in computations with sharp interface models and some of the resultant difficulties in the mathematical and numerical treatment of fracture. Phase field fracture models have proven effective at representing crack propagation, branching, and merging. Specific formulations, beginning with brittle fracture, have also been shown to converge to classical solutions. Extensions to cover the range of material failure, including ductile and cohesive fracture, lead to an array of possible models. There exists a large body of literature focusing on this class of models and on the impact of model form on the predicted crack evolution. However, there have not been systematic studies into how optimal models may be chosen. Here, we take a first step in this direction by developing formal methods for identification of the best parsimonious model of phase field fracture given full-field data on the damage and deformation fields. We consider some of the main models that have been used for the degradation of elastic response due to damage and its propagation. Our approach builds upon Variational System Identification (VSI), a weak form variant of the Sparse Identification of Nonlinear Dynamics (SINDy). Furthermore, in this first communication we focus on synthetically generated data but we also consider central issues associated with the use of experimental full-field data, such as data sparsity and noise.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Modeling heat pipe startup and noncondensable gases in Sockeye

For this work, a one-dimensional gas mixture flow model was developed and implemented in the heat pipe code Sockeye to model the effects of noncondensable gases. Additionally, a startup model based on the dusty gas model was implemented to model the transition from rarefied gas dynamics to continuum flow, which occurs during the frozen startup of high-temperature heat pipes. Multiple startup and noncondensable gas models were tested against experimental data for sodium heat pipes, showing excellent agreement. Additionally, the newly developed gas mixture model for modeling noncondensable gas is further tested with a theoretical case study with arbitrary heating configurations. Finally, several recommendations and conclusions are made from the studies in this work to guide future heat pipe modeling efforts.

97 - MATHEMATICS AND COMPUTING↗

Quick-and-Easy Validation of Protein–Ligand Binding Models Using Fragment-Based Semiempirical Quantum Chemistry

Electronic structure calculations in enzymes converge very slowly with respect to the size of the model region that is described using quantum mechanics (QM), requiring hundreds of atoms to obtain converged results and exhibiting substantial sensitivity (at least in smaller models) to which amino acids are included in the QM region. As such, there is considerable interest in developing automated procedures to construct a QM model region based on well-defined criteria. However, testing such procedures is burdensome due to the cost of large-scale electronic structure calculations. Here, we show that semiempirical methods can be used as alternatives to density functional theory (DFT) to assess convergence in sequences of models generated by various automated protocols. The cost of these convergence tests is reduced even further by means of a many-body expansion. We use this approach to examine convergence (with respect to model size) of protein–ligand binding energies. Fragment-based semiempirical calculations afford well-converged interaction energies in a tiny fraction of the cost required for DFT calculations. Two-body interactions between the ligand and single-residue amino acid fragments afford a low-cost way to construct a “QM-informed” enzyme model of reduced size, furnishing an automatable active-site model-building procedure. This provides a streamlined, user-friendly approach for constructing ligand binding-site models that needs neither a priori information nor manual adjustments. Extension to model-building for thermochemical calculations should be straightforward.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Integrating a Water Tracer Model Into WRF‐Hydro for Characterizing the Effect of Lateral Flow in Hydrologic Simulations

Abstract Most current land models approximate terrestrial hydrological processes as one‐dimensional vertical flow, neglecting lateral water movement from ridges to valleys. Such lateral flow is fundamental at catchment scales and becomes crucial for finer‐scale land models. To test the effect of incorporating lateral flow toward three‐dimensional representations of hydrological processes in the next generation land models, we integrate a water tracer model into the WRF‐Hydro framework to track water movement from precipitation to discharge and evapotranspiration. This hydrologic‐tracer integrated system allows us to identify the key mechanisms by which lateral flow affects the flow paths and transit times in WRF‐Hydro. By comparing modeling experiments with and without lateral routing in two contrasting catchments, we determine the impacts of lateral flow on the transit times of precipitation event‐water. Results show that with limited hydrologic connectivity, lateral flow extends the transit times by reducing (increasing) event‐water drainage loss (accumulation) in ridges (valleys) and allowing reinfiltration of infiltration‐excess flow, which is missing in most land models. On the contrary with high hydrologic connectivity, lateral flow can effectively accelerate the water release to streams and reduce the transit time. However, the transit times are substantially underestimated by the model compared with isotope‐derived estimates, indicating model limitations in representing flow paths and transit times. This study provides some insights on the fundamental differences in terrestrial hydrology simulated by land models with and without lateral flow representation.

54 ENVIRONMENTAL SCIENCES↗

Using Temporal Deep Learning Models to Estimate Daily Snow Water Equivalent Over the Rocky Mountains

Abstract In this study we construct and compare three different deep learning (DL) models for estimating daily snow water equivalent (SWE) from high‐resolution gridded meteorological fields over the Rocky Mountain region. To train the DL models, Snow Telemetry (SNOTEL) station‐based SWE observations are used as the prediction target. All DL models produce higher median Nash‐Sutcliffe Efficiency (NSE) values than a conceptual SWE model and interpolated gridded data sets, although mean squared errors also tend to be higher. Sensitivity of the SWE prediction to the model's input variables is analyzed using an explainable artificial intelligence (XAI) method, yielding insight into the physical relationships learned by the models. This method reveals the dominant role precipitation and temperature play in snowpack dynamics. In applying our models to estimate SWE throughout the Rocky Mountains, an extrapolation problem arises since the statistical properties of SWE (e.g., annual maximum) and geographical properties of individual grid points (e.g., elevation) differ from the training data. This problem is solved by normalizing the SWE with its historical maximum value to alleviate extrapolation for all tested DL models. Our work shows that the DL models are promising tools for estimating SWE, and sufficiently capture relevant physical relationships to make them useful for spatial and temporal extrapolation of SWE values.

54 ENVIRONMENTAL SCIENCES↗

Selecting Appropriate Model Complexity: An Example of Tracer Inversion for Thermal Prediction in Enhanced Geothermal Systems

Abstract A major challenge in the inversion of subsurface parameters is the ill‐posedness issue caused by the inherent subsurface complexities and the generally spatially sparse data. Appropriate simplifications of inversion models are thus necessary to make the inversion process tractable and meanwhile preserve the predictive ability of the inversion results. In this study, we investigate the effect of model complexity on fracture aperture inversion and thermal performance prediction in a field‐scale EGS model. Principal component analysis was used to map the aperture field to a low‐dimensional latent space. The complexity of the inversion model was quantitatively represented by the percentage of total variance in the original aperture fields preserved by the latent space. Tracer, pressure and flow rate data were used to invert for fracture aperture through an ensemble‐based inversion method, and the inferred aperture field was used to predict thermal performance. With an over‐simplified aperture model, ensemble collapse occurred. The inverted aperture models failed to resolve necessary flow and transport features, leading to a biased thermal performance prediction. A complex aperture model involved excessive features and was prone to overinterpreting the inversion data. Both the tracer/pressure/flow rate data reproduction and thermal prediction showed significant uncertainties, making it difficult to properly estimate long‐term thermal performance. Fortunately, our results indicate that there exists an appropriate model complexity which can simultaneously match inversion data and predict thermal performance with an acceptable uncertainty. The quality of the fit of tracer data appears to be a useful indicator of such an appropriate model complexity.

15 GEOTHERMAL ENERGY↗