Search NASASearch

SEARCH · Search NASA

Results for “Error Rate Predictions”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Application of deep learning to single-shot gas-phase laser-induced breakdown spectroscopy

Single-shot fs laser-induced breakdown spectroscopy (LIBS) has the potential to capture ns-scale electrode desorption phenomena in pulsed power fusion drivers. However, the successful implementation of the diagnostic for this purpose is challenging, as it requires interpreting single-shot measurements collected from low-density gas mixtures. In this work, we demonstrate the efficacy of a Bayesian-optimized convolutional neural network (CNN) to interpret these measurements. We generated 256 distinct measurement conditions at relevant gas pressures ranging from 80–530 mTorr by mixing 100–250 sccm H 2 and 50–200 sccm CH 4 in increments of 10 sccm. Despite the considerable overlap between signals separated by 20 sccm, the CNN is able to predict the H 2 flow rate with a root-mean-square error (RMSE) of 15.9 sccm and the CH 4 flow rate with an RMSE of 12.0 sccm. The average relative prediction error is <9% for each gas and largely remains below or near 10%.

Brown, Nathan Parnell [Sandia National Lab. (SNL-N

Toward an improved nonlocal thermodynamic equilibrium model for more predictive simulations of ignition scale hohlraums

Recently, nonlocal thermodynamic equilibrium (NLTE) modeling has been identified as the primary reason for discrepant predictions of the peak neutron production time in indirectly driven inertial confinement fusion (ICF) platforms. It has also been observed that predictions of collisional excitation rates differ by as much as 50% from measurements. Theoretical uncertainties in dielectronic recombination rates have also been posited as possibly contributing to errors in NLTE predictions. This work examines the impact of multipliers on collisional excitation and dielectronic recombination rates on simulations of a directly driven gold sphere and an indirect drive ICF implosion. It is found that multipliers on the collisional excitation rates have a strong impact on radiant intensity and electron temperature and a weaker impact on ionization state, whereas multipliers on dielectronic recombinations rates strongly impact ionization state with a smaller impact on radiant intensity and electron temperature. A self-consistent NLTE model which places multipliers on differing transitions, as motivated by experimental measurements and more detailed atomic physics predictions, improves agreement but does not completely eliminate discrepancies with measurements of the radiant intensity within the 2–4 keV spectral range.

Farmer, W. A. [Lawrence Livermore National Laborat

3D Deep Learning Joint Inversion of Active Seismic Full Waveform and Passive Seismic Traveltime Data for Reservoir Imaging and Uncertainty Quantification

Here, we present deep learning (DL) networks for three-dimensional (3D) joint inversion of active seismic full waveform and passive seismic traveltime data to image reservoirs and their properties and quantify imaging uncertainties. Active seismic full-waveform data can provide high-resolution monitoring images but are collected only intermittently because of their high acquisition cost. In contrast, passive seismic data can be gathered at relatively low cost between regular active surveys, although their imaging quality can be compromised by factors such as low signal-to-noise ratios and limited ray coverage of the target. Although these datasets are routinely acquired together at CO 2 storage sites, their combined inversion within a 3D DL framework has not been previously demonstrated. To our knowledge, this is the first study to address this gap, combining the strength of both data types. For efficient data storage and DL training with large 3D seismic datasets, we use a 3D data matrix in which a random number of passive seismic traveltime data are stored as parabolic envelopes using one-hot encoding and a 3D full-waveform data matrix in which multiple shot gathers are summed. Two network architectures are evaluated: a single-encoder U-Net for single-data type inversion and a dual-encoder U-Net for joint inversion of active and passive seismic data. We also evaluate the single-encoder U-Net for joint inversion by concatenating full-waveform data and traveltime data. We propose a systematic approach for selecting an optimal dropout rate that balances regularization during training and Monte Carlo dropout-based uncertainty quantification during prediction by examining the correlation coefficient between standard deviation and prediction error, along with the training misfit, across a range of dropout rates. 3D DL inversion experiments include five different network configurations, with evaluations under ideal, noisy and dropout-enabled conditions. Both model and data uncertainties are assessed, as well as their combined effects. Across all conditions, the networks consistently predict accurate CO 2 saturation models with low prediction errors, such as a structural similarity index of 0.993 and CO 2 difference of 1.1%. Uncertainty estimates show strong spatial correlation with prediction errors, confirming the effectiveness of the proposed dropout selection approach. The results demonstrate that our DL approach, utilizing compact data representations and appropriate uncertainty quantification, yields accurate subsurface images under various inversion conditions and provides valuable insights into the reliability of predictions.

Um, Evan Schankee [Lawrence Berkeley National Labo

Towards Scaling Law Analysis For Spatiotemporal Weather Data

Compute-optimal scaling laws are relatively well studied for NLP and CV, where objectives are typically single-step and targets are comparatively homogeneous. Weather forecasting is harder to characterize in the same framework: autoregressive rollouts compound errors over long horizons, outputs couple many physical channels with disparate scales and predictability, and globally pooled test metrics can disagree sharply with per-channel, late-lead behavior implied by short-horizon training. We extend neural scaling analysis for autoregressive weather forecasting from single-step training loss to long rollouts and per-channel metrics. We quantify (1) how prediction error is distributed across channels and how its growth rate evolves with forecast horizon, (2) if power law scaling holds for test error, relative to rollout length when error is pooled globally, and (3) how that fit varies jointly with horizon and channel for parameter, data, and compute-based scaling axes. We find strong cross-channel and cross-horizon heterogeneity: pooled scaling can look favorable while many channels degrade at late leads. We discuss implications for weighted objectives, horizon-aware curricula, and resource allocation across outputs.

Kiefer Jr, Alexander [ORNL] (ORCID:000000025398874

Machine learning based prediction of airflow maldistribution in air-to-refrigerant heat exchangers

Flow maldistribution is a common challenge in heat exchanger (HX) design and particularly important for air-to-refrigerant geometries where capacity losses can approach 65%. This has a major impact on central air conditioning systems, as compact duct design motivates the use of A-type HXs which are known to be affected by airflow maldistribution. Because velocity profiles are difficult to predict, components are often oversized leading to increased material cost, system footprint, and refrigerant charge. Several studies detail airflow maldistribution for individual HXs and packages, but findings cannot always be extrapolated to new designs. In this work, a machine learning (ML) based flow profile prediction framework is developed and applied to two common package configurations: (i) A-type and (ii) U-type HXs, across a broad range of HX geometries and flow rates. Porous media CFD simulations are validated against independent data for both package types as well as comprehensive in house measurements for a finless geometry with shape optimized non-round tubes, which validates the framework for new heat transfer surfaces. The ML models are trained on the porous media CFD simulations, predicting volumetric flow rate (VFR) within 1.1% and 1.9% with maximum relative L 2 norm errors of 0.48 and 0.65, respectively, while also delivering 10 5 speed up factor compared to full porous media CFD. HX level simulations show an up to 9% reduction in heat transfer from flow maldistribution, with greater losses occurring at smaller half apex angles. This framework enables rapid and highly accurate prediction of airflow maldistribution induced capacity degradation.

42 ENGINEERING

On the Representativity of Electrode Microstructure Parameters and Their Electrochemical Response for Lithium Ion Batteries

Lithium-ion battery electrochemical models require an accurate description of the electrodes microstructures to be predictive that can be achieved through nanoscale imaging. Such observations are however limited by their field of view (FOV), as they provide only a subset of the whole electrode volume that does not necessarily represent the whole electrode microstructure heterogeneity, and therefore can bias the microstructure analysis. A microstructure scale electrochemical model was used to investigate lithium plating onset, material non-uniform utilization, and in-plane heterogeneities for an NMC-graphite full cell. To evaluate the representativeness, and thus relevance, of these model predictions, a coupled representativity analysis has been performed on the microstructure parameters and, in a novel way, on the full cell electrochemical response. Electrode microstructure parameters representativeness has been first quantified using the representative volume element (RVE) methodology. The RVE major flaw is that ultimately it can only conclude if a FOV contains representative subvolumes of the FOV, but not if the FOV itself is representative of the electrode volume. Analysis can conclude negatively ('FOV is not representative'), but not positively ('FOV is representative'). One major contribution of this work was to quantify the convergence of the RVE size with the FOV, to actually investigate the FOV representativeness and thus partly remedy this intrinsic limitation. The analysis determined that performing a standard RVE calculation, without exploring its FOV convergence, is likely to strongly underestimate the actual RVE size. The new RVE methodology has been automated in the NREL open-source Microstructure Analysis Toolbox (MATBOX) and is available to the battery community. Representativeness of microstructure parameters is however only an intermediate step, as the end-results of an electrochemical model are performances predictions. Indeed, what is the practical consequence of a given deviation for a microstructure parameter? The microstructure parameter deviation propagations to the 3D microstructure scale electrochemical response have been then quantified for different charge rates. This defines a threshold for the microstructure parameters FOV for a desired maximum deviation of the electrochemical response. Such deviation propagation analysis is analogous to error propagation analysis and is necessary to determine the relevance of microstructure scale model predictions for macroscale predictions. Electrochemical model shows cell representative section areas are increasing with C-rate, due to higher in-plane heterogeneities, indicating larger FOVs are required specifically for fast charge modeling. Therefore, we introduced the novel concept of electrochemical RVE (eRVE) that is a function of the operating conditions (thus defined as a dynamic RVE), with an increasing dependence with the C-rate. Representativity analysis of the investigated cell determined a FOV of 144.4 x 54.4 m2 is large enough to establish a convergence on the representative section areas for low to intermediate C-rate (=2.5C), but not large enough to conclude for higher rates. This work aims to emphasize the importance of representativity analysis for LIB electrode microstructures, as it is required to estimate the error, and thus the relevance, of microstructure parameters intended to be used in macroscale models. The methodology and results can help researchers to select the relevant imaging and associated FOV required to provide accurate enough microstructure parameters.

ADVANCED PROPULSION SYSTEMS

Searching for a Pulse: Evaluating the Use of Rapid DC Pulses for Diagnosing Battery Health, State-of-Charge, and Safety

Rapid electrochemical diagnostics, like DC pulse sequences or electrochemical impedance spectroscopy, are known to be useful for capacity prediction. However, it is unclear how previous results will map to different cell types and use cases and whether rapid diagnostics are useful for remaining useful life prediction or for detecting potential safety issues. To that end, we have collected a data set with ∼50,000 DC pulse measurements from four types of commercial lithium-ion batteries to enable training of state-of-charge, health, and safety diagnostic models via machine-learning. We demonstrate that 120-second DC pulse sequences can be used to predict capacity with 2%–9% average error, which can separate high- from low-capacity cells with only a 0.3% false positive rate but is not accurate enough to estimate remaining useful life. We also find that no safety related targets can be accurately predicted, highlighting the critical need for other non-invasive methods to diagnose battery safety.

25 ENERGY STORAGE

Assessing the Role of Hydrodynamics in Enhancing Height-Above-the-Nearest-Drainage Derived Synthetic Rating Curves: A Comparative Study in the Wu River Basin, Taiwan

The conventional approach to generating synthetic rating curves (SRC) using the Height-Above-the-Nearest-Drainage (HAND) method typically relies on the assumption of uniform flow, such as Manning's equation, to establish stage-discharge ratings. The zero-physics application of the uniform flow equation is insufficient for capturing detailed hydraulic features (e.g., backwater effect) and neglects the hydraulic effects from adjacent channels. This lack of hydrodynamic computation can impact the accuracy and effectiveness of riverine flood risk estimation and management. To reduce this foreseeable error, we introduce the HAND-hd workflow, which integrates sophisticated hydrodynamic computations in the production of HAND-based SRC with hydrodynamic features (SRC hd ). The results indicate that SRC hd demonstrates consistent agreement with both gauge observations and benchmark solutions. Additionally, the comparative analysis suggests that SRC hd provides notable improvements in stage-discharge ratings over conventional HAND-based SRCs, particularly in channels with mild bed gradients, where it reduces water stage prediction errors and percent biases. In steeper channel segments, SRC hd maintains comparable accuracy to conventional methods. The comprehensive evaluation in this study emphasizes the potential discrepancies and inaccuracies associated with the adoption of the uniform flow assumption in the conventional HAND-SRCs and addresses the necessity of including hydrodynamic physics in the application of HAND-based SRC (e.g., inundation map) in channels with mild gradients.

54 ENVIRONMENTAL SCIENCES

High-resolution national mapping of natural gas composition substantially updates methane leakage impacts

Methane is emitted from oil and gas operations alongside heavier hydrocarbons and non-hydrocarbon gases, shaping emissions management decision-making, including air quality impacts. Yet, most assessments assume fixed gas composition, overlooking significant spatial and temporal variations. Here, we generate a high-resolution, data-driven map of natural gas composition across the United States, reconstructing methane, heavier hydrocarbons, and non-hydrocarbon species using spatio-temporal interpolation and oil-and-gas production patterns. Our approach is able to reduce composition prediction errors by 39% in terms of Mean Absolute Error (MAE) compared to standard techniques and reveals that methane loss rates have been underestimated by more than 50% in some regions. Beyond methane, we uncover substantial variability in co-emitted gases, exposing blind spots in current emissions inventories and emissions management frameworks. Our work enables more accurate emissions assessments, guides targeted measurement strategies, and informs emissions management decision-making. It also provides a general framework for prediction in environmental applications that integrate sparse measurements with auxiliary variables.

03 NATURAL GAS

HAPPA: A Modular Platform for HPC Application Resilience Analysis with LLMs Embedded

High-performance computing (HPC) systems are increasingly vulnerable to soft errors, which pose significant challenges in maintaining computational accuracy and reliability. Predicting the resilience of HPC applications to these errors is crucial for robust code protection and detailed resilience analysis. In this study, we present HAppA, a modular platform designed for HPC Application Resilience Analysis. Embedding Large Language Models (LLMs), HAppA addresses understanding the context information of long code sequences typical in HPC applications. HAppA implements a novel code representation module that chunks the code into fixed-size segments and aggregates the embeddings of these segments. Three aggregation methods have been explored: MeanPooling, MaxPooling, and LSTM-based techniques. We built a DAtaset for REsilience analysis using Fault Injection (FI), named DARE. Using our DARE dataset, HAppA is trained for regression prediction tasks. Our evaluation results demonstrate the predictive accuracy of HAppA compared to other models, particularly noting that the LSTM-based aggregation method -- HAppA-LSTM -- achieves a mean squared error (MSE) of 0.078 for SDC prediction, surpassing the existing state-of-the-art PARIS model, which recorded an MSE of 0.1172. Additionally, HAppA with the KeyBERT model extracts a list of keywords representing the source code. A comprehensive importance analysis of these keywords further elucidates the code patterns contributing to the error rate. These findings highlight the effectiveness of HAppA in analyzing the resilience of HPC applications and establish a new benchmark for predictive accuracy in resilience.

Jiang, Hailong [Kent State University]

Logical error rates for the surface code under a mixed coherent and stochastic circuit-level noise model inspired by trapped ions

With fault-tolerant quantum computing (FTQC) on the horizon, it is critical to understand sources of logical errors in plausible hardware implementations of quantum error-correcting codes. Detailed error modeling of computational instructions on particular FTQC architectures will enable the better prediction of error propagation in FT-encoded quantum circuits while revealing where greater attention is needed in hardware design. In this work, we consider logical error rates for the surface code implemented on a hypothetical grid-based trapped-ion quantum charge-coupled device architecture. Specifically, we construct logical channels for the idling surface code and examine its diamond error under a mixed coherent and stochastic circuit-level noise model inspired by trapped ions. We include the coherent dephasing noise that is known to accumulate during physical qubit idling and transport in these systems, determining idling and transport durations using the time-resolved output of an open-source trapped-ion surface code compiler. To estimate expectation values of logical Pauli observables following hardware circuits containing non-Clifford sources of noise, we utilize a Monte Carlo technique to sample from an underlying quasiprobability distribution of Clifford circuits that we independently simulate in a phase-sensitive fashion. We verify error suppression up to code distance 𝑑 = 11 at coherent dephasing rates near and below those of current-generation trapped-ion quantum computers and find that logical error rates align with those of analogous fully stochastic simulations in this regime. Exploring higher dephasing rates at 𝑑 = 3−5, we find evidence for growing coherent rotations about all three logical Pauli axes, increased diagonal logical error process matrix elements relative to those of stochastic simulations, and a reduced dephasing rate threshold. Overall, our work paves a way toward realistic hardware emulation of small fault-tolerant quantum processes, e.g., members of an FTQC instruction set.

Quantum benchmarking

Surrogate-driven Variance-based Sensitivity Analysis of Thermal Storage Tanks in Integrated Energy Systems

Sensitivity analysis and uncertainty quantification are essential steps for enhancing the accuracy of computational models by identifying and mitigating uncertainties. This study focuses on these steps for the Thermal Energy Delivery System at Idaho National Laboratory, specifically targeting the thermocline tank. Using a Modelica/Dymola simulation model, the study perturbed various design parameters and boundary conditions, including shape factor, porosity, outlet temperature, inlet mass flow rate, and system pressure, to predict and quantify uncertainty in the tank’s ax- ial temperature. A dataset of over 1,000 simulations was generated, and surrogate models were developed using the pyMAISE (Michigan Artificial Intelligence Standard Environment) library, which is an Automatic Machine Learning library for nuclear engineering applications. The optimal model, a feedforward neural network with two hidden layers, achieved an R2 score above 0.99 and a mean absolute error below 1 Kelvin. Sensitivity analyses using Sobol indices and Fourier amplitude sensitivity testing methods on this surrogate model revealed that the inlet mass flow rate at initial timestamps and porosity significantly impacts predicted temperatures across all sensors and time steps.

22 - GENERAL STUDIES OF NUCLEAR REACTORS

Abatement of ionizing radiation for superconducting quantum devices

Ionizing radiation has been shown to reduce the performance of superconducting quantum circuits. In this report, we evaluate the expected contributions of different sources of ambient radioactivity for typical superconducting qubit experiment platforms. Our assessment of radioactivity inside a typical cryostat highlights the importance of selecting appropriate materials for the experiment components nearest to qubit devices, such as packaging and electrical interconnects. We present a shallow underground facility (30-meter water equivalent) to reduce the flux of cosmic rays and a lead shielded cryostat to abate the naturally occurring radiogenic gamma-ray flux in the laboratory environment. We predict that superconducting qubit devices operated in this facility could experience a reduced rate of correlated multi-qubit errors by a factor of approximately 20 relative to the rate in a typical above-ground, unshielded facility. Finally, we outline overall design improvements that would be required to further reduce the residual ionizing radiation rate, down to the limit of current generation direct detection dark matter experiments.

47 OTHER INSTRUMENTATION

Dynamic Modeling and Simulation of a Subcritical Coal-Fired Power Plant under Load-Following Conditions

Dynamic models for power plants that capture realistic general process trends and effects of manipulated variables are needed to improve load-following, while minimizing carbon footprint. In this work, a dynamic modeling approach and simulation results for subcritical coal-fired power plant components are presented. These encompass simulation of the dynamics in the fireside, including the effects of fuel, air combustion, and the dynamics of the entire waterside and power generation sections. This model development enables the simulation and analysis of the important short and long-time scale dynamics of components such as heaters, evaporative loop, and power generation units. Furthermore, additional variables in the power generation section are introduced to improve model accuracy, extending the prediction capability of subcritical power plant models and opening new opportunities for research in operator training, optimization, and advanced model-based controller design that are based on these models. The change in process gain for different ramp rates associated with disturbance signals that affect process variables is also explored and a correlation developed. This provides opportunities to study disturbance rejection control implementation and adaptation for scenarios with such variations in ramp rates. The prediction capabilities of selected components are compared to data available in literature, with the obtained root mean squared error ranges that reflect the model performance and quality of predictions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Error analysis of low-fidelity models for wake steering based on field measurements

The observations collected by two scanning lidars deployed on the roof of a 2.8-MW turbine undergoing a series of imposed yaw offsets are analyzed. The wake lateral displacement detected by the rear-facing lidar correlates well with the yaw offset sensed by the forward-facing lidar. We find that the high-frequency part of the yaw offset signal is connected to wake meandering, whereas the low frequency component is a good predictor for wake displacement due to yaw misalignment. Conditionally averaged wake velocity data for different yaw offsets are used as benchmarks for the validation of a linearized Reynolds-averaged Navier-Stokes and an empirical wake model. A mean error as low as 2% and a good prediction of the wake trajectory are achieved, provided that the wake recovery rate matches the observations.

17 WIND ENERGY

Empirically-calibrated H100 node power models for accurate AI training energy estimation

Accurately quantifying the energy use of artificial intelligence (AI) training is critical for infrastructure planning, carbon accounting, and sustainable data center operation, but few studies have directly measured the power consumption of production workloads on contemporary hardware. By combining empirical measurements from Brookhaven National Laboratory during AI training on 8-graphics-processing-unit H100 systems with open-source benchmarking data, we develop statistical models relating computational intensity to node-level power consumption. We measure the gap between manufacturer-rated thermal design power (TDP) and actual power demand during AI training. Our analysis reveals that even computationally intensive workloads operate at only 76% of the 10.2 kW TDP rating. Our architecture-specific model, calibrated to floating-point operations, predicts energy consumption with 11.4% mean absolute percentage error, significantly outperforming TDP-based approaches (27%–37% error). We identified distinct power signatures between transformer and convolutional neural network architectures, with transformers showing characteristic fluctuations that may impact grid stability. These results provide a measurement-grounded basis for improving AI training energy estimates, enabling more reliable infrastructure sizing, cost projections, and environmental impact assessments.

Newkirk, Alex C

Evaluating algorithmic bias on biomarker classification of breast cancer pathology reports

Objectives: This work evaluated algorithmic bias in biomarkers classification using electronic pathology reports from female breast cancer cases. Bias was assessed across 5 subgroups: cancer registry, race, Hispanic ethnicity, age at diagnosis, and socioeconomic status. Materials and Methods: We utilized 594 875 electronic pathology reports from 178 121 tumors diagnosed in Kentucky, Louisiana, New Jersey, New Mexico, Seattle, and Utah to train 2 deep-learning algorithms to classify breast cancer patients using their biomarkers test results. We used balanced error rate (BER), demographic parity (DP), equalized odds (EOD), and equal opportunity (EOP) to assess bias. Results: We found differences in predictive accuracy between registries, with the highest accuracy in the registry that contributed the most data (Seattle Registry, BER ratios for all registries >1.25). BER showed no significant algorithmic bias in extracting biomarkers (estrogen receptor, progesterone receptor, human epidermal growth factor receptor 2) for race, Hispanic ethnicity, age at diagnosis, or socioeconomic subgroups (BER ratio <1.25). DP, EOD, and EOP all showed insignificant results. Discussion: We observed significant differences in BER by registry, but no significant bias using the DP, EOD, and EOP metrics for socio-demographic or racial categories. This highlights the importance of employing a diverse set of metrics for a comprehensive evaluation of model fairness. Conclusion: A thorough evaluation of algorithmic biases that may affect equality in clinical care is a critical step before deploying algorithms in the real world. We found little evidence of algorithmic bias in our biomarker classification tool. Artificial intelligence tools to expedite information extraction from clinical records could accelerate clinical trial matching and improve care.

60 APPLIED LIFE SCIENCES

Development and evaluation of a new 4DEnVar-based weakly coupled ocean data assimilation system in E3SMv2

The development, implementation, and evaluation of a new weakly coupled ocean data assimilation (WCODA) system for the fully coupled Energy Exascale Earth System Model version 2 (E3SMv2) utilizing the four-dimensional ensemble variational (4DEnVar) method are presented in this study. The 4DEnVar method, based on the dimension-reduced projection four-dimensional variational (DRP-4DVar) approach, replaces the adjoint model with the ensemble technique, thereby reducing computational demands. Monthly mean ocean temperature and salinity data from the EN4.2.1 reanalysis are integrated into the ocean component of E3SMv2 from 1950 to 2021 with the goal of providing realistic initial conditions for decadal predictions and predictability studies. The performance of the WCODA system is assessed using various metrics, including the reduction rate of the cost function, root mean square error (RMSE) differences, correlation differences, and model biases. Results indicate that the WCODA system effectively assimilates the reanalysis data into the climate model, consistently achieving negative reduction rates of the cost function and notable improvements in RMSE and correlation across various ocean layers and regions. Significant enhancements are observed in the upper ocean layers across the majority of global ocean regions, particularly in the north Atlantic, north Pacific, and Indian Ocean. Model biases in sea surface temperature and salinity are also substantially reduced. For sea surface temperature, cold biases in the north Pacific and north Atlantic are diminished by about 1–2 °C, and warm biases in the Southern Ocean are corrected by approximately 1.5–2.5 °C. In terms of salinity, improvements are observed with bias reductions of about 0.5–1 psu in the north Atlantic and north Pacific and up to 1.5 psu in parts of the Southern Ocean. The ultimate goal of the WCODA system is to advance the predictive capabilities of E3SM for subseasonal to decadal climate predictions, thereby supporting research on strategic energy-sector policies and planning.

54 ENVIRONMENTAL SCIENCES