Search NASASearch

DOE OSTI · 2513960

Moving beyond post hoc explainable artificial intelligence: a perspective paper on lessons learned from dynamical climate modeling

Abstract

AI models are criticized as being black boxes, potentially subjecting climate science to greater uncertainty. Explainable artificial intelligence (XAI) has been proposed to probe AI models and increase trust. In this review and perspective paper, we suggest that, in addition to using XAI methods, AI researchers in climate science can learn from past successes in the development of physics-based dynamical climate models. Dynamical models are complex but have gained trust because their successes and failures can sometimes be attributed to specific components or sub-models, such as when model bias is explained by pointing to a particular parameterization. We propose three types of understanding as a basis to evaluate trust in dynamical and AI models alike: (1) instrumental understanding, which is obtained when a model has passed a functional test; (2) statistical understanding, obtained when researchers can make sense of the modeling results using statistical techniques to identify input–output relationships; and (3) component-level understanding, which refers to modelers' ability to point to specific model components or parts in the model architecture as the culprit for erratic model behaviors or as the crucial reason why the model functions well. We demonstrate how component-level understanding has been sought and achieved via climate model intercomparison projects over the past several decades. Such component-level understanding routinely leads to model improvements and may also serve as a template for thinking about AI-driven climate science. Currently, XAI methods can help explain the behaviors of AI models by focusing on the mapping between input and output, thereby increasing the statistical understanding of AI models. Yet, to further increase our understanding of AI models, we will have to build AI models that have interpretable components amenable to component-level understanding. We give recent examples from the AI climate science literature to highlight some recent, albeit limited, successes in achieving component-level understanding and thereby explaining model behavior. The merit of such interpretable AI models is that they serve as a stronger basis for trust in climate modeling and, by extension, downstream uses of climate model data.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

O'Loughlin, Ryan J. [City Univ. of New York (CUNY), NY (United States). Queens College], Li, Dan [City Univ. of New York (CUNY), NY (United States). Baruch College] (ORCID:0000000257459838), Neale, Richard [National Center for Atmospheric Research (NCAR), Boulder, CO (United States)] (ORCID:0000000342223918), O'Brien, Travis A. [Indiana Univ., Bloomington, IN (United States); Lawrence Berkeley National Laboratory (LBNL), Berkeley, CA (United States)] (ORCID:0000000266431175). 2025-02-11. Moving beyond post hoc explainable artificial intelligence: a perspective paper on lessons learned from dynamical climate modeling. https://doi.org/10.5194/gmd-18-787-2025

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related reports

Water4Energy Step-1 Band-M Ready-to-Train Samples for TVA Weeks-to-Years Prediction, Version 0

AI-ready Band-M (monthly) labelled training pack for the Water4Energy Genesis Task-1 project on weeks-to-years prediction of Tennessee Valley temperature and precipitation. The deposit includes leakage-aware issue-time samples (samples_M_v0.nc; N=486), train-only scalers, issue-time split table, supporting monthly panels, and Python generation scripts to recreate the pack from the companion Tier-1 raw observation collection (https://doi.org/10.13139/ORNLNCCS/3398576). Each sample pairs a 12-month lookback of teleconnection indices and SST box anomalies with TVA-mean ERA5 anomaly targets (t2m, tp, msl) at leads 1–3 months.

54 ENVIRONMENTAL SCIENCES

Multi-Angle Snowflake Camera, particle analysis

The c1 level data product for the Mutli-Angle Snowflake Camera contains snowflake fall speeds and particle size, among other analysis for images associated with each hydrometeor.

54 ENVIRONMENTAL SCIENCES

Multi-Angle Snowflake Camera, time bins

The c1 level data product for the Mutli-Angle Snowflake Camera contains snowflake fall speeds and particle size, among other analysis for images associated with each hydrometeor.

54 ENVIRONMENTAL SCIENCES