Search NASA⌕ Search

SEARCH · Search NASA

Results for “model interpretability”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Use of Physics to Improve Solar Forecast: Part II, Machine Learning and Model Interpretability

Machine learning (ML) models have been applied to forecast solar energy; however, they often lack clarity of interpretability and underlying physics. This work addresses such challenges by developing a hierarchy of ML models that gradually introduce predictors to improve the forecast accuracy based on a physics-based framework. Three ML models (ARIMA, LSTM, and XGBoost) are examined and compared with four physics-informed persistence models reported in Part I and the simple persistence model to assess the improvement of different models. The 7-year measurements at the U.S. Department of Energy's Atmospheric Radiation Measurement's Southern Great Plains Central Facility site are used for forecasts and evaluations. The results reveal that the step-by-step introduction of predictors leads to different improvements for models at different hierarchical levels. Comparison of the ML models with persistence models shows that LSTM and XGBoost outperform all the persistence models, with LSTM having the overall best performance; however, ARIMA underperforms the four physics-informed persistence models. This study demonstrates the importance and utility of incorporating physics into ML models in improving forecast accuracy by introducing a hierarchy of physics-based predictors, distinguishing predictor contributions, and enhancing the ML interpretability. The combined use of Global Horizontal Irradiance (GHI) and Direct Normal Irradiance (DNI) significantly improves the forecast accuracy compared to using individual irradiances alone because the pair contains more information on cloud-radiation interactions.

interpretability↗

Interpretable Models for Workflow Differentiation in High-Performance Scientific Networks

Scientific workflows in high-performance networks spawn hundreds of interdependent flows that must be managed collectively—yet existing network classifiers treat each flow in isolation, leading to fragmented QoS decisions and missed interflow patterns. We present a novel traffic classification solution that operates at the workflow level, distinguishing entire filetransfer operations from streaming analytics by capturing how concurrent flows interact and burst together. We introduce a workflow identification window (WIW) that ingests raw packet headers from parallel flows into unified tensors, preserving the spatial-temporal patterns that differentiate scientific workflows. This approach achieves 98.7% accuracy using CNN, LSTM, and hybrid architectures, while maintaining 84% accuracy on production traffic collected a week later—demonstrating robustness to temporal drift. By integrating SHAP and GradCAM explainability, we reveal that early-packet timing patterns and cross-flow correlations drive classification decisions, providing operators with interpretable insights. Our system enables coherent workflow-level QoS enforcement and dynamic bandwidth allocation in scientific networks, eliminating manual per-flow configuration while maintaining classification latency at millisecond level.

Giannakou, Anna [LBL, Berkeley]↗

Radar imaging of glaciovolcanic stratigraphy, Mount Wrangell caldera, Alaska - Interpretation model and results

Glaciological measurements and an airborne radar sounding survey of the glacier lying in Mount Wrangell caldera raise many questions concerning the glacier thermal regime and volcanic history of Mount Wrangell. An interpretation model has been developed that allows the depth variation of temperature, heat flux, pressure, density, ice velocity, depositional age, and thermal and dielectric properties to be calculated. Some predictions of the interpretation model are that the basal ice melting rate is 0.64 m/yr and the volcanic heat flux is 7.0 W/sq m. By using the interpretation model to calculate two-way travel time and propagation losses, radar sounding traces can be transformed to give estimates of the variation of power reflection coefficient as a function of depth and depositional age. Prominent internal reflecting zones are located at depths of approximately 59-91m, 150m, 203m, and 230m. These internal reflectors are attributed to buried horizons of acidic ice, possibly intermixed with volcanic ash, that were deposited during past eruptions of Mount Wrangell.

Clarke, Garry K. C.↗

Knowledge-guided learning with curated prior genetic biomarkers for robust model interpretation

Abstract Motivation Knowledge-guided learning offers effective and robust model training strategies in data-scarce settings by incorporating established domain knowledge, thereby enhancing generalization, robustness, and interpretability. By contrast, conventional deep learning approaches rely purely on data-driven learning, which can limit robust model interpretability, particularly in high-dimensional settings with limited size samples. In computational biology, knowledge-guided learning has primarily leveraged network- and structural-based knowledge, leading to biologically interpretable representations and enhanced predictive performance compared to conventional approaches. However, curated biomarkers, one of the most accessible forms of biological knowledge, remain largely unexplored within knowledge-guided paradigms. Results In this study, we propose a model-agnostic training paradigm, Biomarker-driven Explainable Prior-guided Learning (BioExPL), that can be applied to any neural networks that incorporates curated prior knowledge. BioExPL enforces neural networks to reflect curated biomarker priors in their latent representations through a novel knowledge-alignment loss. BioExPL consistently demonstrated significantly improved predictive performance and enhanced model interpretability with minimized computational overhead in simulation studies and intensive experiments on multiple cancer datasets. BioExPL not only integrates prior curated knowledge into the model but also accurately identifies unknown associated signals additionally. BioExPL is model-agnostic and domain-independent, enabling its integration into diverse neural network architectures. Availability and implementation The open-source is publicly available at: https://github.com/datax-lab/BioExPL.

Baek, Beomsu [Department of Computer Science, Univ↗

Overview of interpretive modelling of fusion performance in JET DTE2 discharges with TRANSP

In the paper we present an overview of interpretive modelling of a database of JET-ILW 2021 D-T discharges using the TRANSP code. The main aim is to assess our capability of computationally reproducing the fusion performance of various D-T plasma scenarios using different external heating and D-T mixtures, and to understand the performance driving mechanisms. We find that interpretive simulations confirm a general power-law relationship between increasing external heating power and fusion output, which is supported by absolutely calibrated neutron yield measurements. A comparison of measured and computed D-T neutron rates shows that the calculations' discrepancy depends on the absolute neutron yield. The calculations are found to agree well with measurements for higher performing discharges with external heating power above ~20 MW, while low-neutron shots display an average discrepancy of around +40% compared to measured neutron yields. A similar trend is found for the ratio between thermal and beam-target fusion, where larger discrepancies are seen in shots with dominant beam-driven performance. We compare the observations to studies of JET-ILW D discharges, to find that on average the fusion performance is well modelled over a range of heating power, although an increased unsystematic deviation for lower-performing shots is observed. The ratio between thermal and beam-induced D-T fusion is found to be increasing weakly with growing external heating power, with a maximum value of ≳1 achieved in a baseline scenario experiment. An evaluation of the fusion power computational uncertainty shows a strong dependence on the plasma scenario type and fusion drive characteristics, varying between ±25% and 35%. D-T fusion alpha simulations show that the ratio between volume-integrated electron and ion heating from alphas is ≲10 for the majority of analysed discharges. Alphas are computed to contribute between ~15% and 40% to the total electron heating in the core of highest performing D-T discharges. An alternative workflow to TRANSP was employed to model JET D-T plasmas with the highest fusion yield and dominant non-thermal fusion component because of the use of fundamental radio-frequency heating of a large minority in the scenario, which is calculated to have provided ~10% to the total fusion power.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Data from: "Towards CONUS-Wide ML-Augmented Conceptually-Interpretable Modeling of Catchment-Scale Precipitation-Storage-Runoff Dynamics"

This data package was generated to support the manuscript “Towards CONUS-Wide Machine Learning-Augmented Conceptually Interpretable Modeling of Catchment-Scale Precipitation-Storage-Runoff Dynamics.” It provides input files, model outputs, plotting data, scripts, notebooks, and documentation used to develop, evaluate, and reproduce Mass-Conserving Perceptron (MCP)-based hydrologic modeling experiments across 513 selected Catchment Attributes and Meteorology for Large-sample Studies in the United States (CAMELS-US) basins. The files are organized by modeling component and analysis purpose, including rainfall–runoff experiments, snow module experiments, coupled hydrologic-snow experiments, Long Short-Term Memory (LSTM) benchmark results, model skill metrics, initialization and epoch records, cell-state normalization files, Akaike Information Criterion (AIC)-based model comparison files, and data used to generate manuscript figures. Tabular files can be opened using standard spreadsheet software or Python/R data-analysis tools. Python scripts, Jupyter notebooks, and selected MATLAB scripts are included for model execution, postprocessing, plotting, and statistical analysis. Quality assurance and quality control were conducted through the source-data selection and modeling workflow. Meteorological forcing, streamflow, and static catchment attributes were derived from the CAMELS-US dataset, and snow water equivalent data were derived from the University of Arizona (UA) Snow Water Equivalent dataset. Selected basins and time periods were screened during the associated research workflow to avoid missing observations or poor-quality cases. Static geospatial features were processed primarily using Quantum Geographic Information System (QGIS) and Geospatial Data Abstraction Library (GDAL) workflows. Additional details are provided in the associated manuscript and documentation.

ESS-DIVE CSV File Formatting Guidelines Reporting ↗

Exploring the Whole Set of Accurate Sparse Interpretable Models

In data science applications, there are often many models that fit the data well. This phenomenon was called the Rashomon Effect by Leo Breiman. The set of good models is called the Rashomon Set, and the goal of this project is to locate, store, and study the Rashomon sets for classes of interpretable models, including decision trees and generalized additive models.

97 MATHEMATICS AND COMPUTING↗

An interpretable model of pre-mRNA splicing for animal and plant genes

Pre-mRNA splicing is a fundamental step in gene expression, conserved across eukaryotes, in which the spliceosome recognizes motifs at the 3' and 5' splice sites (SSs), excises introns, and ligates exons. SS recognition and pairing is often influenced by protein splicing factors (SFs) that bind to splicing regulatory elements (SREs). Here, we describe SMsplice, a fully interpretable model of pre-mRNA splicing that combines models of core SS motifs, SREs, and exonic and intronic length preferences. We learn models that predict SS locations with 83 to 86% accuracy in fish, insects, and plants and about 70% in mammals. Learned SRE motifs include both known SF binding motifs and unfamiliar motifs, and both motif classes are supported by genetic analyses. Our comparisons across species highlight similarities between non-mammals, increased reliance on intronic SREs in plant splicing, and a greater reliance on SREs in mammalian splicing.

59 BASIC BIOLOGICAL SCIENCES↗

Interpretive modeling of tungsten divertor leakage during experiments with neon gas seeding

Abstract Many existing and future tokamaks with tungsten divertors operate, or will operate, with low- Z impurity seeding, but the direct effect of these seeded impurities on tungsten Scrape-off-Layer (SOL) transport has not been explored in detail. This paper reports on a DIII-D experiment designed to test how tungsten divertor leakage from the Small-Angle Slot V-Shaped, tungsten-coated divertor is impacted by neon seeding at a variety of injection rates and poloidal injection locations. Measurements from the experiment show an inverse relationship between the neon injection rate and the tungsten core penetration factor. Interpretive modeling is performed with a combination of the SOLPS-ITER and DIVIMP codes to assess the underlying tungsten behavior. The modeling results show that the reduction in tungsten divertor leakage is driven by both an increase in the divertor collisionality as well as a reduction in the ion temperature gradient near the divertor target. Collisions between low- Z impurities and tungsten impurities are found to have a significant impact on the tungsten SOL transport, such that ignoring the low- Z impurity collisional effects on the tungsten transport can result in an overestimate of the divertor leakage by an order-of-magnitude. Given the importance of these localized interactions, neon seeding from the closed, slot-like divertor has a clear advantage in being able to reduce tungsten divertor leakage without the high levels of neon core contamination that occur when seeding from other poloidal locations.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Super-X and conventional divertor configurations in MAST-U ohmic L-mode; a comparison facilitated by interpretative modelling

Measurements are presented, alongside corresponding interpretative SOLPS-ITER simulations, of the first MAST-U experiments comparing ohmically heated L-mode fuelling scans in Conventional divertor (CD) and Super-X divertor (SXD) configurations. In experiment, at comparable outer mid-plane separatrix electron density, $n_{e,\textrm{sep,OMP}}$, the maximum lower outer target heat load was found to be a factor 16 $\,\pm\,7$ lower in SXD compared to CD. In simulation, a factor 26.8 reduction was found (slightly higher than the experimental range), suggesting an additional reduction in SXD compared to the factor 9.3 expected from geometric considerations alone. According to the simulations, this additional reduction in the SXD is due to a net radial transport of the energy remaining downstream of the $T_e = 5$ eV location. This energy is carried out of the critical (highest heat load) flux tube by deuterium atoms, demonstrating the importance of a longer legged divertor which provides space for this to occur. Importantly, in both simulation and experiment, the SXD has minimal impact on the upstream n e and T e profiles. Spectral inferences of detachment front movement in SXD compare well between simulation and experiment. In regions of high magnetic field gradient, the parallel movement of the front towards the X-point becomes less sensitive to increasing $n_{e,\textrm{sep,OMP}}$, in qualitative agreement with simplified models and previous predictive simulations. Additional aspects, regarding the target ion flux rollover, upstream separatrix temperature and drift effects, are also presented and discussed.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

A Control Model: Interpretation of Fitts' Law

The analytical results for several models are given: a first order model where it is assumed that the hand velocity can be directly controlled, and a second order model where it is assumed that the hand acceleration can be directly controlled. Two different types of control-laws are investigated. One is linear function of the hand error and error rate; the other is the time-optimal control law. Results show that the first and second order models with the linear control-law produce a movement time (MT) function with the exact form of the Fitts' Law. The control-law interpretation implies that the effect of target width on MT must be a result of the vertical motion which elevates the hand from the starting point and drops it on the target at the target edge. The time optimal control law did not produce a movement-time formula simular to Fitt's Law.

Connelly, E. M.↗

Model interpretation of type III radio burst characteristics. II - Temporal aspects

A model of the radio emission region for kilometric type II bursts is used to interpret systematic variations observed in the temporal behavior of the burst parameters. The temporal behavior of the burst parameters observed by the ISEE-3 spacecraft is reviewed, and it is pointed out that the phase and modulation of the antenna signal vary in a systematic way with both time and observing frequency. The source azimuth is observed to drift with time, with the magnitude and sense of the drift depending on the location of the radio source relative to the observer. The modulation factor usually decreases uniformly with time and is frequently peaked near the burst onset. The model of the radio emission region is developed and used to obtain the intensity, phase, and modulation of the radio signal. Model results are used to show how the behavior of the burst parameters are related to attributes of the source region. It is shown that the temporal behavior of the radio parameters for an observed type II burst is well represented by the model.

Reiner, M. J.↗

An accurate and interpretable model for antimicrobial resistance in pathogenic Escherichia coli from livestock and companion animal species

Understanding the microbial genomic contributors to antimicrobial resistance (AMR) is essential for early detection of emerging AMR infections, a pressing global health threat in human and veterinary medicine. Here we used whole genome sequencing and antibiotic susceptibility test data from 980 disease causing Escherichia coli isolated from companion and farm animals to model AMR genotypes and phenotypes for 24 antibiotics. We determined the strength of genotype-to-phenotype relationships for 197 AMR genes with elastic net logistic regression. Model predictors were designed to evaluate different potential modes of AMR genotype translation into resistance phenotypes. Our results show a model that considers the presence of individual AMR genes and total number of AMR genes present from a set of genes known to confer resistance was able to accurately predict isolate resistance on average (mean F 1 score = 98.0%, SD = 2.3%, mean accuracy = 98.2%, SD = 2.7%). However, fitted models sometimes varied for antibiotics in the same class and for the same antibiotic across animal hosts, suggesting heterogeneity in the genetic determinants of AMR resistance. We conclude that an interpretable AMR prediction model can be used to accurately predict resistance phenotypes across multiple host species and reveal testable hypotheses about how the mechanism of resistance may vary across antibiotics within the same class and across animal hosts for the same antibiotic.

Chung, Henri C.↗

Three-dimensional model interpretation of NO(x) measurements from the lower stratosphere

A three-dimensional off-line chemistry transport model, driven by European Center for Medium-Range Forecasts winds and temperatures, is used to interpret measurements of NO and NO2 taken from the DC-8 during the second Airborne Arctic Stratospheric Expedition. The model was run in three configurations: gas phase chemistry alone, inclusion of the N2O5 aerosol reaction, and inclusion of both N2O5 and ClONO2 aerosol reactions. The run including the N2O5 aerosol reaction alone usually agreed best with measured NO(x)/NO(y) ratios in midlatitude air masses. The NO(x)/NO(y) ratios of the run with both aerosol reactions were always too low, while the gas phase ratios were usually too high, especially during March. All three simulations generated extremely low NO2/NO(y) ratios in air parcels that had spent several days or more in the polar night. Measured NO2/NO(y) ratios in these types of air masses were sometimes equally low but could also be considerably higher. Observed NO/NO2 ratios differed strongly from known theory.

Folkins, Ian↗

Extreme sparsification of physics-augmented neural networks for interpretable model discovery in mechanics

Data-driven constitutive modeling with neural networks has received increased interest in recent years due to its ability to easily incorporate physical and mechanistic constraints and to overcome the challenging and time-consuming task of formulating phenomenological constitutive laws that can accurately capture the observed material response. However, even though neural network-based constitutive laws have been shown to generalize proficiently, the generated representations are not easily interpretable due to their high number of trainable parameters. Sparse regression approaches exist that allow for obtaining interpretable expressions, but the user is tasked with creating a library of model forms which by construction limits their expressiveness to the functional forms provided in the libraries. Here, in this work, we propose to train regularized physics-augmented neural network-based constitutive models utilizing a smoothed version of $L^0$-regularization. This aims to maintain the trustworthiness inherited by the physical constraints, but also enables interpretability which has not been possible thus far on any type of machine learning-based constitutive model where model forms were not assumed a priori but were actually discovered. During the training process, the network simultaneously fits the training data and penalizes the number of active parameters, while also ensuring constitutive constraints such as thermodynamic consistency. We show that the method can reliably obtain interpretable and trustworthy constitutive models for compressible and incompressible hyperelasticity, yield functions, and hardening models for elastoplasticity, using synthetic and experimental data. This work aims to set a new paradigm for interpretable machine learning models in the broad area of solid mechanics where low and limited data is available along with prior knowledge of physical constraints that the learned maps need to obey. This paradigm can potentially be extended to a broader spectrum of scientific exploration.

Data-driven constitutive models↗

Postural control model interpretation of stabilogram diffusion analysis

Collins and De Luca [Collins JJ. De Luca CJ (1993) Exp Brain Res 95: 308-318] introduced a new method known as stabilogram diffusion analysis that provides a quantitative statistical measure of the apparently random variations of center-of-pressure (COP) trajectories recorded during quiet upright stance in humans. This analysis generates a stabilogram diffusion function (SDF) that summarizes the mean square COP displacement as a function of the time interval between COP comparisons. SDFs have a characteristic two-part form that suggests the presence of two different control regimes: a short-term open-loop control behavior and a longer-term closed-loop behavior. This paper demonstrates that a very simple closed-loop control model of upright stance can generate realistic SDFs. The model consists of an inverted pendulum body with torque applied at the ankle joint. This torque includes a random disturbance torque and a control torque. The control torque is a function of the deviation (error signal) between the desired upright body position and the actual body position, and is generated in proportion to the error signal, the derivative of the error signal, and the integral of the error signal [i.e. a proportional, integral and derivative (PID) neural controller]. The control torque is applied with a time delay representing conduction, processing, and muscle activation delays. Variations in the PID parameters and the time delay generate variations in SDFs that mimic real experimental SDFs. This model analysis allows one to interpret experimentally observed changes in SDFs in terms of variations in neural controller and time delay parameters rather than in terms of open-loop versus closed-loop behavior.

NASA Program Biomedical Research and Countermeasur↗

Hypothesis testing via AI: Generating physically interpretable models of scientific data with machine learning (Full Technical Report)

Deep learning has demonstrated an exceptional ability to solve complex tasks (an engineering success); however, it has done so at the expense of the ability to generate new knowledge (a scientific failure). We propose an alternative framework—entitled Deep Symbolic Regression (DSR)—in which artificial neural networks (NNs) rapidly generate hypotheses about physical relationships among inputs. This framework bypasses the need to interpret an NN altogether, while still leveraging the representational power of deep learning. The resulting models are tractable mathematical expressions, which are inherently and readily human interpretable and can provide insights into underlying physical phenomena. Further, we fold this methodology into the scientific process by allowing the scientist to directly integrate a priori knowledge and beliefs to accelerate learning. We demonstrate this methodology on symbolic regression—the problem of rediscovering underlying expressions describing a dataset—and achieve state-of-the-art performance across a wide variety of symbolic regression problems. Further, we generalize our DSR framework to apply to the more general class of symbolic optimization problems, in which one seeks to optimize a sequence of symbols or “tokens” under a black-box reward function. Examples of other symbolic optimization problems include neural architecture search and computational antibody design. Our generalized tool, Deep Symbolic Optimization (DSO), has been demonstrated on the task of learning symbolic control policies for reinforcement learning environments, and has been adopted as an enabling capability for computational antibody design.

97 MATHEMATICS AND COMPUTING↗

Scientific visualization tools for the ISTP project: Mission planning, data analysis and model interpretation

Visualization tools are being developed to meet the challenges of mission planning and data analysis presented by the International Solar-Terrestrial Physics (ISTP) program. ISTP encompasses a large number of spacecraft, multiple ground-based observatories, and several theoretical investigations, with the goal of understanding the global behavior of the solar wind/magnetosphere/ionosphere system. The tools include three-dimensional displays of key boundaries in geospace along with spacecraft trajectories, which can be animated and synchronized to universal time. Magnetic field models and MHD simulation results can be invoked to reveal the magnetic topology or to identify magnetic conjunctions between spacecraft and/or ground-based facilities. Simultaneous displays of satellite trajectories, spacecraft-borne observations, and model predictions are available to facilitate data processing and interpretation efforts. The current status of these tools is described, and their implementation at the ISTP Science Planning and Operations Facility and distribution to the entire ISTP community are discussed.

Peredo, M.↗