Search NASA⌕ Search

SEARCH · Search NASA

Results for “interpretable models”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Interpretable Models for Workflow Differentiation in High-Performance Scientific Networks

Scientific workflows in high-performance networks spawn hundreds of interdependent flows that must be managed collectively—yet existing network classifiers treat each flow in isolation, leading to fragmented QoS decisions and missed interflow patterns. We present a novel traffic classification solution that operates at the workflow level, distinguishing entire filetransfer operations from streaming analytics by capturing how concurrent flows interact and burst together. We introduce a workflow identification window (WIW) that ingests raw packet headers from parallel flows into unified tensors, preserving the spatial-temporal patterns that differentiate scientific workflows. This approach achieves 98.7% accuracy using CNN, LSTM, and hybrid architectures, while maintaining 84% accuracy on production traffic collected a week later—demonstrating robustness to temporal drift. By integrating SHAP and GradCAM explainability, we reveal that early-packet timing patterns and cross-flow correlations drive classification decisions, providing operators with interpretable insights. Our system enables coherent workflow-level QoS enforcement and dynamic bandwidth allocation in scientific networks, eliminating manual per-flow configuration while maintaining classification latency at millisecond level.

Giannakou, Anna [LBL, Berkeley]↗

Radar imaging of glaciovolcanic stratigraphy, Mount Wrangell caldera, Alaska - Interpretation model and results

Glaciological measurements and an airborne radar sounding survey of the glacier lying in Mount Wrangell caldera raise many questions concerning the glacier thermal regime and volcanic history of Mount Wrangell. An interpretation model has been developed that allows the depth variation of temperature, heat flux, pressure, density, ice velocity, depositional age, and thermal and dielectric properties to be calculated. Some predictions of the interpretation model are that the basal ice melting rate is 0.64 m/yr and the volcanic heat flux is 7.0 W/sq m. By using the interpretation model to calculate two-way travel time and propagation losses, radar sounding traces can be transformed to give estimates of the variation of power reflection coefficient as a function of depth and depositional age. Prominent internal reflecting zones are located at depths of approximately 59-91m, 150m, 203m, and 230m. These internal reflectors are attributed to buried horizons of acidic ice, possibly intermixed with volcanic ash, that were deposited during past eruptions of Mount Wrangell.

Clarke, Garry K. C.↗

Knowledge-guided learning with curated prior genetic biomarkers for robust model interpretation

Abstract Motivation Knowledge-guided learning offers effective and robust model training strategies in data-scarce settings by incorporating established domain knowledge, thereby enhancing generalization, robustness, and interpretability. By contrast, conventional deep learning approaches rely purely on data-driven learning, which can limit robust model interpretability, particularly in high-dimensional settings with limited size samples. In computational biology, knowledge-guided learning has primarily leveraged network- and structural-based knowledge, leading to biologically interpretable representations and enhanced predictive performance compared to conventional approaches. However, curated biomarkers, one of the most accessible forms of biological knowledge, remain largely unexplored within knowledge-guided paradigms. Results In this study, we propose a model-agnostic training paradigm, Biomarker-driven Explainable Prior-guided Learning (BioExPL), that can be applied to any neural networks that incorporates curated prior knowledge. BioExPL enforces neural networks to reflect curated biomarker priors in their latent representations through a novel knowledge-alignment loss. BioExPL consistently demonstrated significantly improved predictive performance and enhanced model interpretability with minimized computational overhead in simulation studies and intensive experiments on multiple cancer datasets. BioExPL not only integrates prior curated knowledge into the model but also accurately identifies unknown associated signals additionally. BioExPL is model-agnostic and domain-independent, enabling its integration into diverse neural network architectures. Availability and implementation The open-source is publicly available at: https://github.com/datax-lab/BioExPL.

Baek, Beomsu [Department of Computer Science, Univ↗

Data from: "Towards CONUS-Wide ML-Augmented Conceptually-Interpretable Modeling of Catchment-Scale Precipitation-Storage-Runoff Dynamics"

This data package was generated to support the manuscript “Towards CONUS-Wide Machine Learning-Augmented Conceptually Interpretable Modeling of Catchment-Scale Precipitation-Storage-Runoff Dynamics.” It provides input files, model outputs, plotting data, scripts, notebooks, and documentation used to develop, evaluate, and reproduce Mass-Conserving Perceptron (MCP)-based hydrologic modeling experiments across 513 selected Catchment Attributes and Meteorology for Large-sample Studies in the United States (CAMELS-US) basins. The files are organized by modeling component and analysis purpose, including rainfall–runoff experiments, snow module experiments, coupled hydrologic-snow experiments, Long Short-Term Memory (LSTM) benchmark results, model skill metrics, initialization and epoch records, cell-state normalization files, Akaike Information Criterion (AIC)-based model comparison files, and data used to generate manuscript figures. Tabular files can be opened using standard spreadsheet software or Python/R data-analysis tools. Python scripts, Jupyter notebooks, and selected MATLAB scripts are included for model execution, postprocessing, plotting, and statistical analysis. Quality assurance and quality control were conducted through the source-data selection and modeling workflow. Meteorological forcing, streamflow, and static catchment attributes were derived from the CAMELS-US dataset, and snow water equivalent data were derived from the University of Arizona (UA) Snow Water Equivalent dataset. Selected basins and time periods were screened during the associated research workflow to avoid missing observations or poor-quality cases. Static geospatial features were processed primarily using Quantum Geographic Information System (QGIS) and Geospatial Data Abstraction Library (GDAL) workflows. Additional details are provided in the associated manuscript and documentation.

ESS-DIVE CSV File Formatting Guidelines Reporting ↗

Exploring the Whole Set of Accurate Sparse Interpretable Models

In data science applications, there are often many models that fit the data well. This phenomenon was called the Rashomon Effect by Leo Breiman. The set of good models is called the Rashomon Set, and the goal of this project is to locate, store, and study the Rashomon sets for classes of interpretable models, including decision trees and generalized additive models.

97 MATHEMATICS AND COMPUTING↗

An interpretable model of pre-mRNA splicing for animal and plant genes

Pre-mRNA splicing is a fundamental step in gene expression, conserved across eukaryotes, in which the spliceosome recognizes motifs at the 3' and 5' splice sites (SSs), excises introns, and ligates exons. SS recognition and pairing is often influenced by protein splicing factors (SFs) that bind to splicing regulatory elements (SREs). Here, we describe SMsplice, a fully interpretable model of pre-mRNA splicing that combines models of core SS motifs, SREs, and exonic and intronic length preferences. We learn models that predict SS locations with 83 to 86% accuracy in fish, insects, and plants and about 70% in mammals. Learned SRE motifs include both known SF binding motifs and unfamiliar motifs, and both motif classes are supported by genetic analyses. Our comparisons across species highlight similarities between non-mammals, increased reliance on intronic SREs in plant splicing, and a greater reliance on SREs in mammalian splicing.

59 BASIC BIOLOGICAL SCIENCES↗

Interpretive modeling of tungsten divertor leakage during experiments with neon gas seeding

Abstract Many existing and future tokamaks with tungsten divertors operate, or will operate, with low- Z impurity seeding, but the direct effect of these seeded impurities on tungsten Scrape-off-Layer (SOL) transport has not been explored in detail. This paper reports on a DIII-D experiment designed to test how tungsten divertor leakage from the Small-Angle Slot V-Shaped, tungsten-coated divertor is impacted by neon seeding at a variety of injection rates and poloidal injection locations. Measurements from the experiment show an inverse relationship between the neon injection rate and the tungsten core penetration factor. Interpretive modeling is performed with a combination of the SOLPS-ITER and DIVIMP codes to assess the underlying tungsten behavior. The modeling results show that the reduction in tungsten divertor leakage is driven by both an increase in the divertor collisionality as well as a reduction in the ion temperature gradient near the divertor target. Collisions between low- Z impurities and tungsten impurities are found to have a significant impact on the tungsten SOL transport, such that ignoring the low- Z impurity collisional effects on the tungsten transport can result in an overestimate of the divertor leakage by an order-of-magnitude. Given the importance of these localized interactions, neon seeding from the closed, slot-like divertor has a clear advantage in being able to reduce tungsten divertor leakage without the high levels of neon core contamination that occur when seeding from other poloidal locations.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Super-X and conventional divertor configurations in MAST-U ohmic L-mode; a comparison facilitated by interpretative modelling

Measurements are presented, alongside corresponding interpretative SOLPS-ITER simulations, of the first MAST-U experiments comparing ohmically heated L-mode fuelling scans in Conventional divertor (CD) and Super-X divertor (SXD) configurations. In experiment, at comparable outer mid-plane separatrix electron density, $n_{e,\textrm{sep,OMP}}$, the maximum lower outer target heat load was found to be a factor 16 $\,\pm\,7$ lower in SXD compared to CD. In simulation, a factor 26.8 reduction was found (slightly higher than the experimental range), suggesting an additional reduction in SXD compared to the factor 9.3 expected from geometric considerations alone. According to the simulations, this additional reduction in the SXD is due to a net radial transport of the energy remaining downstream of the $T_e = 5$ eV location. This energy is carried out of the critical (highest heat load) flux tube by deuterium atoms, demonstrating the importance of a longer legged divertor which provides space for this to occur. Importantly, in both simulation and experiment, the SXD has minimal impact on the upstream n e and T e profiles. Spectral inferences of detachment front movement in SXD compare well between simulation and experiment. In regions of high magnetic field gradient, the parallel movement of the front towards the X-point becomes less sensitive to increasing $n_{e,\textrm{sep,OMP}}$, in qualitative agreement with simplified models and previous predictive simulations. Additional aspects, regarding the target ion flux rollover, upstream separatrix temperature and drift effects, are also presented and discussed.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

A Control Model: Interpretation of Fitts' Law

The analytical results for several models are given: a first order model where it is assumed that the hand velocity can be directly controlled, and a second order model where it is assumed that the hand acceleration can be directly controlled. Two different types of control-laws are investigated. One is linear function of the hand error and error rate; the other is the time-optimal control law. Results show that the first and second order models with the linear control-law produce a movement time (MT) function with the exact form of the Fitts' Law. The control-law interpretation implies that the effect of target width on MT must be a result of the vertical motion which elevates the hand from the starting point and drops it on the target at the target edge. The time optimal control law did not produce a movement-time formula simular to Fitt's Law.

Connelly, E. M.↗

Model interpretation of type III radio burst characteristics. II - Temporal aspects

A model of the radio emission region for kilometric type II bursts is used to interpret systematic variations observed in the temporal behavior of the burst parameters. The temporal behavior of the burst parameters observed by the ISEE-3 spacecraft is reviewed, and it is pointed out that the phase and modulation of the antenna signal vary in a systematic way with both time and observing frequency. The source azimuth is observed to drift with time, with the magnitude and sense of the drift depending on the location of the radio source relative to the observer. The modulation factor usually decreases uniformly with time and is frequently peaked near the burst onset. The model of the radio emission region is developed and used to obtain the intensity, phase, and modulation of the radio signal. Model results are used to show how the behavior of the burst parameters are related to attributes of the source region. It is shown that the temporal behavior of the radio parameters for an observed type II burst is well represented by the model.

Reiner, M. J.↗

Three-dimensional model interpretation of NO(x) measurements from the lower stratosphere

A three-dimensional off-line chemistry transport model, driven by European Center for Medium-Range Forecasts winds and temperatures, is used to interpret measurements of NO and NO2 taken from the DC-8 during the second Airborne Arctic Stratospheric Expedition. The model was run in three configurations: gas phase chemistry alone, inclusion of the N2O5 aerosol reaction, and inclusion of both N2O5 and ClONO2 aerosol reactions. The run including the N2O5 aerosol reaction alone usually agreed best with measured NO(x)/NO(y) ratios in midlatitude air masses. The NO(x)/NO(y) ratios of the run with both aerosol reactions were always too low, while the gas phase ratios were usually too high, especially during March. All three simulations generated extremely low NO2/NO(y) ratios in air parcels that had spent several days or more in the polar night. Measured NO2/NO(y) ratios in these types of air masses were sometimes equally low but could also be considerably higher. Observed NO/NO2 ratios differed strongly from known theory.

Folkins, Ian↗

Postural control model interpretation of stabilogram diffusion analysis

Collins and De Luca [Collins JJ. De Luca CJ (1993) Exp Brain Res 95: 308-318] introduced a new method known as stabilogram diffusion analysis that provides a quantitative statistical measure of the apparently random variations of center-of-pressure (COP) trajectories recorded during quiet upright stance in humans. This analysis generates a stabilogram diffusion function (SDF) that summarizes the mean square COP displacement as a function of the time interval between COP comparisons. SDFs have a characteristic two-part form that suggests the presence of two different control regimes: a short-term open-loop control behavior and a longer-term closed-loop behavior. This paper demonstrates that a very simple closed-loop control model of upright stance can generate realistic SDFs. The model consists of an inverted pendulum body with torque applied at the ankle joint. This torque includes a random disturbance torque and a control torque. The control torque is a function of the deviation (error signal) between the desired upright body position and the actual body position, and is generated in proportion to the error signal, the derivative of the error signal, and the integral of the error signal [i.e. a proportional, integral and derivative (PID) neural controller]. The control torque is applied with a time delay representing conduction, processing, and muscle activation delays. Variations in the PID parameters and the time delay generate variations in SDFs that mimic real experimental SDFs. This model analysis allows one to interpret experimentally observed changes in SDFs in terms of variations in neural controller and time delay parameters rather than in terms of open-loop versus closed-loop behavior.

NASA Program Biomedical Research and Countermeasur↗

Scientific visualization tools for the ISTP project: Mission planning, data analysis and model interpretation

Visualization tools are being developed to meet the challenges of mission planning and data analysis presented by the International Solar-Terrestrial Physics (ISTP) program. ISTP encompasses a large number of spacecraft, multiple ground-based observatories, and several theoretical investigations, with the goal of understanding the global behavior of the solar wind/magnetosphere/ionosphere system. The tools include three-dimensional displays of key boundaries in geospace along with spacecraft trajectories, which can be animated and synchronized to universal time. Magnetic field models and MHD simulation results can be invoked to reveal the magnetic topology or to identify magnetic conjunctions between spacecraft and/or ground-based facilities. Simultaneous displays of satellite trajectories, spacecraft-borne observations, and model predictions are available to facilitate data processing and interpretation efforts. The current status of these tools is described, and their implementation at the ISTP Science Planning and Operations Facility and distribution to the entire ISTP community are discussed.

Peredo, M.↗

Model interpretation of type III radio burst characteristics. I - Spatial aspects

The ways that the finite size of the source region and directivity of the emitted radiation modify the observed characteristics of type III radio bursts as they propagate through the interplanetary medium are investigated. A simple model that simulates the radio source region is developed to provide insight into the spatial behavior of the parameters that characterize radio bursts. The model is used to demonstrate that observed radio azimuths are systematically displaced from the geometric centroid of the exciter electron beam in such a way as to cause trajectories of the radio bursts to track back to the observer at low frequencies, rather than to follow expected Archimedean spiral-like paths. The source region model is used to investigate the spatial behavior of the peak intensities of radio bursts, and it is found that the model can qualitatively account for both the frequency dependence and the east-west asymmetry of the observed peak flux densities.

Reiner, M. J.↗

Model Interpretation of Climate Signals: Application to the Asian Monsoon Climate

This is an invited review paper intended to be published as a Chapter in a book entitled "The Global Climate System: Patterns, Processes and Teleconnections" Cambridge University Press. The author begins with an introduction followed by a primer of climate models, including a description of various modeling strategies and methodologies used for climate diagnostics and predictability studies. Results from the CLIVAR Monsoon Model Intercomparison Project (MMIP) were used to illustrate the application of the strategies to modeling the Asian monsoon. It is shown that state-of-the art atmospheric GCMs have reasonable capability in simulating the seasonal mean large scale monsoon circulation, and response to El Nino. However, most models fail to capture the climatological as well as interannual anomalies of regional scale features of the Asian monsoon. These include in general over-estimating the intensity and/or misplacing the locations of the monsoon convection over the Bay of Bengal, and the zones of heavy rainfall near steep topography of the Indian subcontinent, Indonesia, and Indo-China and the Philippines. The intensity of convection in the equatorial Indian Ocean is generally weaker in models compared to observations. Most important, an endemic problem in all models is the weakness and the lack of definition of the Mei-yu rainbelt of the East Asia, in particular the part of the Mei-yu rainbelt over the East China Sea and southern Japan are under-represented. All models seem to possess certain amount of intraseasonal variability, but the monsoon transitions, such as the onset and breaks are less defined compared with the observed. Evidences are provided that a better simulation of the annual cycle and intraseasonal variability is a pre-requisite for better simulation and better prediction of interannual anomalies.

Lau, William K. M.↗

Challenges and approaches to interpretive modeling of boundary plasma and neutral transport in a closed, pumped divertor

An experimental discharge from the DIII-D tokamak is modeled using the SOLPS-ITER code suite and compared against measurements in the pumped and relatively closed upper divertor. Uncertainties of boundary plasma simulations are identified by attempting to match code inputs to experimental conditions, including iteratively solving transport coefficients to match upstream experimental profiles using varying quantities of core particle flux, different pumping models, and various assumptions of ion thermal transport. Simulated boundary conditions for particle injection at the core interface are shown to be relevant to the plasma solution at the divertor targets, even if upstream transport is modified so that plasma profiles are comparatively similar, although seperatrix density is not held constant. When upstream plasma profiles are matched to experimental measurements by varying diffusive transport coefficients, using either poloidally symmetric or ballooning structure, the model finds a majority of injected energy being transported radially off the computational domain, in conflict with experimental radiated power measurements and heat flux measurements at the divertor target. Imposing a maximum thermal diffusivity or radially shifting the experimental separatrix location of the fitted profiles to increase power conducted to the targets by increasing the upstream electron temperature does not significantly modify this result. Including a thermalizing plenum volume in the simulation domain is shown to maintain the experimental volumetric pumping rate without knowing the neutral energy distribution incident on the pump duct a priori. By modifying transport parameters to match different assumptions for ion temperature, downstream neutral pressure changes by more than a factor of two, suggesting that attention to ion thermal transport may be a critical parameter for simulations to accurately resolve recycling and neutral transport, particularly in a closed divertor geometry. In addition to quantifying various modeling uncertainties, this work motivates both further experimental study and modeling improvements to improve predictive capabilities.

divertor↗

A wave model interpretation of the evolution of rotational discontinuities

A hybrid numerical code is employed to trace the evolution of rotational discontinuities (RDs). An extensive parameter variation is carried out, with particular emphasis on beta, Ti/Te, theta sub B (the angle between the normal and total magnetic field), and the helicity of the RD. The RD structure is shown to have features in common with the evolution of both strongly modulated nonlinear wave packets and linear dispersive wave propagation in oblique magnetic fields. For small theta sub B, the RD disperses linearly, giving fast and Alfven waves upstream and downstream, respectively, and the familiar S-shaped hodograms. At larger theta sub B, nonlinearity becomes important and strong coupling to a compressional (sonic) component can occur in the main current layer. The results are applied to RDs observed in the solar wind and at the magnetopause.

Vasquez, Bernard J.↗

Developing interpretable models with optimized set reduction for identifying high risk software components

Applying equal testing and verification effort to all parts of a software system is not very efficient, especially when resources are limited and scheduling is tight. Therefore, one needs to be able to differentiate low/high fault frequency components so that testing/verification effort can be concentrated where needed. Such a strategy is expected to detect more faults and thus improve the resulting reliability of the overall system. This paper presents the Optimized Set Reduction approach for constructing such models, intended to fulfill specific software engineering needs. Our approach to classification is to measure the software system and build multivariate stochastic models for predicting high risk system components. We present experimental results obtained by classifying Ada components into two classes: is or is not likely to generate faults during system and acceptance test. Also, we evaluate the accuracy of the model and the insights it provides into the error making process.

Briand, Lionel C.↗