Search NASA⌕ Search

SEARCH · Search NASA

Results for “Error Modeling”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 289 records · Page 16

Transplatformer: translating toxicogenomic profiles between generations of platforms

Background Transcriptomic profiling technologies have advanced the analysis of biological and toxicological responses. However, substantial differences in probe design, dynamic range, gene coverage, and preprocessing pipelines across platforms introduce artifacts that limit cross-study integration and hinder the reuse of historical datasets. We aim to develop computational methods for accurate cross-platform translation to maximize the value of legacy resources. Results We present TransPlatformer a deep learning framework for translating gene expression profiles across heterogeneous toxicogenomics platforms. TransPlatformer employs a novel attention-based architecture to map high-dimensional fold-change vectors from legacy microarray technologies to current platforms. Models are trained and evaluated using DrugMatrix, spanning three technological generations. We investigate mixed-tissue, single-tissue, and cross-tissue training paradigms and benchmark performance against multilayer perceptron and matrix-completion baselines. In mixed-tissue training, TransPlatformer achieves a greater than 50% reduction in mean absolute error (0.043 vs. 0.09) and nearly doubles Pearson correlation ( ≈ 0.71 vs. 0.37) relative to baseline methods. Importantly, TransPlatformer preserves rare but biologically meaningful over- and under-expressed signals, with mean absolute error below 0.22. Single-tissue models yield further improvements for well-represented organs, such as a 10% reduction in liver mean absolute error, while underscoring the need for data augmentation strategies in low-sample tissues.ra Conclusions TransPlatformer provides an effective and scalable computational solution for cross-platform transcriptomic translation. By enabling biologically faithful harmonization of gene expression data, the proposed approach facilitates the reuse of legacy toxicogenomics datasets, enhances downstream biomarker discovery, and supports more reproducible predictive modeling in toxicology.

59 BASIC BIOLOGICAL SCIENCES↗

An Assessment of the Error Due to Computing Waste Isolation Pilot Plant Porosity Using the Porosity Response Surface Approach

The Waste Isolation Pilot Plant Performance Assessment (WIPP PA) must predict the likelihood that radionuclides will escape into the biosphere via mechanisms that depend on geohydraulic flow. Ideally, one would predict the geohydraulic flow using coupled geohydraulic and geomechanical simulations, but such coupled simulations are not computationally tractable. Instead, Sandia has historically used a look-up table of porosities for a given fluid pressure and time, called the porosity response surface, but this approach can introduce porosity errors because it largely ignores the porosity’s dependence on the past fluid pressure history. This report discusses efforts to quantify these porosity errors for both the legacy and new porosity response surfaces. Six hundred different fluid pressure histories were fed through the legacy/new geomechanical model and the legacy/new porosity response surface to generate six hundred porosity error histories. The error associated with the legacy porosity surface was substantial, while the error associated with the new porosity surface was typically small, except when fluid pressures exceeded the lithostatic pressure at the repository. In response to the errors at high pressures, a preliminary study of the WIPP PA’s sensitivity to these porosity errors was conducted. The study found that reducing the porosity errors at high pressures negligibly affected predictions of radionuclide releases. Finally, an initial machine-learned model for porosity was developed. This ML model significantly reduced the porosity error at high pressures, but sizable errors remained, so more development is necessary before coupling an ML model to the geohydraulic model.

12 MANAGEMENT OF RADIOACTIVE AND NON-RADIOACTIVE W↗

Toward more-robust, AI-enabled subsurface seismic imaging for geotechnical applications

Non-invasive seismic imaging has the potential to cost-effectively evaluate large volumes of subsurface material to inform geotechnical site investigation. However, seismic imaging using full waveform inversion (FWI) requires significant computational time and is dependent on an initial starting model. As a result, FWI has not yet been widely adopted into geotechnical practice. Previous efforts, on relatively simple two-layered models, indicate that data-driven artificial intelligence (AI) models may be as effective as FWI at predicting 2D images of shear wave velocity (V s ). Furthermore, the AI model predictions can be made almost instantaneously after data acquisition and do not require an initial starting model. We examine the generality of these findings by developing a new AI model for subsurface seismic imaging, whereby we make several notable contributions. First, we architect a multimodal AI model that combines time- and frequency-domain representations of the seismic wavefield to predict a 50 m by 20 m subsurface image of V s . Second, we developed a new diverse dataset of 100,000 images with their corresponding seismic wavefields to train the AI model. Third, we propose four physics-informed data augmentations for data-driven seismic imaging. Fourth, we develop two prediction consistency tests to evaluate the model’s performance when the true subsurface is unknown. Our final model, which has been made publicly available, is capable of predicting a subsurface V s image from a single seismic wavefield with an average, mean absolute percent error (MAPE) of 24 %. The predictive model is applied to a field dataset and shown to be consistent with local geology and shear-wave refraction measurements from the same location.

Artificial intelligence↗

Electron-Proton Scattering Event Generation using Structured Tokenization

Recent work such as Omnijet-$\alpha$ has demonstrated that effective tokenization combined with transformer-based architectures can produce effective foundation models for jet physics. While tokenization may help models capture generalizable event characteristics, it also introduces discretization errors that may compromise the precision required for downstream physics analyses. As the number and complexity of the particle features grow, these errors are likely to grow proportionally. In this study, we investigate new tokenization strategies to improve the application of generative transformer models to \textsc{Pythia8} simulations of electron-proton scattering at the Electron-Ion Collider. Specifically, we propose a feature-based structured tokenization approach that utilizes multiple tokens per particle, improving expressivity, while reducing the total number of unique tokens needed. We evaluate this method against grid-based binning, K-means clustering, and vector-quantized variational auto-encoders on the event simulations. Our results show that feature-based structured tokenization reduces discretization error, leading to more accurate generative modeling of particle-level events.

Goldenberg, Steven [Thomas Jefferson National Acce↗

Electron-Proton Scattering Event Generation using Structured Tokenization

Recent work such as Omnijet-$\alpha$ has demonstrated that effective tokenization combined with transformer-based architectures can produce effective foundation models for jet physics. While tokenization may help models capture generalizable event characteristics, it also introduces discretization errors that may compromise the precision required for downstream physics analyses. As the number and complexity of the particle features grow, these errors are likely to grow proportionally. In this study, we investigate new tokenization strategies to improve the application of generative transformer models to \textsc{Pythia8} simulations of electron-proton scattering at the Electron-Ion Collider. Specifically, we propose a feature-based structured tokenization approach that utilizes multiple tokens per particle, improving expressivity, while reducing the total number of unique tokens needed. We evaluate this method against grid-based binning, K-means clustering, and vector-quantized variational auto-encoders on the event simulations. Our results show that feature-based structured tokenization reduces discretization error, leading to more accurate generative modeling of particle-level events.

Goldenberg, Steven [Thomas Jefferson National Acce↗

Cryogenic-Refined MOSFET Modeling for Oscillator, Frequency Divider, and Amplifier Designs Below 4 K

Capturing device characteristic changes at cryogenic temperatures is crucial for cryo-CMOS circuit designs. In this work, we present an isothermal cryogenic-refined modeling approach for CMOS transistors that is simple, low overhead, and easy to implement while offering the required accuracy for predicting circuit performance at the designated temperatures. Guided by die-level measurement data and circuit design principles, the model introduces corrections to only five critical parameters: threshold voltage, carrier mobility, elevated low-frequency flicker noise, dominant high-frequency shot noise, and subthreshold swing (SS). These refinements are implemented around the foundry-provided SPICE model, which is typically validated only down to about 200 K. With these adjustments, the proposed cryogenic-refined model achieves less than 5% error in both large-signal metrics (I–V characteristics) and small-signal parameters (e.g., transconductance) when compared with device measurements at deep-cryogenic temperatures. The methodology is validated in two advanced technologies: TSMC 40-nm CMOS and GlobalFoundries (GF) 45-nm RF-SOI. We further demonstrate its applicability in three representative RF circuits: a 30-GHz LC oscillator, a high-speed current-mode-logic (CML) frequency divider (FD), and a subthreshold Gb/s amplifier, all showing close agreement between simulated predictions and measurements performed at 4 and 2.5 K. Finally, we believe that the proposed approach is implementation-friendly and can significantly accelerate the development of cryo-CMOS integrated circuits.

circuit modeling↗

Validating a Dynamic PWR Safety and Security Model?

Nuclear power plants (NPPs) are assessed for safety and security using separate models that cannot capture how an attacker's decisions and a plant's response unfold together in real time, leaving regulators and operators without a complete picture of true plant vulnerability. Traditional probabilistic risk assessment (PRA) methods treat adversarial events as fixed initiators with predetermined outcomes, and are structurally incapable of representing the time-dependent interplay between physical security events, safety system response, and operator mitigative actions. At Idaho National Laboratory (INL), I contributed to the development and validation of Modeling and Analysis for Safety and Security using the Dynamic EMRALD Framework (MASS-DEF). Where static PRA relies on event-tree logic that cannot evolve mid-scenario, MASS-DEF couples a time-dependent dynamic PRA tool EMRALD (Event Modeling Risk Assessment using Linked Diagrams) with attack simulation software, allowing attacker behavior, plant system states, and operator actions to interact across time. My work focused on validating a general Pressurized Water Reactor (PWR) model. I traced model logic against PWR plant to identified errors in logic and confirm accuracy. I then built and tested attack scenarios against a general PWR model to verify that the model produced expected outcomes across all logical pathways. I also contributed a section to a related technical paper applying the same EMRALD platform to radiation dose modeling. Results show that MASS-DEF can quantitatively demonstrate that many plants exceed their regulatory security thresholds. This demonstrated margin provides a technically defensible basis for reducing the number of guards without compromising regulatory compliance. Physical security costs represent roughly 10% of annual operating budgets, making such reductions directly meaningful to INL's mission of sustaining existing commercial NPPs. This internship strengthened my understanding of nuclear systems, probabilistic modeling, and technical writing, and has solidified my pursuit of a career at a national laboratory.

98 - NUCLEAR DISARMAMENT, SAFEGUARDS, AND PHYSICAL↗

Using Explainable Artificial Intelligence to Predict Perovskite Solar Cell Electrical Metastability from Operando Photoluminescence Images in Accelerated Stress Testing

Metal halide perovskite (MHP) solar cells exhibit a metastable response to bias governed by coupled ionic–electronic processes, complicating the conventional reciprocity relation between luminescence intensity and device open-circuit voltage (V oc ). This limits the use of luminescence as a diagnostic for device screening or accelerated stress testing, motivating new approaches that can interpret photoluminescence (PL) signals under nonequilibrium conditions. From the artificial intelligence perspective, we develop an explainable deep learning framework that integrates convolutional neural networks (CNN), long short-term memory (LSTM) layers, and an attention mechanism to learn spatiotemporal features from operando photoluminescence PL image sequences. The model achieves a mean absolute error of ±0.027 V in predicting open-circuit voltage transients and reduces extreme-tail errors by up to 78% compared to physics-based reciprocity calculations. Gradient-weighted Class Activation Mapping (Grad-CAM) provides interpretability by highlighting physically meaningful regions such as electrode edges and emergent defect features. From the engineering application perspective, this framework enables accurate, contactless prediction of device V oc and identification of degradation-relevant features during accelerated aging of perovskite solar cells. This approach demonstrates how explainable AI can enhance operando diagnostics and reliability analysis in photovoltaic devices under nonequilibrium conditions.

14 SOLAR ENERGY↗

Security of quantum position-verification limits Hamiltonian simulation via holography

We investigate the link between quantum position-verification (QPV) and holography established in [1] using holographic quantum error correcting codes as toy models. By inserting the “temporal” scaling of the AdS metric by hand via the bulk Hamiltonian interaction strength, we recover a toy model with consistent causality structure. This leads to an interesting implication between two topics in quantum information: if position-based verification is secure against attacks with small entanglement then there are new fundamental lower bounds for resources required for one Hamiltonian to simulate another.

AdS-CFT Correspondence↗

Attention-based 3D – convolutional neural network model for mechanical property predictions using visible light images in metal additive manufacturing

Additive manufacturing (AM), while commonly used for rapid prototyping and creating components with complex geometries, has not been widely adopted for critical applications across the aerospace, automotive, defense, energy, and medical industries. This is, in part, due to the challenges of controlling flaws and uncertainty in the mechanical behavior of additively manufactured components. In recent years, there has been an increase in research aimed at predicting the final mechanical properties of additively manufactured components during the printing process. To address these issues, a 3D-CNN model was trained using low-cost in situ visible-light camera data, anomaly classifications, and the chosen process parameters to predict the ultimate tensile strength (UTS), yield strength (YS), total elongation (TE), and uniform elongation (UE). The 3D-CNN layers of the model employed attention mechanisms to prioritize features in the data, thereby improving prediction accuracy. Furthermore, the effect of each process parameter and anomaly class is investigated using attention-based dynamic sigmoid weighted gates to interpret the influence each class has on the final prediction. Different combinations of the in situ data were fed into the 3D-CNN, with varying amounts of image layers, to determine the ideal combination for predicting mechanical properties in situ. Here, the 3D-CNN model achieved mean absolute percentage errors (MAPE) below 5% for both UTS and YS while using only a single camera input and under half of the available image layers.

36 MATERIALS SCIENCE↗

Rapid neutron and gamma-ray source localization using machine learning

Rapid localization of radiation sources is critical for applications including nuclear emergency response, safeguards, and security. However, conventional imaging systems such as neutron scatter cameras and Compton cameras depend on rare coincidence events, which often result in long acquisition times. In this work, we address the challenge of rapid source localization by developing a machine learning approach to predict the direction of a single radiation source using only count rates from an array of neutron and gamma-ray detectors. The proposed model is a fully connected neural network (FCNN) trained using Monte Carlo simulation data from a 252 Cf source. The model hyperparameters are optimized with a small set of routine 252 Cf measurements. We benchmarked the performance of the trained and optimized machine learning model using additional 252 Cf , 137 Cs , and PuBe measurements under laboratory conditions with varying source-detector configurations. For these measurements, the machine learning model achieved a mean localization error smaller than 30° with 3 x 10 3 system counts, corresponding to 8 s measurement time for the imaging system used in this work. In this low-statistics regime, the method outperformed traditional scatter-based imaging by more than 75% in localization accuracy for the evaluated measurement configurations. These results demonstrate that a machine learning-based approach can significantly reduce the time required for accurate single-source localization, providing a robust and computationally efficient alternative to traditional imaging systems in time-critical nuclear security and emergency response scenarios.

Gamma-ray imaging↗

Use of Satellite, Surface Observations and Numerical Weather Prediction Model Data to Improve Cloud Base Height and Cloud Base Vertical Velocity Estimation

Cloud base height (CBH) and cloud base vertical velocity (CBVV) are important variables that impact the overall climate in a region as they influence the formulation, longevity, and evolution of clouds. Retrieval of both parameters have long used ground instrumentation (e.g., Doppler lidar (DL), ground base radar); however, retrieving CBH from satellites is particularly challenging given that space-based instruments only observe cloud tops. In this manuscript, CBH is retrieved using a multi-linear regression equation, while CBVV used a random forests model. Both retrievals combine satellite and numerical weather prediction data. The satellite data used are the Visible Infrared Imaging Radiometer Suite imagery, while measurements of CBH and CBVV include DL and radiosonde data at the Southern Great Plains (SGP) Atmospheric Radiation Measurement observatory. Data from 83 summer days (May-August) in 2018–2021 featuring cumulus clouds forced by solar heating were examined and used to train the models, with years 2022–2023 used for validation. Various spatial domains were defined with one large (2.4° longitude by 2.0° latitude) SGP domain being split into smaller sections (smallest being 0.99° and 0.61° longitude and latitude respectably). CBH and CBVV values obtained from the DL as compared to the models show root mean square errors between 150 and 200 m, with CBVV values between 0.45 and 1 ms -1 . Finally, it was found that the CBH formulation performs well over all domains, while the CBVV retrievals become less accurate due to more turbulence being introduced into the observations as the number of DL stations decreases in the smaller domains.

54 ENVIRONMENTAL SCIENCES↗

Measuring the thermal conductivity of hydrogels with a bidirectional 3w method

Hydrogels are soft, water-absorbing polymer materials with diverse applications in biomedicine and agriculture. Recently, hydrogels have been proposed to encapsulate water-soluble phase change materials which store energy in their latent heat of solidification. In these applications, the thermal conductivity of these materials affects their performance. Few methods exist for measuring the thermal conductivity of small quantities of hydrogels. Here, we describe an implementation of the bidirectional 3w technique to measure the thermal conductivity of hydrogels with particular attention to their moisture content. Our implementation of the technique can probe sample volumes as little as ~20 mL and yields the thermal conductivity without requiring fitting of additional thermal parameters. We numerically simulate 3w sensor designs with frequency-domain 3-D models to quantify and reduce errors introduced by the choice of substrate and insulation layer thickness. Frequencies in the ~1−20 Hz range yield less error for the materials considered here. We verify our setup with measurements on water and report values for polyacrylamide and poly(2-acrylamido-2-methylpropane sulfonic acid) (PAMPS) hydrogels. Our swollen hydrogels exhibited thermal conductivities nearly equivalent to water, 0.6 W m-1 K-1, and we estimate thermal conductivities of 0.43 and 0.42 W m-1 K-1 for neat polyacrylamide and PAMPS, respectively. Finally, we estimate an error of ±7%, consistent with other 3ω methods, with the largest error coming from the sensor calibration. We find our implementation of the bidirectional 3w method gives reasonable results and can be employed for prototyping soft materials relevant for thermal storage.

3-omega, thermal conductivity, hydrogel, moisture ↗

Polarization options in inclusive DIS off tensor polarized deuteron

In the near future, the Jefferson Lab b 1 experiment will provide the second measurement of tensor polarized asymmetries in inclusive DIS on the deuteron. In this asymmetry, 4 independent tensor polarized structure functions contribute. This necessitates systematic approximations in the extraction of the leading twist structure function b 1 from a single tensor asymmetry measurement. Contamination from higher twist structure functions and kinematic effects is discussed here. Using a deuteron convolution model, we quantify the systematic errors from these approximations for two different choices for the target polarization direction (momentum transfer, electron beam direction). For Jefferson Lab 12 GeV kinematics, the systematic error turns out to be comparable between the two polarization options, while at higher Q 2 values the momentum transfer direction is preferred.

Cosyn, Wim [Florida International University, Miam↗

ab initio Sub-Mechanism Development for Cyclopentene Oxidation

To accurately predict low-temperature oxidation behavior, chemical kinetics mechanisms must contain complete reaction networks that include detailed consumption reactions of intermediates produced directly from hydroperoxyalkyl radicals, Q̇OOH, which undergo competing unimolecular reactions and bimolecular reactions with O2. Rates of chain-branching are governed by the flux between the two competing pathways, and inherently depend on temperature, pressure, and oxygen concentration. Neglect of consumption pathways for major oxidation intermediates leads to mechanism truncation error that is ameliorated by expanding the level of detail included in sub-mechanisms and employing ab initio methods for computing rates of elementary reactions and thermochemical properties of species involved. In the present work, an ab initio-derived sub-mechanism is developed using AutoMech to model the chemical kinetics of cyclopentene, a major product of cyclopentane oxidation. The ab initio sub-mechanism builds on a detailed mechanism developed using Reaction Mechanism Generator (RMG) for the specific purpose of determining the extent to which replacing cyclopentene-specific reactions and species with quantum chemical computations reduces model inaccuracies resulting from mechanism truncation error. In an effort to minimize interference from other reactions present during the formation of cyclopentene from cyclopentyl + O2, providing a narrower experimental scope, the model is compared against speciation measurements from jet-stirred reactor (JSR) experiments on cyclopentene oxidation. The experiments utilize vacuum ultraviolet-absorption spectroscopy and mass spectrometry for isomer-resolved speciation of intermediates at 835 Torr from 700 – 950 K. [O2]-dependent experiments were also conducted from 0.057 – 2.01 · 1018 molecules cm–3 at 825 K to examine the influence of oxygen on species profiles. Model predictions using the ab initio-revised mechanism yielded significant improvements in species profiles for both the temperature- and [O2]-dependent measurements, owing in part to increased rates of HOȮ and H2O2 production, which underscores the influence of theoretical calculations of reaction rates involving species produced from Ṙ + O2 such as cyclopentene.

AutoMech↗

Dissipation Scaled Internal Wave Drag in a Global Heterogeneously Coupled Internal/External Mode Total Water Level Model

This study showcases a global, heterogeneously coupled total water level system wherein salinity and temperature outputs from a coarser-resolution (~12 km) ocean general circulation model are used to calculate density-driven terms within a global, higher-resolution (~2.5 km) depth-averaged total water level model. We demonstrate that the inclusion of baroclinic forcing in the barotropic model requires modification of the internal wave drag term to prevent excess degradation of tidal results compared to the barotropic model. By scaling the internal tide dissipation by an easy to calculate dissipation ratio, the resulting heterogeneously coupled model has complex root mean square errors (RMSE) of 2.27 cm in the deep ocean and 12.16 cm in shallow waters for the M 2 tidal constituent. While this represents a 10%–20% deterioration as compared to the barotropic model, the improvements in total water level prediction more than offset this degradation. Global median RMSE compared to observations of total water levels, 30-day sea levels, and non-tidal residuals improve by 1.86 (18.5%), 2.55 (42.5%), and 0.36 (5.3%) cm respectively. The drastic improvement in model performance highlights the importance of including density-driven effects within global hydrodynamic models and will help to improve the results of both hindcasts and forecasts in modeling extreme and nuisance flooding. With only an 11% increase in model run time compared to the fully barotropic total water level model, this approach paves the way for high resolution coastal water level and flood models to be used alongside climate models, improving operational forecasting of total water levels.

Blakely, Coleman Peter [University of Notre Dame, ↗

Resilience Measurement Framework For Post-deployment Artificial Intelligence (ai) Integrated Systems

Resilience is largely defined as the ability to adapt or recover from adverse conditions, stresses, attacks, or compromises on systems that use or are enabled by digital resources. In Artificial Intelligence Management and Research for Advanced Networked Testbed Hub (AMARANTH), resilience is measured in the amount of time it took from the beginning of a testing period for the model to reach predictions outside of the original 95% confidence interval or using the Kullback-Leibler (KL) divergence theorem, the Population Stability Index (PSI), and traditional methods such as root mean squared error (RMSE) threshold. Artificial Intelligence (AI) model drift is of significant concern when deploying AI-integrated systems into critical and/or secure environments. Drift can impact resilience of the AI-integrated system post-deployment and requires consistent maintenance and upkeep to ensure the model is accurate and precise. To quantify model drift and predict the point when a model's drift becomes unacceptable, we describe using Kullback-Leibler (KL) divergence, Population Stability Index (PSI) and/or confidence interval width estimations to determine the point of failure and time to failure of a model post-deployment. Through simple code functions, the KL-divergence, PSI, confidence interval, and root mean squared (RMSE) point of failures can be used to derive when a model needs to be maintained as well as the impact of adversarial action through statistical means.

Yockey, Patience [Idaho National Laboratory (INL),↗

MCP-enabled agentic AI workflow for building energy modelling: framework and use cases

Traditional building energy modelling workflows remain labor-intensive and error-prone, requiring specialized expertise that limits broader adoption. This paper introduces a novel Model Context Protocol (MCP)-enabled framework that connects AI assistants to EnergyPlus through MCP, a standardized interface for tool invocation and context management. Two complementary integration paradigms are presented and compared: conversational integration, where users interact through natural language while an AI assistant orchestrates MCP tools on demand, and agentic workflow integration, where specialized agents coordinate autonomously to complete multi-step tasks. Using an experimental testbed for residential buildings, the end-to-end workflows are demonstrated. The conversational approach reduced typical inspection and modification tasks from 1-2 h to under 15 min, while maintaining full transparency through visible tool invocations. The agentic approach automated parametric analysis. These demonstrations establish MCP as a foundational layer for AI-assisted building energy modelling, enabling natural language interactions with simulation tools while preserving professional oversight and decision-making authority.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗