Search NASASearch

SEARCH · Search NASA

Results for “Models, Statistical”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 records

Hot Droughts and Forest Tree Dynamics in the Amazon - Statistical Models, Scripts, Data, and Outputs

This package contains data, outputs, equations, and R scripts for analyses for manuscript entitled "Hot droughts in the Amazon: A window to a future hypertropical climate" by J. Chambers et al., in particular it contains statistical models and analyses for the INPA BIONTE tree mortality study. The Models folder contains details for all statistical models in PDF files. The Scripts folder contains the R scripts for Bayesian Hierarchical Models (two text files) and SEMs (one text file) are separate and reasonably annotated. All data associated with these scripts are in the data folder. The Data folder contains two of the three CSV files used for the analyses and are called by the R scripts. Two of them are part of published datasets (`BIONTE_mortality-rates.csv` from Lima et al. 2024, DOI:10.15486/ngt/1898910 and `SPEI.csv` from Pastorello et al. 2023 DOI:10.15486/ngt/1958257) and also provided in this package for convenience (please see the corresponding datasets for usage and citation terms). The third dataset (`BIONTE_gapfilled_wd.csv`) contains sensitive information and can be obtained by contacting the manuscript lead author. The Outputs folder contains the two output files that provide extra information about the analyses. The file `figuresFeb2025d.pdf` contains all the figures from the manuscript - captions are in the manuscript. The file `ChambersMS.pdf` contains primary results from Bayesian statistical models, regression analyses, and validation steps applied to the tree mortality data from the INPA experiments. The document includes visual summaries, model diagnostics, and leave-one-out (LOO) validation results. A breakdown of file contents can be found in the README file that is part of this package.

54 ENVIRONMENTAL SCIENCES

Temporal Forecasting of Distributed Temperature Sensing in a Thermal Hydraulic System With Machine Learning and Statistical Models

We benchmark performance of long-short term memory (LSTM) network machine learning model and autoregressive integrated moving average (ARIMA) statistical model in temporal forecasting of distributed temperature sensing (DTS). Data in this study consists of fluid temperature transient measured with two co-located Rayleigh scattering fiber optic sensors (FOS) in a forced convection mixing zone of a thermal tee. We treat each gauge of a FOS as an independent temperature sensor. We first study prediction of DTS time series using Vanilla LSTM and ARIMA models trained on prior history of the same FOS that is used for testing. The results yield maximum absolute percentage error (MaxAPE) and root mean squared percentage error (RMSPE) of 1.58% and 0.06% for ARIMA, and 3.14% and 0.44% for LSTM, respectively. Next, we investigate zero-shot forecasting (ZSF) with LSTM and ARIMA trained on history of the co-located FOS only, which is advantageous when limited training data is available. The ZSF MaxAPE and RMSPE values for ARIMA are comparable to those of the Vanilla use case, while the error values for LSTM increase. We show that in ZSF, performance of LSTM network can be improved by training on most correlated gauges between the two FOS, which are identified by calculating the Pearson correlation coefficient. The improved ZSF MaxAPE and RMSPE for LSTM are 4.4% and 0.33%, respectively. Performance of ZSF LSTM can be further enhanced through transfer learning (TL), where LSTM is re-trained on a subset of the FOS that is the target of forecasting. We show that LSTM pre-trained on correlated dataset and re-trained on 30% of testing target dataset achieves MaxAPE and RMSPE values of 2.32% and 0.28%, respectively.

ARIMA

Approaching hydro-equivalent ignition in laser direct-drive via target design optimization using novel statistical modeling

Laser direct-drive offers significant advantages in terms of target simplicity, improved energy coupling, and large fuel masses over indirect drive. However, performance degradations from hydrodynamic and laser-plasma instabilities seeded and driven by the direct illumination pose limitations on the parameter space available for achieving ignition. In this paper, new design improvements are identified to forge a path forward for a hydro-equivalent ignition demonstration. The first is related to a new formulation of the statistical model (SM) used to accurately predict target performance directly from input parameters such as laser pulse shape and target specifications. This new SM formulation provides direct guidance on target dimensions and laser beam-to-target radius to achieve the highest fusion yield on the OMEGA laser. The second improvement comes from cooling the deuterium–tritium (DT) ice layer below the triple point right before shot time leading to lower DT vapor densities and higher convergence. Guided by these design improvements, a Bayesian optimization algorithm was used to design an implosion that is predicted to closely approach a Lawson triple product that hydrodynamically scales to ignition if equivalent laser–target coupling is achieved at laser energies typical of the National Ignition Facility.

Deuterium

Statistical modelling and Bayesian inversion for a Compton imaging system: application to radioactive source localization

Abstract This paper presents a statistical forward model for a Compton imaging system, called Compton imager. This system, under development at the University of Illinois Urbana Champaign, is a variant of Compton cameras with a single type of sensors which can simultaneously act as scatterers and absorbers. This imager is convenient for imaging situations requiring a wide field of view. The proposed statistical forward model is then used to solve the inverse problem of estimating the location and energy of point-like sources from observed data. This inverse problem is formulated and solved in a Bayesian framework by using a Metropolis within Gibbs algorithm for the estimation of the location, and an expectation-maximization algorithm for the estimation of the energy. This approach leads to more accurate estimation when compared with the deterministic standard back-projection approach, with the additional benefit of uncertainty quantification in the low photon imaging setting.

Tarpau, Cécilia (ORCID:0000000286539490)

Statistical model of the stimulated forward Brillouin scattering driven by a randomized laser beam in plasma

The modeling of a spatially incoherent laser beam remains a central problem of the parametric instabilities in the context of inertial confinement fusion. This letter gives a simplified and comprehensive overview of the recent analytical developments regarding the modeling of these laser beams and a comparison with a dedicated experiment. Our model accounts for the first time for the statistical standard deviation of the gain and accurately captures the entanglement between wave mixing processes and the speckle correlations thus resolving the longstanding contradictions between the random phase approximation and the model of independent speckles. It is successfully compared to a recent laser beam spray experiment and the associated paraxial simulations, demonstrating that backscattering predictions require accounting for the beam spray. Furthermore, our framework thus provides a way to evaluate and guide the analysis of parametric instabilities in high laser energy experiments.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

Enhancing Interpretability in Generative Modeling: Statistically Disentangled Latent Spaces Guided by Generative Factors in Scientific Datasets

This study addresses the challenge of statistically extracting generative factors from complex, high-dimensional datasets in unsupervised or semi-supervised settings. We investigate encoder-decoder-based generative models for nonlinear dimensionality reduction, focusing on disentangling low-dimensional latent variables corresponding to independent physical factors. Introducing Aux-VAE, a novel architecture within the classical Variational Autoencoder framework, we achieve disentanglement with minimal modifications to the standard VAE loss function by leveraging prior statistical knowledge through auxiliary variables. These variables guide the shaping of the latent space by aligning latent factors with learned auxiliary variables. We validate the efficacy of Aux-VAE through comparative assessments on multiple datasets, including astronomical simulations.

97 MATHEMATICS AND COMPUTING

Scenario Planning Management Actions to Restore Cold Water Stream Habitat: Comparing Mechanistic and Statistical Modeling Approaches

ABSTRACT Under the United States Clean Water Act, states are required to periodically assess state waters to determine compliance with water quality criteria (including temperature) and then to develop total maximum daily loads (TMDLs) for impaired waters as necessary to bring them into compliance. We compared the performance of mechanistic stream temperature models (HeatSource, QUAL2K, and QUAL2Kw) applied to the mainstem of three TMDL watersheds (Middle Fork John Day, OR; Wind River, WA; South Fork Nooksack, WA) with that of spatial stream network (SSN) models applied to the full watersheds and used these to evaluate the potential effectiveness of restoration strategies. SSN models performed well with slightly lesser accuracy (RMSE = 0.47–0.87) for mainstem predictions than mechanistic models (RMSE = 0.4) but provided additional benefits to inform management, including information on spatial and temporal heterogeneity of restoration effectiveness throughout the watershed. Of the four scenarios considered (restoration of riparian zones to potential natural vegetation, channel narrowing, increasing flow by restricting irrigation withdrawals, and combined applications), riparian zone restoration was consistently the most effective in reducing temperatures at the outlet, mainstem, and throughout the watersheds. Predicted restoration effectiveness for thermal regimes varied significantly both within and among watersheds. A focus on water quality criteria exceedance only at the watershed outlet or along the mainstem reach can obscure knowledge of restoration potential for fish habitat in tributaries and headwaters, potential for creation of thermal refuge areas along the mainstem critical for maintaining migration corridors, and thermal regime heterogeneity across space and time.

Fuller, M. R.

Spatiotemporal Learning in Power Modules: Wavelet-Enhanced Forecasting of Thermomechanical Degradation

Detecting internal defects in power electronics packages is critical for their performance and reliability, especially under extreme operating conditions, as these defects can lead to catastrophic failure if not properly addressed. Confocal scanning acoustic microscopy (C-SAM) plays a key role in the nondestructive evaluation of bond layer degradation within a power electronics package by detecting defects such as delamination, voids, and cracks. However, accurately quantifying and predicting these defects from C-SAM images remains a significant challenge due to the low noise-to-signal ratio, which typically arises from both imaging process and bond patterns itself. In this paper, we explore machine learning strategies for processing C-SAM images and providing predictive models of defect growth. We use C-SAM images of sintered copper and sintered silver samples, which are obtained under accelerated thermal experiments, as the representative dataset for our study. We investigate the effect of Fourier transforms and wavelet transforms on these datasets to remove high-frequency noise and address noise across multiple scales with histogram equalization to enhance the contrast and improve the visibility of defects. As a result, defect boundaries can be clearly distinguished, enabling more accurate tracking of their growth over time. We then employ different time-series forecasting algorithms on the denoised images to formulate an image-based lifetime prediction model. Statistical models and deep-learning techniques are trained on images obtained in the early stages of thermal shock, and defect growth in the later stages is predicted. Our work serves as a preliminary attempt to improve the accuracy of lifetime prediction models of power electronics packages, which is critical under extreme operating environments.

24 POWER TRANSMISSION AND DISTRIBUTION

Correct Interpretations of ENDF-102 Definitions for Resonance Effects

My Uncle Willie circa 1600 wrote “What’s in a name; a rose by any other name would smell as sweet.” I fear in this case we have a somewhat similar problem in that we may be using the same word but are not using the same definition; specifically, the word Unresolved. The simplest physics definition as it applies to neutron resonances, is the energy point where we can no longer see/measure ALL – let me repeat that – ALL - of the individual resonances. That seems simple and clear, but the question is: how to represent resonances beyond this point in order to accurately reproduce the effects we have seen in measurements and expect/need to reproduce in our applications. We know there are more, unseen resonances, otherwise we wouldn’t say Unresolved. The ENDF approach is well defined in ENDF-102 and simple: for ENDF data the only way to represent Unresolved data is by using a theoretical model to define the distribution of resonances, including those that are too narrow to measure (i.e., are unresolved). It is important to note that in ENDF this is the one and only Unresolved model, e.g., there is no provision in ENDF to accurately define individually ALL resonances above the Resolved energy range – by ALL here I mean both those that we can measure and those that we cannot individually measure, but that theory and integral measurements tells us are present. An alternative approach, which would appear to be equally valid, would be to include the latest measured data as tabulated energy expendent data extending upwards in energy above the Resolved energy range. In this approach the evaluation would not include an ENDF style Unresolved energy range; it would only include a Resolved resonance region, followed by tabulated higher energy points, representing the resonances that could be measured beyond the Resolved range. But an important point to note: By listing these resonances above the resolved energy one admits that at least some resonances in this energy range are missing as Unresolved; i.e., they are too narrow or overlapping to measure. The purpose of this paper is to illustrate that the later approach, while done with good intentions, and appearing to be valid/adequate in plots, does not meet the need of our engineering applications. Why? As we will see below, of these two possible approaches, only the ENDF use of a model to statistically include the missing, i.e., unresolved, resonances, can meet our engineering needs to reproduce the integral effects we have measured and understand. Only with this statistical model can we predict and include in our calculated results the important effects of temperature (Doppler broadening), and energy integrals (self-shielding). Below I will first present results using two ENDF/B-VIII.1 evaluations, U235 and U238, that use the correct ENDF-102 definition of an Unresolved resonance region, using a statistical model to include the effects of resonances that theory predicts are present, but are too narrow to measure. These two evaluations reproduce the expected temperature (Doppler) and energy integral (self-shielding) effects that we expect. Next I will present results using one ENDF/B-VIII.1 evaluation, 26-Fe-56, that does not use an ENDF-102 Unresolved resonance region; instead above its Resolved energy range it lists many tabulated energy points, that look like measured data, but by definition, since they are included above the ENDF Resolved energy range there are missing Unresolved resonances, i.e., there are missing the resonances that are too narrow to resolve, i.e., are unresolved. My conclusion, and I hope yours, is that the below figures illustrate that this approach does not reproduce the temperature and energy integrals that we expect and need to accurately calculate results for our fission reactor calculations. As such this approach should not be used in ENDF formatted evaluations.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

Rapid measurement of soluble xylo-oligomers using near-infrared spectroscopy (NIRS) and multivariate statistics: calibration model development and practical approaches to model optimization

Rapid monitoring of biomass conversion processes using techniques such as near-infrared (NIR) spectroscopy can be substantially quicker and less labor-, resource-, and energy-intensive than conventional measurement techniques such as gas or liquid chromatography (GC or LC) due to the lack of solvents and preparation methods, as well as removing the need to transfer samples to an external lab for analytical evaluation. The purpose of this study was to determine the feasibility of rapid monitoring of a biomass conversion process using NIR spectroscopy combined with multivariate statistical modeling, and to examine the impact of (1) subsetting the samples in the original dataset by process location and (2) reducing the spectral range used in the calibration model on model performance. We develop multivariate calibration models for the concentrations of soluble xylo-oligosaccharides (XOS), monomeric xylose, and total solids at multiple points in a biomass conversion process which produces and then purifies XOS compounds from sugar cane bagasse. A single model using samples from multiple locations in the process stream showed acceptable performance as measured by standard statistical measures. However, compared to the single model, we show that separate models built by segregating the calibration samples according to process location show improved performance. We also show that combining an understanding of the sample spectra with simple multivariate analysis tools can result in a calibration model with a substantially smaller spectral range that provides essentially equal performance to the full-range model. We demonstrate that real-time monitoring of soluble xylo-oligosaccharides (XOS), monomeric xylose, and total solids concentration at multiple points in a process stream using NIR spectroscopy coupled with multivariate statistics is feasible. Segregation of sample populations by process location improves model performance. Models using a reduced spectral range containing the most relevant spectral signatures show very similar performance to the full-range model, reinforcing the importance of performing robust exploratory data analysis before beginning multivariate modeling.

09 BIOMASS FUELS

Attribution of heterogeneous stress distributions in low-grain polycrystals under conditions leading to damage

In high-purity polycrystalline metallic materials, voids tend to favor grain boundaries as nucleation sites due to the elevated stress states produced by granular interactions and the weakened grain boundary from the relative atomic disorder. To quantify the key factors of this elevated stress state, simple compression of a small multi-grain cylinder of body-centered cubic tantalum was simulated using a single crystal plasticity model that incorporates non-Schmid effects. Four increasingly complex synthetic microstructures were created to tractably incorporate grain boundary interactions, and a statistically significant number of combinations were performed by varying the initial crystallographic orientations of the microstructure. Most of these simulations produce the maximum von Mises stress on a grain boundary and less frequently at the multi-grain junctions. To build a statistical model for the maximum von Mises stress at the grain boundary, physically based features that could contribute to the elevated stress state were selected. Then, a learning algorithm based on information theory was used to identify which of these features contributed the most information to the data set. The identified features include a grain’s propensity to accommodate both elastic and plastic deformations and their directional components. The misalignment of the direction of each grain’s mechanical response was found to be strongly correlated to the magnitude of the stress near the grain boundary. For all of the synthetic microstructures, the statistical models produce a residual distribution that is nearly Gaussian with a variance of, at most, 10% of the prior distribution. The successful performance of the statistical model implies the correct identification of the physical features that cause severe stress localization in polycrystalline materials. The statistical models constructed here can be used to formulate a physically motivated void nucleation model which is sensitive to a microstructure’s propensity to produce elevated stress states. As a result, these statistical models also enable the design of material microstructures, in which the crystallographic orientation is chosen to resist void nucleation.

36 MATERIALS SCIENCE

Statistical generic design of glass and optimization: Selective review on oxide glasses

Designing a single glass composition for a multidimensional property space is challenging, and the difficulty increases with the number of design criteria. Traditionally, the task is accomplished using multiple statistical models that describe the relationships between composition (C) and property (P) values, i.e., C-P models. Recently, the structure (S)-property (P) statistical modeling has emerged as a complementary approach. The S-P modeling approach has also been shown to be a preferred method for modeling glass properties, particularly when a small data set is available, such as in single-component studies, or when strong nonlinearities exist between composition and properties. The combined model package, C-S-P, implements the concept of generic glass design, i.e., designing glass for performance by first selecting a specific or optimized set of glass network structural groups using S-P models and then transferring the designed structures (genes) to a particular composition using C-S models. This article reviews a set of supporting cases from the previous C-S-P modeling studies of phosphate, silicate, and borosilicate glasses, which are relevant for many critical commercial applications. The methodology for developing the statistical C-S-P database is presented, enabling the application of P?S?C to achieve a generic glass design and optimization, targeting multiple design criteria for both performance and processing properties simultaneously.

Network structure

Direct cross section measurement of 102 Pd ⁢(𝛾,𝑝) and 102 Pd ⁢(𝛾,𝛼) for the astrophysical 𝑝 process

Background: A handful of neutron-deficient stable nuclei, known as the “p nuclei,” cannot be produced through astrophysical neutron capture processes. Instead, some of these nuclei are proposed to be produced by 𝛾-induced reactions on existing r- and s-process seeds. The specific astrophysical site or sites are not yet identified, however, with uncertainties in the cross sections of these 𝛾-induced reactions playing a role. Databases of reaction rates for astrophysical simulations often rely on theoretical statistical model calculations, such as Hauser-Feshbach, for rates where no experimental information is known. However, reasonable variations in the choice of parametrizations of various nuclear properties can create order-of-magnitude variations in the final predicted cross sections and reaction rates, which are then propagated through the models to the predicted final abundances. Purpose: To better constrain these statistical model calculations and ultimately reduce the uncertainties from the nuclear physics on our understanding of the p nuclei, a measurement of the cross sections of 𝛾-induced reactions on the p-nucleus 102 Pd was undertaken. This work represents the first measurement of its kind, using segmented silicon detectors to measure prompt charged particle emission from 𝛾-induced reactions. Methods: Quasimonoenergetic gamma beams from the High Intensity 𝛾 Source facility bombarded an enriched 102 Pd target. A segmented silicon array was arranged to detect the particles emitted from (𝛾,𝑝) and (𝛾,𝛼) reactions. Results: Reaction cross sections were deduced at multiple 𝛾-beam energies between 10 and 19 MeV, and compared to statistical model calculations using talys-1.96. The 102 Pd ⁢(𝛾,𝑝)⁢ 101 Rh reaction cross section was reasonably well reproduced by a subset of photon strength functions and level densities, though the strength to the ground state of 101 Rh was underestimated at higher incident gamma energies. The 102 Pd ⁢(𝛾,𝛼)⁢ 98 Ru was in general overpredicted by the various alpha-nucleus optical model potentials. Conclusions: While the theoretical cross sections used to model the (𝛾,𝑝) reactions for the p process may be reasonable, a more careful approach is needed in the case of (𝛾,𝛼). Further work to probe gamma-induced reaction cross sections at and near the p nuclei is warranted.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS

coh3

CoH3 (CoH ver.3) is an optical model, exciton pre-equilibrium, and Hauser-Feshbach statistical model code, which calculates nuclear reaction cross sections for medium to heavy targets in the keV to MeV energy region. This program is written in standard C++, divided into approximately 200 source and header files. CoH solves the Schroedinger equation for optical potentials defined in the code, and calculates differential elastic scattering, reaction, and total cross sections, for neutron, proton, deuteron, triton, 3He, and alpha-particle. Deformed optical potentials are solved with the coupled-channels method, in which the ground state rotational band members, or vibrational phonon states are coupled. The optical model gives particle transmission coefficients that are fed into the statistical model calculations. CoH includes the pre-equilibrium model (exciton model), the direct/semidirect capture model, and the multi-stage Hauser-Feshbach statistical decay with width fluctuation correction based on the Gaussian orthogonal ensemble. For weakly coupled levels, the DWBA (distorted wave Born approximation) method is used to calculate the direct inelastic scattering process to the excited states.

Kawano, Toshihiko

Misclassification in Workers’ Telecommuting Frequency Choices Using a Generalized Extreme Value Model

Telecommuting frequency is a response variable collected in travel surveys and is, therefore, prone to errors leading to mismeasurements or misclassification. Misclassification of explanatory variables is a common risk when using statistical modeling techniques. We define “misclassification” as a response reported or recorded in the wrong category; for example, a variable is recorded as a 1 when it should be 0. Here, in this context, this study aims to develop a statistical model to analyze telecommuting data which accounts for potential misclassification errors by building on existing literature in econometrics. The empirical analysis was undertaken using the 2017 National Household Travel Survey (NHTS) and the general extreme value (GEV) models available in the literature. Specifically, the frequency of telecommuting days was analyzed using the negative binomial (NB) model recast as the multinomial logit (MNL) model. By nature—and consistent with other studies—NHTS data are prone to errors that can be classified as intentional or unintentional misinformation provided by the person being interviewed. Ignoring these errors while modeling telecommuting frequencies using standard discrete count models can result in biased parameter estimates. The misclassification parameter was calculated for both over-reporting and under-reporting scenarios. The misclassification errors can be as high as 14% over-reported and 10% under-reported, particularly for the neighboring values. Statistical fit comparison between the models shows that models that ignore misclassification have worse data fit and biased parameter estimates with significant policy implications.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI

First Measurement of 87 Rb( α , xn ) Cross Sections at Weak r -process Energies in Supernova ν -driven Ejecta to Investigate Elemental Abundances in Low-metallicity Stars

Observed abundances of Z ∼ 40 elements in metal-poor stars vary from star to star, indicating that the rapid and slow neutron capture processes may not contribute alone to the synthesis of elements beyond iron. The weak r-process was proposed to produce Z ∼ 40 elements in a subset of old stars. Thought to occur in the ν-driven ejecta of a core-collapse supernova, ( α, xn ) reactions would drive the nuclear flow toward heavier masses at T = 2−5 GK. However, current comparisons between modeled and observed yields do not bring satisfactory insights into the stellar environment, mainly due to the uncertainties of the nuclear physics inputs where the dispersion in a given reaction rate often exceeds 1 order of magnitude. Involved rates are calculated with the statistical model where the choice of an α -optical-model potential ( α OMP) leads to such a poor precision. The first experiment on 87 Rb( α, xn ) reactions at weak r -process energies is reported here. Total inclusive cross sections were assessed at E c.m. = 8.1−13 MeV (3.7−7.6 GK) with the active target MUlti-Sampling Ionization Chamber. With an N = 50 seed nucleus, the measured values agree with statistical model estimates using the α OMP Atomki-V2. A reevaluated reaction rate was incorporated into new nucleosynthesis calculations, focusing on ν-driven ejecta conditions known to be sensitive to this specific rate. These conditions were found to fail to reproduce the lighter heavy element abundances in metal-poor stars.

79 ASTRONOMY AND ASTROPHYSICS