Search NASA⌕ Search

SEARCH · Search NASA

Results for “Data Analysis: Parameter Estimation”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 73 records · Page 4

SAXS Assistant: Automated SAXS analysis for structural discovery in biologics and polymeric nanoparticles

Small-angle x-ray scattering (SAXS) is a powerful technique for assessing macromolecular structure. High-throughput SAXS is limited by the time-consuming and, at times, subjective nature of SAXS data interpretation. Here, we present SAXS Assistant, a Python-based script that streamlines SAXS data analysis to extract features for machine learning (ML) and key structural parameters, including the Guinier radius of gyration (R g ), pair distance distribution function (PDDF)-derived R g , maximum particle dimension (D max ), and Kratky plots. The script builds upon BioXTAS RAW and validates reliability via Guinier/PDDF R g agreement, an important indicator of well-measured data sets. For assistance in D max estimation, a multilayer perceptron regressor was trained with 1940 data files from the Small Angle Scattering Biological Data Bank. The model achieved a test set performance R 2 = 0.90 and mean absolute error = 11.7 Å. Training exclusively with experimental data translates analyses from researchers, including experts in the field, to the ML model, which helps assess D max estimations from PDDF. Gaussian mixture model clustering was implemented to classify profiles into structural classes based on entries in the Small Angle Scattering Biological Data Bank. Users may therefore assess the similarity between experimental samples and known biomolecular shapes within the mapped repository entries. This probabilistic clustering aids in quantifying information from Kratky and generating shape-descriptive features. SAXS Assistant accelerates SAXS data analysis through enforced quality control, ML-ready outputs, and flags for low-confidence results. In addition to providing the ability to analyze large data sets at high throughput, this tool is versatile and may serve researchers in both biological and synthetic polymer research fields.

36 MATERIALS SCIENCE↗

Association Kinetics for Perfluorinated n -Alkyl Radicals

Radical-radical reaction channels are important in the pyrolysis and oxidation chemistry of perfluoroalkyl substances (PFAS). In particular, unimolecular dissociation reactions within unbranched n-perfluoroalkyl chains, and their corresponding reverse barrierless association reactions, are expected to be significant contributors to the gas-phase thermal decomposition of families of species such as perfluorinated carboxylic acids and perfluorinated sulfonic acids. Unfortunately, experimental data for these reactions are scarce and uncertain. Furthermore, obtaining reliable theoretical predictions for such reactions is a laborious and computationally intensive task. Here, in this work, the chemical kinetics of the various association/decomposition reactions producing/decomposing the C 2 -C 4 series of unbranched n-perfluoroalkanes (C 2 F 6 , C 3 F 8 , and C 4 F 10 ) are examined using state-of-the-art ab initio transition-state-theory-based master-equation calculations. The variable-reaction-coordinate transition-state theory (VRC-TST) formalism is employed in computing the microcanonical and canonical rates for the association reactions. Reaction thermochemistry is obtained via composite quantum chemistry calculations and the laddering of error-canceling reaction schemes via a connectivity-based hierarchy approach employing ANL1/ANL0-style reference energies. Lennard-Jones collision model parameters for the considered systems were estimated by a direct dynamics approach, and collisional energy transfer parameters were obtained from analogies to systems of similar size and heavy-atom connectivity. A one-dimensional master equation approach was used to convert the microcanonical rate coefficients from the VRC-TST analysis into temperature- and pressure-dependent rate constants for the association reactions and the reverse dissociation reactions. The data are reported in standardized formats for usage in comprehensive chemical kinetic models for PFAS thermal destruction.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

DESI DR2 Baryon Acoustic Oscillations from the Lyman Alpha Forest Multipoles

We present an alternative measurement of the Baryon Acoustic Oscillation (BAO) using the Legendre multipole representation of the Ly$α$ forest correlation functions from the second data release (DR2) of the Dark Energy Spectroscopic Instrument survey. Compressing the auto- and cross-correlation functions into Legendre multipoles yields a positive-definite covariance matrix without any smoothing -- unlike the baseline DR2 analysis -- thanks to a significantly reduced data vector size. We introduce the statistical corrections required to debias the finite-sample covariance matrix estimate and demonstrate that monopole and quadrupole terms for both auto- and cross-correlations can be used even when the correlation functions are distorted by continuum errors and contaminated by metals. This formalism has slightly diminished the constraining power of the BAO scale, while considerably weakening constraints on nuisance parameters. We measure the isotropic BAO scale with $0.93\%$ precision at $z_\mathrm{eff}=2.35$, the Hubble parameter $H(z_\mathrm{eff})=(239.5\pm3.4)~(147.09~\mathrm{Mpc}/r_d) ~\mathrm{km~s}^{-1}~\text{Mpc}^{-1}$, and the transverse comoving distance $D_M(z_\mathrm{eff})=(5.80 \pm 0.10)~(r_d/147.09~\mathrm{Mpc})$~Gpc for a given value of the sound horizon ($r_d$). Our BAO results are entirely consistent with the baseline DR2 analysis.

Karaçaylı, Naim Göksel [Chicago U., KICP; Ohio Sta↗

Parametrization of Generalized Parton Distributions from 𝑡-Channel String Exchange in AdS Spaces

We introduce a string-based parametrization for nucleon quark and gluon generalized parton distributions (GPDs) that is valid for all skewness. Our approach leverages conformal moments, representing them as the sum of spin-𝑗 nucleon 𝐴-form factor and skewness-dependent spin-𝑗 nucleon 𝐷-form factor, derived from 𝑡-channel string exchange in AdS spaces consistent with Lorentz invariance and unitarity. This model-independent framework, satisfying the polynomiality condition due to Lorentz invariance, uses Mellin moments from empirical data to estimate these form factors. With just five Regge slope parameters, our method accurately produces various nucleon quark GPD types and symmetric nucleon gluon GPDs through pertinent Mellin-Barnes integrals. Our isovector nucleon quark GPD is in agreement with existing lattice data, promising to improve the empirical extraction and global analysis of nucleon GPDs in exclusive processes, by avoiding the deconvolution problem at any skewness, for the first time.

QCD phenomenology↗

Testing the ΛCDM Cosmological Model with Forthcoming Measurements of the Cosmic Microwave Background with SPT-3G

We forecast constraints on cosmological parameters enabled by three surveys conducted with SPT-3G, the third-generation camera on the South Pole Telescope. The surveys cover separate regions of 1500, 2650, and 6000 deg$^{2}$ to different depths, in total observing 25% of the sky. These regions will be measured to white noise levels of roughly 2.5, 9, and 12μK-armin, respectively, in cosmic microwave background (CMB) temperature units at 150 GHz by the end of 2024. The survey also includes measurements at 95 and 220 GHz, which have noise levels a factor of ∼1.2 and 3.5 times higher than 150 GHz, respectively, with each band having a polarization noise level ∼2times higher than the temperature noise. We use a novel approach to obtain the covariance matrices for jointly and optimally estimated gravitational lensing potential band powers and unlensed CMB temperature and polarization band powers. We demonstrate the ability to test the ΛCDM model via the consistency of cosmological parameters constrained independently from SPT-3G and Planck data, and consider the improvement in constraints on ΛCDM extension parameters from a joint analysis of SPT-3G and Planck data. The ΛCDM cosmological parameters are typically constrained with uncertainties up to ∼2 times smaller with SPT-3G data, compared to Planck, with the two data sets measuring significantly different angular scales and polarization levels, providing additional tests of the standard cosmological model.

79 ASTRONOMY AND ASTROPHYSICS↗

Optimizing Batch Crystallization with Model-based Design of Experiments

Adaptive and self-optimizing intelligent systems such as digital twins are increasingly important in science and engineering. Digital twins utilize mathematical models to provide added precision to decision-making. However, physics-informed models are challenging to build, calibrate, and validate with existing data science methods. Model-based design of experiments (MBDoE) is a popular framework for optimizing data collection to maximize parameter precision in mathematical models and digital twins. In this work, we apply MBDoE, facilitated by the open-source package Pyomo.DoE, to train and validate mathematical models for batch crystallization. We quantitatively examined the estimability of the model parameters for experiments with different cooling rates. This analysis provides a quantitative explanation for the heuristic of using multiple experiments at different cooling rates.

Lynch, Hailey↗

Multi-facility analysis using metered power data to quantify MRI energy use and utility bill costs across scanner operating modes

This study quantifies the energy consumption of magnetic resonance imaging (MRI) scanners across discrete operating modes during routine clinical workflows, based solely on electrical power measurements. Although previous studies have investigated MRI energy consumption within single hospitals or specific clinical settings, this research provides a broader and more systematic analysis. Researchers analyzed electrical power data and applied a previously developed semi-automatic method for identifying MRI operating modes using load duration curves for 20 MRI scanners across four different U.S. healthcare facilities, encompassing outpatient, inpatient, and mixed-use clinical settings. A key innovation is the inclusion of localized hourly utility rates to estimate costs, a parameter absent in prior literature. Key findings indicate significant variability in energy and cost profiles between weekdays and weekends. Scanner characteristics, including magnet strength, manufacturer, vintage, location, and clinical setting, influenced average daily energy consumption and power thresholds for operating modes. Notably, the clinical setting of a scanner predominantly determines its energy use. For example, the scanners in outpatient facilities consumed more energy. The breakdown of energy usage and costs by operating modes showed scanners spend between 61% and 93% of their time in nonproductive modes, with one outlier spending 34%. Average daily energy use for the scanners in the study ranged from 160 to 1069 kWh, with energy costs ranging from $\$$9 to $\$$149. This study uses an existing framework to quantify MRI energy behavior, leading to insights that can enable improved performance and cost savings across different healthcare environments.

24 POWER TRANSMISSION AND DISTRIBUTION↗

Discovering the Unknowns: A First Step

This article aims at discovering the unknown variables in the system through data analysis. The main idea is to use the time of data collection as a surrogate variable and try to identify the unknown variables by modeling gradual and sudden changes in the data. We use Gaussian process modeling and a sparse representation of the sudden changes to efficiently estimate the large number of parameters in the proposed statistical model. The method is tested on a realistic dataset generated using a one-dimensional implementation of a Magnetized Liner Inertial Fusion (MagLIF) simulation model, and encouraging results are obtained.

42 ENGINEERING↗

Polaris-PARCS Sensitivity Study on LWR Fuel Cycles: Polaris Input Options

This study is the first of a multi-phase effort to assess the sensitivity of light-water reactor (LWR) core-level prediction biases to changes in lattice-level calculation parameters. Prediction bias is the measured-to-predicted difference in a core-level quantity of interest (QOI) which can be estimated by comparing the simulation results with the plant-measured data for key nuclear parameters. The LWR two-step neutronics codes employed herein are the SCALE–Polaris lattice physics code (v6.3.1) and the Purdue Advanced Reactor Core Simulator (PARCS) nodal diffusion simulator (v3.4.2), both funded and used for confirmatory analysis to support licensing by the US Nuclear Regulatory Commission (NRC). Polaris–PARCS is used to model Watts Bar Unit 1 cycles 1–3 and Peach Bottom Unit 2 cycles 1–3. This study focuses on the impact of changes to Polaris input options such as scattering treatment or quadrature settings and how these input options induce changes in core-level quantities of interest (QOIs)bias. The report documents multiple bias assessments for different modeling choices and compares the bias magnitude to the QOI measurement uncertainties. Future companion reports will investigate the sensitivity of core-level LWR prediction bias to Polaris input options and Polaris-computed QOIs such as few-group assembly-homogenized cross sections to gain an understanding of the key drivers of prediction bias at lattice and core levels for application of a two-step LWR neutronics procedure in a licensing scenario.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Bayesian calibration and uncertainty quantification of a rate-dependent cohesive zone model for polymer interfaces

In this work we present a rate-dependent cohesive zone model for the fracture of polymeric interfaces and performs a Bayesian calibration, an uncertainty quantification, and a sensitivity analysis for the model. The proposed cohesive zone model accounts for both reversible elastic and irreversible rate-dependent separation sliding deformation at the interface. The viscous dissipation due to the irreversible opening at the interface is modeled using elastic-viscoplastic kinematics that incorporates the effects of strain rate. Inverse calibration of parameters for such complex models through trial and error is challenging due to the large number of parameters of the model. Moreover, the calibrated parameter values are often non-unique and uncertain when the available experimental data is limited. To tackle this challenge, we employ a Bayesian calibration approach to identify parameters from experimental data, the resulting parameters significantly enhance the accuracy of the model. To quantify the uncertainty associated with the inverse parameter estimation, a modular Bayesian approach is employed to calibrate the unknown model parameters, accounting for the parameter uncertainty of the cohesive zone model. The advantages of the Bayesian calibration over a deterministic parameter fit are demonstrated. Further, to quantify the model uncertainties, such as incorrect assumptions or missing physics, a discrepancy function is introduced, which significantly improves the model’s prediction. Finally, the total uncertainty of the model is quantified in a predictive setting. A sensitivity analysis is performed to assess how changes in the input variables of the model affect the peak load, facilitating the identification of a concise set of highly influential parameters. The present approach can be used for calibration and uncertainty quantification for other complex computational mechanics models. It should also facilitate the designing of interface materials under uncertainty.

42 ENGINEERING↗

Uncertainty Propagation from Experiment Measurements to Modeling Approaches: A Case for SMR Steam Entrainment Testing

To license new and advanced reactor designs, regulators must be convinced that their unique safety cases—relative to existing large scale reactors—have been adequately addressed by the designed reactor protection systems. In water cooled small modular reactors (SMRs), droplet entrainment in steam flow has significant implications on the progression of accident scenarios due to its compact design features, which requires representative test data applicable to SMR designs. Computer code, modeling and simulation (M&S) tools and models require adequate verification, assessment, and qualification. This includes M&S results validation against scaled empirical data within allowable uncertainty bands to gain regulatory approvals during the various stages of reactor system design, demonstration, and commercialization. However, measurement uncertainty within the empirical datasets and test data applicability ranges requires careful consideration of M&S inputs (i.e., boundary conditions, and initial conditions), and verification and validation efforts. This study focuses on uncertainty quantification in designing scaled test facilities for SMR applications with appropriate measurements and a standard data-reduction method to estimate thermal hydraulics characteristics parameters that incorporate physics phenomena of interest. In addition, this study supports the evaluation model development and assessment process using M&S that interfaces with advanced computing tools and digital twin capabilities. This will allow synchronization between experiment and modeling approaches for droplet entrainment testing and analysis, improving diagnostics, prognostics, and decision-making to accelerate regulatory approval.

21 SPECIFIC NUCLEAR REACTORS AND ASSOCIATED PLANTS↗

PV-Finder: ML Based Algorithm for Primary Vertex Identification

he CMS detector at the High-Luminosity Large Hadron Collider (HL-LHC) will operate in challenging conditions with expected pile-up of up to 200 collisions per bunch crossing, necessitating the development of a more resilient primary vertex (PV) reconstruction method to ensure the integrity of data analysis and the efficiency of the CMS triggering system. This contribution describes preliminary studies on a new ML based PV-Finder method for PV identification. The method is based on a model trained using Kernel Density Estimations (KDEs) derived from the positions of reconstructed tracks at the beamline, incorporating uncertainties from track parameters. It also utilizes target histograms, modeled as Gaussian distributions centered on the actual ground truth values of specific primary vertices.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

eDNAjoint: An R package for interpreting paired or semi‐paired environmental DNA and traditional survey data in a Bayesian framework

Abstract Environmental DNA (eDNA) sampling is increasingly used in surveys of species distribution as a potentially sensitive and efficient monitoring method. Yet access to modelling tools designed specifically for interpreting this new data type lags behind its ubiquity. While occupancy modelling software has dominated the analytical landscape for eDNA data analysis of single species, this type of model may not always be the most appropriate. The rate of eDNA detection often corresponds to species density, rather than just occupancy, and researchers often have access to observations from non‐genetic sampling methods at the same sites. To provide users access to a modelling framework designed to maximize the use of all available data, we developed an R package, eDNAjoint . The package provides an easy‐to‐use interface for fitting a ‘joint’ model that integrates data from paired or semi‐paired eDNA and traditional surveys in a Bayesian framework. The model can be used to estimate parameters like the probability of a false positive eDNA detection and mean catch rate at a site, and the package allows access to multiple model variations and Bayesian prior customization. Additional functionality can be used for model selection, summarising posteriors and comparing the relative sensitivities of the two survey methods. We demonstrate the use of eDNAjoint by fitting a variation of the model with site‐level covariates that scale the sensitivity of eDNA sampling relative to traditional sampling. The example workflow uses binary eDNA and seine count data for the endangered tidewater goby ( Eucyclogobius newberryi ) from a study by Schmelzle and Kinziger (2016). This use case includes a prior sensitivity analysis and an evaluation of the relationship between detection rates and environmental variables. eDNAjoint has the potential to greatly increase the range of users who will be able to rigorously analyse eDNA and traditional survey data in a Bayesian framework, understand if and how eDNA can improve monitoring practices, and gain confidence in the interpretability of eDNA data.

Keller, Abigail G. [Department of Environment Scie↗

An implementation of neural simulation-based inference for parameter estimation in ATLAS

Neural simulation-based inference (NSBI) is a powerful class of machine-learning-based methods for statistical inference that naturally handles high-dimensional parameter estimation without the need to bin data into low-dimensional summary histograms. Such methods are promising for a range of measurements, including at the Large Hadron Collider, where no single observable may be optimal to scan over the entire theoretical phase space under consideration, or where binning data into histograms could result in a loss of sensitivity. This work develops a NSBI framework for statistical inference, using neural networks to estimate probability density ratios, which enables the application to a full-scale analysis. It incorporates a large number of systematic uncertainties, quantifies the uncertainty due to the finite number of events in training samples, develops a method to construct confidence intervals, and demonstrates a series of intermediate diagnostic checks that can be performed to validate the robustness of the method. As an example, the power and feasibility of the method are assessed on simulated data for a simplified version of an off-shell Higgs boson couplings measurement in the four-lepton final states. This approach represents an extension to the standard statistical methodology used by the experiments at the Large Hadron Collider, and can benefit many physics analyses.

frequentist statistics↗

Improving Trustworthiness of Data-Driven Power Grid Contingency Analysis With Bayesian Residual Graph Neural Networks

The evolving energy landscape requires novel tools to efficiently perform contingency analysis and reliability assessment of power grids, potentially in real-time. The high computational cost of traditional power flow solvers limits their applicability in practice. Machine learning (ML) surrogates such as deep neural networks (NNs) accelerate power flow solvers computations, enabling high-order contingency analysis and real-time decision-making by learning highly nonlinear functions and integrating grid topology via graph architectures. However, (graph) NNs lack predictive power away from training data and do not provide predictive confidence estimates. Here, we present a Bayesian residual graph NN that integrates knowledge from low-fidelity data via residual training and embeds granular quantification of uncertainties, improving trustworthiness critical for high-consequence decision-making. Applying Bayesian concepts to NNs is challenging due to the high-dimensionality of both the parameter space, complicating derivation of a meaningful prior, and the output space in large grid systems, requiring enhanced techniques to assess the predicted high-dimensional uncertainties. Our contributions include: (1) Deriving a prior for fully connected and graph NNs that leverages low-fidelity data to guide mean predictions and appropriately control prior predictive uncertainty. (2) Integrating this prior within an ensembling with anchoring scheme for efficient approximate posterior inference. (3) Deriving enhanced metrics to assess accuracy of both the mean and uncertainty predictions in high dimensions, appropriately accounting for correlations propagated through graph layers. The resulting Bayesian residual graph NN is tested on a contingency analysis task for 14-bus and 118-bus grids.

24 - POWER TRANSMISSION AND DISTRIBUTION↗

Precision measurements of EFT parameters and BAO peak shifts for the Lyman- α forest

We present precision measurements of the bias parameters of the one-loop power spectrum model of the Lyman- α (Ly- α ) forest, derived within the effective field theory (EFT) of large-scale structure. We fit our model to the three-dimensional flux power spectrum measured from the ACCEL 2 hydrodynamic simulations. The EFT model fits the data with an accuracy of below 2% up to k = 2 h Mpc − 1 . Further, we analytically derive how nonlinearities in the three-dimensional clustering of the Ly- α forest introduce biases in measurements of the baryon acoustic oscillations (BAOs) scaling parameters in radial and transverse directions. From our EFT parameter measurements, we obtain a theoretical error budget of Δ α ∥ = − 0.2 % ( Δ α ⊥ = − 0.3 % ) for the radial (transverse) parameters at redshift z = 2.0 . This corresponds to a shift of − 0.3 % (0.1%) for the isotropic (anisotropic) distance measurements. We provide an estimate for the shift of the BAO peak for Ly- α -quasar cross-correlation measurements assuming analytical and simulation-based scaling relations for the nonlinear quasar bias parameters resulting in a shift of − 0.2 % ( − 0.1 % ) for the radial (transverse) dilation parameters, respectively. This analysis emphasizes the robustness of Ly- α forest BAO measurements to the theory modeling. We provide informative priors and an error budget for measuring the BAO feature—a key science driver of the currently observing Dark Energy Spectroscopic Instrument (DESI). Our work paves the way for full-shape cosmological analyses of Ly- α forest data from DESI and upcoming surveys such as the Prime Focus Spectrograph, WEAVE-QSO, and 4MOST. Published by the American Physical Society 2025

de Belsunce, Roger (ORCID:0000000336604028)↗

Hiperclust

This software leverages transfer learning to analyze atom probe tomography (APT) data. It is trained on synthetic data and then applies this knowledge to predict the optimal number of clusters for a given APT dataset. Initially, the software used preliminary clustering to estimate the general structure of the data. Based on this, it provides suggestions for key parameters like minimum cluster size and minimum number of points. These parameters are critical for algorithms like HDBSCAN, ensuring accurate cluster formation without the need for trial-and-error testing. The software runs on High-Performance computing (HPC) systems, enabling fast, scalable analysis of large APT datasets, ultimately saving time and improving the reliability of clustering outcomes.

Tang, Yalei [Idaho National Laboratory (INL), Idah↗

System Study: Emergency Power System 1998-2024

This report presents an unreliability evaluation of the emergency power system (EPS) at 93 U.S. commercial operating nuclear reactors. New Standardized Plant Analysis Risk (SPAR) models with the most recent SPAR parameter update results were used in this report. Demand, run hour, and failure data from 1998 to 2024 for selected components were obtained from the Institute of Nuclear Power Operations Industry Reporting and Information System. The unreliability results are trended for the most recent 10 year period while yearly estimates for system unreliability are provided for the entire active period. No statistically significant increasing or decreasing trends were identified in the industry-wide estimates of EPS system start-only unreliability, but a statistically significant decreasing trend was identified in the industry-wide estimates of EPS system 24-hour mission unreliability.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗