Search NASA⌕ Search

SEARCH · Search NASA

Results for “statistical model”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 415 records · Page 23

Constraining primordial non-Gaussianity from the large scale structure two-point and three-point correlation functions

Surveys of cosmological large-scale structure (LSS) are sensitive to the presence of local primordial non-Gaussianity (PNG), and may be used to constrain models of inflation. Local PNG, characterized by f NL ⁠, the amplitude of the quadratic correction to the potential of a Gaussian random field, is traditionally measured from LSS two-point and three-point clustering via the power spectrum and bi-spectrum. We propose a framework to measure f NL using the configuration space two-point correlation function (2pcf) monopole and three-point correlation function (3pcf) monopole of survey tracers. Our model estimates the effect of the scale-dependent bias induced by the presence of PNG on the 2pcf and 3pcf from the clustering of simulated dark matter haloes. We describe how this effect may be scaled to an arbitrary tracer of the cosmological matter density. The 2pcf and 3pcf of this tracer are measured to constrain the value of f NL ⁠. In LSS surveys, the effect of imaging systematics on two-point statistics is often degenerate with the PNG signal. Our proposed model employs three-point statistics primarily to break this degeneracy. Using simulations of luminous red galaxies observed by the Dark Energy Spectroscopic Instrument (DESI), we demonstrate the accuracy and constraining power of our method. Our forecast indicates the ability to constrain f NL to a precision of σf NL ≈ 22 with one year of DESI survey data, as well as the ability to constrain the imaging systematic weights in situ.

early Universe↗

Nonparametric Multiparticle Set Methods for Interpreting Environmental Samples

Collection and analysis of environmental samples is commonly used by a range of stakeholders in nuclear safeguards and security contexts. While the ubiquity of samples and their transport in the environment allow regular collection, developing and demonstrating methods for analyzing these samples is difficult. In this work, an environmental sample consists of a set of one or more individual particles. Recent advances in reactor simulation have allowed us to generate data that are more representative of real-world environmental samples, enabling statistically defensible method development and testing. The most notable of these advances is a drastic increase in the number of material depletion regions, which allows our simulations to capture the variation in isotopic composition seen at length scales consistent with environmental samples. Traditional approaches for handling multiparticle samples treat each particle in the sample individually, estimating the quantity of interest (e.g., core-average burnup) resulting from measurement and analysis of signatures (e.g., nuclide assays) from each individual particle. Individual estimates are then averaged to generate a single estimate of the quantity of interest over the entire sample. In this presentation, we introduce two novel approaches for interpreting environmental samples that comprise of multiple particles: (1) the Quantile-Quantile Comparator, which uses a multivariate generalization of quantile-quantile plots for comparing unknown statistical distributions, and (2) the Set Transformer, an attention-based neural network module designed to model interactions among elements (particles) in the input set (sample). Statistically representative sampling cannot be guaranteed as samples are passively collected and are beholden to what particles are available in the environment. These new analysis methods for set-input problems are expected to be more robust than traditional approaches to issues of sampling bias where particles are not uniformly distributed throughout regions of interest, as well as generally outperform traditional approaches by jointly considering all elements in the set. We will present results comparing the performance of traditional single particle approaches and the novel Quantile-Quantile Comparator and Set Transformer for interpretation of simulated environmental samples.

Phathanapirom, Birdy↗

Surrogate model evaluation and building energy benchmarking for commercial buildings

Building energy consumption benchmarking involves challenges associated with various energy patterns for different building types; heating, ventilating, and air-conditioning (HVAC) system types; and climates. Given significant variation in energy use patterns, accurate prediction of long-term energy use using surrogate models remains challenging. Multiple linear regression (MLR) is commonly used for building energy benchmarking because of its simple structure; however, it lacks accuracy compared to other black-box models. Although many studies have compared surrogate models and offer guidance on model selection based on metrics, they do not provide detailed analysis on improving the surrogate model accuracy. In this paper, we implement a surrogate model using polynomial ridge regression (i.e., MLR with interaction terms combined with ridge regularization) for small office and retail strip mall buildings across six HVAC system types and all climate zones, for electricity and natural gas in baseline and proposed scenarios. A simulation workflow is developed using OpenStudio TM /EnergyPlus TM to generate simulation data using measures over a wide range of efficiency inputs. Enhancements based on statistical insights are used for improving the model accuracy using filters, input transformations, and change points. Surrogate models achieved average coefficient of variation of the root mean squared error (CVRMSE) values of 2.17, 1.06, 2.05, and 3.26 for proposed electricity, proposed natural gas, baseline electricity, and baseline natural gas, respectively, with enhancements reducing CVRMSE by an average of 14.9% across all combinations. We provide model interpretation via Shapley additive explanations to determine which input variables most influence energy consumption and provide supportive arguments for enhancements.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Nonlinearities in Magnetic Confinement, Ionospheric Physics, and Population Explosion Leading to Profile Resilience Нелінійності в магнетному утриманні, фізиці іоносфери та процесі демографічного вибуху, які приводять до стійкості профілю

Nonlinearities play an important role in many fields. In the field of thermonuclear fusion, they are involved in questions such as profile resilience and fluid closure. A nonlinear phenomenon common to both fusion and astrophysical planets is the generation of zonal flows. These flows play a significant role in determining the level of turbulence and fluid closure in fusion. The effects of resonance broadening and nonlinearities are investigated, specifically focusing on the case of nonlinear instability that has appeared in drift waves. Similarities and differences between our systems are discussed, with population explosion and the dynamics of nonlinear systems for drift waves by different states in profile resilience described with great precision. The aim of our study is to put our fluid model for drift waves in tokamaks within the wider framework of statistical physics principles. This reinforces our belief in the broad application of our drift wave model, which encompasses current tokamaks, ITER, and the fusion pilot plant.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Ranking Biological Features in Soil-Based Microbial Multi-Omics Data with Integration Modeling

Distinguishing the most important features (e.g. proteins, metabolites, etc.) per group (e.g. control and treatment) is a critical challenge in feature-rich multi-omics experiments, especially in soil data. Traditional feature identification and ranking approaches, such as differential expression, are based on single omics and thus not directly translatable to multi-omics experiments. Here, 5 multi-omics integration models (DIABLO, JACA, MOFA, MultiMLP, and SLIDE) that were not explicitly built for soil data applications were tested using a soil-based multi-omics experiment. The data were obtained from an experimental setup of an autoclaved soil system inoculated with 8 bacteria and using chitin as the carbon source and including samples collected at 0- (control), 4-, 8-, and 12-weeks post-inoculation. The omics data included metaproteomics, 16S rRNA sequencing, and LC-MS/MS metabolomics (in positive and negative mode). Each multi-omics integration model was implemented, and top features were compared to differential univariate statistics per omic type, demonstrating that integration approaches cut the potential number of top features from 2957 identified by differential statistics to 13-224 (a 99.6% to 92.4% reduction). Interestingly, most top features across integration models were not shared; though, scaling and averaging ranks across models shared similar patterns. This work highlights the usefulness of multi-omics integration models in soil-based microbial studies and the power of using multiple integration models together to interpret results.

54 ENVIRONMENTAL SCIENCES↗

Statistical Uncertainty of Inhalation Dose Coefficients: Impact of Particle Deposition in ICRP 66 Human Respiratory Tract Model

Inhaled radioactive materials can pose a long-term health concern, as the material can be incorporated into the body’s metabolic pathways and remain in organs and tissues for extended durations. During the retention period, the radioactive material may localize in a source organ and irradiate adjacent target organs and tissues. Distribution of these materials changes over time, requiring biokinetic modeling to evaluate their movement through various tissues and organs. The evolving distribution depends on multiple inputs characterizing the inhaled material, such as particle size and size distribution, particle density, aspect ratio, specific radionuclide, the chemical form, and solubility. In addition, biological parameters such as breathing rate, breathing type (nasal or nasal/oral), respiratory system morphometry, tidal volume, functional residual capacity, and anatomical dead space all influence material transport. These aerosol properties and physiological characteristics of the respiratory tract jointly define a range of initial conditions that influence the time-dependent distribution of radioactive material. To evaluate both uncertainty in the initial conditions of inhalation exposure and the final output (committed effective dose) from biokinetic models, a Python-based software tool, Radiological Exposure Dose Calculator (REDCAL), was developed to propagate uncertainty within the human respiratory tract model. Focusing on deposition fraction uncertainty, the primary objective was to characterize the initial activity distribution across respiratory regions as a function of anticipated particle sizes and distributions. The impact of the deposition fraction uncertainty was propagated to committed effective dose coefficients for selected radionuclides in a companion publication. For each particle size, a lognormal distribution, characterized by its geometric mean as defined within ICRP Publication 66, serves as the basis for introducing uncertainty into the physical processes governing deposition in various lung regions. Finally, this study addresses the deposition process and examines how uncertainty in deposition mechanisms affects activity distribution in the airways, ultimately presenting the expected range and standard deviation of deposited activity as a function of particle size.

International Commission on Radiological Protectio↗

Stochastic Modeling of the Joint Neutron Number-Cumulative Fission Fragment Kinetic Energy Deposition Distribution and its Statistical Moments [Slides]

We investigate the joint distribution of the neutron number and cumulative fission-fragment kinetic energy (FKE) deposition, with a specific focus on low-order statistical moments: the mean, variance, and correlation. Starting from a point-kinetic framework, we derive a forward Master equation (FME) for the joint distribution and develop the corresponding moment equations.

42 ENGINEERING↗

Analyzing historical snow trends in interior Alaska

Study region The Chena River watershed in Interior Alaska, USA Study focus This study examines 40 years (water years 1982–2021) of snowpack characteristics to consider its hydrological implications in the 5350 km² Chena River basin. Using observations and a fine-scale physics model, we analyzed trends of snow water equivalent (SWE), snow onset and disappearance, and snow cover duration (SCD). New hydrological insights for the region Results indicate a decline in SWE across the modeled domain, averaging a decrease of 3 mm per decade, with larger decreases (up to 10 mm per decade) at lower elevations. While domain-averaged SWE trends were not statistically significant, observed SCD showed statistically significant decreases: −5.2, −5.0, and −4.4 days per decade at Teuchet Creek, Fairbanks F.O., and Little Chena Ridge, respectively. Notably, observations at SNOTEL stations and modeling revealed no statistically significant change in domain-averaged Rain-on-Snow (ROS) events over the 40-year period, contrasting some regional future estimates of increased ROS frequency. Peak streamflow did not consistently correlate with peak SWE levels, suggesting that other environmental factors such as ROS events and rapid temperature increases (e.g., a 10°C spike observed in 1992) are key drivers of hydrological outcomes. These findings improve understanding of complex subarctic hydrological processes impacting permafrost and highlight the need for adaptive water resource management to mitigate multi-factor risks like flooding and wildfire, requiring proactive planning.

54 ENVIRONMENTAL SCIENCES↗

ELM2.1-XGBfire1.0: improving wildfire prediction by integrating a machine learning fire model in a land surface model

Wildfires have shown increasing trends in both frequency and severity across the contiguous United States (CONUS). However, process-based fire models have difficulties in accurately simulating the burned area over the CONUS due to a simplification of the physical process and cannot capture the interplay among fire, ignition, climate, and human activities. The deficiency of burned area simulation deteriorates the description of fire impact on energy balance, water budget, and carbon fluxes in the Earth system models (ESMs). Alternatively, fire models based on machine learning (ML), which capture statistical relationships between the burned area and environmental factors, have shown promising burned area predictions and corresponding fire impact simulation. We develop a hybrid framework (ELM2.1-XGBFire1.0) that integrates an eXtreme Gradient Boosting (XGBoost) wildfire model with the Energy Exascale Earth System Model (E3SM) land model (ELM) version 2.1. A Fortran–C–Python deep learning bridge is adapted to support online communication between ELM and the ML fire model. Specifically, the burned area predicted by the ML-based wildfire model is directly passed to ELM to adjust the carbon pool and vegetation dynamics after disturbance, which are then used as predictors in the ML-based fire model in the next time step. Evaluated against the historical burned area from Global Fire Emissions Database 5 from 2001–2019, the ELM2.1-XGBFire1.0 outperforms process-based fire models in terms of spatial distribution and seasonal variations. The ELM2.1-XGBFire1.0 has proven to be a new tool for studying vegetation–fire interactions and, more importantly, enables seamless exploration of climate–fire feedback, working as an active component of E3SM.

54 ENVIRONMENTAL SCIENCES↗

Topology-Dependent Performance of Free-Space Photonic Quantum Networks Under Noise

Photonic quantum communication enables secure and high-fidelity information transfer beyond classical limits, with direct relevance to emerging quantum networks operating in free-space environments. While physical-layer models of depolarizing noise, Gamma–Gamma turbulence statistics, entanglement swapping, and decoy-state QKD security bounds are individually well established, prior work typically treats these components in isolation or under fixed network assumptions. In this work, we develop a unified topology-aware analytical framework that simultaneously integrates free-space optical link budgets, turbulence-induced visibility degradation, depolarizing qubit noise, multi-hop entanglement cascade dynamics, teleportation fidelity thresholds, CHSH nonlocality certification, and asymptotic decoy-state secret key rate bounds across star, mesh, and ring graph structures. Rather than introducing new physical channel models, we demonstrate that identical physical links exhibit fundamentally different end-to-end performance once embedded within different network topologies. Mesh architectures minimize visibility cascade through hop-count reduction but incur quadratic hardware scaling. Star topologies minimize link count but concentrate noise and synchronization overhead at the hub. Ring configurations offer linear hardware scaling with multiplicative fidelity degradation. The results establish topology as a first-order design parameter in near-term free-space quantum networks operating without full quantum repeater infrastructures. While motivated by distributed multi-agent architectures, the framework applies broadly to terrestrial, airborne, and satellite-assisted photonic quantum communication systems.

QKD↗

Bubble Point Measurements of cis-1,1,1,4,4,4-Hexafluorobutene [R-1336mzz(Z)] + trans-1,2-Dichloroethene [R-1130(E)] mixtures

Saturation pressures of pure R-1336mzz(Z) and R-1130(E) and bubble point pressures of three R-1336mzz(Z)/1130(E) blends were measured from 265 K to 360 K. For each pure refrigerant or refrigerant blend, a total of twenty unique saturation pressures or bubble points were measured. In total 100 unique state points were obtained. Presently, no Helmholtz-energy-explicit type equation of state (EoS) is available for R-1130(E). While an extended corresponding states EoS for R-1130(E) is available to estimate the properties of the R-1336mzz(Z)/1130(E) blend, this model does not resolve the azeotropic behavior of the mixture. Therefore, the perturbed-chain statistical-associating fluid theory (PC-SAFT) EoS is used to model the vapor–liquid equilibria of the R-1336mzz(Z)/1130(E) blend. PC-SAFT model parameters are reported, and the overall performance of the model is characterized by deviations from the experimental data.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Quantitative radiography for determining density fluctuations in HED experiments

We have developed a method to extract density fluctuation measurements from x-ray radiographs of high-energy density (HED) instability growth and turbulence experiments. We use this information to calculate density fluctuation statistics for constraining the performance of turbulent mix models in HED systems. The density calculation combines image filtering, removal of systemic effects such as backlighter variation, calculation of transmission across multiple materials, and use of tracer materials to generate an approximate single-material density field. From the density map, we calculate both average density and a variance-like moment b (density-specific-volume covariance), which we compare to our models. We infer both quantities from a single image, which is significantly more information than the historic single scalar mix width measurements. We also develop a method of analyzing simulation outputs that incorporate both the density fluctuation metric from a turbulence model and the bulk material maps from the hydrodynamic code. This analysis helps address the question of how to initialize the simulations for best comparison to data from systems with large separations of scale in the mixing perturbation initial condition. We find that our data analysis method yields 1D average density and b curves with similar morphology and amplitudes as those from preliminary simulation comparisons.

47 OTHER INSTRUMENTATION↗

Cosmological constraints from density-split clustering in the BOSS CMASS galaxy sample

We present a clustering analysis of the BOSS DR12 CMASS galaxy sample, combining measurements of the galaxy two-point correlation function and density-split clustering down to a scale of $1 \, h^{-1}\, \text{Mpc}$. Our theoretical framework is based on emulators trained on high-fidelity mock galaxy catalogues that forward model the cosmological dependence of the clustering statistics within an extended-ΛCDM framework, including redshift-space and Alcock–Paczynski distortions. Our base-ΛCDM analysis finds ω cdm = 0.1201 ± 0.0022, σ 8 = 0.792 ± 0.034, and n s = 0.970 ± 0.018, corresponding to fσ 8 = 0.462 ± 0.020 at z ≈ 0.525, which is in agreement with Planck 2018 predictions and various clustering studies in the literature. We test single-parameter extensions to base-ΛCDM, varying the running of the spectral index, the dark energy equation of state, and the density of mass-less relic neutrinos, finding no compelling evidence for deviations from the base model. We model the galaxy–halo connection using a halo occupation distribution framework, finding signatures of environment-based assembly bias in the data. We validate our pipeline against mock catalogues that match the clustering and selection properties of CMASS, showing that we can recover unbiased cosmological constraints even with a volume 84 times larger than the one used in this study.

79 ASTRONOMY AND ASTROPHYSICS↗

Dark energy survey: Modeling strategy for multiprobe cluster cosmology and validation for the full six-year dataset

Here, we introduce an updated To&Krause2021 model for joint analyses of cluster abundances and large-scale two-point correlations of weak lensing and galaxy and cluster clustering (termed CL+3×2 pt analysis) and validate that this model meets the systematic accuracy requirements of analyses with the statistical precision of the final Dark Energy Survey (DES) Year 6 (Y6) dataset. The validation program consists of two distinct approaches, (i) identification of modeling and parametrization choices and impact studies using simulated analyses with each possible model misspecification and (ii) end-to-end validation using mock catalogs from customized Cardinal simulations that incorporate realistic galaxy populations and DES-Y6-specific galaxy and cluster selection and photometric redshift modeling, which are the key observational systematics. In combination, these validation tests indicate that the model presented here meets the accuracy requirements of DES-Y6 for CL+3×2 pt based on a large list of tests for known systematics. In addition, we also validate that the model is sufficient for several other data combinations: the CL+GC subset of this data vector (excluding galaxy–galaxy lensing and cosmic shear two-point statistics) and the CL+3×2 pt+BAO+SN (combination of CL+3×2 pt with the previously published Y6 DES baryonic acoustic oscillation and Y5 supernovae data).

79 ASTRONOMY AND ASTROPHYSICS↗

Dark Energy Survey: Modeling strategy for multiprobe cluster cosmology and validation for the Full Six-year Dataset

We introduce an updated To&Krause2021 model for joint analyses of cluster abundances and large-scale two-point correlations of weak lensing and galaxy and cluster clustering (termed CL+3x2pt analysis) and validate that this model meets the systematic accuracy requirements of analyses with the statistical precision of the final Dark Energy Survey (DES) Year 6 (Y6) dataset. The validation program consists of two distinct approaches, (1) identification of modeling and parameterization choices and impact studies using simulated analyses with each possible model misspecification (2) end-to-end validation using mock catalogs from customized Cardinal simulations that incorporate realistic galaxy populations and DES-Y6-specific galaxy and cluster selection and photometric redshift modeling, which are the key observational systematics. In combination, these validation tests indicate that the model presented here meets the accuracy requirements of DES-Y6 for CL+3x2pt based on a large list of tests for known systematics. In addition, we also validate that the model is sufficient for several other data combinations: the CL+GC subset of this data vector (excluding galaxy--galaxy lensing and cosmic shear two-point statistics) and the CL+3x2pt+BAO+SN (combination of CL+3x2pt with the previously published Y6 DES baryonic acoustic oscillation and Y5 supernovae data).

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

A Parameter-masked Mock Data Challenge for Beyond-two-point Galaxy Clustering Statistics

The past few years have seen the emergence of a wide array of novel techniques for analyzing high-precision data from upcoming galaxy surveys, which aim to extend the statistical analysis of galaxy clustering data beyond the linear regime and the canonical two-point (2pt) statistics. We test and benchmark some of these new techniques in a community data challenge named “Beyond-2pt,” initiated during the Aspen 2022 Summer Program “Large-Scale Structure Cosmology beyond 2-Point Statistics,” whose first round of results we present here. The challenge data set consists of high-precision mock galaxy catalogs for clustering in real space, in redshift space, and on a light cone. Participants in the challenge have developed end-to-end pipelines to analyze mock catalogs and extract unknown (“masked”) cosmological parameters of the underlying ΛCDM models with their methods. The methods represented are density-split clustering, nearest neighbor statistics, BACCO power spectrum emulator, void statistics, LEFTfield field-level inference using effective field theory (EFT), and joint power spectrum and bispectrum analyses using both EFT and simulation-based inference. In this work, we review the results of the challenge, focusing on problems solved, lessons learned, and future research needed to perfect the emerging beyond-2pt approaches. The unbiased parameter recovery demonstrated in this challenge by multiple statistics and the associated modeling and inference frameworks supports the credibility of cosmology constraints from these methods. The challenge data set is publicly available, and we welcome future submissions from methods that are not yet represented.

Krause, Elisabeth [Univ. of Arizona, Tucson, AZ (U↗

Applications of emulation and Bayesian methods in heavy-ion physics

Abstract Heavy-ion collisions provide a window into the properties of many-body systems of deconfined quarks and gluons. Understanding the collective properties of quarks and gluons is possible by comparing models of heavy-ion collisions to measurements of the distribution of particles produced at the end of the collisions. These model-to-data comparisons are extremely challenging, however, because of the complexity of the models, the large amount of experimental data, and their uncertainties. Bayesian inference provides a rigorous statistical framework to constrain the properties of nuclear matter by systematically comparing models and measurements. This review covers model emulation and Bayesian methods as applied to model-to-data comparisons in heavy-ion collisions. Replacing the model outputs (observables) with Gaussian process emulators is key to the Bayesian approach currently used in the field, and both current uses of emulators and related recent developments are reviewed. The general principles of Bayesian inference are then discussed along with other Bayesian methods, followed by a systematic comparison of seven recent Bayesian analyses that studied quark-gluon plasma properties, such as the shear and bulk viscosities. The latter comparison is used to illustrate sources of differences in analyses, and what it can teach us for future studies.

Paquet, Jean-François (ORCID:0000000187368171)↗

Stochastic fracture generation and thermo-hydro-mechanical modeling in an equivalent continuum framework for enhanced geothermal systems

Enhanced geothermal systems (EGS) involve fracturing low permeability material to establish well connectivity and then injecting and circulating fluid into the fractured subsurface for geothermal power production. Changes in fracture aperture from contraction of the cooling matrix rock may alter network connectivity and risk thermal short-circuiting. Thermo-hydro-mechanical (THM) models are a useful tool to study these processes. However, as fracture networks are complex, and data may be limited, fracture networks in THM models are often stochastically generated. Given reliance on stochastic fracture networks and THM modeling to represent the subsurface and assess productivity of EGS, increased understanding of the influence of such statistically derived fracture networks on flow and heat transport in THM models is needed. Here, a new fracture process model is developed in the reactive transport code PFLOTRAN to stochastically generate fracture families and simulate changes in fracture aperture over time due to temperature changes of the rock matrix. Sixty-four different fracture networks ranging from well to poorly-connected, are modeled in PFLOTRAN with and without mechanical processes (THM vs TH). Results indicate that for well-connected fracture networks, thermal short-circuiting is less of a concern due to the abundance of available alternative flowpaths. For poorly-connected fracture networks, inclusion of mechanical processes showed steep thermal drawdown coincident with increase in fracture aperture along developing colder flowpaths, demonstrating the risk of thermal short-circuiting. Simulations with additional, larger fractures engineered to establish connectivity in a poorly-fractured subsurface, indicate that while stochastic variation of fracture orientation of the background network had limited influence, such variation in the engineered fractures significantly affected flow and heat transport.

Discrete fracture networks (DFN)↗