Search NASA⌕ Search

SEARCH · Search NASA

Results for “statistical model”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 289 records · Page 16

Pandora-based muon neutrino disappearance search in the SBN program with 1muonNp quasi-elastic like channel

The Short-Baseline Neutrino (SBN) program at Fermilab aims to perform a definitive search for light sterile neutrinos using multiple liquid argon time projection chamber detectors. We present the status of an analysis of muon neutrino charged-current interactions in two SBN detectors (SBND and ICARUS), selecting fully contained events with one muon and at least one proton in the final state, inspired by the recent ICARUS standalone results. This topology-driven selection reduces dependence on neutrino interaction cross-section modeling while preserving high statistics and good neutrino energy resolution, enabled by robust reconstruction and particle identification. Fully contained events enable precise kinematic reconstruction and support relative measurements between detectors, directly addressing the core goals of the SBN program. A detailed evaluation of systematic uncertainties is currently underway, with the objective of reducing the dominant systematics to the percent level. The current status of this analysis, with a particular focus on the event selection performance, will be presented in this poster.

Artero Pons, Maria [Padua U.; INFN, Padua]↗

From Ensemble Climate to Ensemble Impacts

Many climate-risk tools rely on ensemble mean projections or endpoint climate snapshots to characterize future hazards. Although convenient for communication, these representations remove the statistical, temporal, and physical information that real infrastructure systems respond to. Infrastructure degradation and failure arise from extremes, sequences, cumulative stress, compound hazards, and nonlinear fragility relationships, none of which survive ensemble averaging or temporal compression. Power-system failure statistics and cascading failure models further show that infrastructure risk is dominated by tail events and path-dependent dynamics rather than by mean conditions. This paper demonstrates why ensemble mean or endpoint-only climate representations are mathematically and physically inconsistent with engineering-grade risk analysis. We outline a model-resolved, time-series-based workflow that preserves extremes, variability, and sequencing by propagating each climate-model realization independently through hazard formation, exposure, fragility, and cascading failure mechanisms. Taking the ensemble of impacts—rather than the ensemble of climate—provides a defensible, physically coherent foundation for infrastructure resilience planning, regulatory compliance, and long-term investment decisions.

54 - ENVIRONMENTAL SCIENCES/GLOBAL CLIMATE CHANGE ↗

Data from: Understanding the biogeochemical and spatial drivers of methane and carbon dioxide fluxes in a large temperate reservoir

This dataset contains spatially resolved measurements of CO₂ and CH₄ fluxes and associated environmental variables collected across 200 sites in Douglas Reservoir (Tennessee, USA) between July 29-August 2, 2024. Measurements include diffusive fluxes of CO₂ and CH₄, CH₄ ebullition, and biogeochemical and spatial variables such as dissolved oxygen, temperature, conductivity, chlorophyll-a, pH, water depth, and distance from the dam. Sampling was conducted using a spatially balanced design to capture longitudinal and depth-related gradients throughout the reservoir. The dataset is structured to support analyses of spatial variability, flux pathway comparisons, and modeling approaches (e.g., spatial statistics and machine learning) aimed at understanding controls on reservoir CO₂ and CH₄ fluxes and improving upscaling to whole-reservoir and regional estimates.

Neeper, Jamie [ORNL] (ORCID:0009000842342101)↗

Characterizing Turbulence at a Forest Edge: Comparing Sub-Filter Scale Turbulence Models in Simulations of Flow over a Canopy

In wildfires, atmospheric turbulence plays a major role in the transfer of turbulent kinetic energy. Understanding how turbulence feeds back into a dynamical system is important, down to the varying small scales of fuel structures (i.e. pine needles, grass). Large eddy simulations (LES) are a common way of numerically representing turbulence. The Smagorinsky model (1963) serves as one of the most studied sub-grid scale representations in LES. In this investigation, the Smagorinsky model was implemented in HIGRAD/FIRETEC, LANL’s coupled fire-atmosphere model. This study was motivated by the need to quantitatively investigate the vorticity budget equation in HIGRAD/FIRETEC. The Smagorinsky turbulent kinetic energy (TKE) was compared to FIRETEC’s 1.5-order TKE eddy-viscosity subgrid-scale model, known as the Linn turbulence model. This was done in simulations of flow over flat terrain with a homogeneous, cuboidal canopy in the center of the domain. Examinations of the modeled vertical TKE profile and turbulent statistics at the leading edge, and throughout the canopy, show that the Smagorinsky model provides comparable results to that of the original closure model posed in FIRETEC.

58 GEOSCIENCES↗

Data and scripts associated with “Moisture content modulates DOM thermodynamic regulation of oxygen consumption in drying streambed sediments”

This data package is associated with the publication “Moisture content modulates DOM thermodynamic regulation of oxygen consumption in drying streambed sediments” published in Scientific Reports (Garayburu-Caruso et al., 2026). The package contains processed data products and scripts used to quantify how drying and re-inundation of riverbed sediments influence dissolved organic matter (DOM) thermodynamic properties and their relationship with sediment oxygen (O₂) consumption across 33 stream sites in the contiguous United States. The data package contains DOM thermodynamic metrics (e.g., Gibbs free energy of carbon oxidation and thermodynamic efficiency), and O₂ consumption along with watershed-scale climate and land-cover metrics used as explanatory variables in the analyses. Underlying unprocessed and processed ultrahigh-resolution mass spectrometry data, oxygen consumption rates from laboratory moisture-manipulation experiments, within-sample environmental properties, sediment moisture content and contextual field measurements are archived separately at https://data.ess-dive.lbl.gov/datasets/doi:10.15485/2428003 (Laan et al., 2024) and https://data.ess-dive.lbl.gov/datasets/doi:10.15485/1923689 (Forbes et al.,2023). A preliminary version of this data package was published in February 2026 at the time of manuscript submission. It was updated in June 2026, at the time of manuscript acceptance, to include the finalized data and additional metadata (readme, data dictionary, and file level metadata). For details on how to navigate data packages generated by this project, see https://data.ess-dive.lbl.gov/portals/PNNLRiverCorridorSFA/About. In addition to a readme, this data package also includes a file-level metadata (FLMD) file that describes each file and a data dictionary (DD) that describes all column/row headers and variable definitions. At the top level, the data package is organized into five main folders: (1) Data, (2)Figures, (3) Map, (4) GAM_Reulsts, and (5) src. The Data folder contains analysis-ready tabular files with oxygen consumption rates, DOM thermodynamic properties by site and treatment, site-level environmental variables, watershed-scale metrics, and other derived variables referenced in the manuscript. The Figures folder contains static image files associated with the main text and supplemental figures, while the Map folder includes spatial data and map-layer files used to create the sampling-location map. The GAM results folder contains the results for each of the general additive model (GAM).The src folder contains R scripts used to perform data processing, statistical analyses (including clustering, generalized additive models, and threshold analysis), and figure generation. This data package is associated with a GitHub repository found at https://github.com/WHONDRS-Hub/ECA_DOM_Thermodynamics.

Dissolved organic matter↗

Cosmological limits on the neutrino mass sum for beyond-Λ⁢ CDM models

The sum of neutrino masses can be measured cosmologically, as the sub-eV particles behave as “hot” dark matter whose main effect is to suppress the clustering of matter compared to a universe with the same amount of purely cold dark matter. Current astronomical data provide an upper limit on ∑𝑚 𝜈 between 0.07–0.12 eV at 95% confidence, depending on the choice of data. This bound assumes that the cosmological model is Λ Cold Dark Matter (Λ⁢ CDM), where dark energy is a cosmological constant, the spatial geometry is flat, and the primordial fluctuations follow a pure power law. Here, we update studies on how the mass limit degrades if we relax these assumptions. To existing data from the Planck satellite we add new gravitational lensing data from the Atacama Cosmology Telescope, the new Type Ia supernova sample from the Pantheon+survey, and baryonic acoustic oscillation (BAO) measurements from the Sloan Digital Sky Survey and the Dark Energy Spectroscopic Instrument. Using our fiducial data combination, described in the appendix, we find the neutrino mass limit is stable to most model extensions, with such extensions degrading the limit by less than 10%. We find a broadest bound of ∑𝑚 𝜈 < 0.19 eV at 95% confidence for a model with dynamical dark energy, although this scenario is not statistically preferred over the simpler Λ ⁢CDM model.

79 ASTRONOMY AND ASTROPHYSICS↗

Cosmology with second- and third-order shear statistics for the Dark Energy Survey: Methods and simulated analysis

We present a new pipeline designed for the robust inference of cosmological parameters using both second- and third-order shear statistics. We build a theoretical model for rapid evaluation of three-point correlations using our fastnc code and integrate it into the cosmosis framework. We measure the two-point functions 𝜉 ± and the full configuration-dependent three-point shear correlation functions across all auto- and cross-redshift bins. We compress the three-point functions into the mass aperture statistic ⟨ℳ$^{3}_{ap}$⟩ for a set of 796 simulated shear maps designed to model the Dark Energy Survey Year 3 data. We estimate from it the full covariance matrix and model the effects of intrinsic alignments, shear calibration biases and photometric redshift uncertainties. We apply scale cuts to minimize the contamination from the baryonic signal as modeled through hydrodynamical simulations. We find a significant improvement of 83% on the figure of merit in the Ω m − 𝑆 8 plane when we add the ⟨ℳ$^{3}_{ap}$⟩ data to 𝜉 ± . Here, we present our findings for all relevant cosmological and systematic uncertainty parameters and discuss the complementarity of third-order and second-order statistics.

79 ASTRONOMY AND ASTROPHYSICS↗

Combining physics-based and data-driven models for quantitatively accurate plasma profile prediction that extrapolates well; with application to DIII-D, AUG, and ITER tokamaks

For design, scenario planning, and control, ITER and all other envisioned tokamaks rely on a variety of statistical and physics-based models to extrapolate to unseen regimes; most notably from low plasma current to high. A 'meta-learning' methodology for combining the accuracy of data-driven models with the generalizability of physics-based models is described and tested, yielding a 5–10 percent improvement in performance beyond either alone for the task of extrapolating time-dependent plasma profile prediction from low- to high- plasma current DIII-D tokamak discharges. Meanwhile, it is shown that both machine learning models extrapolated far-distribution and state-of-the-art 'physics-based' profile predictors fare worse than merely assuming plasma profiles do not change from their initial values. Finally, a variety of other mechanisms for helping data-driven models generalize—transfer learning, adding contextual information from physics simulators, and adding data from the ASDEX Upgrade tokamak—are attempted for similar extrapolation tasks but, in the methodology used in this paper, yield no significant improvement beyond simple data-driven models. Results are summarized in figures 15 and 16.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Information theory optimization of signals from small-angle scattering measurements

Small-angle X-ray scattering (SAXS) of particles in solution informs on the conformational states and assemblies of biological macromolecules (bioSAXS) outside of cryo- and solid-state conditions. In bioSAXS, the SAXS measurement under dilute conditions is resolution limited, and through an inverse Fourier transform, the measured SAXS intensities directly relate to the physical space occupied by the particles via the P (r)-distribution. Yet, this inverse transform of SAXS data has been historically cast as an ill-posed, ill-conditioned problem requiring an indirect approach. Here, we show that through the applications of matrix and information theories, the inverse transform of SAXS intensity data is a well-conditioned problem. The so-called ill-conditioning of the inverse problem is directly related to the Shannon number. By exploiting the oversampling enabled by modern detectors, a direct inverse Fourier transform of the SAXS data is possible, provided the recovered information does not exceed the Shannon number. The Shannon limit corresponds to the maximum number of significant singular values that can be recovered in a SAXS experiment, suggesting this relationship is a fundamental property of band-limited inverse integral transform problems. This correspondence reduces the complexity of the inverse problem to the Shannon limit and maximum dimension. We propose a hybrid scoring function using an information theory framework that assesses both the quality of the model-data fit as well as the quality of the recovered P (r)-distribution. The hybrid score utilizes the Akaike information criteria and Durbin-Watson statistic that considers parameter-model complexity, i.e., degrees of freedom, and the randomness of the model-data residuals. The described tests and findings extend the boundaries for bioSAXS by completing the information theory formalism initiated by Peter B. Moore to enable a quantitative measure of resolution in SAXS, robustly determine maximum dimension, and more precisely define the best parameter model appropriately representing the observed scattering data.

Rambo, Robert P. [Science and Technology Facilitie↗

Search for a heavy neutral lepton with the MAGNETO-𝜈 experiment using 241 Pu 𝛽 − decays

The MAGNETO-𝜈 experiment searches for keV-scale heavy neutral leptons (HNLs) through precise measurements of the 𝛽 − -decay spectrum of 241 Pu. We present spectra comprising a total of 194 million 𝛽 − decays recorded using decay energy spectrometry with metallic magnetic calorimeters, representing the most statistically precise measurement of 241 Pu 𝛽 − decay to date. The 𝛽-endpoint energy was determined using 𝛾 rays and x-rays from an external 133 Ba calibration source, yielding 𝑄 𝛽 = 22.273⁢ (33) ⁢keV. The measured spectrum shows no statistically significant deviation from the allowed 𝛽-decay model. From a subset of the high-statistics data, we set an upper limit on the mixing of an 11.5-keV HNL with the electron neutrino, |𝑈 𝑒⁢4 | 2 < 1.31 × 10 −3 at the 95% confidence level.

A ≥ 220↗

Predicting U.S. federal fleet electric vehicle charging patterns using internal combustion engine vehicle fueling transaction statistics

Utilizing fueling transactions from internal combustion engine vehicles (ICEVs), the authors estimated how frequently midday public charging would be required for U.S. federal fleet battery electric vehicles (BEVs). Fueling transaction summary statistics are more widely available than trip-level telematics data, making this methodology more accessible and transferable to other researchers and fleet managers considering BEV replacements. For example, readers can easily apply a linear model using only the count of back-to-back fueling events at gas stations over 57 straight-line miles apart to predict days exceeding range. This linear regression predicted binned days exceeding 250 miles at 80% accuracy on a hold-out test set from the same fleet as the training data and 66 % accuracy on a new fleet displaying different driving behaviors. The authors additionally provide linear equations for days exceeding 200 and 300 miles as alternative range estimates to account for differences in BEV range and temperature impacts. Beyond the single-feature linear models which readers can apply, the authors tuned and trained other machine learning models on a variety of fueling transaction statistics including consecutive transaction distances, transaction distance from garage, estimated miles traveled from fuel economy and fuel quantity, and transaction periodicity. Utilizing a subset of 1678 light-duty federal fleet vehicles which contained daily vehicle miles traveled (VMT) in addition to fueling statistics, the authors determined which fueling transaction statistics were most relevant in predicting driving days exceeding 250 miles (an approximation of BEV rated driving range). In support of the U.S. federal fleet transition to zero-emission vehicles (ZEVs), the authors used these statistics and machine learning models to predict the frequency of BEV midday charging. After training models on the subset with VMT, the authors predicted days exceeding rated range for 112,902 light-duty vehicles operating in similar circumstances in the federal fleet using a Support Vector Regressor (SVR). In conclusion, they then used the projections as part of the ZEV Planning and Charging (ZPAC) tool to identify optimal candidates for BEVs for the federal fleet. An anonymized version of ZPAC is included in the supplementary materials.

25 ENERGY STORAGE↗

The DESI-Lensing Mock Challenge: large-scale cosmological analysis of 3x2-pt statistics

The current generation of large galaxy surveys will test the cosmological model by combining multiple types of observational probes. Realising the statistical promise of these new datasets requires rigorous attention to all aspects of analysis including cosmological measurements, modelling, covariance and parameter likelihood. In this paper we present the results of an end-to-end simulation study designed to test the analysis pipeline for the combination of the Dark Energy Spectroscopic Instrument (DESI) Year 1 galaxy redshift dataset and separate weak gravitational lensing information from the Kilo-Degree Survey, Dark Energy Survey and Hyper-Suprime-Cam Survey. Our analysis employs the 3x2-pt correlation functions including cosmic shear and galaxy-galaxy lensing, together with the projected correlation function of the spectroscopic DESI lenses. We build realistic simulations of these datasets including galaxy halo occupation distributions, photometric redshift errors, weights, multiplicative shear calibration biases and magnification. We calculate the analytical covariance of these correlation functions including the Gaussian, noise and super-sample contributions, and show that our covariance determination agrees with estimates based on the ensemble of simulations. We use a Bayesian inference platform to demonstrate that we can recover the fiducial cosmological parameters of the simulation within the statistical error margin of the experiment, investigating the sensitivity to scale cuts. This study is the first in a sequence of papers in which we present and validate the large-scale 3x2-pt cosmological analysis of DESI-Y1.

79 ASTRONOMY AND ASTROPHYSICS↗

Enhanced climate reproducibility testing with false discovery rate correction

Simulating the Earth's climate is an important and complex problem, thus climate models are similarly complex, comprised of millions of lines of code. In order to appropriately utilize the latest computational and software infrastructure advancements in Earth system models running on modern hybrid computing architectures to improve their performance, precision, accuracy, or all three; it is important to ensure that model simulations are repeatable and robust. This introduces the need for establishing statistical or non-bit-for-bit reproducibility, since bit-for-bit reproducibility may not always be achievable. Here, we propose a short-simulation ensemble-based test for an atmosphere model to evaluate the null hypothesis that modified model results are statistically equivalent to that of the original model. We implement this test in version 2 of the US Department of Energy's Energy Exascale Earth System Model (E3SM). The test evaluates a standard set of output variables across the two simulation ensembles and uses a false discovery rate correction to account for multiple testing. The false positive rates of the test are examined using re-sampling techniques on large simulation ensembles and are found to be lower than the currently implemented bootstrapping-based testing approach in E3SM. We also evaluate the statistical power of the test using perturbed simulation ensemble suites, each with a progressively larger magnitude of change to a tuning parameter. The new test is generally found to exhibit more statistical power than the current approach, being able to detect smaller changes in parameter values with higher confidence.

Kelleher, Michael E. [Oak Ridge National Laborator↗

Performance Evaluation of Weather@home2 Simulations over West African Region

Weather and climate forecasting, using climate models, have become essential tools and life-savers in the West African region; in spite of the fact that climate models do not fully comply with attributes of forecast qualities—RASAP: reliability, association, skill, accuracy, and precision. The objective of this paper is to quantitatively evaluate, in comparison to CRU and ERA5 datasets, the RASAP compliance-level of the weather@home2 modeling system (w@h2). Findings from some statistical evaluations show that, to a moderately significant extent, w@h2 model provides useful information during the monsoon seasons; skills to capture the Little Dry Season over the Guinea zone; predictive skills for the onset season; ability to reproduce all the annual characteristics of the surface maximum air temperature over the region; as well as skill to detect heat waves that usually ravage West Africa during the boreal spring. The model displays traces of attributes that are needed for seasonal climate predictions and applications. Deficiencies in the quantitative reproducibility point to the facts that the model does provide a reliability akin to that of regional climate models. This paper further furnishes a prospective user with information on whether the model might be “useful or not” for a particular application.

West Africa↗

Hypertriton puzzle in relativistic heavy-ion collisions

The yields of hadrons and light nuclei in relativistic collisions of heavy-nuclei at a center of mass energy of $\sqrt{s_{NN}}$ = 2.6 TeV can be described remarkably well by a thermal distribution of an ideal gas of hadrons and light nuclei interacting only via the decay of resonances. Given the particularly small binding energy of hypertritons relative to the temperature describing the yields (about 156 MeV), one might naturally expect hypertrions to dissociate in medium, making the agreement of hypertriton yields with thermal predictions highly puzzling. The puzzle is compounded by the fact that small binding energy is associated with the large size of the hypertriton. This size is on a similar scale to the overall size of the fireball and much larger than the length scale over which temperatures in the fireball vary over phenomenologically relevant amounts. Here, this paper quantifies the tension this effect causes and shows that it is sufficiently large to render the thermal model inconsistent: its natural assumptions are in conflict with its outputs. The possibility that hypertritons are formed at freeze out as compact objects, quark droplets, that subsequently evolve into hypertritons is considered as a way to resolve the puzzle. It is noted that beyond making the assumption that compact quark droplets form, additional detailed dynamical assumptions which have not been justified are needed to make the thermal model work. The issue of why, despite these issues, the hypertriton is well described by a simple statistical description at freeze out is unresolved. Resolving the hypertriton puzzle is important as it may clarify whether the phenomenological success of the simple thermal model for yields accurately reflects the simple picture of the underlying physics on which it is based.

hydrodynamic models↗

Trust Model Measurements for the Energy Grid of Things

Information security is essential for the reliable operation of an Energy Grid of Things (EGoT). In addition to basic information security protocols as defined by published standards, there is a need for a monitoring function that measures the trustworthiness of the various actors participating in an EGoT. We describe in this paper the implementation and evaluation of a Distributed Trust Model that was developed specifically for monitoring communication within an EGoT. We then show how the model parameters are set using statistical measures for hypothesis testing.

Energy Grid of Things, EGoT, Smart Grid Security, ↗

Continuum contribution to charged-current absorption of low-energy $ν_e$ on $^{40}$Ar

Accurate modeling of the absorption of tens-of-MeV $ν_e$ on $^{40}$Ar is needed to enable measurements of astrophysical neutrinos using large liquid argon time projection chamber (LArTPC) detectors, such as those planned for the Deep Underground Neutrino Experiment (DUNE). We revisit the MARLEY neutrino interaction model used in present estimates of DUNE sensitivity to supernova and solar neutrino signals. Multiple theoretical refinements are pursued, especially in the unbound continuum region of nuclear excitation energy. Inclusive charged-current neutrino-argon cross sections are calculated using a hybrid strategy. Nuclear transitions to unbound states are treated using a Hartree-Fock Continuum Random Phase Approximation (HF-CRPA) model, including forbidden contributions. Allowed transitions to low-lying discrete levels are also included using indirect measurements and approximate corrections for the momentum transfer dependence. Exclusive predictions are obtained by coupling these calculations with a statistical nuclear de-excitation model. The impact on observables of interest for DUNE and similar experiments is examined in terms of both total and differential cross sections. Our refined calculations predict a lower allowed portion of the cross section relative to the prior MARLEY model. At neutrino energies appreciably below 100 MeV, the inclusion of forbidden transitions does not fully compensate for the loss of allowed strength. For a representative neutrino burst from a galactic core-collapse supernova, our results suggest that MARLEY 1.2.0 overestimates the event yield in a DUNE-like detector by approximately 20%. However, because this overestimation is more severe at backwards angles, use of the charged-current $ν_e$-$^{40}$Ar reaction for supernova pointing may be more feasible than previously expected.

Gardiner, Steven [Fermilab]↗

Out-of-Distribution Detection and Radiological Data Monitoring Using Statistical Process Control

Abstract Machine learning (ML) models often fail with data that deviates from their training distribution. This is a significant concern for ML-enabled devices as data drift may lead to unexpected performance. This work introduces a new framework for out of distribution (OOD) detection and data drift monitoring that combines ML and geometric methods with statistical process control (SPC). We investigated different design choices, including methods for extracting feature representations and drift quantification for OOD detection in individual images and as an approach for input data monitoring. We evaluated the framework for both identifying OOD images and demonstrating the ability to detect shifts in data streams over time. We demonstrated a proof-of-concept via the following tasks: 1) differentiating axial vs. non-axial CT images, 2) differentiating CXR vs. other radiographic imaging modalities, and 3) differentiating adult CXR vs. pediatric CXR. For the identification of individual OOD images, our framework achieved high sensitivity in detecting OOD inputs: 0.980 in CT, 0.984 in CXR, and 0.854 in pediatric CXR. Our framework is also adept at monitoring data streams and identifying the time a drift occurred. In our simulations tracking drift over time, it effectively detected a shift from CXR to non-CXR instantly, a transition from axial to non-axial CT within few days, and a drift from adult to pediatric CXRs within a day—all while maintaining a low false positive rate. Through additional experiments, we demonstrate the framework is modality-agnostic and independent from the underlying model structure, making it highly customizable for specific applications and broadly applicable across different imaging modalities and deployed ML models.

Zamzmi, Ghada↗