Search NASA⌕ Search

SEARCH · Search NASA

Results for “Statistical methods”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 199 records · Page 11

Field testing and validation of a low-cost MPC for demand flexibility for grid-interactive K-12 schools

K-12 school buildings account for the highest energy consumption within the public sector. Implementing advanced HVAC controls in grid-interactive K-12 schools could bring substantial economic advantages and grid flexibility. Our previous study demonstrated that a low-cost model predictive control (MPC) solution, which coordinates multiple packaged units, can enable demand flexibility without major hardware upgrades. However, a significant gap remains between academic pilots and market-ready scalable solutions. This paper extends the previous single-site pilot to a multi-site demonstration involving three school campuses (95 total units) through a commercial technology transfer process. Addressing the challenge of verifying performance with sparse field data, we present a new statistical approach using Bayesian methods to estimate the MPC’s effect on peak demand. Unlike traditional methods, this approach robustly quantifies uncertainty in non-normal, limited datasets. The results confirm the solution’s replicability, achieving a 21.6–38.9% reduction in HVAC peak demand (10.8–22.1% at the site-level) with > 98% probability across diverse locations. Finally, we document critical barriers to scaling software-as-a-service (SaaS) solutions–such as API instability and diverse legacy systems–and offer practical strategies to accelerate the commercial adoption of grid-interactive efficient buildings.

Ham, Sang Woo↗

Quantitative Infrared-to-Terahertz Nanospectroscopy of Semiconductors

Semiconductor technology now employs few-nanometer features, necessitating tools probing electronic properties on the same length scale. While the concentration of free charge carriers is routinely measured, the scattering rate remains challenging to access at the nanoscale. Here, we present ultrabroadband (5–50 THz) synchrotron infrared nanospectroscopy as a quantitative metrology tool for semiconductors. This technique can determine both the charge carrier concentration and scattering rate with percent-level accuracy, and it is inherently capable of ∼10 nm spatial resolution. We study silicon with different doping levels and confirm the method’s accuracy by statistical analysis and comparison with established far-field infrared spectroscopy. Near-field measurements systematically reveal charge-carrier concentrations ∼30% lower than far-field values, consistent with increased surface sensitivity and surface depletion. Our work establishes synchrotron infrared nanospectroscopy as a precise tool for quantitative nanoscale semiconductor characterization and paves the way toward all-optical characterization of surface depletion effects.

36 MATERIALS SCIENCE↗

Optimizing time integration for accurate recovery of shockwave interface location in radiography

We present simulations and experiments of time integrated radiographic imaging of a moving 1D shock wave front and a quantitative method for determining the statistical error in locating the shock front as a function of integration time and noise in the radiograph. We discuss the trade-off between increasing motion blur, which leads to decreased shock front location certainty, and increasing signal-to-noise, which leads to improved image quality with increasing integration time. We find an optimum integration time between a short integration time, where noise limits the error, and a long integration time, where motion blurring limits the error. This methodology can be used to tune experimental configurations to obtain the highest quality radiograph for a given experimental configuration.

Bremsstrahlung↗

Transverse Momentum Distributions from Lattice QCD without Wilson Lines

The transverse-momentum-dependent distributions (TMDs), which are defined by gauge-invariant 3D parton correlators with staple-shaped lightlike Wilson lines, can be calculated from quark and gluon correlators fixed in the Coulomb gauge on a Euclidean lattice. These quantities can be expressed gauge invariantly as the correlators of Coulomb-gauge-dressed fields, which reduce to the standard TMD correlators under principal-value prescription in the infinite boost limit. In the framework of large-momentum effective theory, a quasi-TMD defined from such correlators in a large-momentum hadron state can be matched to the TMD via a factorization formula, whose exact form is derived using soft collinear effective theory and verified at one-loop order. Compared to the currently used gauge-invariant correlators, this new method can substantially improve statistical precision and simplify renormalization for the time-reversal-even TMDs, which will greatly enhance the predicative power of lattice QCD in the nonperturbative region.

Effective field theory↗

Generalized Tensor-on-Tensor Regression (GToTR)

SAND2026-23069O Generalized Tensor-on-Tensor Regression (GToTR) is a Python-based tool for conducting generalized tensor-on-tensor regression. It provides Canonical Polyadic (CP)-based generalized tensor regression models, support for generalized linear model-like families and links, alternating-optimization model fitting methods, and a standard statistics software interface. The tool supports tensor-valued responses and covariates using the open-source Python Tensor Toolbox (pyttb) software package. Sandia National Laboratories is a multimission laboratory managed and operated by National Technology & Engineering Solutions of Sandia, LLC, a wholly owned subsidiary of Honeywell International Inc., for the U.S. Department of Energy’s National Nuclear Security Administration under contract DE-NA0003525.

Dunlavy, Daniel [Sandia National Lab. (SNL-CA), Li↗

Hazard Detection Detector Cards

This report presents a comprehensive summary of five advanced anomaly detection tools developed and deployed by Oak Ridge National Laboratory in support of the VA’s Health Information Technology modernization. These detectors—Order Path Tracker, Trend Watcher, Pain Pointer, Performance Monitor, and Patient Record Flag Detector—leverage statistical and machine learning methods to monitor workflow disruptions, detect anomalies in care sequences and volumes, identify bottlenecks, and track system-level performance metrics across VistA and Millennium systems. All detectors have been integrated into the Health Data Analytics Platform (HDAP), with most having completed deployment and testing using live data from targeted stations in cardiology and oncology domains. This work enhances VA’s capacity for proactive system surveillance, promotes patient safety, and informs data-driven operational improvements across the EHR ecosystem.

97 MATHEMATICS AND COMPUTING↗

Virtual refrigerant charge sensing algorithm for residential CO₂ heat pumps

Natural refrigerants are increasingly adopted in next-generation heat pump systems, among which CO₂ heat pumps have attracted significant attention. However, due to their high operating pressures, the leakage risk is higher, resulting in undercharge conditions and degraded heat pump performance. Thus, developing an accurate refrigerant charge level detection technique is necessary to guarantee safe and efficient operation. Although virtual refrigerant charge (VRC) level calculation algorithms for CO₂ heat pumps exist, they typically rely on empirically selected features without a systematic selection framework, leading to multicollinearity and potential overfitting, which limit their prediction accuracy and generalizability. To address these issues, this study proposes a VRC algorithm framework with a systematic feature selection method that identifies physically meaningful and statistically significant features, and is applied using a residential CO₂ heat pump as a case study. The method is extended from previous work on conventional refrigerants to account for charge behavior in CO₂ gas coolers. The selected features include gas cooler outlet density, evaporator pressure, and superheat temperature. The results demonstrate that the proposed feature selection method significantly improves prediction accuracy compared to existing VRC approaches. A relatively small training dataset (∼30 samples) is sufficient for feature identification and model development. The developed algorithm achieves less than 3% prediction error under both undercharge and overcharge conditions, representing reductions of 46.7% and 35.3% compared to two recent reference VRC algorithms for transcritical CO₂ heat pumps reported in the literature. The proposed algorithm and feature selection method enhance leakage detection capability, facilitate the deployment of CO₂ heat pump systems, and contribute to reduced energy waste and maintenance costs.

Guo, Fangzhou [Lawrence Berkeley National Laborato↗

A Causal Approach to Model Validation and Calibration

This poster presents a novel method for validation and verification that focuses on identifying causal relationships between data elements, moving beyond traditional statistical and machine learning approaches. These methods employ causal discovery techniques to reveal the underlying mechanisms of data generation. The research utilizes structural causal models and directed acyclic graphs to depict causal relationships. This approach assists in achieving alignment between simulation models and reality.

97 MATHEMATICS AND COMPUTING↗

Investigating event-shape methods in the search for the chiral magnetic effect in relativistic heavy ion collisions

The chiral magnetic effect (CME) is a phenomenon in which electric charge is separated by a strong magnetic field from local domains of chirality imbalance and parity violation in quantum chromodynamics. The CME-sensitive observable, the charge-dependent three-point azimuthal correlator Δ⁢𝛾 , is contaminated by a major physics background proportional to the particle's elliptic flow anisotropy 𝑣 2 . Event-shape engineering (ESE) binning events in dynamical fluctuations of 𝑣 2 and event-shape selection (ESS) binning events in statistical fluctuations of 𝑣 2 are two methods to search for the CME by projecting Δ⁢𝛾 to the measured anisotropy 𝑣 2 = 0 intercept. Here, we conduct a systematic study of these two methods using physics models as well as toy model simulations. It is observed that the ESE method fulfills the general premise of measuring the CME but is statistically hungry. It is found that the intercept from the ESS method depends on the details of the event content, such as the mixtures of background-contributing sources, because of statistical fluctuations of intertwining variables used in the method, and is thus not practically useful to measure the CME.

Relativistic heavy-ion collisions↗

Causal relationship between mitochondrial-associated proteins and cerebral aneurysms: a Mendelian randomization study

Background Cerebral aneurysm is a high-risk cerebrovascular disease with a poor prognosis, potentially linked to multiple factors. This study aims to explore the association between mitochondrial-associated proteins and the risk of cerebral aneurysms using Mendelian randomization (MR) methods. Methods We used GWAS summary statistics from the IEU Open GWAS project for mitochondrial-associated proteins and from the Finnish database for cerebral aneurysms (uIA, aSAH). The association between mitochondrial-associated exposures and cerebral aneurysms was evaluated using MR-Egger, weighted mode, IVW, simple mode and weighted median methods. Reverse MR assessed reverse causal relationship, while sensitivity analyses examined heterogeneity and pleiotropy in the instrumental variables. Significant causal relationship with cerebral aneurysms were confirmed using FDR correction. Results Through MR analysis, we identified six mitochondrial proteins associated with an increased risk of aSAH: AIF1 (OR: 1.394, 95% CI: 1.109–1.752, p = 0.0044), CCDC90B (OR: 1.318, 95% CI: 1.132–1.535, p = 0.0004), TIM14 (OR: 1.272, 95% CI: 1.041–1.553, p = 0.0186), NAGS (OR: 1.219, 95% CI: 1.008–1.475, p = 0.041), tRNA PusA (OR: 1.311, 95% CI: 1.096–1.569, p = 0.003), and MRM3 (OR: 1.097, 95% CI: 1.016–1.185, p = 0.0175). Among these, CCDC90B, tRNA PusA, and AIF1 demonstrated a significant causal relationship with an increased risk of aSAH (FDR q < 0.1). Three mitochondrial proteins were associated with an increased risk of uIA: CCDC90B (OR: 1.309, 95% CI: 1.05–1.632, p = 0.0165), tRNA PusA (OR: 1.306, 95% CI: 1.007–1.694, p = 0.0438), and MRM3 (OR: 1.13, 95% CI: 1.012–1.263, p = 0.0303). In the reverse MR study, only one mitochondrial protein, TIM14 (OR: 1.087, 95% CI: 1.004–1.177, p = 0.04), showed a causal relationship with aSAH. Sensitivity analysis did not reveal heterogeneity or pleiotropy. The results suggest that CCDC90B, tRNA PusA, and MRM3 may be common risk factors for cerebral aneurysms (ruptured and unruptured), while AIF1 and NAGS are specifically associated with an increased risk of aSAH, unrelated to uIA. TIM14 may interact with aSAH. Conclusion Our findings confirm a causal relationship between mitochondrial-associated proteins and cerebral aneurysms, offering new insights for future research into the pathogenesis and treatment of this condition.

Wang, Shuai↗

Model orthogonalization and Bayesian forecast mixing via principal component analysis

One can improve predictability in the unknown domain by combining forecasts of imperfect complex computational models using a Bayesian statistical machine learning framework. In many cases, however, the models used in the mixing process are similar. In addition to contaminating the model space, the existence of such similar, or even redundant, models during the multimodeling process can result in misinterpretation of results and deterioration of predictive performance. In this paper we describe a method based on the principal component analysis that eliminates model redundancy. We show that by adding model orthogonalization to the proposed Bayesian model combination framework, one can arrive at better prediction accuracy and reach excellent uncertainty quantification performance.

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Optimal experimental design: Formulations and computations

Questions of ‘how best to acquire data’ are essential to modelling and prediction in the natural and social sciences, engineering applications, and beyond. Optimal experimental design (OED) formalizes these questions and creates computational methods to answer them. This article presents a systematic survey of modern OED, from its foundations in classical design theory to current research involving OED for complex models. We begin by reviewing criteria used to formulate an OED problem and thus to encode the goal of performing an experiment. We emphasize the flexibility of the Bayesian and decision-theoretic approach, which encompasses information-based criteria that are well-suited to nonlinear and non-Gaussian statistical models. We then discuss methods for estimating or bounding the values of these design criteria; this endeavour can be quite challenging due to strong nonlinearities, high parameter dimension, large per-sample costs, or settings where the model is implicit. A complementary set of computational issues involves optimization methods used to find a design; we discuss such methods in the discrete (combinatorial) setting of observation selection and in settings where an exact design can be continuously parametrized. Finally we present emerging methods for sequential OED that build non-myopic design policies, rather than explicit designs; these methods naturally adapt to the outcomes of past experiments in proposing new experiments, while seeking coordination among all experiments to be performed. Throughout, we highlight important open questions and challenges.

97 MATHEMATICS AND COMPUTING↗

Coordinated JWST Imaging of Three Distance Indicators in a Supernova Host Galaxy and an Estimate of the Tip of the Red Giant Branch Color Dependence

Boasting a 6.5 m mirror in space, JWST can increase by several times the number of supernovae (SNe) to which a redshift-independent distance has been measured with a precision distance indicator (e.g., tip of the red giant branch (TRGB) or Cepheids); the limited number of such SN calibrators currently dominates the uncertainty budget in distance ladder Hubble constant (H 0 ) experiments. JWST/NIRCAM imaging of the Virgo Cluster galaxy NGC 4536 is used here to preview JWST program GO-1995, which aims to measure H 0 using three stellar distance indicators (Cepheids, TRGB, and J-branch asymptotic giant branch/carbon stars). Each population of distance indicator was here successfully detected—with sufficiently large number statistics, well-measured fluxes, and characteristic distributions consistent with ingoing expectations—so as to confirm that we can acquire distances from each method precise to about 0.05 mag (statistical uncertainty only). We leverage overlapping Hubble Space Telescope imaging to identify TRGB stars, crossmatch them with the JWST photometry, and present a preliminary constraint on the slope of the TRGB's F115W versus (F115W – F444W) relation equal to -0.99 ± 0.16 mag mag -1 . This slope is consistent with prior slope measurements in the similar Two Micron All-Sky Survey J band, as well as with predictions from the BaSTI isochrone suite. We use the new TRGB slope estimate to flatten the 2D TRGB feature and measure a (blinded) TRGB distance relative to a set of fiducial TRGB colors, intended to represent the absolute fiducial calibrations expected from geometric anchors such as NGC 4258 and the Magellanic Clouds. In doing so, we empirically demonstrate that the TRGB can be used as a standardizable candle at the IR wavelengths accessible with JWST.

79 ASTRONOMY AND ASTROPHYSICS↗

Maximum a posteriori Ly α estimator (MAPLE): band power and covariance estimation of the 3D Ly α forest power spectrum

We present a novel maximum a posteriori estimator to jointly estimate band powers and the covariance of the three-dimensional power spectrum (P3D) of Ly $\alpha$ forest flux fluctuations, called MAPLE. Our Wiener-filter based algorithm reconstructs a window-deconvolved P3D in the presence of complex survey geometries typical for Ly $\alpha$ surveys that are sparsely sampled transverse to and densely sampled along the line of sight. We demonstrate our method on idealized Gaussian random fields with two selection functions: (i) a sparse sampling of 30 background sources per square degree designed to emulate the current Dark Energy Spectroscopic Instrument; (ii) a dense sampling of 900 background sources per square degree emulating the upcoming Prime Focus Spectrograph Galaxy Evolution Survey. Our proof-of-principle shows promise, especially since the algorithm can be extended to marginalize jointly over nuisance parameters and contaminants, i.e. offsets introduced by continuum fitting. Our code is implemented in JAX and is publicly available on GitHub.

79 ASTRONOMY AND ASTROPHYSICS↗

Modern chemical graph theory

Abstract Graph theory has a long history in chemistry. Yet as the breadth and variety of chemical data is rapidly changing, so too do graph encoding methods and analyses that yield qualitative and quantitative insights. Using illustrative cases within a basic mathematical framework, we showcase modern chemical graph theory's utility in Chemists' analysis and model development toolkit. The encoding of both experimental and simulation data is discussed at various levels of granularity of information. This is followed by a discussion of the two major classes of graph theoretical analyses: identifying connectivity patterns and partitioning methods. Measures, metrics, descriptors, and topological indices are then introduced with an emphasis upon enhancing interpretability and incorporation into physical models. Challenging data cases are described that include strategies for studying time dependence. Throughout, we incorporate recent advancements in computer science and applied mathematics that are propelling chemical graph theory into new domains of chemical study. This article is categorized under: Molecular and Statistical Mechanics > Molecular Dynamics and Monte‐Carlo Methods Structure and Mechanism > Computational Materials Science Structure and Mechanism > Molecular Structures

Leite, Leonardo S. G.↗

Bootstrap-determined p values in lattice QCD

We present a general method to determine the probability that stochastic Monte Carlo data, in particular those generated in a lattice QCD calculation, would have been obtained were that data drawn from the distribution predicted by a given theoretical hypothesis. Such a probability, or p -value, is often used as an important heuristic measure of the validity of that hypothesis. The proposed method offers the benefit that it remains usable in cases where the standard Hotelling T 2 methods based on the conventional χ 2 statistic do not apply, such as for uncorrelated fits. Specifically, we analyze q 2 , defined as the correlated χ 2 statistic obtained using an arbitrary covariance matrix estimator, and show how to use the bootstrap as a data-driven method to determine the expected distribution of q 2 for a given hypothesis with minimal assumptions. This distribution can then be used to determine the p -value for a fit to the data. We also describe a bootstrap approach for quantifying the impact upon this p -value of estimating population parameters from a single ensemble of N samples. The overall method is accurate up to a 1 / N bias which we do not attempt to quantify. Published by the American Physical Society 2025

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Determination of proton PDF uncertainties with Markov chain Monte Carlo

We present an analysis of parton distribution functions (PDFs) of the proton using Markov chain Monte Carlo (MCMC) methods. The MCMC approach naturally implements Bayes’ theorem and, thus, provides a means to directly sample the underlying probability distribution—in this case, the probability distribution of the PDF parameters. This allows for a straightforward propagation of the resulting uncertainties into any PDF-dependent observable, preserving their simple probabilistic interpretation. In our analysis we include a broad set of deep inelastic scattering data from HERA, BCDMS and NMC experiments along with the Drell-Yan, 𝑊 and 𝑍 boson data from LHC and Tevatron experiments, which combined with theoretical calculations at next-to-next-to-leading order in QCD allow for realistic determination of PDFs. The main focus of this analysis is to explore alternative methods for PDF uncertainty estimation that are more firmly grounded in statistical principles. We show that the flexibility of the Bayes framework, allowing one, e.g., to account for non-Gaussianity or inconsistencies of datasets, is crucial to extract realistic uncertainties when such assumptions are not fulfilled. We also demonstrate that MCMC allows one to determine the Δ⁢𝜒 2 value corresponding to a given confidence level in the sample, which can, in turn, be used as a statistically well-founded tolerance criterion used in the Hessian method, thus addressing one of its main long-standing drawbacks.

Risse, Peter Clemens [Universität Münster (Germany↗