Search NASASearch

SEARCH · Search NASA

Results for “Statistical error”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 55 records · Page 3

Online and Offline Data Quality Monitoring for the Mu2e Calorimeter

This thesis presents the design, implementation, and validation of a calorimeter Data Quality Monitoring (DQM) toolchain for the Mu2e experiment at Fermilab. Mu2e searches for charged lepton flavor violation via coherent muon-to-electron conversion in the field of an aluminum nucleus, $\mu^- Al \rightarrow e^-Al$, a process whose observation would constitute clear evidence of physics beyond the Standard Model. Achieving target sensitivity requires stringent control of detector performance and data integrity during acquisition, as subtle issues in readout configuration, data formatting, or electronics behavior can compromise reconstruction and bias downstream analyzes. To address these challenges, this work develops a multi-layer DQM approach spanning both raw data validation and reconstructed digi-level diagnostics. At the low level, a fragment analysis component performs word- and bit-field decoding of calorimeter readout blocks, enabling sanity checks of the expected structure and producing detailed error and integrity statistics useful for commissioning and troubleshooting. At the digi level, the CaloDigiDQM analyzer is implemented within the art framework and transforms each CaloDigiCollection into a structured hierarchy of ROOT histograms designed for fast drill-down diagnostics. The module generates coherent monitoring views at global, disk, board, and channel granularity, including occupancy, waveform-derived features (baseline, RMS, peak amplitude and position), and left-right sensor consistency metrics. Detector-aware channel-to-electronics mapping is performed through the conditions system (CaloDAQMap), ensuring that diagnostics remain aligned with hardware identifiers used in operations. For end-to-end testing without reliance on live DAQ data, a synthetic CaloDigi producer is developed to generate realistic waveforms with controlled noise and pulse shapes. The resulting system supports both offline ROOT-file production and online operation, including optional histogram streaming through otsdaq via ots::HistoSender. This toolchain provides a practical and scalable foundation for calorimeter commissioning and stable data collection, enabling early detection of anomalies and reducing operational risk for Mu2e.

Vakulenko, Mark [Drew U.] (ORCID:0009000276197818)

Type Ia Supernova Growth-rate Measurement with LSST Simulations: Intrinsic Scatter Systematics

Measurement of the growth rate of structures (fσ 8 ) with Type Ia supernovae (SNe Ia) will improve our understanding of the nature of dark energy and enable tests of general relativity. In this paper, we generate simulations of the 10 yr SN Ia data set of the Rubin-LSST survey, including a correlated velocity field from an N-body simulation and realistic models of SNe Ia properties and their correlations with host-galaxy properties. We find, similar to SN Ia analyses that constrain the dark energy equation-of-state parameters w 0 w a , that constraints on fσ 8 can be biased depending on the intrinsic scatter of SNe Ia. While for the majority of intrinsic scatter models we recover fσ 8 with a precision of ∼13%–14%, for the most realistic dust-based model, we find that the presence of non-Gaussianities in Hubble diagram residuals leads to a bias on fσ 8 of ∼ −20%. When trying to correct for the dust-based intrinsic scatter, we find that the propagation of the uncertainty on the model parameters does not significantly increase the error on fσ 8 . We also find that while the main component of the error budget of fσ 8 is the statistical uncertainty (>75% of the total error budget), the systematic error budget is dominated by the uncertainty on the damping parameter, σ u , that gives an empirical description of the effect of redshift space distortions on the velocity power spectrum. Our results motivate a search for new methods to correct for the non-Gaussian distribution of the Hubble diagram residuals, as well as an improved modeling of the damping parameter.

Carreres, Bastien [Duke Univ., Durham, NC (United

Data Set Analysis to Reduce Uncertainty in Formula Assignments of Ultrahigh Resolution Mass Spectra

Environmental samples contain a vast array of organic compounds with diverse elemental compositions and heteroatom content. Molecular formula assignments of ultrahigh resolution mass spectra (HRMS) hold promise for elucidating the molecular composition of these compounds. However, the need to account for an assortment of heteroatoms increases the uncertainty associated with individual assignments – and ultimately the ecological, biological, and biogeochemical insights gleaned from the assignments. To address this challenge, we introduce a formula assignment strategy that leverages HRMS data sets to improve assignment confidence, filter false assignments, and mitigate bias in assignment routines. The strategy, implemented using CoreMS, first identifies the highest confidence assignment for a recurring ion in a data set by assessing the mass accuracy and isotopologue similarity of all assignments to the ion across the data set. The second component of the strategy examines the consistency of mass errors for an assigned ion throughout a data set and flags formulas with statistically unlikely deviations in mass error. Here, we illustrate the application and utility of the strategy by comparing its results against documented misassignment patterns within a set of oceanographic samples that were measured with 21 T Fourier Transform Ion Cyclotron Resonance Mass Spectrometry. Because the efficacy of our strategy improves with data set size, it is particularly useful for enhancing assignment confidence in large HRMS data sets common in studies of environmental systems.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

The DESI-Lensing Mock Challenge: large-scale cosmological analysis of 3x2-pt statistics

The current generation of large galaxy surveys will test the cosmological model by combining multiple types of observational probes. Realising the statistical promise of these new datasets requires rigorous attention to all aspects of analysis including cosmological measurements, modelling, covariance and parameter likelihood. In this paper we present the results of an end-to-end simulation study designed to test the analysis pipeline for the combination of the Dark Energy Spectroscopic Instrument (DESI) Year 1 galaxy redshift dataset and separate weak gravitational lensing information from the Kilo-Degree Survey, Dark Energy Survey and Hyper-Suprime-Cam Survey. Our analysis employs the 3x2-pt correlation functions including cosmic shear and galaxy-galaxy lensing, together with the projected correlation function of the spectroscopic DESI lenses. We build realistic simulations of these datasets including galaxy halo occupation distributions, photometric redshift errors, weights, multiplicative shear calibration biases and magnification. We calculate the analytical covariance of these correlation functions including the Gaussian, noise and super-sample contributions, and show that our covariance determination agrees with estimates based on the ensemble of simulations. We use a Bayesian inference platform to demonstrate that we can recover the fiducial cosmological parameters of the simulation within the statistical error margin of the experiment, investigating the sensitivity to scale cuts. This study is the first in a sequence of papers in which we present and validate the large-scale 3x2-pt cosmological analysis of DESI-Y1.

79 ASTRONOMY AND ASTROPHYSICS

Generalized master equation for particle transport in binary random media with renewal statistics

Particle transport in binary stochastic mixtures is classically modeled assuming Markovian or exponential mixing statistics but in many applications material memory invalidates the Markov assumption. For non-Markovian mixing characterized by alternating renewal processes, a transport-theoretic framework is presented that provides an exact description of transport in nonscattering random binary media with general non-exponential statistics. Our approach is to Markovianize the problem by augmenting the {material type, particle flux} state space with the age or distance from the last interface. A Chapman-Kolmogorov equation is formulated for the joint probability density of the material type, particle flux, and age, and subsequently reduced to a generalized Master equation (GME) in differential form. This constitutes the primary result of this work. A state-updating Monte Carlo algorithm consistent with the GME is developed and benchmarked against analytical solutions for multiple chord-length laws. For purely absorbing renewal statistical media, the GME reproduces analytical benchmarks for the equilibrium age distribution, interior mean/variance of material-conditioned fluxes, and boundary transmittance. Simulations further demonstrate that a Markov (exponential) approximation of non-exponential statistics can introduce large errors in transmittance and interior flux profiles. Lastly, the reintroduction of memory due to scattering is briefly addressed through heuristic considerations.

Fluctuations & noise

Parton physics from a heavy-quark operator product expansion: Lattice QCD calculation of the fourth moment of the pion distribution amplitude

The pion light-cone distribution amplitude (LCDA) is an essential nonperturbative input for a range of high-energy exclusive processes in quantum chromodynamics. Building on our previous work, the continuum limit of the fourth Mellin moment of the pion LCDA is determined in quenched QCD using quark masses which correspond to a pion mass of 𝑚 𝜋 = 550 MeV. This calculation finds ⟨𝜉 2 ⟩ = 0.202⁢(8)⁢(9) and ⟨𝜉 4 ⟩ = 0.039⁢(28)⁢(11) where the first error indicates the combined statistical and systematic uncertainty from the analysis and the second indicates the uncertainty from working with Wilson coefficients computed to next-to-leading order. These results are presented in the $\overline{\textrm{MS}}$ scheme at a renormalization scale of 𝜇 = 2 GeV.

Detmold, William [Massachusetts Inst. of Technolog

Lanczos Algorithm, the Transfer Matrix, and the Signal-to-Noise Problem

This Letter introduces a method for determining the energy spectrum of lattice quantum chromodynamics by applying the Lanczos algorithm to the transfer matrix and using a bootstrap generalization of the Cullum-Willoughby method to filter out spurious eigenvalues. Proof-of-principle analyses of the simple harmonic oscillator and the lattice quantum chromodynamics proton mass demonstrate that this method provides faster ground-state convergence than the “effective mass,” which is related to the power-iteration algorithm. Lanczos provides more accurate energy estimates than multistate fits to correlation functions with small imaginary times while achieving comparable statistical precision. Two-sided error bounds are computed for Lanczos results and guarantee that excited-state effects cannot shift Lanczos results far outside their statistical uncertainties.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS

Distance Estimate Method for Asymptotic Giant Branch Stars Using Infrared Spectral Energy Distributions

We present a method to estimate distances to asymptotic giant branch (AGB) stars in the Galaxy, using spectral energy distributions (SEDs) in the near- and mid-infrared. By assuming that a given set of source properties (initial mass, stellar temperature, composition, and evolutionary stage) will provide a typical SED shape and brightness, sources are color matched to a distance-calibrated template and thereafter scaled to extract the distance. The method is tested by comparing the distances obtained to those estimated from very long baseline interferometry or Gaia parallax measurements, yielding a strong correlation in both cases. Additional templates are formed by constructing a source sample likely to be close to the Galactic center, and thus with a common, typical distance for calibration of the templates. These first results provide statistical distance estimates to a set of almost 15,000 Milky Way AGB stars belonging to the Bulge Asymmetries and Dynamical Evolution (BAaDE) survey, with typical distance errors of ±35%. With these statistical distances, a map of the intermediate-age population of stars traced by AGBs is formed, and a clear bar structure can be discerned, consistent with the previously reported inclination angle of 30° to the GC–Sun direction vector. These results motivate deeper studies of the AGB population to tease out the intermediate-age stellar distribution throughout the Galaxy, as well as determining statistical properties of the AGB population luminosity and mass-loss-rate distributions.

79 ASTRONOMY AND ASTROPHYSICS

Robust measurement of microbial reduction of graphene oxide nanoparticles using image analysis

ABSTRACT Shewanella oneidensis ( S. oneidensis ) has the capacity to reduce electron acceptors within a medium and is thus used frequently in microbial fuel generation, pollutant breakdown, and nanoparticle fabrication. Microbial fuel setups, however, often require costly or labor-intensive components, thus making optimization of their performance onerous. For rapid optimization of setup conditions, a model reduction assay can be employed to allow simultaneous, large-scale experiments at lower cost and effort. Since S. oneidensis uses different extracellular electron transfer pathways depending on the electron acceptor, it is essential to use a reduction assay that mirrors the pathways employed in the microbial fuel system. For microbial fuel setups that use nanoparticles to stimulate electron transfer, reduction of graphene oxide provides a more accurate model than other commonly used assays as it is a bulk material that forms flocculates in solutions with a large ionic component. However, graphene oxide flocculates can interfere with traditional absorbance-based measurement techniques. This study introduces a novel image analysis method for quantifying graphene oxide reduction, showing improved performance and statistical accuracy over traditional methods. A comparative analysis shows that the image analysis method produces smaller errors between replicates and reveals more statistically significant differences between samples than traditional plate reader measurements under conditions causing graphene oxide flocculation. Image analysis can also detect reduction activity at earlier time points due to its use of larger solution volumes, enhancing color detection. These improvements in accuracy make image analysis a promising method for optimizing microbial fuel cells that use nanoparticles or bulk substrates. IMPORTANCE Shewanella oneidensis ( S. oneidensis ) is widely used in reduction processes such as microbial fuel generation due to its capacity to reduce electron acceptors. Often, these setups are labor-intensive to operate and require days to produce results, so use of a model assay would reduce the time and expenses needed for optimization. Our research developed a novel digital analysis method for analysis of graphene oxide flocculates that may be utilized as a model assay for reduction platforms featuring nanoparticles. Use of this model reduction assay will enable rapid optimization and drive improvements in the microbial fuel generation sector.

Bennett, Danielle T. (ORCID:0009000188748827)

Short-term electricity load forecasting: Application-driven evaluation of machine learning models across spatial and temporal scales

As we transition towards a decarbonized economy, the integration of variable renewable energy resources and new demands (e.g., electric vehicles, heat pumps) into the electricity grid places unprecedented pressure on grid operators to effectively anticipate and manage peak load. In this context, machine learning algorithms are proving to be indispensable for accurate short-term load forecasting, a crucial task to address these challenges. This study benchmarks 6 machine learning algorithms, including three neural networks and three tree-based algorithms, across various levels of spatial aggregation and time horizons (1, 4, 8, 24, and 48 h). The central contribution of this work is the comparison and analysis of load forecasting models not only based on statistical metrics, but also based on a novel error metric, which evaluates the cost implications of forecast errors for power system stakeholders. Results show that tree-based models outperform neural networks, based on statistical metrics, and yield less skewed error distributions for most spatial scales. However, through the lens of the novel error metric, neural networks are the more competitive choice, especially for forecast horizons that exceed 8 h. The study concludes with actionable recommendations to grid operators and highlights the need for the development of error metrics that link forecasting accuracy to operational costs. To promote transparency and open science, the datasets and Python code are open-sourced via a supplementary repository.

Houben, Nikolaus

Model validation and error attribution for a drifting qubit

Qubit performance is often reported in terms of a variety of single-value metrics, each providing a facet of the underlying noise mechanism limiting performance. However, the value of these metrics may drift over long timescales, and reporting a single number for qubit performance fails to account for the low-frequency noise processes that give rise to this drift. Here, in this work, we demonstrate how we can use the distribution of these values to validate or invalidate candidate noise models. We focus on the case of randomized benchmarking (RB), where typically a single error rate is reported but this error rate can drift over time when multiple passes of RB are performed. We show that using a statistical test as simple as the Kolmogorov-Smirnov statistic on the distribution of RB error rates can be used to rule out noise models, assuming the experiment is performed over a long enough time interval to capture relevant low frequency noise. With confidence in a noise model, we show how care must be exercised when performing error attribution using the distribution of drifting RB error rate.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND

Shuttling Majorana zero modes in disordered and noisy topological superconductors

The braiding of Majorana zero modes (MZMs) forms the fundamental building block for topological quantum computation. Braiding protocols which involve the physical exchange of MZMs are typically envisioned on a network of topological superconducting wires. An important component of these protocols is the transport of MZMs, which can be performed by using electric gates to locally tune sections of the wire between topologically trivial and nontrivial phases. In this work, we numerically simulate this transport by tuning a single section of a superconducting wire which contains either disorder (uncorrelated and correlated) or noise. We focus on the impact of these additional effects on the diabatic error, which describes unwanted transitions between the ground state and excited states. We show that the behavior of the average diabatic error is predominantly controlled by the statistics of the minimum bulk energy gap which is suppressed in the presence of disorder. The increase in diabatic error can be several orders of magnitude and is most deleterious when the disorder correlation length is a finite fraction of the transport distance and negligible when these lengths are far apart. In the presence of noise, the diabatic error is significantly enhanced due to optical transitions which depend on the minimum bulk energy gap as well as the frequency modes present in the noise. The results presented here serve to further characterize the diabatic error in disordered and noisy settings, which are important considerations in practical implementations of physical braiding schemes.

75 CONDENSED MATTER PHYSICS, SUPERCONDUCTIVITY AND

Rare events and Griffiths phases in topological quantum error correction

The performance of quantum error correcting (QEC) codes is often studied under the assumption of spatiotemporally uniform error rates. On the other hand, experimental implementations almost always produce heterogeneous error rates, in either space or time, as a result of effects such as imperfect fabrication and/or cosmic rays. It is therefore important to understand if and how their presence can affect the performance of QEC in qualitative ways. Here, in this work, we study the effects of nonuniform error rates in the representative examples of the 1D repetition code and the 2D toric code, focusing on when they have extended spatiotemporal correlations; these may arise, for instance, from rare events (such as cosmic rays) that temporarily elevate error rates over the entire code patch. These effects can be described in the corresponding statistical mechanics models for decoding, where long-range correlations in the error rates lead to extended rare regions of weaker coupling. For the 1D repetition code where the rare regions are linear, we find two distinct decodable phases: a conventional ordered phase in which logical failure rates decay exponentially with the code distance, and a rare-region dominated Griffiths phase in which failure rates are parametrically larger and decay as a stretched exponential. In particular, the latter phase is present when the error rates in the rare regions are above the bulk threshold. For the 2D toric code where the rare regions are planar, we find no decodable Griffiths phase: rare events which boost error rates above the bulk threshold lead to an asymptotic loss of threshold and failure to decode. Unpacking the failure mechanism implies that techniques for suppressing extended sequences of repeated rare events (which, without intervention, will be statistically present with high probability) will be crucial for QEC with the toric code.

classical statistical mechanics

Error field predictability and consequences for ITER

Abstract ITER coil tolerances are re-evaluated using the modern understanding of coupling to least-stable plasma modes and an updated center-line-traced model of ITER’s coil windings. This reassessment finds the tolerances to be conservative through a statistical, linear study of n = 1 error fields (EFs) due to tilted, shifted misplacements and nominal windings of central solenoid and poloidal field coils within tolerance. We also show that a model-based correction scheme remains effective even when metrology quality is sub-optimal, and compare this to projected empirical correction schemes. We begin with an analysis of the necessity of error field correction (EFC) for daily operation in ITER using scalign laws for the EF penetration threshold. We then consider the predictability of EF dominant mode overlap across early planned ITER scenarios and, as measuring EFs in high power scenarios can pose risks to the device, the potential for extrapolation to the ITER Baseline Scenario (IBS). We find that carefully designing a scenario matching currents proportionally to those of the IBS is far more important than plasma shape or profiles in accurately measuring an optimal correction current set.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY

How Representative Are Uncrewed Aircraft System Measurements of the Convective Boundary Layer?

Abstract Uncrewed aircraft systems (UAS) demonstrate significant potential for filling data gaps in the atmospheric boundary layer. However, the extent to which UAS observations—typically vertical profiles taken over 15 min—are representative of the boundary layer as a whole remains poorly characterized. Using large eddy simulations (LES) of the daytime convective boundary layer (CBL), we quantify random errors in UAS measurements that occur due to insufficient statistical convergence of the time average to the true ensemble mean. Random errors in first‐order moments increase as the CBL becomes increasingly unstable, and are largest near the surface for most quantities. Errors are on the order of 2–6 m for wind speed, 15–60 for wind direction, 0.2–3 K for potential temperature, and 0.1–1 g for specific humidity, with errors in turbulent fluxes on the order of 50%–100%. Sampling strategies that mitigate random errors are discussed in light of our results.

Greene, Brian R. [Now at Verisk Extreme Event Solu

Full-stack Quantification of Variability in Predicting Ion Transport Properties using Machine-learned Interatomic Potentials

Machine-learned interatomic potentials (MLIPs) have become the state-of-the-art for performing accurate, scalable molecular dynamics (MD) simulations. It is therefore crucial to understand and quantify the reliability of MLIPs for downstream property predictions. Uncertainty in predicted properties can arise from limitations in first-principles training data, intrinsic MLIP model errors in representing the data, and the statistical noise introduced during subsequent MD simulations. Using ion transport in Li7P3S11 as a case study, we systematically assess the impact of training set size and selection, neural network stochasticity, and MD sampling statistics on predicted diffusivity and activation energy. We find that when using equivariant MLIP architectures with standard MD protocols, uncertainty arising from MD sampling dominates over model-induced errors. In contrast, MLIP errors relative to the underlying first-principles data are consistently minor. Given this, there are two main routes to improving the accuracy of predictions based on MLIP potentials: adopting higher accuracy reference data generation methods, and improving the MD sampling statistics.

36 MATERIALS SCIENCE

An evaluation of multi-fidelity methods for quantifying uncertainty in projections of ice-sheet mass change

Abstract. This study investigated the computational benefits of using multi-fidelity statistical estimation (MFSE) algorithms to quantify uncertainty in the mass change of Humboldt Glacier, Greenland, between 2007 and 2100 using a single climate change scenario. The goal of this study was to determine whether MFSE can use multiple models of varying cost and accuracy to reduce the computational cost of estimating the mean and variance of the projected mass change of a glacier. The problem size and complexity were chosen to reflect the challenges posed by future continental-scale studies while still facilitating a computationally feasible investigation of MFSE methods. When quantifying uncertainty introduced by a high-dimensional parameterization of the basal friction field, MFSE was able to reduce the mean-squared error in the estimates of the statistics by well over an order of magnitude when compared to a single-fidelity approach that only used the highest-fidelity model. This significant reduction in computational cost was achieved despite the low-fidelity models used being incapable of capturing the local features of the ice-flow fields predicted by the high-fidelity model. The MFSE algorithms were able to effectively leverage the high correlation between each model's predictions of mass change, which all responded similarly to perturbations in the model inputs. Consequently, our results suggest that MFSE could be highly useful for reducing the cost of computing continental-scale probabilistic projections of sea-level rise due to ice-sheet mass change.

54 ENVIRONMENTAL SCIENCES

Object-Based Evaluation of Dynamical and Statistical Downscaled Precipitation Products over CONUS

High-resolution precipitation data, generated through dynamical downscaling (DD) or statistical downscaling (SD) of global climate model output, provide critical information for regional climate assessment and adaptation planning. Most downscaling development and validation have focused on accurate gridscale precipitation construction and ignored the spatial structure of precipitation across model grids and at the event scale. However, many applications, e.g., hydrologic modeling and the analysis using the downscaled precipitation, require a reasonable representation of the spatial structure of precipitation within watersheds. Therefore, a set of standard metrics to evaluate the representation of the spatial structure of individual storms across diverse downscaled precipitation products is desired. To address this need, we conducted an object-based evaluation of precipitation in decades-long DD and SD products over the contiguous United States (CONUS). Specifically, we evaluate their ability to reproduce various features of precipitation objects in the observations: total volume, precipitation area, peak intensity, and spatial structure. Multiple metrics (bias, Perkins score, and nonparametric statistical tests) are used to quantify model performance. Our evaluation reveals notable variations in performance among individual products across different climate zones and seasons, as well as between extreme and nonextreme events. In general, most DD products exhibit balanced performance across the four precipitation object features, while SD products vary more significantly in their performance across products. Based on this comprehensive evaluation, we provide guidance on choosing downscaled products for specific regions, seasons, and precipitation object features. These findings and recommendations can inform precipitation-relevant modeling and analysis over CONUS, guide future downscaling technique developments, and provide actionable information for climate impact assessment and adaptation.

Downscaling