Search NASA⌕ Search

SEARCH · Search NASA

Results for “Statistical methods”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 307 records · Page 17

Going Off Grid: A Comparative Study of the Lagrangian and Eulerian Perspectives of New Particle Formation Events

New particle formation and growth (NPF&G) is the process by which ultrafine particles are formed from gas-phase precursors. NPF&G is the dominant source of global aerosol number with important influences on climate. Most observations of NPF&G events are conducted at stationary sites; however, NPF&G observed from stationary sites is influenced by gradual or rapid changes in the air masses passing over the site, complicating NPF&G analysis. In this work, we use observations and a 3D aerosol model to compare aerosol size distributions at a stationary site (Southern Great Plains [SGP] observatory, Oklahoma, USA) and along Lagrangian trajectories crossing the site. The model simulates the NPF&G events reasonably well at SGP. Using the model to compare the Lagrangian and stationary perspectives, we can explain previously unanalyzable days with some evidence of NPF&G as either non-event or analyzable NPF&G days. We find most of the unanalyzable NPF&G days are due to isolated and inhomogeneous NPF&G occurring upwind of the stationary site, often in the outflow of urban regions. Finally, we compare formation rates of 3 nm particles, growth rates, and the survival probability of 3 nm particles growing to 25 nm between the stationary and Lagrangian perspectives. Because of the much larger number of analyzable days along the Lagrangian trajectories, this perspective potentially provides more robust statistics and better characterization of NPF&G event extremes. Our method for extracting chemical/physical properties along Lagrangian trajectories from 3D models can be applied to a wide range of science questions.

O’Donnell, Samuel E. [Colorado State Univ., Fort C↗

Improving tropical cyclone rapid intensification forecasts with satellite measurements of sea surface salinity and calibrated machine learning

Forecasting rapid intensification (RI) of tropical cyclones (TC) is a mission known for large errors. One under-researched factor that affects TC intensification is salinity, which is important for density stratification in certain ocean regions and can affect the surface enthalpy flux under a strengthening hurricane. To investigate the impact and efficacy of using salinity information in state-of-the-art forecasting, we use a statistical model consisting of a variety of machine learning (ML) methods. For salinity data, we use satellite measurements of pre-storm sea surface salinity (SSS) as a proxy for the salinity stratification. We train and test the model on various ocean basins, including the Atlantic, eastern North Pacific and western North Pacific. A calibrator is trained on top of the ML models to correct and enhance probability forecasts. The calibrator significantly improves probability forecasts relative to recent works. The ML model performance is improved with the addition of SSS in the Eastern North Pacific, western North Pacific, and the Caribbean subregion of the North Atlantic, and the overall model performance is better than previous studies. SSS decreases model skill for a model trained on the full Atlantic basin. In the Indian Ocean, SSS is also notably correlated with RI occurrence, but the TC samples are not sufficient to train ML models.

hurricane↗

Machine Learned Force Field Modeling of Metal Organic Frameworks for CO2 Direct Air Capture

Metal organic frameworks (MOFs) are a large class of porous materials and have garnered significant interest due to their large surface areas and their tunable physical and chemical properties. Numerous prior studies have been performed to screen large databases of this material class for promising DAC sorbent materials. These studies have often relied on classical model potentials. While density functional theory (DFT) calculations have been shown to be very accurate for modeling the interaction of CO2 with MOFs, such calculations are too computationally demanding for statistically significant adsorption predictions. To overcome this barrier, we developed methods for training models to achieve DFT-level accuracy for the forces and energies associated with MOF flexibility and CO2 adsorption using machine learned force fields (MLFFs). These methods were parametrized based on DFT calculations of CO2 in a flexible MOF and used to predict MOF structural properties as well as CO2 adsorption in several MOFs.

Findley, John↗

169 Tm ( n , γ ) cross section and statistical decay properties from measurements at the DANCE facility

Background: Radiative neutron capture on thulium, which is a monoisotopic element, plays a role in different applications such as nuclear astrophysics or nuclear burning environments. Considerable discrepancies—reaching 20%—exist between evaluations in the unresolved-resonance region. Furthermore, experimental data on statistical 𝛾 decay in odd-odd rare-earth nuclei is scarce. There are still open questions about the systematics of the so-called scissors mode in the 𝑀⁢1 photon strength function, especially in odd-odd nuclei. Purpose: This work is focused on two main topics—deriving experimental 169 Tm ⁢(𝑛,𝛾) cross section and studying statistical 𝛾 decay of 170 Tm, in particular properties of the scissors mode. Methods: The capture experiments to obtain experimental cross section were performed at the Los Alamos Neutron Science Center using the time-of-flight technique and employing the Detector for Advanced Neutron Capture Experiments. Measured coincident 𝛾-ray spectra were also compared with statistical simulations using the dicebox code to test different models of level density and photon strength functions. Results: The capture cross section was determined from 1.8 eV to 0.97 MeV, the broadest neutron-energy range ever measured for this isotope. Several new resonances have been observed. The statistical 𝛾 decay of 170 Tm cannot be reproduced without a scissors mode resonance centered at ≈ 3.3MeV. Conclusions: The measured cross section in the unresolved-resonance region is generally lower than the latest evaluations. The derived 169 Tm 𝑠-process abundance is expected to increase by a factor of 1.26, while the changes of the abundances of elements heavier than 169 Tm are in the order of 0.2%. The scissors mode properties in 170 Tm are similar to those deduced in previous analyses of neighboring nuclei 168 Er and 166 Ho .

73 NUCLEAR PHYSICS AND RADIATION PHYSICS↗

Multivariate Analysis as a Tool for Validating Tester Matching

A method of applying Principal Component Analysis, Soft Independent Modeling of Class Analysis, and statistical analysis is described that can be applied to many types of testers to ascertain how well matched the performance of the testers in the analysis are to one another or how well matched a tester is to itself at a later time. This method is most useful for situations for which the same units have not been run across the testers being analyzed for matched performance.

Multari, Rosalie A [Sandia National Laboratories (↗

Quantitative measurements of dislocations in metals for advancing predictive simulations

LLNL applications require scientists to predict how materials evolve under various thermomechanical conditions. While this is achieved through physics-based simulations, uncertainty in the predictions of mechanical properties remains a serious challenge that limits the predictive capabilities of models because we lack methods to compare predictions of atomic-scale defects (dislocations) with experimental measurements. High energy X-ray diffraction (HEXRD) is the most relevant technique that can provide the necessary statistical information on dislocations. However, this technique is not yet quantitative because we lack a precise understanding of the relationship between X-ray diffraction patterns and the underlying material dislocation content and arrangements. To address this need, we used our novel computational X-ray diffraction method to simulate the effect of dislocations on the diffraction patterns. We compared virtual and experimental diffraction patterns. Results allowed us to clearly establish the relationship between X-ray diffraction patterns and the underlying dislocation structures, proving that it is feasible to quantitatively measure dislocation statistics with HEXRD. This project delivered a method that can provide the missing piece to LLNL’s mechanical property simulations in advanced metals by obtaining experimentally long-needed quantitative dislocation data, which could fully enable predictive capabilities.

36 MATERIALS SCIENCE↗

Estimating an executive summary of a time series: the tendency

In this paper, we revisit the problem of decomposing a signal into a tendency and a residual. The tendency describes an executive summary of a signal that encapsulates its notable characteristics while disregarding seemingly random, less interesting aspects. Building upon the Intrinsic Time Decomposition (ITD) and information-theoretical analysis, we introduce two alternative procedures for selecting the tendency from the ITD baselines. The first is based on the maximum extrema prominence, namely the maximum difference between extrema within each baseline. Specifically this method selects the tendency as the baseline from which an ITD step would produce the largest decline of the maximum prominence. The second method uses the rotations from the ITD and selects the tendency as the last baseline for which the associated rotation is statistically stationary. We delve into a comparative analysis of the information content and interpretability of the tendencies obtained by our proposed methods and those obtained through conventional low-pass filtering schemes, particularly the Hodrik–Prescott (HP) filter. Our findings underscore a fundamental distinction in the nature and interpretability of these tendencies, highlighting their context-dependent utility with emphasis in multi-scale signals. Through a series of real-world applications, we demonstrate the computational robustness and practical utility of our proposed tendencies, emphasizing their adaptability and relevance in diverse time series contexts.

Time series analysis↗

Phase-space methods for neutrino oscillations: Extension to multibeams

The phase-space approach (PSA), which was originally introduced in Lacroix [] to describe neutrino flavor oscillations for interacting neutrinos emitted from stellar objects is extended to describe arbitrary numbers of neutrino beams. The PSA is based on mapping the quantum fluctuations into a statistical treatment by sampling initial conditions followed by independent mean-field evolution. A new method is proposed to perform this sampling that allows treating an arbitrary number of neutrinos in each neutrino beams. We validate the technique successfully and confirm its predictive power on several examples where a reference exact calculation is possible. We show that it can describe many-body effects, such as entanglement and dissipation induced by the interaction between neutrinos. Due to the complexity of the problem, exact solutions can only be calculated for rather limited cases, with a limited number of beams and/or neutrinos in each beam. The PSA approach considerably reduces the numerical cost and provides an efficient technique to accurately simulate arbitrary numbers of beams. Examples of PSA results are given here, including up to 200 beams with time-independent or time-dependent Hamiltonians. We anticipate that this approach will be useful to bridge exact microscopic techniques with more traditional transport theories used in neutrino oscillations. It will also provide important reference calculations for future quantum computer applications where other techniques are not applicable to classical computers. Published by the American Physical Society 2024

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Machine learning method for enforcing variable independence in background estimation with LHC data: ABCDisCoTEC

A novel solution is presented for the problem of estimating the backgrounds of a signal search using observed data while simultaneously maximizing the sensitivity of the search to the signal. The 'ABCD method' provides a reliable framework for background estimation by partitioning events into one signal-enhanced region (A) and three background-enhanced control regions (B, C, and D) via two smoothly varying, statistically independent variables. In practice, even slight correlations between the two variables can significantly undermine the method's performance. Thus, choosing appropriate variables by hand can present a formidable challenge, especially when background and signal differ only subtly. To address this issue, the ABCD with distance correlation (ABCDisCo) method was developed to construct two learned variables via a neural network trained to provide strong signal-background discrimination with small values of the distance correlation (DisCo) measure between the two learned variables. However, relying solely on minimizing the DisCo can result in learned variables that may not have distributions of background events that are smoothly varying and localized at extreme values, as necessary for the validity of the background estimation. The ABCDisCo training enhanced with closure (ABCDisCoTEC) method is introduced to solve this issue by directly minimizing the nonclosure, expressed as a dedicated differentiable loss term. This extended method is applied to a data set of proton-proton collisions at a center-of-mass energy of 13 TeV recorded by the CMS detector at the CERN Large Hadron Collider. Additionally, given the complexity of the minimization problem with constraints on multiple loss terms, the modified differential method of multipliers is applied and shown to greatly improve the stability and robustness of the ABCDisCoTEC method, compared to grid search hyperparameter optimization procedures.

Hayrapetyan, Aram [Yerevan Phys. Inst.]↗

Robust Design Under Uncertainty in Quantum Error Mitigation

Error mitigation techniques are crucial to achieving near-term quantum advantage. Classical postprocessing of quantum computation outcomes is a popular approach for error mitigation, which includes methods, such as zero noise extrapolation, virtual distillation, and learning-based error mitigation. However, these techniques have limitations due to the propagation of uncertainty resulting from the finite shot number of a quantum measurement. In this work, we introduce general and unbiased methods for quantifying the uncertainty and error of error-mitigated observables based on the strategic sampling of error mitigation outcomes. We then extend our approach to demonstrate the optimization of performance and robustness of error mitigation under uncertainty. To illustrate our methods, we apply them to zero noise extrapolation and Clifford date regression in the ground state of the XY model simulated using depolarizing and International Business Machines Corporation (IBM) Toronto noise models, respectively. In particular, we optimize the choice of noise levels and the allocation of shots for zero noise extrapolation and the distribution of the training circuits for Clifford data regression. While our methods are readily applicable to any postprocessing-based error mitigation approach, in practice they must not be prohibitively expensive—even though they perform optimizations of the error mitigation hyperparameters requiring sampling of a statistical distribution of error mitigation outcomes. By leveraging surrogate-based optimization, we show that our methods can efficiently perform optimal design for a zero noise extrapolation implementation. We then further demonstrate the transferability of learned zero noise extrapolation hyperparameters to other similar circuits.

97 MATHEMATICS AND COMPUTING↗

A kinetic-based regularization method for data science applications

We propose a physics-based regularization technique for function learning, inspired by statistical mechanics. By drawing an analogy between optimizing the parameters of an interpolator and minimizing the energy of a system, we introduce corrections that impose constraints on the lower-order moments of the data distribution. This minimizes the discrepancy between the discrete and continuum representations of the data, in turn allowing to access more favorable energy landscapes, thus improving the accuracy of the interpolator. Our approach improves performance in both interpolation and regression tasks, even in high-dimensional spaces. Unlike traditional methods, it does not require empirical parameter tuning, making it particularly effective for handling noisy data. We also show that thanks to its local nature, the method offers computational and memory efficiency advantages over Radial Basis Function interpolators, especially for large datasets.

97 MATHEMATICS AND COMPUTING↗

Comparing multi-source urban flood indicators: satellite, simulation, and citizen-reported data

Urban flooding arises from complex mechanisms, making it challenging to capture accurately with a single detection method. This study evaluates three complementary approaches to detect flooding across three Chicago neighborhoods: (i) Sentinel-1 synthetic aperture radar (SAR), offering weather-independent, high-resolution (10 m) imagery of surface inundation; (ii) the storm water management model (SWMM), simulating combined sewer overflow and drainage performance; and (iii) citizen-generated 311 service requests, capturing observed flooding impacts. By analyzing six storms ranging from severe to mild, we examine how each source uniquely contributes to identifying urban flood events. SAR imagery effectively identifies standing water but can miss brief flooding due to satellite revisit constraints. SWMM provides detailed insights into system-wide drainage behavior yet may underestimate localized street-level flooding. Meanwhile, 311 calls reflect real-world flooding impacts but are vulnerable to underreporting. Statistical overlap analysis highlights chronic flood hotspots repeatedly identified across multiple detection methods, indicating persistent infrastructure and topographic vulnerabilities. Temporal analysis further reveals that while SWMM flooding aligns closely with rainfall peaks, 311 calls typically precede or persist beyond these peaks. Our findings emphasize the value of using satellite observations, hydrological modeling, and resident-reported data in a complementary manner to better interpret patterns in flood timing, severity, and spatial distribution—providing insights that can inform targeted infrastructure improvements and contribute to urban flood resilience planning.

311↗

Segmentation and Classification of Fission as Pores in Reactor Irradiated Annular U–10Zr Metallic Fuel Using Machine Learning Models

Metallic fuels, particularly U—10Zr, are promising candidates for next-generation sodium-cooled fast reactors. Irradiation of nuclear fuels in reactors can lead to the formation of solid and gas fission product which subsequently forms microstructural pores, deteriorating fuel performance. Due to the massive amount of pores and complex phases formed, a quantitative description of fission gas pores is not yet available, preventing the development of microstructure-informed fuel performance modeling for fuel qualification. This paper applied a pre-trained deep learning model to ~10,260 high magnification scanning electron microscopy images. This method increased the accuracy of fission gas pore segmentation and allows statistical features to be extracted which cannot be achieved manually. A pre-trained decision tree model worked on the segemenation results and further classified the pores into different categories to produce a correlation between the pores, movement of lanthanides, and temperature gradient during irradiation. Finally, this paper emphasizes the potentials of machine learning models to accelerate fuel research, development, and qualification for advanced reactors.

36 MATERIALS SCIENCE↗

Biopolymer-Templated Titania Film Formation for Nanostructured Coatings Revealed by Machine Learning-Supported Time-Resolved Analysis

This study presents a machine learning approach to derive the film formation of biopolymer-templated titania nanostructures during spray deposition, in combination with in situ grazing-incidence small-angle X-ray scattering (GISAXS). A neural network trained on synthetic GISAXS data directly predicts domain-size distributions from experimental two-dimensional scattering patterns, capturing the full kinetics of nanostructure evolution with high temporal resolution. The predictions reveal hierarchical size distributions and periodic growth features, consistent with layer-by-layer spray deposition and validated by complementary scanning electron microscopy (SEM) imaging. Quantitative comparison with conventional parametric GISAXS fits shows good qualitative agreement, with systematic differences explained by domain-shape assumptions and resolved by applying a geometric scaling factor. Simulated SEM-like surfaces derived from neural network outputs reproduce the porous, foam-like nanoscale morphology observed experimentally, reinforcing the method’s credibility. This integrated approach enables real-time, nondestructive, statistically averaged monitoring of bulk nanostructure development in functional coatings, offering a scalable methodology to accelerate the characterization and process control of sustainably manufactured nanostructured titania films for energy-related applications such as photocatalysis and photovoltaics.

Heger, JulianEliah↗

Optimization of foreground moment deprojection for semi-blind CMB polarization reconstruction

Abstract Upcoming Cosmic Microwave Background (CMB) experiments, aimed at measuring primordial CMB polarization B-modes, require exquisite control of instrumental systematics and Galactic foreground contamination. Blind minimum-variance techniques, like the Needlet Internal Linear Combination (NILC), have proven effective in reconstructing the CMB polarization signal and mitigating foregrounds and systematics across diverse sky models without suffering from foreground mismodelling errors. Still, residual foreground contamination from NILC may bias the recovered CMB polarization at large angular scales when confronted with the most complex foreground scenarios.By adding constraints to NILC to deproject statistical moments of the Galactic emission, the Constrained Moment ILC (cMILC) method has been demonstrated to further enhance foreground subtraction, albeit with an associated increase in overall noise variance. Faced with this trade-off between foreground bias reduction and overall variance minimization, there is still no recipe on which moments to deproject and which are better suited for blind variance minimization. To address this, we introduce the optimized cMILC (ocMILC) pipeline, which performs full automated optimization of the required number and set of foreground moments to deproject, pivot parameter values, and deprojection coefficients across the sky and angular scales, depending on the actual sky complexity, available frequency coverage, and experiment sensitivity. The optimal number of moments for deprojection, before paying significant noise penalty, is determined through a data diagnosis inspired by the Generalized NILC (GNILC) method.Validated on B-mode simulations of thePICOspace mission concept with four challenging foreground models, ocMILC exhibits lower Galactic foreground contamination compared to NILC and cMILC at all angular scales, with limited noise penalty. This multi-layer optimization enables the ocMILC pipeline to achieve unbiased posteriors of the tensor-to-scalar ratio, regardless of foreground complexity.

Astronomy & Astrophysics↗

Dark Energy Survey Year 3 results: $w$CDM cosmology from simulation-based inference with persistent homology on the sphere

We present cosmological constraints from Dark Energy Survey Year 3 (DES Y3) weak lensing data using persistent homology, a topological data analysis technique that tracks how features like clusters and voids evolve across density thresholds. For the first time, we apply spherical persistent homology to galaxy survey data through the algorithm TopoS2, which is optimized for curved-sky analyses and HEALPix compatibility. Employing a simulation-based inference framework with the Gower Street simulation suite, specifically designed to mimic DES Y3 data properties, we extract topological summary statistics from convergence maps across multiple smoothing scales and redshift bins. After neural network compression of these statistics, we estimate the likelihood function and validate our analysis against baryonic feedback effects, finding minimal biases (under $0.3σ$) in the $Ω_\mathrm{m}-S_8$ plane. Assuming the $w$CDM model, our combined Betti numbers and second moments analysis yields $S_8 = 0.821 \pm 0.018$ and $Ω_\mathrm{m} = 0.304\pm0.037$-constraints 70% tighter than those from cosmic shear two-point statistics in the same parameter plane. Our results demonstrate that topological methods provide a powerful and robust framework for extracting cosmological information, with our spherical methodology readily applicable to upcoming Stage IV wide-field galaxy surveys.

Prat, J. [Nordita; Royal Inst. Tech., Sodertalje; ↗

Stacked reverberation mapping of high-redshift quasars in DESI. I. Feasibility analysis

The broad-line region of quasars has long been probed by reverberation mapping techniques that measure time lags between continuum and broad emission-line variations. Stacked reverberation mapping has been proposed as a less observationally expensive alternative to traditional methods. This ensemble approach also reduces biases from small-number statistics. The Dark Energy Spectroscopic Instrument (DESI) is conducting the most extensive spectroscopic survey of quasars to date. We create mock light curves emulating expected DESI quasar observations at redshifts $1.48\lt z\lt 5.2$ and luminosities $44.68 \le \log \lambda L_{1350 \mathring{\rm A}{}} / \mathrm{erg\, s^{-1}} \le 45.99$ to test stacked reverberation mapping feasibility using sparse spectroscopic data paired with well-sampled photometric data. The pipeline, using the lag estimation code JAVELIN (Just Another Vehicle for Estimating Lags In Nuclei), successfully recovers the simulated C IV lags within 1σ of the true values using spectroscopic light curves composed of only a few spectral epochs (2–10) with irregular cadences. We investigate how observational factors, including C IV flux error magnitude, number of stacked quasars, and spectral epoch count, affect performance. This work motivates a pathway for future stacked reverberation mapping projects with large-scale spectroscopic surveys of quasars having $\ge 2$ spectroscopic observations. Our results suggest an economical alternative for constraining and extending the radius–luminosity relation to higher redshifts and luminosities. Subsequently, this relation can be employed more reliably in single-epoch black hole mass measurements and quasar cosmology in these distant regimes.

quasars: general, quasars: supermassive black hole↗

Labeling sequential data from noisy annotations

Crowdsourcing algorithms often work under the assumption that the data samples are independent. Recent work has shown that data dependence, such as temporal correlations in sequential data, can be leveraged to improve the label quality. Existing methods that exploit this special structure rely on third-order statistics of the annotator outputs to ensure the identifiability of key latent parameters, which are costly to acquire. This work proposes an approach for integrating crowdsourced annotations under the Dawid-Skene/Hidden Markov Model (DS-HMM) for sequential data based on second-order statistics, which naturally enjoys a lower sample complexity. An effective algorithm is proposed to tackle the challenging optimization problem associated with the proposed estimator. Numerical experiments showcase the effectiveness of the data labeling paradigm.

Marrinan, Timothy P.↗