Search NASA⌕ Search

SEARCH · Search NASA

Results for “statistical analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 397 records · Page 22

The National Climate Database (NCDB): An Unbiased 100-Year Dataset for PV Modeling

In this study, we develop a statistical technique to downscale the future projection of solar irradiance for photovoltaics (PV) energy-related applications. A set of Regional Climate Model (RCM)-based projections obtained from the North American Coordinated Regional Climate Downscaling Experiment (NA-CORDEX) are used as inputs to statistical methods to generate high-resolution global horizontal irradiance (GHI) over the contiguous United States (CONUS). The main steps of the statistical downscaling method include (1) regridding RCM output (0.22 degree and daily resolutions) to handle the modeled-observed data sets on a common grid, (2) correcting bias of RCM GHI using satellite-derived observation, and (3) implementing temporal and spatial downscaling to generate GHI at 8-km and hourly resolution. Basically, complex physical processes and interactions between solar radiation and various atmospheric constituents lead solar irradiance to be highly variable and uncertain. Underrepresentation of clouds from the RCM parameterizations is the main source of error and uncertainty in modeling solar irradiance. Thus, we adapt and use the high-quality satellite-derived data from the National Solar Radiation Database (NSRDB) to analyze the bias and error of RCM GHI as well as estimate the statistical parameters for spatial and temporal downscaling. This presentation will summarize the comprehensive analysis conducted to produce and assess the results under two climate scenarios (RCP4.5 and RCP8.5). We will also present a detailed validation demonstrating the strengths of the downscaling method, a summary of the 100-year dataset from 2001-2100, and future extension of this research.

bias correction↗

Measurement of the branching fraction, polarization, and time-dependent C P asymmetry in B 0 → ρ + ρ − decays and constraint on the CKM angle ϕ 2

We present a measurement of the branching fraction and fraction of longitudinal polarization of B 0 → ρ + ρ − decays, which have two π 0 ’s in the final state. We also measure time-dependent C P violation parameters for decays into longitudinally polarized ρ + ρ − pairs. This analysis is based on a data sample containing ( 387 ± 6 ) × 10 6 ϒ ( 4 S ) mesons collected with the Belle II detector at the SuperKEKB asymmetric-energy e + e − collider in 2019–2022. We obtain B ( B 0 → ρ + ρ − ) = ( 2.8 9 − 0.22 + 0.23 − 0.27 + 0.29 ) × 10 − 5 , f L = 0.92 1 − 0.025 + 0.024 − 0.015 + 0.017 , S = − 0.26 ± 0.19 ± 0.08 , and C = − 0.02 ± 0.1 2 − 0.05 + 0.06 , where the first uncertainties are statistical and the second are systematic. We use these results to perform an isospin analysis to constrain the Cabibbo-Kobayashi-Maskawa angle ϕ 2 and obtain two solutions; the result consistent with other Standard Model constraints is ϕ 2 = ( 92.6 − 4.7 + 4.5 ) ° . Published by the American Physical Society 2025

Adachi, I. (ORCID:0000000322870173)↗

A data-driven framework for predicting machining stability: employing simulated data, operational modal analysis, and enhanced transfer learning

Chatter, a self-excited vibration phenomenon, presents a significant challenge in machining operations, particularly in high-speed milling, where it can degrade tool life, reduce material removal efficiency, and compromise workpiece quality. Addressing this challenge requires a reliable predictive model that can accommodate the complex dynamics of various machining scenarios. This study introduces a novel, data-driven approach to predicting machining stability, leveraging over 140,000 simulated datasets and employing advanced techniques such as operational modal analysis (OMA), enhanced transfer learning (TL), and receptance coupling substructure analysis (RCSA). By integrating these methodologies, the framework effectively classifies and predicts chatter across diverse operational modes, achieving robust and accurate outcomes. Our model utilizes a Random Forest (RF) classifier trained with the comprehensive dataset, which demonstrates substantial improvements in both predictive accuracy and robustness. Specifically, the RF model achieved an accuracy rate of 85%, an area under the curve (AUC) of 0.90, and an F1 score of 0.88, underscoring its capability to adapt to varying machining configurations. These results highlight the framework’s potential to enhance operational efficiency and machining quality by providing reliable chatter predictions across a broad range of machining parameters. In conclusion, this research thus offers a significant advancement in predictive maintenance for machining processes, enabling more stable and efficient manufacturing operations.

42 ENGINEERING↗

How are Heterogeneous Nucleation Rate Observations Influenced by Instrument Resolution?

Experimental measurements of the heterogeneous nucleation rate rely on counting the number of nuclei with time. However, the size of a thermodynamically stable nucleus is often a few nanometers in diameter and is below the resolution of most (in situ) measurement techniques that provide a statistically valid sample. Due to the finite resolution of the instruments and analysis methods, it is challenging to capture the incipient nuclei and the subsequent evolution of nuclei density over time. In this work, we demonstrate the impact of instrument resolution on observed nuclei densities by comparing numerical modeling with experimental results. Further, to achieve this, we implemented heterogeneous nucleation within the pore-scale reactive transport modeling framework using classical nucleation theory (CNT). We compared the modeling results with nucleation rates measured using X-ray nanotomography (XnT) and evaluated how these impact the apparent values of the prefactor and interfacial energy based on CNT and the crystal growth rate. Specifically, we applied a resolution threshold (artificial resolution limit) in the model during nuclei counting to resemble an experimental resolution, ranging from 15 to 500 nm. The findings reveal that the instrument resolution significantly impacts the apparent prefactor and interfacial energy. Both apparent prefactor and interfacial energy decrease with a decrease in the instrument resolution. While deviation in the prefactor due to resolution is anticipated, those in the interfacial energy are unexpected. The approach described here allows one to correct apparent nucleation rates that depend on the instrument’s resolution to derive “intrinsic” CNT parameters for the prefactor and interfacial energy.

47 OTHER INSTRUMENTATION↗

Causal Directions Matter: How Environmental Factors Drive Convective Cloud Detrainment Heights

This study investigates how environmental factors influence the level of maximum detrainment (LMD) in deep convective clouds. Through a novel application of the Linear Non‐Gaussian Acyclic Model (LiNGAM), we discover causal structures between environmental variables and LMD, observed at six tropical sites operated by the Atmospheric Radiation Measurement (ARM) user facility. LiNGAM effectively identifies causal directions among variables of interest, revealing robust relationships such as those among the lifting condensation level (LCL), level of free convection (LFC), and convective inhibition (CIN), aligning with prior knowledge. Relative humidity is shown to directly influence LMD; however, this relationship exhibits strong nonlinearity and becomes difficult to detect when the contrast between oceanic and continental environments is excluded from the analysis. This study highlights the importance of establishing causal relationships before performing statistical inference.

54 ENVIRONMENTAL SCIENCES↗

Are light curve classification metrics good proxies for SN Ia cosmological constraining power?

Context. When selecting a light curve classifier for use as part of a photometric supernova Ia (SN Ia) cosmological analysis, it is common to make decisions based on metrics of classification performance, such as the contamination within the photometrically classified SN Ia sample, rather than a measure of cosmological constraining power. If the former is an appropriate proxy for the latter, this practice would eliminate the computational expense of a full cosmology forecast in the analysis pipeline design process. Aims. This study tests the assumption that light curve classification metrics are an appropriate proxy for cosmology metrics. Methods. We emulated photometric SN Ia cosmology light curve samples with controlled contamination rates of individual contaminant classes and evaluated each of them under a set of classification metrics. We then derived cosmological parameter constraints from all samples under two common analysis approaches and quantified the impact of contamination by each contaminant class on the resulting cosmological parameter estimates. Results. We observe that cosmology metrics are sensitive to both the contamination rate and the class of the contaminating population, whereas the classification metrics are shown to be insensitive to the latter. Conclusions. Based on these findings, we discourage any exclusive reliance on light curve classification-based metrics for analysis design decisions, which (counterintuitively) include but are not limited to the classifier choice. Instead, we recommend optimising science analysis pipeline design choices using a metric of the information gained about the physical parameters of interest.

79 ASTRONOMY AND ASTROPHYSICS↗

Robust error calibration for serial crystallography

Serial crystallography is an important technique with unique abilities to resolve enzymatic transition states, minimize radiation damage to sensitive metalloenzymes and perform de novo structure determination from micrometre-sized crystals. This technique requires the merging of data from thousands of crystals, making manual identification of errant crystals unfeasible. cctbx.xfel.merge uses filtering to remove problematic data. However, this process is imperfect, and data reduction must be robust to outliers. We add robustness to cctbx.xfel.merge at the step of uncertainty determination for reflection intensities. This step is a critical point for robustness because it is the first step where the data sets are considered as a whole, as opposed to individual lattices. Robustness is conferred by reformulating the error-calibration procedure to have fewer and less stringent statistical assumptions and incorporating the ability to down-weight low-quality lattices. We then apply this method to five macromolecular XFEL data sets and observe the improvements to each. The appropriateness of the intensity uncertainties is demonstrated through internal consistency. This is performed through theoretical CC 1/2 and I /σ relationships and by weighted second moments, which use Wilson's prior to connect intensity uncertainties with their expected distribution. This work presents new mathematical tools to analyze intensity statistics and demonstrates their effectiveness through the often underappreciated process of uncertainty analysis.

Mittan-Moreau, David W.↗

CORRLA-RS

The CORRLA-RS package provides a suite of statistical methods for sampling multidimensional distributions and to conduct sensitivity and correlation analysis of large scale data in the Rust programming language. The software provides a unique solution to multidimensional constrained sampling problems utilizing a combination of parallelized Markov Chain Monte Carlo methods and traditional rejection sampling. The sensitivity and correlation analysis methods are backed by a high performance randomized singular value decomposition implementation which enables datasets larger than the random access memory (RAM) size to be analyzed. Additionally, CORRLA-RS implements the active subspace identification method using a KD-Tree and the randomized singular value decomposition acting in concert.

Gurecky, William [Oak Ridge National Laboratory (O↗

A Unified Photometric Redshift Calibration for Weak Lensing Surveys Using the Dark Energy Spectroscopic Instrument

The effective redshift distribution n(z) of galaxies is a critical component in the study of weak gravitational lensing. Here, we introduce a new method for determining n(z) for weak lensing surveys based on high-quality redshifts and neural-network-based importance weights. Additionally, we present the first unified photometric redshift calibration of the three leading stage-III weak lensing surveys, the Dark Energy Survey (DES), the Hyper Suprime-Cam (HSC) survey, and the Kilo-Degree Survey (KiDS), with state-of-the-art spectroscopic data from the Dark Energy Spectroscopic Instrument (DESI). We verify our method using a new, data-driven approach and obtain n(z) constraints with statistical uncertainties of the order of $σ_z$ ~ 0.01 and smaller. Our analysis is largely independent of previous photometric redshift calibrations and, thus, provides an important cross-check in light of recent cosmological tensions. Overall, we find excellent agreement with previously published results on the DES Y3 and HSC Y1 data sets, while there are some differences on the mean redshift with respect to the previously published KiDS-1000 results. We attribute the latter to mismatches in photometric noise properties in the COSMOS field compared to the wider KiDS self-organizing map-gold catalog. At the same time, the new n(z) estimates for KiDS do not significantly change estimates of cosmic structure growth from cosmic shear. Finally, we discuss how our method can be applied to future weak lensing calibrations with DESI data.

Lange, J. U. [American Univ., Washington, DC (Unit↗

Smokescreen: A Python package for data vector blinding and encryption in cosmological analyses

Smokescreen is an open-source Python library for data-vector concealment (blinding) in cosmological analyses. Data-vector blinding works by applying cosmology-dependent shifts to the observed data vector, moving it away from the true cosmological signal without affecting its statistical properties, so that analysts cannot infer the true result until the analysis is frozen and the blinding is lifted. The package computes these shifts using Firecrown likelihoods applied to data vectors stored in the SACC format, ensuring that the theoretical model used for blinding is identical to that used for inference whilst remaining agnostic to the specific observable being blinded. To prevent accidental unblinding, the original SACC file, containing the true cosmology, is encrypted. Although developed for the Vera C. Rubin Observatory Legacy Survey of Space and Time (LSST), Smokescreen is applicable to any experiment using Firecrown likelihoods and the SACC data format.

Loureiro, Arthur [Stockholm U., OKC; Imperial Coll↗

Fuel performance analysis of fully-resolved TRISO compact

The TRi-structural ISOtropic (TRISO) fuel multilayered coating structure offers multiple barriers to fission product release, enhancing safety and performance. The heterogeneous nature of TRISO fuel compacts, comprising thousands of randomly distributed coated fuel particles embedded in a graphite matrix, creates intricate stress fields and thermal gradients that cannot be accurately modeled using simplified one-dimensional or homogenized approaches. Consequently, three-dimensional modeling enables the prediction of fuel compact dimensional changes, internal pressure buildup, and fission product transport pathways under diverse irradiation and thermal conditions. This capability facilitates detailed analysis of particle-to-particle interactions, matrix cracking mechanisms, and the statistical distribution of coating failures, which directly impact fuel performance and safety margins. This capability is particularly critical for advanced reactors, such as high-temperature gas-cooled reactors and other Generation IV reactor designs where TRISO fuel operates at elevated temperatures and burn-up levels. This work introduces a novel method to generate an optimized packing of TRISO compacts and a complete 3D mesh with random distribution of TRISO particles, which are discretized into each coating component layer.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

3D TRISO particle-explicit compact meshing

The TRI-structural ISOtropic (TRISO) layered fuel particle is a robust nuclear fuel form offering enhanced safety and performance for advanced reactor concepts, including high-temperature gas-cooled reactors and other Generation IV designs. These poppy-seed-sized particles are embedded in a graphite matrix to form fuel elements that must withstand elevated temperatures and high burn-up levels. The heterogeneous nature of these fuel elements — comprising thousands of randomly distributed TRISO particles — produces complex stress fields and thermal gradients that one- and two-dimensional models cannot accurately capture. While three-dimensional modeling has improved predictions of dimensional changes, internal pressure buildup, and fission product transport under irradiation, current approaches rely on homogenized material properties that are known to have considerable divergence from experimental observations. This work presents a methodology for optimized random packing of TRISO fuel compacts and full three-dimensional mesh generation within the BISON fuel performance code, with each particle coating layer individually discretized. The resulting mesh was demonstrated through heat conduction simulations under representative in-reactor operating conditions, showing strong agreement with expected behavior. This capability enables detailed analysis of particle-to-particle interactions, matrix cracking mechanisms, and the statistical distribution of coating layer failures — all of which directly govern fuel performance and safety margins.

11 - NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

Data-Driven Insights into the Structural Essence of Plasticity in High-Entropy Alloys

The heterogeneous mechanical response of a crystalline alloy with multiple principal elements was investigated using molecular dynamics simulations. The local configuration of the alloy in its quiescent state was characterized by the variables derived from the gyration tensor and the atomic electronegativity. A multivariate analysis identified the geometric and chemical factors that influenced the atomic packing variations. Further, upon straining, the non-affine displacement exhibited spatial heterogeneity. A statistical correlation was established between the local yield events and the specific features of the local configuration. Our findings, validated by the performance metrics analysis, provided a structural criterion for the instability mechanisms in high-entropy alloys (HEAs) and enhanced the understanding of their plasticity.

36 MATERIALS SCIENCE↗

Measuring Neutron Polarisation in Deuteron Photo-disintegration with the CLAS Start Counter [Thesis]

Deuteron photo-disintegration (γd → γp) is a reaction that represents the simplest case in which nuclear and hadron physics models can be tested. Despite this, associated polarization analyses are limited in terms of angular coverage and energy ranges, especially in observables related to the recoil neutron. This is largely due to a lack in dedicated polarimetry equipment, and represents a roadblock in global progress to understand high-energy phenomena such as hexaquarks, and quark-gluon degrees of freedom. To address this problem, this PhD thesis pioneers a new methodology for the parasitic measurement of nucleon polarization using kinematic reconstruction of (spin-dependent) nucleon-nucleus scattering of reaction products, prior to their detection in large acceptance particle detector apparatus. Following this novel approach, which requires no dedicated polarimeter, a determination of the double polarization observable, $C^n_{x'}$, from deuteron photo-disintegration is presented, using Jefferson Lab’s CLAS detector. The analysis utilizes the (n,p) charge exchange reaction in CLAS’s "start counter" (plastic scintillator) to determine the final state neutron polarizations. The results present the first ever data for this observable above 0.7 GeV (photon beam energy) and significantly extend the angular range of the world data set. This new data is largely statistically consistent with the previous measurement of $C^n_{x'}$ by Bashkanov et al . in the overlapping energy range of 0.4-0.7 GeV. It is planned for the statistical accuracy of the presented result to be increased by the inclusion of additional data. The analysis herein serves as a key proof of concept for future applications, including a recommended similar analysis to be implemented with data from the more modern CLAS12 detector. This paves the way for a plethora of additional analyses using existing data sets that would provide crucial new constraints for hadron and nuclear physics.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Detecting Masquerade Attacks in Controller Area Networks Using Graph Machine Learning

Modern vehicles rely on a myriad of electronic control units (ECUs) interconnected via controller area networks (CANs) for critical operations. Despite their ubiquitous use and reliability, CANs are susceptible to sophisticated cyberattacks, particularly masquerade attacks, which inject false data that mimic legitimate messages at the expected frequency. These attacks pose severe risks such as unintended acceleration, brake deactivation, and rogue steering. Traditional intrusion detection systems (IDS) often struggle to detect these subtle intrusions due to their seamless integration into normal traffic. This paper introduces a novel framework for detecting masquerade attacks in the CAN bus using graph machine learning (ML). We hypothesize that the integration of shallow graph embeddings with time series features derived from CAN frames enhances the detection of masquerade attacks. We show that by representing CAN bus frames as message sequence graphs (MSGs) and enriching each node with contextual statistical attributes from time series, we can enhance detection capabilities across various attack patterns compared to using graph-based features only. Our method ensures a comprehensive and dynamic analysis of CAN frame interactions, improving robustness and efficiency. Extensive experiments on the ROAD dataset validate the effectiveness of our approach, demonstrating statistically significant improvements in the detection rates of masquerade attacks compared to a baseline that uses graph-based features only as confirmed by Mann-Whitney U and Kolmogorov-Smirnov tests (p < 0.05) .

Marfo, William [Univ. of Texas, El Paso, TX (Unite↗

Experimental and statistical study on the effect of process parameters on the quality of continuous fiber composites made via additive manufacturing

Ongoing research in additive manufacturing towards structural and industrial application has led to the use of commingled roving as a manufacturing feedstock for printing high fiber volume fraction composites. The prospects of using this technology for high performance applications necessitates the need for a comprehensive experimental investigation into the effects of processing parameters on the quality of an additively manufactured composite printed from commingled roving feedstock. Here, in this work, transverse flexure and void fraction matrix pyrolysis testing are both performed to evaluate composite quality. The transverse flexure test is a testing approach that evaluates the quality of the interfacial fiber-matrix bond while the void fraction test estimates the void content in the printed composite. A full observational study consisting of 27 different test combinations is done to investigate the effects of three different process parameters namely, temperature, pressure, and print speed across three different levels. Composite samples were made from commingled roving of E-glass and amorphous PET using an in-house built continuous fiber composite digital manufacturing system. Least squares regression analysis is performed to study the main, interaction and quadratic effects of process parameters. A statistical regression model having an R2 adjusted value of 80.1% is generated from the transverse flexure study, which is used to explain main and interaction effects and also predict performance. Response surface plots are also generated and are used to optimize process parameters which can subsequently be of help in scaling up composite manufacturing. Results show that all three process parameters are highly statistically significant at the 0.01 level of significance. Pressure * Temperature and Pressure * Printspeed are significant interaction terms. Pressure plays a weightier role when print speed is increased or temperature is decreased as it closes more voids that would ordinarily have been introduced because of drop in polymer melt viscosity. Micrographic analysis is also performed.

36 MATERIALS SCIENCE↗

Data and scripts associated with a manuscript on a meta-analysis synthesizing stream biogeochemical response to wildfires across space and time (v2)

This data package is associated with the publication “Catchment characteristics modulate the influence of wildfires on nitrate and dissolved organic carbon in lotic systems across space and time: A meta-analysis” submitted to Global Biogeochemical Cycles (Cavaiani et al. 2025). This study uses meta-analytical techniques to evaluate the effect of wildfire on in-stream responses in burned and unburned watersheds. The study aims to provide additional insight into the range of responses and net influences that wildfires have on hydro-biogeochemistry across broad spatial scales, burn extents, and the persistence of water-quality change. This study compiles data and metadata from 18 total publications that includes 1) surface water geochemistry data (dissolved organic carbon; nitrate), 2) climate classifications, 3) year of the wildfire, 4) the time lag between when the fire occurred and when the sampling occurred, and 5) study design of the publication. In total, this meta-analysis draws data that spans 8 climate guilds, 3 biomes, 62 watersheds, and 20 unique wildfires. See Sites_meta_data.csv for citations of the papers used in this meta-analysis. All R scripts and the associated data can also be found on GitHub at This data package was originally published in March 2024. It was updated in April 2025 (v2; new and modified files). See the change history section in the readme for more details. This data package contains five primary folders that include the following: (1) inputs; (2) output for analysis; (3) initial plots; (4) R scripts; and (5) GIS data. The data package also contains a data dictionary (dd) that provides column header definitions and a file-level metadata (flmd) file that describes every file. The “inputs” folder contains a list of all publications identified during the formal web search and an indication of whether each publication was included in the final analysis. Additionally, it includes site-level metadata, catchment characteristics, and GIS data for all publications included in the final analysis. The “Output_for_analysis” folder contains all data frames and figures generated from each R script used for additional data analysis. The “initial_plots” folder includes all exploratory figures that will be included in a supplemental and figures that will be submitted with the manuscript for publication. The “R_scripts” folder contains the scripts that perform all the data manipulations, statistical analyses, and plots. The “gis_data” folder includes shape files for each fire included in this meta-analysis. This data package contains the following file types: csv, pdf, jpeg, cpg, dbf, prj, shp, shp.ea.iso.xml, shp.iso.xml, shx.

54 ENVIRONMENTAL SCIENCES↗

Compressed baryon acoustic oscillation analysis is robust to modified-gravity models

Abstract We study the robustness of the baryon acoustic oscillation (BAO) analysis to the underlying cosmological model. We focus on testing the standard BAO analysis that relies on the use of a template. These templates are constructed assuming a fixed fiducial cosmological model and used to extract the location of the acoustic peaks. Such “compressed analysis” had been shown to be unbiased when applied to the ΛCDM model and some of its extensions. However, it has not been known whether this type of analysis introduces biases in a wider range of cosmological models where the template may not fully capture relevant features in the BAO signal. In this study, we apply the compressed analysis to noiseless mock power spectra that are based on Horndeski models, a broad class of modified-gravity theories specified with eight additional free parameters. We study the precision and accuracy of the BAO peak-location extraction assuming DESI, DESI II, and MegaMapper survey specifications. We find that the bias in the extracted peak locations is negligible; for example, it is less than 10% of the statistical error for even the proposed future MegaMapper survey. Our findings indicate that the compressed BAO analysis is remarkably robust to the underlying cosmological model.

Astronomy & Astrophysics↗