Search NASA⌕ Search

SEARCH · Search NASA

Results for “statistical analysis”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 433 records · Page 24

Cosmological constraints from density-split clustering in the BOSS CMASS galaxy sample

We present a clustering analysis of the BOSS DR12 CMASS galaxy sample, combining measurements of the galaxy two-point correlation function and density-split clustering down to a scale of $1 \, h^{-1}\, \text{Mpc}$. Our theoretical framework is based on emulators trained on high-fidelity mock galaxy catalogues that forward model the cosmological dependence of the clustering statistics within an extended-ΛCDM framework, including redshift-space and Alcock–Paczynski distortions. Our base-ΛCDM analysis finds ω cdm = 0.1201 ± 0.0022, σ 8 = 0.792 ± 0.034, and n s = 0.970 ± 0.018, corresponding to fσ 8 = 0.462 ± 0.020 at z ≈ 0.525, which is in agreement with Planck 2018 predictions and various clustering studies in the literature. We test single-parameter extensions to base-ΛCDM, varying the running of the spectral index, the dark energy equation of state, and the density of mass-less relic neutrinos, finding no compelling evidence for deviations from the base model. We model the galaxy–halo connection using a halo occupation distribution framework, finding signatures of environment-based assembly bias in the data. We validate our pipeline against mock catalogues that match the clustering and selection properties of CMASS, showing that we can recover unbiased cosmological constraints even with a volume 84 times larger than the one used in this study.

79 ASTRONOMY AND ASTROPHYSICS↗

Hyperspectral Detection of the Fluorescence Shift between Chirality-Sorted Empty and Water-Filled Single-Wall Carbon Nanotube Enantiomers

Single-wall carbon nanotubes (SWCNTs) have extraordinary electronic and optical properties that depend strongly on their exact chiral structure and their interaction with their inner and outer environment. The fluorescence (PL) of semiconducting SWCNTs, for instance, will shift depending on the molecules with which the SWCNT’s hollow core is filled. These interaction-induced shifts are challenging to resolve on the ensemble level in samples containing a mixture of different filling contents due to the relatively large inhomogeneous line width of the ensemble SWCNT PL compared to the size of these shifts. To circumvent this inhomogeneous broadening, single-tube spectroscopy and hyperspectral imaging are often applied, which until now required time-consuming statistical studies. Here, we present hyperspectral PL microscopy combined with automated SWCNT segmenting based on either principal component analysis or a convolutional neural network, capable of both spatially and spectrally resolving the PL along the length of many individual SWCNTs at the same time and automatically fitting peak positions and line widths of individual SWCNTs. The methodology is demonstrated by accurately determining the emission shifts and line widths of thousands of left- and right-handed empty and water-filled SWCNTs coated with a chiral surfactant, resulting in four statistical distributions which cannot be resolved in ensemble spectroscopy of unsorted samples. The results demonstrate a robust method to quickly probe ensemble properties with single-enantiomer spectral resolution. Moreover, it promises to be an absolute quantitative method to characterize the relative abundances of SWCNTs with different handedness or filling content in macroscopic samples, simply by counting individual species.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Dark energy survey: Modeling strategy for multiprobe cluster cosmology and validation for the full six-year dataset

Here, we introduce an updated To&Krause2021 model for joint analyses of cluster abundances and large-scale two-point correlations of weak lensing and galaxy and cluster clustering (termed CL+3×2 pt analysis) and validate that this model meets the systematic accuracy requirements of analyses with the statistical precision of the final Dark Energy Survey (DES) Year 6 (Y6) dataset. The validation program consists of two distinct approaches, (i) identification of modeling and parametrization choices and impact studies using simulated analyses with each possible model misspecification and (ii) end-to-end validation using mock catalogs from customized Cardinal simulations that incorporate realistic galaxy populations and DES-Y6-specific galaxy and cluster selection and photometric redshift modeling, which are the key observational systematics. In combination, these validation tests indicate that the model presented here meets the accuracy requirements of DES-Y6 for CL+3×2 pt based on a large list of tests for known systematics. In addition, we also validate that the model is sufficient for several other data combinations: the CL+GC subset of this data vector (excluding galaxy–galaxy lensing and cosmic shear two-point statistics) and the CL+3×2 pt+BAO+SN (combination of CL+3×2 pt with the previously published Y6 DES baryonic acoustic oscillation and Y5 supernovae data).

79 ASTRONOMY AND ASTROPHYSICS↗

Benchmarking large language models for materials synthesis: The case of atomic layer deposition

In this work, we introduce an open-ended question benchmark, ALDbench, to evaluate the performance of large language models (LLMs) in materials synthesis, and, in particular, in the field of atomic layer deposition, a thin film growth technique used in energy applications and microelectronics. Our benchmark comprises questions with a level of difficulty ranging from the graduate level to domain expert current with the state of the art in the field. Human experts reviewed the questions along the criteria of difficulty and specificity, and the model responses along four different criteria: overall quality, specificity, relevance, and accuracy. We ran this benchmark on an instance of OpenAI’s GPT-4o. The responses from the model received a composite quality score of 3.7 on a 1–5 scale, consistent with a passing grade. However, 36% of the questions received at least one below average score. An in-depth analysis of the responses identified at least five instances of suspected hallucination. Finally, we observed statistically significant correlations between the difficulty of the question and the quality of the response, the difficulty of the question and the relevance of the response, the specificity of the question, and the accuracy of the response as graded by the human experts. Furthermore, this emphasizes the need to evaluate LLMs across multiple criteria beyond difficulty or accuracy.

Artificial intelligence↗

Highly Resolved Reference Projections of Building Energy Use for the Contiguous United States: Building Sector Energy Baselines, Projection Methods, and Results

This report describes one methodology of projecting energy consumption of the US residential and commercial building sectors using NREL's ResStock™ and ComStock™ as well as growth rates derived from EIA's Annual Energy Outlook (AEO). The impetus for this work is to provide an intermediate method for compiling demand-side sectoral energy projections that is suitable for grid-scale analysis, such as NREL's Standard Scenarios. ResStock and ComStock are physics-based and statistically representative building stock models of the US residential and commercial sector, respectively. Using the 2012 actual meteorological year (AMY) weather data, the sectoral energy baselines are simulated and then segmented along key dimensions (e.g., geography, dwelling/building type). The segmented results are then scaled using the corresponding annual growth rates derived from the 2021 AEO reference case to produce energy projections out to 2050. The compiled result is a demand-side grid model (dsgrid) data set suitable for use in NREL's large-scale grid models, such as the Regional Energy Deployment System (ReEDS). This simple projection method does not endogenously represent how the building stock could evolve through time. Most notably, it does not reflect large-scale electrification, for example, the conversion of space heating, water heating, clothes drying, and cooking from primary fossil fuels to electricity, as this is not part of AEO's reference case assumptions. Nonetheless this approach is more resolved and potentially extensible compared to the current method used by Standard Scenarios's reference case, which augments a sector's total load based on a single growth rate from AEO.

32 ENERGY CONSERVATION, CONSUMPTION, AND UTILIZATI↗

Dark Energy Survey: Modeling strategy for multiprobe cluster cosmology and validation for the Full Six-year Dataset

We introduce an updated To&Krause2021 model for joint analyses of cluster abundances and large-scale two-point correlations of weak lensing and galaxy and cluster clustering (termed CL+3x2pt analysis) and validate that this model meets the systematic accuracy requirements of analyses with the statistical precision of the final Dark Energy Survey (DES) Year 6 (Y6) dataset. The validation program consists of two distinct approaches, (1) identification of modeling and parameterization choices and impact studies using simulated analyses with each possible model misspecification (2) end-to-end validation using mock catalogs from customized Cardinal simulations that incorporate realistic galaxy populations and DES-Y6-specific galaxy and cluster selection and photometric redshift modeling, which are the key observational systematics. In combination, these validation tests indicate that the model presented here meets the accuracy requirements of DES-Y6 for CL+3x2pt based on a large list of tests for known systematics. In addition, we also validate that the model is sufficient for several other data combinations: the CL+GC subset of this data vector (excluding galaxy--galaxy lensing and cosmic shear two-point statistics) and the CL+3x2pt+BAO+SN (combination of CL+3x2pt with the previously published Y6 DES baryonic acoustic oscillation and Y5 supernovae data).

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

An R Shiny graphical user interface for analyzing, visualizing, and interpreting high precision mass spectrometric data

There is currently a lack of software that meets the needs for the analysis of raw data produced by modern isotope ratio mass spectrometers for both R&D and routine use at SRNL and other US national labs • Needs to accommodate multiple isotope systems, instruments, and manufacturers • Include modern statistical methods and handling/visualization of uncertainty • Flexible software with transparent (no “black box”) and reproducible methods • This project is inspired by existing discipline-specific data analysis software (e.g., Tripoli1 , ET_Redux2 , IsoplotR3) used in the geochemical community • Our goal is to build an open source data analysis software package that focuses on flexibility, transparency, and reproducibility

Labone, Elizabeth↗

An R shiny graphical user interface for highprecision mass spectrometric data analysis

• There is currently a lack of software that meets the needs for the analysis of raw data produced by modern isotope ratio mass spectrometers for both R&D and routine use at SRNL and other US national labs • Needs to accommodate multiple isotope systems, instruments, and manufacturers • Include modern statistical methods and handling/visualization of uncertainty • Flexible software with transparent (no “black box”) and reproducible methods • This project is inspired by existing discipline-specific data analysis software (e.g., Tripoli1 , ET_Redux2, IsoplotR3) used in the geochemical community • Our goal is to build an open source data analysis software package that focuses on flexibility, transparency, and reproducibility

LABONE, ELIZABETH↗

High-count-rate effects in event processing for the XRISM/Resolve X-ray microcalorimeter. II. Energy scale and resolution in orbit

The Resolve instrument on the X-ray Imaging and Spectroscopy Mission (XRISM) uses a 36 pixel microcalorimeter designed to deliver high-resolution, non-dispersive X-ray spectroscopy. Although it is optimized for extended sources with low count rates, Resolve observations of bright point sources are still able to provide unique insights into the physics of these objects, as long as high-count-rate effects are addressed in the analysis. These effects include the loss of exposure time for each pixel, changes in the energy scale, and changes in the energy resolution. To investigate these effects under realistic observational conditions, we observed the bright X-ray source, the Crab Nebula, with XRISM at several offset positions with respect to the Resolve field of view and with continuous illumination from 55 Fe sources on the filter wheel. For the spectral analysis, we excluded data where exposure-time loss was too significant to ensure reliable spectral statistics. The energy scale at 6 keV shows a slight negative shift in the high-count-rate regime. The energy resolution at 6 keV worsens as the count rate in electrically neighboring pixels increases, but can be restored by applying a nearest-neighbor coincidence cut (“cross-talk cut”). We examined how these effects influence the observation of bright point sources, using GX 13+1 as a test case, and identified an eV-scale energy offset at 6 keV between the inner (brighter) and outer (fainter) pixels. Users who seek to analyze velocity structures on the order of tens of km s–1 should account for such high-count-rate effects. These findings will aid in the interpretation of Resolve data from bright sources and provide valuable considerations for designing and planning for future microcalorimeter missions.

X-rays: general↗

Impacts of PV Module Connector Failures on Cost and Performance of Utility Scale Photovoltaic Systems

The reliability, cost and performance of electrical connectors are a concern in all types of electrical systems, and demands on connectors used on photovoltaic (PV) systems include that connectors maintain electrical conductivity and physical strength, endure ultraviolet sunlight and high ambient temperature, and resist moisture and chemical intrusion over a very long (>25 year) performance period. Connector failures increase operation and maintenance (O&M) costs and reduce plant production, but connector failure can also cause safety and liability problems, which are of greater concern. This work results from a three-year collaboration between Sandia National Laboratories (SNL), the Electric Power Research Institute (EPRI), and the National Renewable Energy Laboratory (NREL) and funded by the U.S. Department of Energy (DOE) Solar Energy Technology Office (SETO) under Agreements #39035 and #38531 "Connector Reliability Across the US Solar Sector." a multi-pronged investigation of PV connector health across the US (see https://energy.sandia.gov/pvconnectors/). This report presents derivation of a Techno-Economic Analysis (TEA) that models failure modes and frequencies (how often failure occurs), estimates O&M costs and lost production associated with connector failures, and then calculates the effect that PV module connectors can have on Levelized Cost of Energy (LCOE). The model is informed with initial data from quantitative assessment of failure rates, root causes and mechanisms, in-situ diagnostics and data collection, lab-based forensics, and interviews with PV connector manufacturers and plant operators. SNL conducted site inspections at multiple utility-scale sites in different climates and subjected field samples of new, used, and degraded connectors to visual and electrical characterization. EPRI conducted metallurgical analysis of the pin and sleeve conductors to study failure-induced morphological and compositional changes. There is in general a shortage of statistically valid data, but data from PVROM database maintained by SNL was sufficient to ascertain failure rates and lost production as well as provide qualitative insight in its curated maintenance records. This report details the structure of the mathematical model but the sources of data to inform the model will continue to evolve. Analysis of a 100 MW PV plant is provided as an example of the use of the model, with results indicating that connectors are responsible for Annualized O&M Costs of $\$$71,933/year; Annualized Unit O&M Costs of $\$$0.72/kW/year; that a Reserve Account of $\$$187,220 should be available to fund repairs related to connectors; that connectors add $\$$1,494,004 to the Net Present Value of the O&M Costs (project life); and that O&M related to connectors adds about $\$$0.00088/kWh to the Levelized Cost of Energy. The impact of this model is to provide a tool to make the US solar sector more robust by quantifying and monetizing the reliability risks to utility-scale PV systems posed by poorly installed, mismatched and/or poorly designed and manufactured connectors. The TEA provides a model incorporating failure statistics, O&M cost data, and lost production into a single figure of merit, informing decisions and enabling practitioners to optimize cost and performance trade-offs. Stakeholders include connector manufacturers, system designers and equipment specifiers, standards bodies, installers and O&M providers, investors and insurance underwriters. This report supports continued growth of PV predicated on assurances that properly installed and maintained PV system connectors are safe and reliable. The project team is proposing future work including accelerated testing of connectors and expanding the approach taken here to other PV system components, such as TEA for rapid shut-down devices.

14 SOLAR ENERGY↗

An implementation of neural simulation-based inference for parameter estimation in ATLAS

Neural simulation-based inference (NSBI) is a powerful class of machine-learning-based methods for statistical inference that naturally handles high-dimensional parameter estimation without the need to bin data into low-dimensional summary histograms. Such methods are promising for a range of measurements, including at the Large Hadron Collider, where no single observable may be optimal to scan over the entire theoretical phase space under consideration, or where binning data into histograms could result in a loss of sensitivity. This work develops a NSBI framework for statistical inference, using neural networks to estimate probability density ratios, which enables the application to a full-scale analysis. It incorporates a large number of systematic uncertainties, quantifies the uncertainty due to the finite number of events in training samples, develops a method to construct confidence intervals, and demonstrates a series of intermediate diagnostic checks that can be performed to validate the robustness of the method. As an example, the power and feasibility of the method are assessed on simulated data for a simplified version of an off-shell Higgs boson couplings measurement in the four-lepton final states. This approach represents an extension to the standard statistical methodology used by the experiments at the Large Hadron Collider, and can benefit many physics analyses.

frequentist statistics↗

Molecular Dynamics Simulation of Complex Reactivity with the Rapid Approach for Proton Transport and Other Reactions (RAPTOR) Software Package

Simulating chemically reactive phenomena such as proton transport on nanosecond to microsecond and beyond time scales is a challenging task. Ab initio methods are unable to currently access these time scales routinely, and traditional molecular dynamics methods feature fixed bonding arrangements that cannot account for changes in the system’s bonding topology. The Multiscale Reactive Molecular Dynamics (MS-RMD) method, as implemented in the Rapid Approach for Proton Transport and Other Reactions (RAPTOR) software package for the LAMMPS molecular dynamics code, offers a method to routinely sample longer time scale reactive simulation data with statistical precision. RAPTOR may also be interfaced with enhanced sampling methods to drive simulations toward the analysis of reactive rare events, and a number of collective variables (CVs) have been developed to facilitate this. Key advances to this methodology, including GPU acceleration efforts and novel CVs to model water wire formation are reviewed, along with recent applications of the method which demonstrate its versatility and robustness.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Determination of proton PDF uncertainties with Markov chain Monte Carlo

We present an analysis of parton distribution functions (PDFs) of the proton using Markov chain Monte Carlo (MCMC) methods. The MCMC approach naturally implements Bayes’ theorem and, thus, provides a means to directly sample the underlying probability distribution—in this case, the probability distribution of the PDF parameters. This allows for a straightforward propagation of the resulting uncertainties into any PDF-dependent observable, preserving their simple probabilistic interpretation. In our analysis we include a broad set of deep inelastic scattering data from HERA, BCDMS and NMC experiments along with the Drell-Yan, 𝑊 and 𝑍 boson data from LHC and Tevatron experiments, which combined with theoretical calculations at next-to-next-to-leading order in QCD allow for realistic determination of PDFs. The main focus of this analysis is to explore alternative methods for PDF uncertainty estimation that are more firmly grounded in statistical principles. We show that the flexibility of the Bayes framework, allowing one, e.g., to account for non-Gaussianity or inconsistencies of datasets, is crucial to extract realistic uncertainties when such assumptions are not fulfilled. We also demonstrate that MCMC allows one to determine the Δ⁢𝜒 2 value corresponding to a given confidence level in the sample, which can, in turn, be used as a statistically well-founded tolerance criterion used in the Hessian method, thus addressing one of its main long-standing drawbacks.

Risse, Peter Clemens [Universität Münster (Germany↗

Final Search for Short-Baseline Neutrino Oscillations with the PROSPECT-I Detector at HFIR

The PROSPECT experiment is designed to perform precise searches for antineutrino disappearance at short distances (7–9 m) from compact nuclear reactor cores. This Letter reports results from a new neutrino oscillation analysis performed using the complete data sample from the PROSPECT-I detector operated at the High Flux Isotope Reactor in 2018. The analysis uses a multiperiod selection of inverse beta decay neutrino interactions with reduced backgrounds and enhanced statistical power to set limits on electron neutrino disappearance caused by mixing with sterile neutrinos with 0.2–20 eV 2 mass splittings. Inverse beta decay positron energy spectra from six different reactor-detector distance ranges are found to be statistically consistent with one another, as would be expected in the absence of sterile neutrino oscillations. The data excludes at 95% confidence level the existence of sterile neutrinos in regions above 3 eV 2 previously unexplored by terrestrial experiments, including all space below 10 eV 2 suggested by the recently strengthened Gallium Anomaly. The best-fit point of the Neutrino-4 reactor experiment’s claimed observation of short-baseline oscillation is ruled out at more than 5 standard deviations.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

Automated analysis of unlabeled PV data with Solar Data Tools software: Overview and feature updates

Distributed rooftop PV systems: ubiquitous, yet commonly have unlabeled data Difficult or impossible to form a performance index We developed Solar Data Tools (SDT), an open-source Python library for analyzing PV power (and irradiance) time-series data SDT enables analysis of unlabeled PV data—no model, no meteorological data, no performance index required Takes a statistical signal processing approach Data processing steps are largely pre-defined and automatic regardless of system type—from utility tracking systems to multi-pitch rooftop systems

Meyers-Im, Bennet E↗

Field Test Report Neutron Scintillator Array Dry Storage Cask Scanner FY2024

During two weeks of Field Testing at the Idaho National Laboratory INTEC Cask Farm in July and August 2024, the LLNL Dry Storage Cask Scanner Array was lifted on top of an MC-10 dry storage fuel cask and operated to acquire neutron and gamma-ray data from the 24 fuel bundle positions. Neutron and gamma-ray data acquisition scans across the top of the cask of varying dwell times were performed July 15-18, 2024 and August 19-22, 2024 to evaluate the ability of the scanner data to reveal asymmetries in the fuel positions that reflect asymmetries in the MC-10 cask fuel bundle loading. The MC-10 cask 24 position fuel bundle loading at the INTEC Cask Farm is well documented, including the locations of six empty fuel bundle positions. This loading presents an opportunity to test the ability of the scanner system to detect diversion of spent fuel bundles as well as to validate the MC-10 cask MCNP modeling. The cask scanner array consists of six Stilbene crystal scintillator detectors and a linear actuator frame that moves the six detectors across the MC-10 dry storage cask to obtain data above each of the 24 fuel bundle positions. The detectors are connected to a pulse-shape discrimination data acquisition system capable of generating separate neutron and gamma-ray spectra for each detector and for each scan position. From the prior single detector Field Test in 2021 and iteration with MCNP modeling, the neutron and gamma-ray data were analyzed in multiple energy regions to identify an analysis method that would provide the strongest and most consistent signature of the asymmetric MC-10 cask fuel loading1 . From both the 2021 Field Test and the current Field Test results, the neutron capture gamma-ray count rate around 2.2 MeV provides the strongest signature of the asymmetric MC-10 cask fuel loading and has qualitative agreement with MCNP calculations. Counting all gamma-rays produces a similar signature. Neutrons emerging from the cask top are moderated and captured by the hydrogen in the polyethylene moderator and scintillator detector, producing a 2.2 MeV gamma ray which is seen in the scintillator gamma-ray spectrum. The count rate in the 2.2 MeV gamma-ray region is ~50 c/s, which is ~1000x higher than the ~0.05 n/s rate in the > 4MeV neutron region, and ~50x greater than the ~1 n/s rate in the neutrons > 500 keV region. Analysis of the 2.2 MeV neutron-capture Compton-scattered gamma-rays produces a statistically significant signature of the INTEC Cask Farm MC-10 asymmetric fuel loading. MCNP simulations indicate that the average neutron energy spectrum offers the potential to detect a large asymmetry from several missing bundles as well as individual missing fuel bundles. Testing this feature will require measurements on a cask with single missing elements.

11 NUCLEAR FUEL CYCLE AND FUEL MATERIALS↗

The composition of gases from a diffusion flame above longleaf pine needle fuel beds

The gas and tar composition of a diffusion flame from longleaf pine needles is currently poorly understood and more data are needed to fill in the gap between pyrolysis data and smoke plume data, thus improving physical and chemical modeling of wildland smoke formation. A pilot experiment to measure light gas and tar composition of such a flame is described for three flame regions: persistent flame (flame base), intermittent flame, and smoke plume. Flame gases from 24 experimental fires were collected in canisters and analyzed using EPA method TO-14A for CO 2 , CO, H 2 , CH 4 , and C 2 to C 7 hydrocarbon gases. Condensed gas (tar) samples were collected and analyzed using GC/MS. Other light gases were measured using FTIR spectroscopy. Results from compositional data analysis suggest significant differences in (relative) concentration of compounds detected in the three regions of the flame. Statistical tests for differences in flame zones were performed using the canister data: Concentration of hydrocarbons relative to CO and CO 2 decreased from the persistent flame zone above the pyrolyzing needles through the intermittent flame region into the flame-free plume. This was likely due to both chemical reactions (oxidation) occurring in the flame as well as the introduction of air into the flame/plume by entrainment.

Biomass↗

Link statistics of dislocation network during strain hardening

Dislocations are line defects in crystals that multiply and self-organize into a complex network during strain hardening. The length of dislocation links, connecting neighboring nodes within this network, contains crucial information about the evolving dislocation microstructure. By analyzing data from Discrete Dislocation Dynamics (DDD) simulations in face-centered cubic (fcc) Cu, we characterize the statistical distribution of link lengths of dislocation networks during strain hardening on individual slip systems. Here, our analysis reveals that link lengths on active slip systems follow a double-exponential distribution, while those on inactive slip systems conform to a single-exponential distribution. The distinctive long tail observed in the double-exponential distribution is attributed to the stress-induced bowing out of long links on active slip systems, a feature that disappears upon removal of the applied stress. We further demonstrate that both observed link length distributions can be explained by extending a one-dimensional Poisson process to include different growth functions. Specifically, the double-exponential distribution emerges when the growth rate for links exceeding a critical length becomes super-linear, which aligns with the physical phenomenon of long links bowing out under stress. This work advances our understanding of dislocation microstructure evolution during strain hardening and elucidates the underlying physical mechanisms governing its formation.

Crystal plasticity↗