Search NASA⌕ Search

SEARCH · Search NASA

Results for “Statistical Methods”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 685 records · Page 38

Towards a data-driven model of hadronization using normalizing flows

We introduce a model of hadronization based on invertible neural networks that faithfully reproduces a simplified version of the Lund string model for meson hadronization. Additionally, we introduce a new training method for normalizing flows, termed MAGIC, that improves the agreement between simulated and experimental distributions of high-level (macroscopic) observables by adjusting single-emission (microscopic) dynamics. Our results constitute an important step toward realizing a machine-learning based model of hadronization that utilizes experimental data during training. Finally, we demonstrate how a Bayesian extension to this normalizing-flow architecture can be used to provide analysis of statistical and modeling uncertainties on the generated observable distributions.

72 PHYSICS OF ELEMENTARY PARTICLES AND FIELDS↗

CdTe Core: Final Technical Report (FTR)

CdTe is presently the cost-leading thin-film PV technology, directly competing with Si at scale, even when domestically manufactured. While an impressive technology, its efficiency remains much below the detailed balance limit with the largest cause due to its low photovoltage and fill factor. To realize gains, the carrier concentration, minority carrier lifetime, and interface recombination all need to be improved simultaneously over historic levels. Using a new defect chemistry (group V doping instead of copper) has been identified as a viable route using single crystal systems. This project focused on implementing this new defect chemistry in scalable, polycrystalline thin-film photovoltaic CdTe devices with tasks focusing improvements to the front interface, absorber, and rear interface as well as capability development & stakeholder engagement. The goal of the project was to establish a strategy using devices, test structures, detailed characterization, and modeling to quantify the sources of losses in state-of-the-art CdTe photovoltaic devices. Using this strategy and advanced synthesis, losses at the front interface, absorber, and rear interface were worked on in parallel. The final objective was to significantly improve the voltage deficit in CdTe devices to enable improvements in photovoltage and efficiency that can be implemented by industry in the near-term. Over the course of the project, the team developed new characterization techniques, analysis, and modeling which were then applied to state-of-the-art materials generated internally and collaboratively. In particular to enable rapid progress, NREL worked closely with First Solar where NREL grew complete devices as well as ones that interleaved process steps where First Solar had completed different steps such as absorber growth or absorber growth and activation using their baseline methods. Using detailed characterization and analysis including photoemission, photoluminescence, and scanning probe techniques enabled understanding of the loss pathways and area for improvements in our own and First Solar s materials. Ultimately, this contributed to the first series of new world record CdTe efficiencies since 2016, culminating in a 23.1% certified cell that was P-doped along with As-doped cells of similar performance. Internally, NREL improved the statistical variation in baseline As-doped devices and improved average photovoltage by over 100 mV. This was done through an improvement in absorber quality, changed front interface, and improved back contact. In addition to materially improving the fabrication processes at NREL, characterization, analysis, and modeling were developed and disseminated. NREL also played a pivotal role in community building over the course of this project working closely with the Cadmium Telluride Accelerator Consortium. NREL worked in a series of collaborations with academic and industry partners, leveraging knowledge and innovations from this project, as well as helped organize a series of workshops to ensure rapid progress in the field. Working closely with the academic community has led to a dissemination of knowledge; working with First Solar as increased US competitiveness First Solar expanded domestic production to ~10 GW and opened new facilities.

14 SOLAR ENERGY↗

Determining Exact RANS Operators with the Macroscopic Forcing Method (Final Report)

This report contains a compilation of key results and findings from the ACT project “Determining Exact RANS Operators with the Macroscopic Forcing Method.” The Macroscopic Forcing Method (MFM), a numerical tool for determining closure operators, is used to measure eddy diffusivity moments in Rayleigh-Taylor (RT) instability. It is first applied to low-Atwood 2D RT instability; that work is then extended to 3D RT at different finite Atwood numbers. It is found that nonlocality is important for modeling the mean scalar transport closure operator in RT mixing. Additionally, work is done to improve the statistical convergence of MFM for chaotic problems like RT mixing.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

A flexible class of priors for orthonormal matrices with basis function-specific structure

Statistical modeling of high-dimensional matrix-valued data motivates the use of a low-rank representation that simultaneously summarizes key characteristics of the data and enables dimension reduction. Low-rank representations commonly factor the original data into the product of orthonormal basis functions and weights, where each basis function represents an independent feature of the data. However, the basis functions in these factorizations are typically computed using algorithmic methods that cannot quantify uncertainty or account for basis function correlation structure a priori. While there exist Bayesian methods that allow for a common correlation structure across basis functions, empirical examples motivate the need for basis function-specific dependence structure. We propose a prior distribution for orthonormal matrices that can explicitly model basis function-specific structure. The prior is used within a general probabilistic model for singular value decomposition to conduct posterior inference on the basis functions while accounting for measurement error and fixed effects. We discuss how the prior specification can be used for various scenarios and demonstrate favorable model properties through synthetic data examples. Finally, we apply our method to two-meter air temperature data from the Pacific Northwest, enhancing our understanding of the Earth system’s internal variability.

97 MATHEMATICS AND COMPUTING↗

Uncertainties in the production of iron-group nuclides in core-collapse supernovae from Monte Carlo variations of reaction rates

Core-collapse supernovae, occurring at the end of massive star evolution, produce heavy elements, including those in the iron peak. Although the explosion mechanism is not yet fully understood, theoretical models can reproduce optical observations and observed elemental abundances. However, many nuclear reaction rates involved in explosive nucleosynthesis have large uncertainties, impacting the reliability of abundance predictions. To address this, we have previously developed a Monte Carlo-based nucleosynthesis code that accounts for reaction rate uncertainties and has been applied to nucleosynthesis processes beyond iron. Our framework is also well suited for studying explosive nucleosynthesis in supernovae. In this paper, we investigate 1D explosion models using the ‘PUSH method’ , focusing on progenitors with varying metallicities and initial masses around $M_{\rm ZAMS} = 16\, {\rm M}_{\odot }$. Detailed post-process nucleosynthesis calculations and Monte Carlo analyses are used to explore the effects of reaction rate uncertainties and to identify key reaction rates in explosive nucleosynthesis. We find that many reactions have little impact on the production of iron-group nuclei, as these elements are primarily synthesized in the nuclear statistical equilibrium. However, we identify a few ‘key reactions’ that significantly influence the production of radioactive nuclei, which may affect astrophysical observables. In particular, for the production of ${}^{44}{\rm Ti}$, we confirm that several traditionally studied nuclear reactions have a strong impact. However, determining a single reaction rate is insufficient to draw a definitive conclusion.

79 ASTRONOMY AND ASTROPHYSICS↗

Accessing the gluon momentum fraction of nucleons through the gradient flow

We calculate the gluon momentum fraction of the nucleon using lattice QCD, with a nonperturbative renormalization technique based on the gradient flow. The gluon momentum fraction is determined on a single Wilson-clover ensemble using 𝑁 𝑓 =2 +1 flavors with pion mass 358 MeV and lattice spacing 0.094 fm. We employ the variational method to reduce excited-state contamination and apply the distillation framework to ensure a large operator basis. To reduce systematic uncertainties, we apply Bayesian model averaging to all fit procedures. We apply matching coefficients to the flow-time dependent lattice results to recover the gluon momentum fraction in the $\overline{MS}$-scheme at 2 GeV. Our final result is ⟨𝑥⟩ 𝑔 ⁢(𝜇 =2 GeV) =0.482⁢(35), where we quote only statistical uncertainties.

Lattice QCD↗

Adaptive Client Selection in Federated Learning: A Network Anomaly Detection Use Case

Federated Learning (FL) has become a ubiquitous approach for training machine learning models on decentralized data, addressing the myriad privacy concerns inherent in traditional centralized methods. However, the efficiency of FL depends on effective client selection and robust privacy preservation mechanisms. Inadequate client selection may lead to suboptimal model performance, while insufficient privacy measures risk exposing sensitive data. This paper proposes a client selection framework for FL that integrates differential privacy and fault tolerance. Our adaptive approach dynamically adjusts the number of selected clients based on model performance and system constraints, ensuring privacy through calibrated noise addition. We evaluate our method on a network anomaly detection use case using the UNSW-NB15 and ROAD datasets. Results show up to a 7% increase in accuracy and a 25% reduction in training time compared to FedL2P. Moreover, we highlight the trade-offs between privacy budgets and model performance, with higher privacy budgets reducing noise and improving accuracy. Our fault tolerance mechanism, while causing a slight performance drop, enhances robustness to client failures. Statistical validation using Mann-Whitney U tests confirms the significance of these improvements (p < 0.05).

Marfo, William [University of Texas at El Paso,Dep↗

Strong Coupling of Hydrodynamics and Reactions in Nuclear Statistical Equilibrium for Modeling Convection in Massive Stars

We build on the simplified spectral deferred corrections (SDC) coupling of hydrodynamics and reactions to handle the case of nuclear statistical equilibrium (NSE) and electron/positron captures/decays in the cores of massive stars. Our approach blends a traditional reaction network on the grid with a tabulated NSE state from a very large, ${\mathcal O }(100)$ nuclei network. We demonstrate how to achieve second-order accuracy in the simplified-SDC framework when coupling NSE to hydrodynamics, with the ability to evolve the star on the hydrodynamics time step. We discuss the application of this method to convection in massive stars leading up to core collapse. We also show how to initialize the initial convective state from a 1D model in a self-consistent fashion. All of these developments are done in the publicly available Castro simulation code and the entire simulation methodology is fully GPU-accelerated.

Explosive nucleosynthesis↗

Basic Research Needs for Inverse Methods for Complex Systems under Uncertainty [Brochure]

The four priority research directions outlined in this brochure represent a cohesive vision for advancing the science of inverse problems for complex systems under uncertainty. Together, they address the critical challenges of: discovering, exploiting, and preserving physical and problem structure; overcoming model limitations; integrating disparate, multimodal, and/or dynamic data; and tailoring the solution of inverse problems to downstream tasks. While each PRD focuses on a distinct aspect of inverse-problem research, their interconnected nature highlights the importance of a holistic approach that leverages progress across all areas to achieve transformative solutions. This agenda calls for research across mathematics, statistics, and computer science disciplines, which are guided and complemented by rapid advances in artificial intelligence, high-performance computing, and experimental facilities, to unlock new capabilities, maximize scientific impact, and meet the growing demands of inverse problems that arise across applications that are critical to DOE's mission.

97 MATHEMATICS AND COMPUTING↗

The Thermal and Kinematic Sunyaev–Zeldovich Effect in Galaxy Clusters and Filaments Using Multifrequency Temperature Maps of the Cosmic Microwave Background: A399–A401 Cluster Pair Case Study

We present a multifrequency and multi-instrument methodology to study the physical properties of galaxy clusters and cosmic filaments using cosmic microwave background observations. Our approach enables simultaneous measurement of both the thermal (tSZ) and kinematic Sunyaev–Zeldovich (kSZ) effects, incorporates relativistic corrections, and models astrophysical foregrounds such as thermal dust emission. We do this by jointly fitting a single physical model across multiple maps from multiple instruments at different frequencies, rather than fitting a model to a single Compton-y map. We demonstrate the success of this method by fitting the A399–A401 galaxy cluster pair and filament system using archival data from the Planck satellite and new, targeted deep data from the Atacama Cosmology Telescope, covering 11 different frequencies over 14 maps from 30 GHz to 545 GHz. Our tSZ results are consistent with previous work using Compton-y maps. We measure the line-of-sight peculiar velocities of the cluster–filament system using the kSZ effect and find statistical uncertainties on individual cluster peculiar velocities of ≲600 km s −1 , which are competitive with current state-of-the-art measurements. Additionally, we measure the optical depth of the filament component with a signal-to-noise of 8.5σ and reveal hints of its morphology. This modular approach is well-suited for application to future instruments across a wide range of millimeter and submillimeter wavebands.

Gill, Ajay S. [National Research Council of Canada↗

Stochastic GW -GPU: Rapid Quasi-Particle Energies for Molecules beyond 10,000 Atoms

StochasticGW is a code for computing accurate quasi-particle (QP) energies of molecules and material systems in the GW approximation. StochasticGW utilizes the stochastic Resolution of the Identity (sROI) technique to enable a massively parallel implementation with computational costs that scale semilinearly with system size, allowing the method to access systems with tens of thousands of electrons. Here, we introduce a new implementation, StochasticGW-GPU, for which the main bottleneck steps have been ported to GPUs and give substantial performance improvements over previous versions of the code. We showcase the new code by computing band gaps of hydrogenated silicon clusters (Si x H y ) containing up to 10,001 atoms and 35,144 electrons, and we obtain individual QP energies with a statistical precision of better than ±0.03 eV with times-to-solution of less than 1 h.

Thomas, Phillip S. [Lawrence Berkeley National Lab↗

High-throughput micro-scale bandgap mapping for perovskite-inspired materials with complex composition space

Abstract To realize the full promise of high-throughput experimental workflows, the rate of sample synthesis must be matched by that of characterization. Of growing interest are contactless optical techniques that can rapidly measure material homogeneity and properties. Here, we present a hyperspectral imaging method to measure local optical bandgap distributions within samples, utilizing spatially-resolved reflectance spectra coupled with automated data analysis. We collect approximately one million optical bandgap data across the compositional space of Cs 3 (Bi x Sb 1-x ) 2 (Br y I 1-y ) 9 perovskite-inspired materials. Our results show non-monotonic bandgap variations (i.e., bandgap bowing) along six composition gradient sequences, in addition to identifying samples with multiple bandgaps in statistics. High-throughput transient absorption spectroscopy reveals that within these compositions, the depletion of the ground state carriers to excited states occurred at discrete energy levels with independent carrier dynamics, consistent with the bandgap observation and indicative of phase separation. This work demonstrates the potential for rapid optical measurements to assess material quality and homogeneity in a high-throughput experimental setting, supporting screening and recipe optimization of optoelectronic material candidates with desired carrier dynamics and optical properties.

Science & Technology - Other Topics↗

Mesoscale Convective Systems Tracking Method Intercomparison (MCSMIP): Application to DYAMOND Global km‐Scale Simulations

Abstract Global kilometer‐scale models represent the future of Earth system modeling, enabling explicit simulation of organized convective storms and their associated extreme weather. Here, we comprehensively evaluate tropical mesoscale convective system (MCS) characteristics in the DYAMOND (DYnamics of the atmospheric general circulation modeled on non‐hydrostatic domains) simulations for both summer and winter phases. Using 10 different feature trackers applied to simulations and satellite observations, we assess MCS frequency, precipitation, and other key characteristics. Substantial differences (a factor of 2–3) arise among trackers in observed MCS frequency and their precipitation contribution, but model‐observation differences in MCS statistics are more consistent across trackers. DYAMOND models are generally skillful in simulating tropical mean MCS frequency, with multi‐model mean biases ranging from −2%–8% over land and −8%–8% over ocean (summer vs. winter). However, most DYAMOND models underestimate MCS precipitation amount (23%) and their contribution to total precipitation (17%). Biases in precipitation contributions are generally smaller over land (13%) than over ocean (21%), with moderate inter‐model variability. While models better simulate MCS diurnal cycles and cloud shield characteristics, they overestimate MCS precipitation intensity and underestimate stratiform rain contributions (up to a factor of 2), particularly over land, albeit observational uncertainties exist. Additionally, models exhibit a wide range of precipitable water in the tropics compared to reanalysis and satellite observations, with many models showing exaggerated sensitivity of MCS precipitation intensity to precipitable water. The MCS metrics developed here provide process‐oriented diagnostics to guide future model development.

54 ENVIRONMENTAL SCIENCES↗

Advanced Signal Decomposition Analysis and Anomaly Detection in Photovoltaic Systems

With the rapid expansion of large-scale photovoltaic (PV) plants, it is paramount for solar stakeholders to understand the reliability and efficiency of their plants to inform maintenance decisions, increase production, and understand the design factors that impact performance. Diagnosing underperformance in PV plants is challenging due to the relatively few monitoring points with respect to the large geographic footprint of the plant. This work introduces a cutting-edge method that transforms the analysis and management of key factors influencing PV plant performance, including performance loss rate (PLR), recoverable soiling, and major system changes. Identifying these factors is critical for deriving actionable insights. Leveraging advanced analytical techniques such as wavelet transformation, robust regression, and extreme point analysis, this approach provides a nuanced understanding of these factors. This method has been tested across two synthetic datasets and one real dataset, consistently surpassing existing benchmarks by achieving a lower median mean absolute error and reduced error variability across all comparable components.

14 SOLAR ENERGY↗

Development of an Unbiased Future Solar Dataset for Solar Resource Adequacy Research Over CONUS

A high-resolution, long-term solar dataset is essential for capturing the variability of solar energy resources and informing strategies to ensure grid reliability and resilience in systems with high levels of solar energy integration. This study focuses on generating unbiased, high-resolution projections of solar irradiance through a statistical downscaling framework, using Earth system model (ESM) simulations obtained from the North American Coordinated Regional Climate Downscaling Experiment (NA-CORDEX). The National Solar Radiation Database (NSRDB) is used to calibrate statistical downscaling models. The newly developed dataset provides solar irradiance, surface air temperature, and surface wind speed at 4-km and hourly resolutions across the contiguous United States (CONUS), based on two future scenarios (RCP4.5 and RCP8.5). This study outlines key steps in developing the high-resolution future solar dataset, including (1) regridding ESM data to a common 20-km resolution grid, (2) correcting ESM biases using the NSRDB, and (3) applying temporal and spatial downscaling methods to generate high-resolution (4-km, hourly) solar projections. Preliminary results indicate that downscaled projections (4-km) captured reasonable spatial patterns when compared to observations across CONUS for four variables. On average across all pixels, 4-km daily-total GHI and DNI projections showed normalized bias (nBias) less than 1% and 6% for GHI and DNI against NSRDB, respectively (nBias less than 1% and 5% for daily-average surface air temperature and surface wind speed). In terms of long-term trend for GHI and DNI, there was no strong increasing or decreasing trend (when compared to surface air temperature), but it showed a very weak decreasing trend.

14 SOLAR ENERGY↗

Divide and conquer: Learning chaotic dynamical systems with multistep penalty neural ordinary differential equations

Forecasting high-dimensional dynamical systems is a fundamental challenge in various fields, such as geosciences and engineering. Neural Ordinary Differential Equations (NODEs), which combine the power of neural networks and numerical solvers, have emerged as a promising algorithm for forecasting complex nonlinear dynamical systems. However, classical techniques used for NODE training are ineffective for learning chaotic dynamical systems. In this work, we propose a novel NODE-training approach that allows for robust learning of chaotic dynamical systems. Here, our method addresses the challenges of non-convexity and exploding gradients associated with underlying chaotic dynamics. Training data trajectories from such systems are split into multiple, non-overlapping time windows. In addition to the deviation from the training data, the optimization loss term further penalizes the discontinuities of the predicted trajectory between the time windows. The window size is selected based on the fastest Lyapunov time scale of the system. Multi-step penalty(MP) method is first demonstrated on Lorenz equation, to illustrate how it improves the loss landscape and thereby accelerates the optimization convergence. MP method can optimize chaotic systems in a manner similar to least-squares shadowing with significantly lower computational costs. Our proposed algorithm, denoted the Multistep Penalty NODE, is applied to chaotic systems such as the Kuramoto-Sivashinsky equation, the two-dimensional Kolmogorov flow, and ERA5 reanalysis data for the atmosphere. It is observed that MP-NODE provide viable performance for such chaotic systems, not only for short-term trajectory predictions but also for invariant statistics that are hallmarks of the chaotic nature of these dynamics.

Chaotic dynamical systems↗

Understanding Process–Structure Relationships during Lamination of Halide Perovskite Interfaces

Fabrication of halide perovskite (HP) solar cells typically involves the sequential deposition of multiple layers to create a device stack, which is limited by the thermal and chemical incompatibility of top contact layers with the underlying HP semiconductor. One emerging strategy to overcome these restrictions on material selection and processing conditions is lamination, where two half-stacks are independently processed and then diffusion bonded to complete the device. Lamination reduces the processing constraints on the top side of the solar cell to allow new device designs, expanded use of deposition methods, and self-encapsulation of devices. While laminated perovskite solar cells with high efficiencies and novel interlayer combinations have been demonstrated, there is a limited understanding of how the lamination process parameters affect the diffusion-bond quality and material properties of the resulting HP layer. In this study, we systematically vary temperature, pressure, and time during lamination and quantify the resulting impacts on bonded area, grain domain size, and photoluminescence. A design of experiments is performed, and statistical analysis of the experimental results is used to quantitatively evaluate the resulting process–structure–property relationships. The lamination temperature is found to be the key parameter controlling these properties. Furthermore, a temperature of 150 °C enables successful bonding over 95% of the substrate area and also results in increases in apparent grain domain size and photoluminescence intensity. Based on these insights, the lamination temperature of functional perovskite solar cell devices is varied, demonstrating the importance of the resulting bond quality on device performance metrics.

14 SOLAR ENERGY↗

Velocity reconstruction in the era of DESI and Rubin/LSST. I. Exploring spectroscopic, photometric, and hybrid samples

Peculiar velocities of galaxies and halos can be reconstructed from their spatial distribution alone. This technique is analogous to the baryon acoustic oscillations reconstruction, using the continuity equation to connect density and velocity fields. The resulting reconstructed velocities can be used to measure imprints of galaxy velocities on the cosmic microwave background like the kinematic Sunyaev-Zel’dovich effect or the moving lens effect. As the precision of these measurements increases, characterizing the performance of the velocity reconstruction becomes crucial to allow unbiased and statistically optimal inference. In this paper, we quantify the relevant performance metrics: the variance of the reconstructed velocities and their correlation coefficient with the true velocities. We show that the relevant velocities to reconstruct for kSZ and moving lens are actually the halo—rather than galaxy—velocities. We quantify the impact of redshift-space distortions, photometric redshift errors, satellite galaxy fraction, incorrect cosmological parameter assumptions and smoothing scale on the reconstruction performance. Here, we also investigate hybrid reconstruction methods, where velocities inferred from spectroscopic samples are evaluated at the positions of denser photometric samples. We find that using exclusively the photometric sample is better than performing a hybrid analysis. The 2 Gpc/ℎ length simulations from abacussummit with realistic galaxy samples for DESI and Rubin LSST allow us to perform this analysis in a controlled setting. In the companion paper [B. Hadzhiyska, S. Ferraro, B. Ried Guachalla, and E. Schaan, companion paper, Phys. Rev. D 109, 103534 (2024).], we further include the effects of evolution along the light cone and give realistic performance estimates for DESI luminous red galaxies, emission line galaxies, and Rubin LSST-like samples.

79 ASTRONOMY AND ASTROPHYSICS↗