Search NASA⌕ Search

SEARCH · Search NASA

Results for “MATHEMATICAL STATISTICS”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 109 records · Page 6

Material Control & Accounting Modeling Developments for a Generic TRISO Fuel Fabrication Facility

The Material Protection, Accounting, and Control Technologies (MPACT) program utilizes modeling and simulation to assess Material Control and Accountability (MC&A) concerns for a variety of nuclear facilities. The Sandia National Laboratories (SNL)-developed Fissile Facility Flow Modeler (F3M) and the Material Accountancy Performance Indicator Toolkit (MAPIT) have historically provided MPACT with the capability to analyze MC&A approaches for nuclear facilities to determine that these facilities meet regulatory requirements. In FY25, improvements on the application of the F3M and MAPIT tools to simulate a generic TRi-structural ISOtropic (TRISO) fuel fabrication facility were successfully completed. The generic TRISO fuel fabrication F3M model captures the entire TRISO fuel fabrication process and is adaptable to any final TRISO fuel form, including spherical pebbles and cylindrical compacts loaded into graphite prismatic blocks. Comprehensive F3M/MAPIT functionality for the generic TRISO fuel fabrication model has been demonstrated. This modeling framework can be applied to support the U.S. Department of Energy and domestic nuclear industry stakeholders in developing MC&A approaches for advanced fuel fabrication facilities via statistical tests that demonstrate compliance to regulatory requirements.

97 MATHEMATICS AND COMPUTING↗

Polynomial Chaos Surrogate Construction for Random Fields with Parametric Uncertainty

Engineering and applied science rely on computational experiments to rigorously study physical systems. The mathematical models used to probe these systems are highly complex, and sampling-intensive studies often require prohibitively many simulations for acceptable accuracy. Surrogate models provide a means of circumventing the high computational expense of sampling such complex models. In particular, polynomial chaos expansions (PCEs) have been successfully used for uncertainty quantification studies of deterministic models where the dominant source of uncertainty is parametric. We discuss an extension to conventional PCE surrogate modeling to enable surrogate construction for stochastic computational models that have intrinsic noise in addition to parametric uncertainty. We develop a PCE surrogate on a joint space of intrinsic and parametric uncertainty, enabled by Rosenblatt transformations, which are evaluated via kernel density estimation of the associated conditional cumulative distributions. Furthermore, we extend the construction to random field data via the Karhunen–Loève expansion. We then take advantage of closed-form solutions for computing PCE Sobol indices to perform a global sensitivity analysis of the model which quantifies the intrinsic noise contribution to the overall model output variance. Additionally, the resulting joint PCE is generative in the sense that it allows generating random realizations at any input parameter setting that are statistically approximately equivalent to realizations from the underlying stochastic model. The method is demonstrated on a chemical catalysis example model and a synthetic example controlled by a parameter that enables a switch from unimodal to bimodal response distributions.

97 MATHEMATICS AND COMPUTING↗

Diagnostics, Prognostics, and Optimization for Lithium-Ion Battery Systems

Health management of lithium-ion battery systems presents a host of challenges due to their complex physics, large numbers of components, and a wide variety of degradation behaviors across different battery types. Dr. Paul Gasper will present on research from the Electrochemical Energy Storage Group on Lithium-ion battery diagnostics, prognostics, and optimization. Diagnostics research, including state-estimation via machine-learning from electrochemical impedance spectroscopy and DC pulses as well as continuous state-estimation via Kalman filters, will highlight the ongoing challenges for accurately measuring the state of batteries without performing time-consuming characterization tests. NLR's industry-recognized battery prognostics work, which predicts real-world battery degradation by identifying degradation rate models from accelerated aging data using statistical modeling and machine-learning, will be used to demonstrate the critical impact of battery controls, thermal management, and operating strategy on durability and lifetime. Finally, the use of prognostic models for financial or lifetime optimization will be discussed.

25 ENERGY STORAGE↗

V-HAMSTeR v1.0.0

V-HAMSTeR is a bioinformatics software tool designed to predict the hosts of viruses directly from genomic sequences. It can be used by researchers to predict animal, prokaryotic, plant, protist or fungal viral hosts including viruses that may be fragmented or discovered in environmental metagenomic datasets. Features & Uses: The software employs a novel dual-stream deep learning architecture that dynamically fuses implicit sequence embeddings from a genomic foundation model with 13 explicit, handcrafted biological features (e.g., coding density and strand switch rates). To ensure maximum reliability, V=HAMSTeR deploys a 5-fold deep ensemble calibrated via Joint Temperature Scaling, providing users with statistically rigorous confidence probabilities. It also features an automated sequence chunking and mean-pooling module to seamlessly process variable-length contigs. Advantages Over Similar Technologies: Existing tools (e.g., IPEV, RNAVirHost) typically rely on either basic k-mers or isolated neural networks. V-HAMSTeR's hybrid architecture captures both broad genomic context and specific biological motifs that standalone foundation models often miss. Furthermore, unlike competitor tools that struggle with incomplete data or exhibit extreme overconfidence, V-HAMSTeR is explicitly benchmarked and mathematically calibrated for fragmented assemblies (1kb–10kb). This makes it uniquely robust, accurate, and trustworthy for the messy reality of real-world environmental viromics.

Grigson, Susie [Lawrence Berkeley National Laborat↗

Simulating Wind-Driven Loading on PV Systems

As PV modules continue to trend toward larger, thinner, and more flexible forms they grow more susceptible to damage from dynamic wind loading. As a result, understanding the impact of wind on PV systems, particularly when mounted on compliant solar-tracking hardware, and identifying robust, stable array layouts and stow strategies is becoming increasingly important for the PV community. We are developing an open-source software package, PVade (PV aerodynamic design engineering), to simulate the cascading fluid-structure interaction that occurs within single-axis, solar-tracking arrays to enable researchers to test hardware, layout, and tracker control changes, leading to enhanced stability and a reduction in wind-driven damage. We will give an overview of the PVade software and present the latest outcomes from our ongoing validation campaign in which we compare time series and statistical structural responses with field data. From there, we will present simulated results from a larger, multi-row array and highlight the effect of varying tracker angles on stability and the differences between positive and negative tilt angles.

fluid↗

Neural network based emulation of galaxy power spectrum covariances: A reanalysis of BOSS DR12 data

We train neural networks to quickly generate redshift-space galaxy power spectrum covariances from a given parameter set (cosmology and galaxy bias). This covariance emulator utilizes a combination of traditional fully connected network layers and transformer architecture to accurately predict covariance matrices for the high redshift, north galactic cap sample of the BOSS DR12 galaxy catalog. We run simulated likelihood analyses with emulated and brute-force computed covariances, and we quantify the network’s performance via two different metrics: (1) difference in Χ 2 and (2) likelihood contours for simulated BOSS DR 12 analyses. We find that the emulator returns excellent results over a large parameter range. We then use our emulator to perform a reanalysis of the BOSS HighZ NGC galaxy power spectrum, and find that varying covariance with cosmology along with the model vector produces Ω m = $0.27⁢6$$^{+0.013}_{–0.015}$, H 0 = 70.2 ± 1.9 km/s/Mpc, and σ 8 = $0.67⁢4$$^{+0.058}_{–0.077}$. These constraints represent an average 0.46⁢σ shift in best-fit values and a 5% increase in constraining power compared to fixing the covariance matrix (Ω m = 0.293 ± 0.017, H 0 = 70.3 ± 2.0 km/s/Mpc, σ 8 = $0.70⁢2$$^{+0.063}_{–0.075}$). As a result, this work demonstrates that emulators for more complex cosmological quantities than second-order statistics can be trained over a wide parameter range at sufficiently high accuracy to be implemented in realistic likelihood analyses.

79 ASTRONOMY AND ASTROPHYSICS↗

Stochastic Modeling of the Joint Neutron Number-Cumulative Fission Fragment Kinetic Energy Deposition Distribution and its Statistical Moments [Slides]

We investigate the joint distribution of the neutron number and cumulative fission-fragment kinetic energy (FKE) deposition, with a specific focus on low-order statistical moments: the mean, variance, and correlation. Starting from a point-kinetic framework, we derive a forward Master equation (FME) for the joint distribution and develop the corresponding moment equations.

42 ENGINEERING↗

Surface Pressure Fluctuations Induced by a Hypersonic Turbulent Boundary Layer on a Sharp Cone at Angle of Attack

High-fidelity simulations are performed to characterize the turbulence-induced wall pressure fluctuations on a sharp cone at a 5.5° angle-of-attack in a Mach 8 flow. Wall-resolved large-eddy simulation (LES) and wall-modeled large-eddy simulation (WMLES) results are compared to measurements at several locations on the cone body, where comparisons with high-frequency PCB sensors are good, while comparison with low-frequency kulite sensors varies depending on the location. For a given streamwise location, simulation results show significant azimuthal variation in the wall pressure fluctuation statistics. Comparisons between LES and WMLES results indicate that WMLES can accurately reproduce the low-frequency component of the autospectra quite well, while slight sensitivity is apparent in the coherence. A grid sensitivity study for WMLES further demonstrated sensitivity in the wall pressure fluctuation coherence, but not in the autospectra.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Enhancing Interpretability in Generative Modeling: Statistically Disentangled Latent Spaces Guided by Generative Factors in Scientific Datasets

This study addresses the challenge of statistically extracting generative factors from complex, high-dimensional datasets in unsupervised or semi-supervised settings. We investigate encoder-decoder-based generative models for nonlinear dimensionality reduction, focusing on disentangling low-dimensional latent variables corresponding to independent physical factors. Introducing Aux-VAE, a novel architecture within the classical Variational Autoencoder framework, we achieve disentanglement with minimal modifications to the standard VAE loss function by leveraging prior statistical knowledge through auxiliary variables. These variables guide the shaping of the latent space by aligning latent factors with learned auxiliary variables. We validate the efficacy of Aux-VAE through comparative assessments on multiple datasets, including astronomical simulations.

97 MATHEMATICS AND COMPUTING↗

Classification of Cloud Particle Imagery and Thermodynamics (COCPIT): A New Databasing Tool for the Characterization of Cloud Particle Images Captured During DOE Field Campaigns

The Department of Energy for decades has explored the earth system and atmosphere through research and deployment of in-situ and remote sensing platforms during field campaigns. Among these datasets exists a vast supply of cloud particle images that provide visual insight into the complex microphysics in the clouds that span our globe. The millions of images collected over decades of deployments provides a unique opportunity to further our understanding of our atmosphere down to the crystal size. This work over the past 5 years has sought to organize these images into digestible datasets that can then be used by scientists to further our understanding of microphysics. A machine learning model was developed that categorizes over 1.5 million images across 11 weather events with over 90% accuracy according to particle type. The database was then extended to include dimensional characteristics of the particle as well as co-location of environmental properties, such as temperature and water content. Then, to initialize the connection between these data and our understanding of how crystals form and grow, weather research and forecasting simulations were run to generate the growth histories of the classified crystals. This research culminates with 2 databases per event: (1) a database of all classified crystals and their dimensional and environmental properties and (2) simulated growth histories of each crystal. Finally, a user interface was created to allow researchers to explore data statistics.

54 ENVIRONMENTAL SCIENCES↗

Analysis of Warped April Tag Impacts on Detection and Pose Estimation

This report evaluates the impact of geometric deformation on an April Tag, particularly when warped due to attachment on a curved surface, on its detectability and pose estimation performance. A comparative analysis was conducted using a flat April Tag as a control under identical experimental conditions, which involved recording video sequences with varying viewing angles. For detectability, the warped tag exhibited consistent detection failures at viewing angles beyond 40° and complete failures beyond 60°, whereas the flat tag maintained reliable detection across all angles. For pose estimation, measured by pose jitter (variation in rotation and translation), differences between the warped and flat tags were minimal and statistically insignificant, indicating robust performance even for the deformed tag. These findings suggest that while geometric warping reduces an April Tag’s detectability, its pose estimation accuracy remains relatively unaffected under the tested conditions.

42 ENGINEERING↗

Machine learning BPS spectra and the gap conjecture

We explore statistical properties of Bogomol’nyi-Prasad-Sommerfield q-series for strongly coupled supersymmetric theories that correspond to a particular family of three-manifolds. We discover that gaps between exponents in the -series are statistically more significant at the beginning of the -series compared to gaps that appear in higher powers of. Our observations are obtained by calculating saliencies of -series features used as input data for principal component analysis, which is a standard example of an explainable machine learning technique that allows for a direct calculation and a better analysis of feature saliencies.

97 MATHEMATICS AND COMPUTING↗

Benchmark Tracking System for Performance Monitoring

Benchmarking is essential for high-performance software development, particularly for monitoring performance across code iterations. This project focused on enhancing the benchmarking process for Lamellar, an asynchronous runtime for High-Performance Computing (HPC) systems developed at Pacific Northwest National Laboratory. Prior to this work, benchmark results were difficult to track and compare across code versions, presenting significant challenges in identifying performance regressions and long-term trends. The primary objective was to establish a systematic, reproducible approach for measuring performance and detecting regressions following code commits. Our methodology involved three key components: standardizing benchmark outputs, implementing data versioning, and developing analysis tools. We standardized the benchmark output format to JSON Line records containing specific fields (execution time, hardware specifications, and environmental variables). To address data management challenges, we evaluated several options and eventually chose a git repository dedicated to benchmark data. We developed a suite of Python tools that processed benchmark results, enriched them with metadata, and facilitated search in the repository. The resulting system enables more efficient filtering and comparison of performance metrics across commit histories, hardware configurations, and benchmark variants through a unified query interface. Our implementation reduces computational overhead by first checking for existing results through configuration matching before initiating new benchmark runs, thereby conserving resources. The system has been validated by Lamellar developers. It organizes results by benchmark type and build configurations for efficient retrieval. Future developments include a planned Large Language Model interface for predicting benchmark performance, incorporating the criterion package for statistical analysis, which will enable automated detection of statistically significant performance changes, and integration with continuous integration pipelines. Despite these enhancements being reserved for future work, this project has successfully provided the Lamellar development team with a framework for maintaining consistent performance standards and identifying optimization opportunities across workloads and hardware environments.

97 MATHEMATICS AND COMPUTING↗

A Simulation and Optimization Framework for Managing Wind-Driven Loading on PV Systems

As PV modules continue to trend toward larger, thinner, and more flexible forms they grow more susceptible to damage from dynamic wind loading. As a result, understanding the impact of wind on PV systems, particularly when mounted on solar-tracking hardware, and identifying robust, stable array layouts and stow strategies is becoming increasingly important for the PV community. We are developing an open-source software package, PVade (PV aerodynamic design engineering), to simulate the cascading fluid-structure interaction that occurs within solar-tracking arrays to enable researchers to test hardware, layout, and tracker control changes, leading to enhanced stability and a reduction in wind-driven damage. We will give an overview of the PVade software and present the latest outcomes from our ongoing validation campaign in which we compare statistical structural responses with field data. From there, we will present simulated results from a larger, multi-row array and highlight the effect of varying tracker angles on stability as measured by both acceleration and deformation, focusing on the stability differences between positive and negative tilt angles.

fluid structure interaction↗

Semi-Analytical Hierarchical Bayesian Inference of Nonlinear Model Structure in Stochastic Dynamics: Applied to Compartmental Models of Infectious Diseases

A Bayesian computational framework for parsimonious inference in stochastic nonlinear dynamical systems is presented. This framework enables the concurrent estimation of system states, time-varying parameters, time-invariant parameters, and the optimal sparsity structure of the model parameters. Because differential equation-based models are often simplified mechanistic or phenomenological representations, robust inference from noisy measurement data requires explicit treatment of model error and uncertainty. Model error and time-varying parameters can be represented as random processes, enabling inference while making minimal assumptions about the underlying sources of discrepancy and variability. Adopting stochastic differential equation representations affords the model significant flexibility, but can also render it susceptible to overfitting during statistical inversion, where the inferred model may track noise rather than the underlying signal. To alleviate the effects of overfitting and to enable the discovery of the optimal sparse representation of the time-invariant parameters, a Bayesian sparse learning algorithm is embedded within the framework. This sparse learning framework adopts an approximate hierarchical Bayesian setting defined by a series of semi-analytical expressions. The model structure inference framework is validated using a stochastic compartmental model for tracking and forecasting active cases of an infectious disease. Compartmental models describe population-level infectious disease dynamics through interactions among population fractions grouped by disease state. Mathematically, such models consist of a system of coupled ordinary differential equations. This example adopts an expressive compartmental model that includes multiple possible interactions between disease states, motivated by early uncertainty surrounding COVID-19 reinfection dynamics and their implications for long-term epidemic forecasting. The sparse learning exercise permits the inference of a priori unknown epidemiological dynamics from simulated public health data, discovering the nested compartmental model that optimizes the trade-off between average data-fit and model complexity. It is shown that inducing sparsity among the model parameters eliminates redundant interactions between compartments, equivalently revealing the optimal coupling structure between differential equations.

97 MATHEMATICS AND COMPUTING↗

Generating multi-scale Li-ion battery cathode particles with radial grain architectures using stereological generative adversarial networks

Abstract Understanding structure-property relationships of Li-ion battery cathodes is crucial for optimizing rate-performance and cycle-life resilience. However, correlating the morphology of cathode particles, such as in LiNi0.8Mn0.1Co0.1O2 (NMC811), and their inner grain architecture with electrode performance is challenging, particularly, due to the significant length-scale difference between grain and particle sizes. Experimentally, it is not feasible to image such a high number of particles with full granular detail. A second challenge is that sufficiently high-resolution 3D imaging techniques remain expensive and are sparsely available at research institutions. Here, we present a stereological generative adversarial network-based model fitting approach to tackle this, that generates representative 3D information from 2D data, enabling characterization of materials in 3D using cost-effective 2D data. Once calibrated, this multi-scale model can rapidly generate virtual cathode particles that are statistically similar to experimental data, and thus is suitable for virtual characterization and materials testing through numerical simulations. A large dataset of simulated particles with inner grain architecture has been made publicly available.

25 ENERGY STORAGE↗

Effect of Solvent on the Local Structure, Dynamics, and Vibrational Density of States in Sn-BEA Zeolite

Lewis acid zeolites are attractive catalysts for epoxidation and biomass valorization, as they are highly active and selective in the liquid phase and can operate at or near ambient conditions. While a rich experimental literature exists on liquid-phase Lewis acid zeolite catalysis, our understanding of the molecular organization and solvent dynamics in the vicinity of Lewis acid sites with differing metal site speciation remains limited. In this work, we investigate the molecular coordination and diffusion of two common solvents (methanol and water) around the closed and open Sn-BEA zeolite active sites using molecular dynamics simulations with a machine-learned interatomic potential trained on ab initio molecular dynamics trajectories. Molecular dynamics simulations reveal that introducing active sites significantly enhances local order in the first and second solvation shells compared to the pure silica case. For methanol, both closed and open active sites are singly coordinated, while more than two water molecules coordinate the open site. In contrast to methanol, we observed that water molecules dissociate, leading to the formation of additional Sn-OH and silanol groups away from the active site. The diffusion coefficients of water and methanol are functions of the solvent population in the pore. Here, our work provides insights into how active site speciation in Lewis acid zeolites affects solvent coordination, diffusion, and vibrational signature. This information is foundational for catalyst design and optimization of liquid-phase catalytic processes in zeolites. It also demonstrates the suitability of machine-learned interatomic potentials for modeling reactive systems, enabling sufficiently long trajectories for appropriate statistical averaging.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Enhancing ZFP: A Statistical Approach to Understanding and Reducing Error Bias in a Lossy Floating-Point Compression Algorithm

The amount of data generated and gathered in scientific simulations and data collection applications is continuously growing, putting mounting pressure on storage and bandwidth concerns. A means of reducing such issues is data compression; but, lossless data compression is typically ineffective when applied to floating-point data. Thus, users tend to apply a lossy data compressor, which allows for small deviations from the original data. It is essential to understand how the error from lossy compression impacts the accuracy of the data analytics. Thus, we must analyze not only the compression properties but the error as well. In this paper, we provide a statistical analysis of the error caused by ZFP compression, a state-of-the-art, lossy compression algorithm explicitly designed for floating-point data. We show that the error is indeed biased and propose simple modifications to the algorithm to neutralize the bias and further reduce the resulting error.

97 MATHEMATICS AND COMPUTING↗