Search NASASearch

SEARCH · Search NASA

Results for “Statistical methods”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 37 records · Page 2

Report on the AAPM grand challenge on deep generative modeling for learning medical image statistics

Abstract Background The findings of the 2023 AAPM Grand Challenge on Deep Generative Modeling for Learning Medical Image Statistics are reported in this Special Report. Purpose The goal of this challenge was to promote the development of deep generative models for medical imaging and to emphasize the need for their domain‐relevant assessments via the analysis of relevant image statistics. Methods As part of this Grand Challenge, a common training dataset and an evaluation procedure was developed for benchmarking deep generative models for medical image synthesis. To create the training dataset, an established 3D virtual breast phantom was adapted. The resulting dataset comprised about 108 000 images of size 512 512. For the evaluation of submissions to the Challenge, an ensemble of 10 000 DGM‐generated images from each submission was employed. The evaluation procedure consisted of two stages. In the first stage, a preliminary check for memorization and image quality (via the Fréchet Inception Distance [FID]) was performed. Submissions that passed the first stage were then evaluated for the reproducibility of image statistics corresponding to several feature families including texture, morphology, image moments, fractal statistics, and skeleton statistics. A summary measure in this feature space was employed to rank the submissions. Additional analyses of submissions was performed to assess DGM performance specific to individual feature families, the four classes in the training data, and also to identify various artifacts. Results Fifty‐eight submissions from 12 unique users were received for this Challenge. Out of these 12 submissions, 9 submissions passed the first stage of evaluation and were eligible for ranking. The top‐ranked submission employed a conditional latent diffusion model, whereas the joint runners‐up employed a generative adversarial network, followed by another network for image superresolution. In general, we observed that the overall ranking of the top 9 submissions according to our evaluation method (i) did not match the FID‐based ranking, and (ii) differed with respect to individual feature families. Another important finding from our additional analyses was that different DGMs demonstrated similar kinds of artifacts. Conclusions This Grand Challenge highlighted the need for domain‐specific evaluation to further DGM design as well as deployment. It also demonstrated that the specification of a DGM may differ depending on its intended use.

Radiology, Nuclear Medicine & Medical Imaging

Dynamic data-driven multiscale modeling for predicting the degradation of a 316L stainless steel nuclear cladding material

Here, we have developed a long short-term memory stacked ensemble (LSTM-SE) surrogate modeling approach that can provide rapid predictions of microstructural evolution and the resultant mechanical properties of American Iron and Steel Institute (AISI) 316L series stainless steel (316LSS) fuel cladding under conditions of varying temperature and radiation dose rate. To acquire training data, we developed and implemented a kinetic Monte Carlo (KMC) model to simulate precipitation kinetics of M 23 C 6 , γ', and G phases within SS316L cladding. Experimentally reported precipitation kinetics of SS316L in literature were linked to the kinetic parameters of the simulated precipitation in our KMC model. The model was then used to simulate microstructure evolution under synthetically generated treatments of varying temperature and radiation dose rate, for periods of up to 3000 hours. Changes in volume fraction, number density, and particle size of precipitates were recorded, and particle area fractions were correlated using statistical methods to develop the surrogate model. Simultaneously, the mechanical properties of the simulated microstructures were evaluated using microstructure-based finite element method (FEM) analysis to determine the elastic modulus, yield stress, ultimate tensile strength, and elongation to failure of the aged microstructures. Using this approach, our surrogate model can predict precipitation behavior within 0.25% volume fraction and mechanical properties within 6% relative error from the values predicted by the KMC and FEM models using 50 training simulations as input. The trained recurrent neural network-based model can return estimations of precipitation kinetics and mechanical properties ~1000 times faster than the physics-based codes. This work demonstrates, as a proof of concept, that reactor material service lifetimes under variable service conditions can be predicted for a statistics-based model from a practicably obtainable dataset.

36 MATERIALS SCIENCE

Data-driven upper bounds and event attribution for unprecedented heatwaves

The last decade has seen numerous record-shattering heatwaves in all corners of the globe. In the aftermath of these devastating events, there is interest in identifying worst-case thresholds or upper bounds that quantify just how hot temperatures can become. Generalized Extreme Value theory provides a data-driven estimate of extreme thresholds; however, upper bounds may be exceeded by future events, which undermines attribution and planning for heatwave impacts. Here, we show how the occurrence and relative probability of observed yet unprecedented events that exceed a priori upper bound estimates, so-called “impossible” temperatures, has changed over time. We find that many unprecedented events are actually within data-driven upper bounds, but only when using modern spatial statistical methods. Furthermore, there are clear connections between anthropogenic forcing and the “impossibility” of the most extreme temperatures. Robust understanding of heatwave thresholds provides critical information about future record-breaking events and how their extremity relates to historical measurements.

54 ENVIRONMENTAL SCIENCES

Analysis of differential scanning calorimetry data for aged plutonium

Differential scanning calorimetry data for samples of a 52 year old plutonium alloy with 3.3 at. % Ga that were heated beyond the melting point is analyzed using transition state theory to find activation energies for the δ to ε and ε to liquid phase transitions. A Bayesian statistical method involving a Gaussian process model is used to find mean values and confidence intervals for the activation energies. The activation energy for the δ to ε phase transition increases by 3.3 ± 3.8% per decade, relative to the case when all age related plutonium lattice point defects have been removed through annealing. The corresponding increase in activation energy for the ε to liquid transition is shown to be 7.1 ± 1.8% per decade. It is postulated that the change in activation energy with age for both phase transitions is caused, in part, by the accumulation of the same type of lattice point defects associated with the observed increase in elastic bulk modulus over time.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH

Quantifying biases in stellar masses of JWST high- z quasar host galaxies caused by quasar subtraction

The James Webb Space Telescope (JWST) has enabled dozens of high-z quasar host galaxy detections. Many of these observations imply galaxies with black holes that are overmassive compared to their low-z counterparts. However, the bright quasar point source removal can cause significant biases in recovered host magnitudes and stellar mass measurements due to the degeneracy in host galaxy and quasar light. We develop a statistical method to disentangle the quasar host galaxy stellar mass measurements from observational biases during the point source removal assuming the PSF is modelled perfectly. We use the BlueTides simulation to generate mock images and perform point source removal on thousands of simulated high-z quasar host galaxies, constructing corrected host magnitude posteriors. We find that removing a bright quasar in JWST photometry tends to either correctly recover or modestly misestimate host magnitudes, with a maximum magnitude underestimate of 0.2 mag. With our corrected magnitude posteriors, we perform SED fitting on each quasar host galaxy and compare the stellar mass measurement before and after the correction. We find that stellar mass estimates are generally robust, or misestimated by $<$ 0.3 dex. We also find that the stellar masses of a subset of hosts (J0844−0132, J0911+0152, and J1146−0005) remain unconstrained, as key photometric bands provide only flux upper limits. Accounting for observational biases does not resolve the apparent mismatch between black hole and host galaxy growth at high-z, where some quasars appear to host overmassive black holes while others reside in relatively massive galaxies.

79 ASTRONOMY AND ASTROPHYSICS

Predicting RNA structure and dynamics with deep learning and solution scattering

Advanced deep learning and statistical methods can predict structural models for RNA molecules. However, RNAs are flexible, and it remains difficult to describe their macromolecular conformations in solutions where varying conditions can induce conformational changes. Small-angle x-ray scattering (SAXS) in solution is an efficient technique to validate structural predictions by comparing the experimental SAXS profile with those calculated from predicted structures. There are two main challenges in comparing SAXS profiles to RNA structures: the absence of cations essential for stability and charge neutralization in predicted structures and the inadequacy of a single structure to represent RNA’s conformational plasticity. We introduce a solution conformation predictor for RNA (SCOPER) to address these challenges. This pipeline integrates kinematics-based conformational sampling with the innovative deep learning model, IonNet, designed for predicting Mg 2+ ion binding sites. Validated through benchmarking against 14 experimental data sets, SCOPER significantly improved the quality of SAXS profile fits by including Mg 2+ ions and sampling of conformational plasticity. We observe that an increased content of monovalent and bivalent ions leads to decreased RNA plasticity. Therefore, carefully adjusting the plasticity and ion density is crucial to avoid overfitting experimental SAXS data. SCOPER is an efficient tool for accurately validating the solution state of RNAs given an initial, sufficiently accurate structure and provides the corrected atomistic model, including ions.

59 BASIC BIOLOGICAL SCIENCES

Leveraging interpolation models and error bounds for verifiable scientific machine learning

Effective verification and validation techniques for modern scientific machine learning workflows are challenging to devise. Statistical methods are abundant and easily deployed, but often rely on speculative assumptions about the data and methods involved. Error bounds for classical interpolation techniques can provide mathematically rigorous estimates of accuracy, but often are difficult or impractical to determine computationally. Here, in this work, we present a best-of-both-worlds approach to verifiable scientific machine learning by demonstrating that (1) multiple standard interpolation techniques have informative error bounds that can be computed or estimated efficiently; (2) comparative performance among distinct interpolants can aid in validation goals; (3) deploying interpolation methods on latent spaces generated by deep learning techniques enables some interpretability for black-box models. We present a detailed case study of our approach for predicting lift-drag ratios from airfoil images. Code developed for this work is available in a public Github repository.

97 MATHEMATICS AND COMPUTING

Formation of 1 H -Phenalene (C 13 H 10 ) in the Taurus Molecular Cloud via Methylidyne Addition-Cyclization-Aromatization (MACA)

The formation of 1H-phenalene (C 13 H 10 ) in cold molecular clouds, such as the Taurus Molecular Cloud-1 (TMC-1), presents a significant challenge to traditional astrochemical models, which predominantly suggest high-temperature pathways for polycyclic aromatic hydrocarbon (PAH) formation. In this study, we explore computationally the Methylidyne Addition-Cyclization-Aromatization (MACA) mechanism as a viable, barrierless pathway for phenalene synthesis under low-temperature conditions. Through electronic structure calculations and Rice–Ramsperger–Kassel–Marcus (RRKM) statistical methods, we demonstrate that the reaction of 1-vinylnaphthalene (C 10 H 7 C 2 H 3 ) with the methylidyne radical (CH) leads to the formation of 1H-phenalene via a bimolecular reaction, a process that is exoergic and without entrance barrier. The MACA mechanism facilitates the growth of the aromatic carbon backbone via a [5 + 1] ring annulation, providing a new insight into PAH formation in cold molecular clouds. Notably, the MACA mechanism has previously been shown to form indene (C9H8), which was detected in TMC-1 as well, via a [4 + 1] annulation, demonstrating its potential to produce a variety of complex PAHs by addition of a five- and six-membered ring to a benzene moiety via [4 + 1] and [5 + 1] annulation, respectively. As a result, this work highlights the importance of barrierless, exoergic reactions involving MACA in the synthesis of complex aromatic molecules in space, expanding our physicochemical understanding of carbon-rich chemistry in cold molecular clouds.

Aromatic compounds

Effect of adaptive cruise control on fuel consumption in real-world driving conditions

This paper presents a comprehensive analysis of the impact of adaptive cruise control on energy consumption in real-world driving conditions based on a natural experiment: a large-scale observational dataset of driving data from a diverse fleet of vehicles and drivers. The analysis is conducted at two different fidelity levels: (1) a macroscopic trip-level benefit estimate that compares trips with and without cruise control in a counterfactual way using statistical methods, and (2) a situation-based comparison achieved through the segmentation of trips into distinct driving situations such as acceleration, braking, cruising, and other maneuvers. The results of this research show that the effect of cruise control on energy consumption varies across different driving situations and levels of analysis. In a macroscopic trip-level analysis, cruise control engagement is associated with a slight increase in fuel consumption across the fleet. As revealed later by the situation-based analysis, this result can be attributed to the negative impact of cruise control on energy consumption in cruising mode, which is the most common driving situation. However, the situation-based comparison demonstrates that cruise control can provide fuel consumption benefits in situations involving acceleration and braking, particularly when a preceding vehicle is present. The study also emphasizes the importance of controlling for various factors that can influence both fuel consumption and the likelihood of cruise control engagement to properly evaluate its effects.

33 ADVANCED PROPULSION SYSTEMS

Autonomous platform for solution processing of electronic polymers

The manipulation of electronic polymers’ solid-state properties through processing is crucial in electronics and energy research. Yet, efficiently processing electronic polymer solutions into thin films with specific properties remains a formidable challenge. We introduce Polybot, an artificial intelligence (AI) driven automated material laboratory designed to autonomously explore processing pathways for achieving high-conductivity, low-defect electronic polymers films. Leveraging importance-guided Bayesian optimization, Polybot efficiently navigates a complex 7-dimensional processing space. In particular, the automated workflow and algorithms effectively explore the search space, mitigate biases, employ statistical methods to ensure data repeatability, and concurrently optimize multiple objectives with precision. The experimental campaign yields scale-up fabrication recipes, producing transparent conductive thin films with averaged conductivity exceeding 4500 S/cm. Feature importance analysis and morphological characterizations reveal key design factors. This work signifies a significant step towards transforming the manufacturing of electronic polymers, highlighting the potential of AI-driven automation in material science.

Wang, Chengshi [Argonne National Laboratory (ANL),

Anticipating decoherence in quantum systems

Large-scale quantum technologies require coherence across distant nodes, necessitating indistinguishable quantum states. However, environmental disorder, including dephasing, spectral diffusion, and spin-bath interactions, undermines coherence. Using statistical methods, we uncover correlations in decoherence channels induced by slowly varying environments. Spectral diffusion serves as a representative demonstration case that can be extended to other remote, disordered systems such as spins in nitrogen-vacancy centers and quantum-dot spin qubits, as well as flux noise in superconducting qubits. In this work, we employ replica-theory-inspired trajectory analysis to reveal predictable temporal structures in decoherence dynamics, and validate these through an anticipatory systems framework with internal prediction of unseen spectral dynamics in multiple quantum systems, showing that this framework could, if implemented, reduce spectral shift by average factors of approximately 2 to 19, depending on emitter stability, thereby enabling enhanced coherence and multi-node synchronization for scalable quantum communication, computation, imaging, and sensing.

Maan, Pranshu [Purdue University]

Bayesian stability and force modeling for uncertain machining processes

Accurately simulating machining operations requires knowledge of the cutting force model and system frequency response. However, this data is collected using specialized instruments in an ex-situ manner. Bayesian statistical methods instead learn the system parameters using cutting test data, but to date, these approaches have only considered milling stability. This paper presents a physics-based Bayesian framework which incorporates both spindle power and milling stability. Initial probabilistic descriptions of the system parameters are propagated through a set of physics functions to form probabilistic predictions about the milling process. The system parameters are then updated using automatically selected cutting tests to reduce parameter uncertainty and identify more productive cutting conditions, where spindle power measurements are used to learn the cutting force model. The framework is demonstrated through both numerical and experimental case studies. Results show that the approach accurately identifies both the system natural frequency and cutting force model.

42 ENGINEERING

Granger causal inference for climate change attribution

Abstract Climate change detection and attribution (D&A) is concerned with determining the extent to which anthropogenic activities have influenced specific aspects of the global climate system. D&A fits within the broader field of causal inference, the collection of statistical methods that identify cause and effect relationships. There are a wide variety of methods for making attribution statements, each of which require different types of input data and focus on different types of weather and climate events and each of which are conditional to varying extents. Some methods are based on Pearl causality (direct experimental interference) while others leverage Granger (predictive) causality, and the causal framing provides important context for how the resulting attribution conclusion should be interpreted. However, while Granger-causal attribution analyses have become more common, there is no clear statement of their strengths and weaknesses relative to Pearl-causal attribution and no clear consensus on where and when Granger-causal perspectives are appropriate. In this prospective paper, we provide a formal definition for Granger-based approaches to trend and event attribution and a clear comparison with more traditional methods for assessing the human influence on extreme weather and climate events. Broadly speaking, Granger-causal attribution statements can be constructed quickly from observations and do not require computationally-intesive dynamical experiments. These analyses also enable rapid attribution, which is useful in the aftermath of a severe weather event, and provide multiple lines of evidence for anthropogenic climate change when paired with Pearl-causal attribution. Confidence in attribution statements is increased when different methodologies arrive at similar conclusions. Moving forward, we encourage the D&A community to embrace hybrid approaches to climate change attribution that leverage the strengths of both Granger and Pearl causality.

Risser, Mark D. (ORCID:0000000319561783)

Analysis and optimization of seismic monitoring networks with Bayesian optimal experimental design

SUMMARY Monitoring networks increasingly aim to assimilate data from a large number of diverse sensors covering many sensing modalities. Bayesian optimal experimental design (OED) seeks to identify data, sensor configurations or experiments which can optimally reduce uncertainty and hence increase the performance of a monitoring network. Information theory guides OED by formulating the choice of experiment or sensor placement as an optimization problem that maximizes the expected information gain (EIG) about quantities of interest given prior knowledge and models of expected observation data. Therefore, within the context of seismo-acoustic monitoring, we can use Bayesian OED to configure sensor networks by choosing sensor locations, types and fidelity in order to improve our ability to identify and locate seismic sources. In this work, we develop the framework necessary to use Bayesian OED to optimize a sensor network’s ability to locate seismic events from arrival time data of detected seismic phases at the regional-scale. This framework requires five elements: (i) A likelihood function that describes the distribution of detection and traveltime data from the sensor network, (ii) A prior distribution that describes a priori belief about seismic events, (iii) A Bayesian solver that uses a prior and likelihood to identify the posterior distribution of seismic events given the data, (iv) An algorithm to compute EIG about seismic events over a data set of hypothetical prior events, (v) An optimizer that finds a sensor network which maximizes EIG. Once we have developed this framework, we explore many relevant questions to monitoring such as: how to trade off sensor fidelity and earth model uncertainty; how sensor types, number and locations influence uncertainty; and how prior models and constraints influence sensor placement.

58 GEOSCIENCES

The 3D Lyman- α forest power spectrum from eBOSS DR16

We measure the three-dimensional power spectrum (P3D) of the transmitted flux in the Lyman-α (Ly α) forest using the complete extended Baryon Oscillation Spectroscopic Survey data release 16 (eBOSS DR16). This sample consists of ~205 000 quasar spectra in the redshift range 2 ≤ z ≤ 4 at an effective redshift z = 2.334. We propose a pair-count spectral estimator in configuration space, weighting each pair by exp( i k ∙ r), for wave vector k and pixel pair separation r, effectively measuring the anisotropic power spectrum without the need for fast Fourier transforms. This accounts for the window matrix in a tractable way, avoiding artefacts found in Fourier-transform based power spectrum estimators due to the sparse sampling transverse to the line of sight of Ly α skewers. We extensively test our pipeline on two sets of mocks: (i) idealized Gaussian random fields with a sparse sampling of Ly α skewers, and (ii) log-normal LyaCoLoRe mocks including realistic noise levels, the eBOSS survey geometry and contaminants. On eBOSS DR16 data, the Kaiser formula with a non-linear correction term obtained from hydrodynamic simulations yields a good fit to the power spectrum data in the range $(0.02 ≤ k ≤ 0.35)$ h Mpc -1 at the 1–2σ level with a covariance matrix derived from LyaCoLoRe mocks. We demonstrate a promising new approach for full-shape cosmological analyses of Ly α forest data from cosmological surveys such as eBOSS, the currently observing Dark Energy Spectroscopic Instrument and future surveys such as the Prime Focus Spectrograph, WEAVE-QSO, and 4MOST.

79 ASTRONOMY AND ASTROPHYSICS

Cluster spin glass correlations and dynamics in Zn 0.5⁢ Mn 0.5⁢ Te

Here, we present a combined magnetometry, muon spin-relaxation (𝜇⁢SR), and neutron-scattering study of the insulating spin glass Zn 0.5 ⁢Mn 0.5 ⁢Te, for which magnetic Mn 2+ and nonmagnetic Zn 2+ ions are randomly distributed on a face-centered cubic lattice. The magnetometry and 𝜇⁢SR results confirm a spin freezing transition around 𝑇 𝑓 ≈ 23 K, with the spin-fluctuation rate decreasing gradually and somewhat inhomogeneously through the sample volume as the temperature decreases toward 𝑇 𝑓 . Characteristic spin-correlation times well above 𝑇 𝑓 are on the order of 10 −10 s, much slower than typically observed in canonical spin glasses but in line with expectations for a cluster spin glass. Using magnetic pair distribution function (mPDF) analysis and reverse Monte Carlo (RMC) modeling of the magnetic diffuse neutron-scattering data, we show that the spin-glass ground state consists of clusters of spins exhibiting short-range-ordered type-III antiferromagnetic correlations with a locally ordered moment of 3.1⁢(1)⁢𝜇 B between nearest-neighbor spins. The type-III correlations decay exponentially as a function of spin separation distance with a correlation length of approximately 5 Å. The diffuse magnetic scattering and corresponding mPDF show no significant changes across 𝑇 𝑓 , indicating that the dynamically fluctuating short-range spin correlations in the paramagnetic state retain the same basic type-III configuration that characterizes the spin-glass state; the only change apparent from the neutron-scattering data is a gradual reduction of the correlation length and locally ordered moment with increasing temperature. Taken together, these results paint a unique and detailed picture of the local magnetic structure and dynamics in Zn 0.5 ⁢Mn 0.5⁢ Te and provide strong evidence that this material is best described as a cluster spin glass. In addition, this work showcases a statistical method for extracting diffuse scattering signals from neutron powder diffraction data, which we developed to facilitate the mPDF and RMC analysis of the neutron data. This method has the potential to be broadly useful for neutron powder diffraction experiments on a variety of materials with short-range atomic or magnetic order.

magnetism

Coherency-Constrained Spectral Clustering for Power Network Reduction

This paper presents a methodology for reducing the complexity of large-scale power network models using spectral clustering, aggregation of electrical components, and cost function approximation. Two approaches are explored using unconstrained and constrained spectral clustering to determine areas for effective system reduction. Once the system areas are determined, both loads and generators by type are aggregated, and their new cost function is approximated through polynomial curve-fitting or statistical methods. The performance of reduced networks is evaluated in terms of their ability to follow the true daily cost of the original system over a 24-hour period considering a set of several days. Two test systems are taken as test beds. Application of the methodology to a modified version of the IEEE 39-bus system reduces it from 17 generators to a 4-bus system and 9 generators with about 93% of accuracy. Similarly, the IEEE 118-bus system is reduced from 19 generators to a 3-bus system with three aggregated units achieving over 99% of accuracy. These findings address scalability challenges and enhance accuracy for high and mid-loading level conditions, and by aggregating thermal units with similar cost functions.

42 ENGINEERING

Unraveling design principles of protein landscapes in photosynthetic membranes in plant chloroplasts

The supramolecular organization of proteins within photosynthetic membranes is crucial for energy conversion in plants. Here, we introduce an analytical and computational pipeline that integrates high-resolution cryo–scanning electron microscopy, biochemical quantification, advanced Monte Carlo computer simulations, and statistical methods to elucidate the elusive protein landscapes of grana membranes in intact Arabidopsis leaves. Our integrated analysis challenges the prevailing view that particles on the exoplasmic fracture faces in freeze-fracture samples represent photosystem II exclusively. Instead, these particles also include cytochrome b 6 f complexes. Furthermore, our steric clash analysis demonstrates that stacked membranes contain a mixture of larger PSII supercomplexes (C 2 S 2 M 2 and C 2 S 2 ) in addition to a smaller complex (C 2 ). This suggests that in vivo PSII supercomplexes exist in an equilibrium distribution of differing sizes. Furthermore, we discovered that, although size exclusion effects govern the global protein arrangement, local packing exhibits orientational order indicative of lateral attractive protein-protein interactions.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH