Search NASA⌕ Search

SEARCH · Search NASA

Results for “Bayesian sampling”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 127 records · Page 7

Statistical data analysis of x-ray spectroscopy data enabled by neural network accelerated Bayesian inference

Bayesian inference applied to x-ray spectroscopy data analysis enables uncertainty quantification necessary to rigorously test theoretical models. However, when comparing to data, detailed atomic physics and radiation transfer calculations of x-ray emission from non-uniform plasma conditions are typically too slow to be performed in line with statistical sampling methods, such as Markov Chain Monte Carlo sampling. Furthermore, differences in transition energies and x-ray opacities often make direct comparisons between simulated and measured spectra unreliable. Here, we present a spectral decomposition method that allows for corrections to line positions and bound–bound opacities to best fit experimental data, with the goal of providing quantitative feedback to improve the underlying theoretical models and guide future experiments. In this work, we use a neural network (NN) surrogate model to replace spectral calculations of isobaric hot-spots created in Kr-doped implosions at the National Ignition Facility. The NN was trained on calculations of x-ray spectra using an isobaric hot-spot model post-processed with Cretin, a multi-species atomic kinetics and radiation code. The speedup provided by the NN model to generate x-ray emission spectra enables statistical analysis of parameterized models with sufficient detail to accurately represent the physical system and extract the plasma parameters of interest.

47 OTHER INSTRUMENTATION↗

A copula-based rank histogram ensemble filter

Serial ensemble filters implement triangular probability transport maps to reduce high-dimensional inference problems to sequences of state-by-state univariate inference problems. The univariate inference problems are solved by sampling posterior probability densities obtained by combining constructed prior densities with observational likelihoods according to Bayes' rule. Many serial filters in the literature focus on representing the marginal posterior densities of each state. However, rigorously capturing the conditional dependencies between the different univariate inferences is crucial to correctly sampling multidimensional posteriors. This work proposes a new serial ensemble filter, called the copula rank histogram filter (CoRHF), that seeks to capture the conditional dependency structure between variables via empirical copula estimates; these estimates are used to rigorously implement the triangular (state-by-state univariate) Bayesian inference. The success of the CoRHF is demonstrated on two-dimensional examples and the Lorenz'63 problem. A practical extension to the high-dimensional setting is developed by localizing the empirical copula estimation, and is demonstrated on the Lorenz'96 problem.

97 MATHEMATICS AND COMPUTING↗

Investigating Kinetic Mechanisms of Soot Formation in Plasma Pyrolysis of Methane via Active Learning (Final Technical Report)

Plasma pyrolysis of methane is an effective route for zero-carbon hydrogen production. Yet, soot generated from pyrolysis of hydrocarbons is detrimental to the climate and human health. There is ample experimental and theoretical evidence that suggests polycyclic aromatic hydrocarbons (PAHs) are the molecular precursors to soot particles. The reaction pathways of PAH formation are intricately dependent on a multitude of process parameters, whose kinetic mechanisms are not well-understood in plasma pyrolysis. This project aims to leverage advances in the kinetic modeling of soot formation in combustion, as well as in surrogate modeling and active learning, to systematically investigate the effects of process parameter on the kinetics of PAH formation in plasma pyrolysis of methane. To this end, we propose to use the PAH formation kinetics model developed by the PPPL/PU group based on the well-established ABF and HACA mechanisms, coupled with low-temperature plasma models. We will develop an active learning (AL) framework based on Bayesian optimization to systematically and data-efficiently explore the complex and multivariable parameter space of plasma pyrolysis in order to quantify the effects of plasma and feed parameters on the ABF and HACA kinetic pathways. AL is the branch of machine learning concerned with systematically querying samples from a system (experimental or computational) to train a data-driven model that maps design parameters to a performance criterion. We will use the data generated via AL to perform global sensitivity analysis, combined with uncertainty quantification, to elucidate the impact of different reaction pathways on minimizing formation of soot precursors. This study will result in an improved understanding of kinetics of PAH formation in plasma pyrolysis and can pave the way for more advanced mechanistic studies (e.g., soot nucleation mechanisms). Additionally, the findings will be useful for establishing practical strategies for increasing the pyrolysis efficiency and producing high-grade carbon for synthesis of nanomaterials.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Bayesian Fit to NOvA Data Subsamples for Three Flavor Oscillation Analysis

NOvA (NuMI Off-Axis $\nu_e$ Appearance) is a long baseline neutrino experiment designed to measure the oscillation of muon neutrinos to electron neutrinos over a distance of 810 km. NOvA uses a near and far detector to observe $\nu_\mu$ disappearance and $\nu_e$ appearance of neutrinos produced by the NuMI beam at Fermilab. NOvA uses a Bayesian analysis framework in addition to its Frequentist method to measure neutrino oscillation parameters such as the mixing angles, mass ordering, and CP-violating phase. We report preliminary results of Bayesian fits to representative NOvA datasets. Comparison of fits to $\nu_\mu$ disappearance and $\nu_e$ appearance enables a cross-check of NOvA results with reactor $\bar{\nu_e}$ disappearance measurements. NOvA also searches for violation of Lorentz invariance by analyzing fits of forward horn current (FHC) versus reverse horn current (RHC) samples. The results validate and advance NOvA's contributions to precision measurements of neutrino properties.

Zhao, Larry [Fermilab]↗

Optimal experimental design: Formulations and computations

Questions of ‘how best to acquire data’ are essential to modelling and prediction in the natural and social sciences, engineering applications, and beyond. Optimal experimental design (OED) formalizes these questions and creates computational methods to answer them. This article presents a systematic survey of modern OED, from its foundations in classical design theory to current research involving OED for complex models. We begin by reviewing criteria used to formulate an OED problem and thus to encode the goal of performing an experiment. We emphasize the flexibility of the Bayesian and decision-theoretic approach, which encompasses information-based criteria that are well-suited to nonlinear and non-Gaussian statistical models. We then discuss methods for estimating or bounding the values of these design criteria; this endeavour can be quite challenging due to strong nonlinearities, high parameter dimension, large per-sample costs, or settings where the model is implicit. A complementary set of computational issues involves optimization methods used to find a design; we discuss such methods in the discrete (combinatorial) setting of observation selection and in settings where an exact design can be continuously parametrized. Finally we present emerging methods for sequential OED that build non-myopic design policies, rather than explicit designs; these methods naturally adapt to the outcomes of past experiments in proposing new experiments, while seeking coordination among all experiments to be performed. Throughout, we highlight important open questions and challenges.

97 MATHEMATICS AND COMPUTING↗

Official data release for Bayes 2022 paper, arXiv:2311.07835

Official data release to accompany the first NOvA Bayesian results paper, https://arxiv.org/abs/2311.07835. The included `README.md` below describes the contents more fully, but in short, included here are: * A `README.md` describing the files * Four `.root` files containing marginal posterior densities in neutrino oscillation parameters, and associated 1D and 2D credible regions derived from them * One `.root` file containing a `TTree` with Markov Chain Monte Carlo samples, which can be used to recreate the posteriors above. **Please use the URL below to download this large (2GB) file** https://mod.fnal.gov/cwsmod/n/nova_docdb_ext/arxiv-2311.07835.data-release.mcmcsamples.root * One Jupyter notebook giving extensive examples how the MCMC samples above can be used

Collaboration, NOvA↗

Active oversight and quality control in standard Bayesian optimization for autonomous experiments

The fusion of experimental automation and machine learning has catalyzed a new era in materials research, prominently featuring Gaussian Process (GP) Bayesian Optimization (BO) driven autonomous experiments. Here we introduce a Dual-GP approach that enhances traditional GPBO by adding a secondary surrogate model to dynamically constrain the experimental space based on real-time assessments of the raw experimental data. This Dual-GP approach enhances the optimization efficiency of traditional GPBO by isolating more promising space for BO sampling and more valuable experimental data for primary GP training. We also incorporate a flexible, human-in-the-loop intervention method in the Dual-GP workflow to adjust for unanticipated results. We demonstrate the effectiveness of the Dual-GP model with synthetic model data and implement this approach in autonomous pulsed laser deposition experimental data. This Dual-GP approach has broad applicability in diverse GPBO-driven experimental settings, providing a more adaptable and precise framework for refining autonomous experimentation for more efficient optimization.

36 MATERIALS SCIENCE↗

Bayesian estimation of HIV acquisition dates for prevention trials

Accurate timing estimates of when participants acquire HIV in HIV prevention trials are necessary for determining antibody levels at acquisition. The Antibody-Mediated Prevention (AMP) Studies showed that a passively administered broadly neutralizing antibody can prevent the acquisition of HIV from a neutralization-sensitive virus. We developed a pipeline for estimating the date of detectable HIV acquisition (DDA) in AMP Study participants using diagnostic and viral sequence data. Using a Bayesian strategy that combines three streams of data (REN [rev/vpu/env/Δnef] sequence, GP [gag/Δpol] sequence, and diagnostic) where their 95% credible intervals overlap based on pre-specified criteria and decision rules. We evaluated the performance of our AMP pipeline using PacBio viral sequence data from 41 participants across two prospective acute HIV acquisition cohort studies, FRESH and RV217, with twice-weekly sampling. These cohort studies enrolled young women in South Africa and men and women in Kenya and Thailand, respectively, with a high likelihood of HIV acquisition. In evaluating performance, “true DDA” was the center of bounds between last-negative and first-positive RNA diagnostic tests (median time 4 days, range 2–7 days); bias was the mean difference between estimated and true DDA. Using diagnostic data alone yielded timing estimates with a bias of 2.4 days and root mean square error (RMSE) of 7.9 days. These results were improved using sequence + diagnostic data (bias 1.5 days, RMSE 6.9 days), as well as by restricting sequence-based estimation to samples from ≤5 weeks post-DDA (bias 0.2 days, RMSE 7.8 days).

59 BASIC BIOLOGICAL SCIENCES↗

California Trees Seasonally Use Augmented Water Sources: Water Isotope Tracking in a Groundwater‐Dependent Ecosystem

Sustainable groundwater management must account for the needs of groundwater dependent ecosystems. To understand the relationship of ecosystems and seasonal water use, we studied the stable isotope composition (δ 18 O and δ 2 H) of water in streamside trees in a semi-arid streamside environment (Livermore, California, USA). We sampled seven trees at two sites every other month from April 2024 through April 2025 for tree xylem stable water isotope signatures. These data were compared to potential source waters: precipitation, imported surface water, soil water and regional groundwaters. Large daily precipitation events were found to be isotopically similar to regional groundwater and were thus treated as one water source. A Bayesian mixing model using stable water isotopes was used to determine the ratios of these three potential source waters (small daily precipitation events, groundwater/large daily precipitation events and imported water) present in tree xylem water. On average, tree water sources include 32% imported water (SD = 9%), 35% small daily precipitation events (SD = 10%) and 33% groundwater (SD = 3%), with significant seasonal variation (t summer-winter = 30.8, p < 0.01), particularly drawing more imported water (more than 55%) in the summer. While small daily precipitation events contribute only 10% of the total precipitation in our dataset, it represents a third of water used by trees. In addition, while these ecosystems are designated as groundwater dependent ecosystems, the trees use approximately one third imported water and even more during dry summer months. This approach provides water managers with a practical tool for quantifying ecosystem water needs, supporting data-driven decisions and regulatory compliance. While California's Sustainable Groundwater Management Act emphasizes supporting groundwater dependent ecosystems, it allows flexibility in demonstrating benefits. In conclusion, our methodology provides a way to document that management actions (such as managed aquifer recharge with imported water) deliver measurable co-benefits to GDEs.

Geosciences↗

Deep inference of simulated strong lenses in ground-based surveys

The large number of strong lenses discoverable in future astronomical surveys will likely enhance the value of strong gravitational lensing as a cosmic probe of dark energy and dark matter. However, leveraging the increased statistical power of such large samples will require further development of automated lens modeling techniques. We show that deep learning and simulation-based inference (SBI) methods produce informative and reliable estimates of parameter posteriors for strong lensing systems in ground-based surveys. We present the examination and comparison of two approaches to lens parameter estimation for strong galaxy-galaxy lenses — Neural Posterior Estimation (NPE) and Bayesian Neural Networks (BNNs). We perform inference on 1-, 5-, and 12-parameter lens models for ground-based imaging data that mimics the Dark Energy Survey (DES). We find that NPE outperforms BNNs, producing posterior distributions that are more accurate, precise, and well-calibrated for most parameters. For the 12-parameter NPE model, the calibration is consistently within <10% of optimal calibration for all parameters, while the BNN is rarely within 20% of optimal calibration for any of the parameters. Similarly, residuals for most of the parameters are smaller (by up to an order of magnitude) with the NPE model than the BNN model. This work takes important steps in the systematic comparison of methods for different levels of model complexity.

46 INSTRUMENTATION RELATED TO NUCLEAR SCIENCE AND ↗

Greedy emulators for nuclear two-body scattering

Applications of reduced basis method emulators are increasing in low-energy nuclear physics because they enable fast and accurate sampling of high-fidelity calculations, enabling robust uncertainty quantification. Here, in this paper, we develop, implement, and test two model-driven emulators based on the (Petrov-)Galerkin projection using the prototypical test case of two-body scattering with the Minnesota potential and a more realistic local chiral potential. The high-fidelity scattering equations are solved with the matrix Numerov method, a reformulation of the popular Numerov recurrence relation for solving special second-order differential equations as a linear system of coupled equations. A novel error estimator based on reduced-space residuals is applied to an active learning approach (a greedy algorithm) to choosing training samples (“snapshots”) for the emulator and contrasted with a proper orthogonal decomposition (POD) approach. Both approaches allow for computationally efficient offline-online decompositions, but the greedy approach requires many fewer snapshot calculations. These developments set the groundwork for emulating scattering observables based on chiral nucleon-nucleon and three-nucleon interactions and optical models, where computational speed-ups are necessary for Bayesian uncertainty quantification. Our emulators and error estimators are widely applicable to linear systems.

Bayesian methods↗

Performance Evaluation of an Offshore Wave Measurement Buoy in Monochromatic Waves

The accurate measurement of waves underpins marine energy resource characterization, device design, and project development. Datawell wave buoys are widely deployed and have long served as a trusted standard for wave measurements. We quantify the measurement performance, including wave elevation and energy flux estimation, of a Datawell DWR-MkIII buoy using prescribed monochromatic heave motions on a large-amplitude six-degree-of-freedom motion platform at the National Laboratory of the Rockies, assuming the buoy behaves as an ideal wave follower. Commanded motions were validated with an optical motion tracking system while buoy elevation and raw acceleration were recorded. Wave elevations were propagated to wave energy flux estimation using four methods, including one frequency-domain method and three time-domain methods. The Bayesian optimization was applied for design of experiments, and records from three test sites were also applied and evaluated in the present study. Results show two error regions within the nominal period range of 1.6 s to 30 s. For wave periods between 5 s and 25 s, the buoy provides accurate wave height measurements. For short periods less than 5 s, the 1.28 Hz sampling frequency induces sub-Nyquist artifacts that bias elevation and can drive maximum energy flux estimation errors above 100%. For long periods exceeding 25 s, the buoy reported elevation is underpredicted with error depending on period but relatively independent of wave height, with maximum wave height and wave energy flux errors reaching 64% and 87%, respectively. Furthermore, analysis of three field-derived cases shows that frequency-domain estimates at 1.28 Hz agree within 2% of the corresponding 100 Hz estimates, while larger method-dependent differences are observed for the Hilbert method.

16 TIDAL AND WAVE POWER↗

PhaseT3M: 3D imaging at 1.6 Å resolution via electron cryo-tomography with nonlinear phase retrieval

Electron cryo-tomography (cryo-ET) enables 3D imaging of complex, radiation-sensitive structures with molecular detail. However, image contrast from the interference of scattered electrons is nonlinear with atomic density and multiple scattering further complicates interpretation. These effects degrade resolution, particularly in conventional reconstruction algorithms, which assume linearity. Particle averaging can reduce such issues but is unsuitable for heterogeneous or dynamic samples ubiquitous in biology, chemistry, and materials sciences. Here, we develop a phase retrieval-based cryo-ET method, PhaseT3M. We experimentally demonstrate its application to an approximately 7 nm Co3O4 nanoparticle on an approximately 30 nm carbon substrate, achieving a maximum resolution of 1.6 Å, surpassing conventional limits using standard cryo-TEM equipment. PhaseT3M uses a multislice model for multiple scattering and Bayesian optimization for alignment and computational aberration correction, with a positivity constraint to recover ‘missing wedge’ information. Applied directly to biological particles, it enhances reconstruction quality and reduces artifacts, establishing a standard for routine 3D imaging with phase contrast.

Biophysics↗

Open World Dempster-Shafer Theory/The Transferable Belief Model with Intervals: A Practitioner's Guide to DST and TBM

Dempster-Shafer theory (DST) is a mathematical framework that allows for uncertainty or ignorance to be quantified and included when making predictions from evidence. This is in contrast to Bayesian theory, which does not allow for any quantification of ignorance. The framework is described in great detail in [7]. DST is particularly useful for problems where the inclusion of additional evidence (for example, data from another sensor) could lead to a different conclusion. Thus, it is a useful data fusion method, especially in applications not suited to maximum likelihood or maximum a posteriori estimations due to limited samples or incomplete prior knowledge.

97 MATHEMATICS AND COMPUTING↗

Measurement of the top-quark pole mass in dileptonic $t\overline{t}$ + 1-jet events at $\sqrt{s}=13$ TeV with the ATLAS experiment

A measurement of the top-quark pole mass $m$$^{pole}_{t}$ is presented in $t\bar{t}$ events with an additional jet, $t\bar{t}$+ 1-jet, produced in pp collisions at $\sqrt{s} = 13 TeV. The data sample, recorded with the ATLAS experiment during Run 2 of the LHC, corresponds to an integrated luminosity of 140 fb −1 . Events with one electron and one muon of opposite electric charge in the final state are selected to measure the $t\bar{t}$ + 1-jet differential cross-section as a function of the inverse of the invariant mass of the $t\bar{t}$ + 1-jet system. Iterative Bayesian Unfolding is used to correct the data to enable comparison with fixed-order calculations at next-to-leading-order accuracy in the strong coupling. The process pp → $t\bar{t}$j(2 → 3), where top quarks are taken as stable particles, and the process pp → $b\bar{b}$l + vl – $\overline{ν}$j (2 → 7), which includes top-quark decays to the dilepton final state and off-shell effects, are considered. The top-quark mass is extracted using a χ 2 fit of the unfolded normalized differential cross-section distribution. The results obtained with the 2 → 3 and 2 → 7 calculations are compatible within theoretical uncertainties, providing an important consistency check.

Hadron-Hadron Scattering↗

Targeted Adaptive Design

Modern advanced manufacturing and advanced materials design often require searches of relatively high-dimensional process control parameter spaces for settings that result in optimal structure, property, and performance parameters. The mapping from the former to the latter must be determined from noisy experiments or from expensive simulations. Here, we abstract this problem to a mathematical framework in which an unknown function from a control space to a design space must be ascertained by means of expensive noisy measurements, which locate control settings generating desired design features within specified tolerances, with quantified uncertainty. We describe targeted adaptive design (TAD), a new algorithm that performs this sampling task efficiently. TAD creates a Gaussian process surrogate model of the unknown mapping at each iterative stage, proposing a new batch of control settings to sample experimentally and optimizing the updated expected log-predictive probability density of the target design. TAD either stops upon locating a solution with uncertainties that fit inside the tolerance box or uses a measure of expected future information to determine that the search space has been exhausted with no solution. TAD thus embodies the exploration-exploitation tension in a manner that recalls, but is essentially different from, Bayesian optimization and optimal experimental design.

97 MATHEMATICS AND COMPUTING↗

Description of Pegethrix niliensis sp. nov., a Novel Cyanobacterium from the Nile River Basin, Egypt: A Polyphasic Analysis and Comparative Study of Related Genera in the Oculatellales Order

In this paper, we examine the filamentous cyanobacterial strain NILCB16 and describe it as a new species within the genus Pegethrix. The original population was sampled from a mat growing in an irrigation canal in the Nile River, Egypt. Initially classified under Plectonema or Planktolyngbya, the strain is a potential producer of the toxins microcystin and β-N-Methylamino-L-Alanine (BMAA). Additionally, we reviewed the taxonomic relationships between the Oculatellales genera. To describe the new species, we conducted a polyphasic study, encompassing 16S rRNA gene phylogenetic analyses performed using both Maximum Likelihood and Bayesian methods, sequence identity (p-distance) analysis, 16S-23S ITS secondary structures, and morphological and habitat comparisons. The phylogenetic analysis revealed that strain NILCB16 clustered within the Pegethrix clade with strong phylogenetic support, but in a distinct position from other species in the genus. The strain shared a maximum 16S rRNA gene identity of 97.3% with P. qiandaoensis and 96.1% with the type species, P. bostrychoides. Morphologically, NILCB16 can be differentiated from other species in the genus by its lack of false branching. Our phylogenetic analyses also show that Pegethrix, Cartusia, Elainella, and Maricoleus are clustered with strong phylogenetic support. They exhibit high 16S rRNA gene identity and are morphologically indistinguishable, suggesting they could potentially be merged into a single genus in the future.

Hentschke, Guilherme Scotta (ORCID:000000034396024↗

DESI Spectroscopy of HETDEX Emission-line Candidates. I. Line Discrimination Validation

The Hobby–Eberly Dark Energy Experiment (HETDEX) is an untargeted spectroscopic galaxy survey that uses Lyα-emitting galaxies (LAEs) as tracers of 1.9 < z < 3.5 large-scale structure. Most detections consist of a single emission line, whose identity is inferred via a Bayesian analysis of ancillary data. To determine the accuracy of these line identifications, HETDEX detections were observed with the Dark Energy Spectroscopic Instrument (DESI). In two DESI pointings, high-confidence spectroscopic redshifts are obtained for 1157 sources, including 982 LAEs. The DESI spectra are used to evaluate the accuracy of the HETDEX object classifications and tune the methodology to achieve the HETDEX science requirement of ≲2% contamination of the LAE sample by low-redshift emission-line galaxies, while still assigning 96% of the true Lyα emission sample with the correct spectroscopic redshift. We compare emission-line measurements between the two experiments assuming a simple Gaussian line fitting model. Fitted values for the central wavelength of the emission line, the measured line flux, and line widths are consistent between the surveys within uncertainties. Derived spectroscopic redshifts, from the two classification pipelines, when both agree as an LAE classification, are consistent to within $\langle$Δz/(1 + z)$\rangle$ = 6.9 × 10 −5 with an rms scatter of 3.3 × 10 −4 . Data are available at https://data.desi.lbl.gov/desi/public/dr1/vac/dr1/hetdex.

79 ASTRONOMY AND ASTROPHYSICS↗