Search NASA⌕ Search

SEARCH · Search NASA

Results for “Statistical methods”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 415 records · Page 23

Statistical generic design of glass and optimization: Selective review on oxide glasses

Designing a single glass composition for a multidimensional property space is challenging, and the difficulty increases with the number of design criteria. Traditionally, the task is accomplished using multiple statistical models that describe the relationships between composition (C) and property (P) values, i.e., C-P models. Recently, the structure (S)-property (P) statistical modeling has emerged as a complementary approach. The S-P modeling approach has also been shown to be a preferred method for modeling glass properties, particularly when a small data set is available, such as in single-component studies, or when strong nonlinearities exist between composition and properties. The combined model package, C-S-P, implements the concept of generic glass design, i.e., designing glass for performance by first selecting a specific or optimized set of glass network structural groups using S-P models and then transferring the designed structures (genes) to a particular composition using C-S models. This article reviews a set of supporting cases from the previous C-S-P modeling studies of phosphate, silicate, and borosilicate glasses, which are relevant for many critical commercial applications. The methodology for developing the statistical C-S-P database is presented, enabling the application of P?S?C to achieve a generic glass design and optimization, targeting multiple design criteria for both performance and processing properties simultaneously.

Network structure↗

Pentaquarks made of light quarks and their admixture to baryons

This paper is a continuation of our studies of multiquark hadrons. The antisymmetrization of their wave functions required by Fermi statistics is nontrivial, as it mixes orbital, color, spin, and flavor structures. In our previous papers we developed a method to find them based on the representations of the permutation group, and derived the explicit wave functions for baryons excited to the first and second shells (L = 1, 2), tetraquarks $qq$$\overline{q}$$\overline{q}$ and hexaquarks (6q). Now we apply it to light pentaquarks ($qqq$$\overline{q}$), in the S- and P-shells (L = 0, 1). Using Jacobi coordinates, one can use the hyperdistance approximation in 12-dimensional space. We further address the issue of “unquenching” of baryons, by considering their mixing with pentaquarks, via two channels, through the addition of σ-like or π-like $\overline{q}$$q$ pairs. This mixing is central for understanding of the observed flavor asymmetry of the antiquark sea, the amount of orbital motion issue as well as other nucleon properties.

Baryons↗

Boundary-induced classical generalized Gibbs ensemble with angular momentum

We investigate how confinement geometry leads to the emergence of a Generalized Gibbs Ensemble (GGE) in classical systems. Unlike the standard Gibbs ensemble, the GGE includes additional conserved quantities, such as angular momentum, that arise from boundary-induced symmetries. Using analytical arguments based on the maximum entropy principle, we show that circular boundaries preserve angular momentum and drive the system toward a chiral, non-ergodic GGE that violates time-reversal symmetry. This ensemble differs fundamentally from the Gibbs case, producing near-boundary condensation and revealing how geometry alone can alter thermal equilibration. To quantify these effects, we introduce an order parameter measuring deviations from Gibbs behavior and demonstrate that conventional Monte Carlo methods must incorporate angular momentum conservation under such conditions. Our study highlights how geometric constraints shape non-equilibrium statistical ensembles and lead to subtle departures from the Bohr-van Leeuwen theorem. These predictions are validated through detailed simulations of confined classical hard-disk gases.

71 CLASSICAL AND QUANTUM MECHANICS, GENERAL PHYSIC↗

Theoretical studies of chemical reactions related to the formation and growth of polycyclic aromatic hydrocarbons (PAH) and molecular properties of their key intermediates (Final Progress Report)

The formation mechanisms of polycyclic aromatic hydrocarbons, (PAHs) – organic molecules carrying fused benzene rings – are of great interest to scientists and engineers due to their importance in combustion chemistry and astrochemistry. On Earth, PAHs are largely produced in incomplete combustion of fossil fuel and are considered as critical precursors to unwanted soot particles leading to combustion inefficiency and causing air pollution along with detrimental health effects. Simple PAH molecules initially formed in the gas phase, are further involved in a build-up process in combustion flames leading to larger PAH, bowl-shaped nanostructures, fullerenes, and solid-phase species including carbonaceous dust, graphene particles, and soot. In deep space, PAH and their derivatives are potential key intermediates and nucleation sites leading eventually to carbonaceous nanoparticles (“interstellar grains”). Therefore, the understanding of the key processes in the synthesis of PAHs along with their precursors and their degradation mechanisms in combustion systems and in interstellar, circumstellar, and planetary atmospheric environments will provide critical insights into how complex aromatic structures, carbonaceous nanoparticles, and fullerenes are formed and destroyed. Achieving this understanding is an important step in the development of the efficient combustion processes and of the ecofriendly devices with reduced environmental pollution as well as technological strategies for the production of hydrogen and solid carbon through thermal or plasma-assisted pyrolysis of natural gas and biomass. Also, the understanding of the key processes of PAH and soot growth will help in our comprehension of chemical evolution in the universe. Detailed information on the mechanisms and reliable rate constants of the key elementary chemical reactions involved in PAH formation and destruction processes and in inception of soot particles is often missing, with the main deficiencies being the absence of temperature- and pressure-dependent rate constants for the broad range of conditions occurring in various terrestrial and interstellar processes and the lack of data on the reaction products and their branching ratios. Complementary to experimental studies, these gaps in knowledge can be filled by using quantum chemical calculations of reaction potential energy surfaces providing us with accurate energies of reaction products, intermediates, and transition states, revealing the reaction mechanism, and giving the molecular properties required to compute rate constants for relevant reaction steps and product branching ratios using the RRKM-Master Equation (ME) method. Molecular dynamics (MD) simulations can be used in cases when a reaction rate cannot be properly described by statistical theories. During the terminal renewal project period we employed these ab initio/RRKM-ME and MD approaches to complete our studies on several key reactions relevant to the formation/growth of PAH and inception of soot particles including (1) the reaction mechanism and kinetics of the resonance stabilized fulvenallenyl radical with propargyl and C 3 H 4 isomers; (2) the reaction mechanism and kinetics for the C + indene and C 2 + styrene reactions producing naphthyl or azulenyl radicals in low-temperature environments; (3) the MD study of non-equilibrium dimerization of acepyrene and coronene and its radical. The information derived from our theoretical calculations contributed to a better fundamental understanding of the reaction mechanisms and provide missing critical kinetic data to improve combustion models of hydrocarbon fuels and astrochemical models of the growth of carbonaceous molecules and particles in cold molecular clouds, circumstellar envelopes, and planetary atmospheres.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Excited-state uncertainties in lattice-QCD calculations of hadron masses and scattering phase shifts

Lattice QCD has historically produced energy results interpretable as either estimates relying on implicit assumptions about asymptotic behavior or one-sided upper bounds. New Lanczos methods providing two-sided bounds with less-restrictive assumptions are introduced and quantified in a high-statistics calculation with unphysical quark masses. Two-sided bounds without spectral assumptions provide sub-percent constraints on the nucleon mass. Other bounds, which assume all states in a given energy window are resolved, provide meaningful two-sided constraints on nucleon-nucleon scattering phase shifts.

Detmold, William [MIT, Cambridge, CTP]↗

Statistical White-Line Analysis in High-Throughput TXM-XANES for Chemical State Quantification

The transmission X-ray microscopy (TXM) based X-ray absorption near-edge structure (XANES) technique provides three-dimensional mapping of element-specific chemical states at nanometer-scale spatial resolution and micrometer-scale fields of view. However, compared to conventional volume-averaged XANES (VA-XANES) measurements, the inherently small voxel size in TXM-XANES leads to a lower signal-to-noise ratio, making full-spectrum analysis computationally demanding and less robust. Here, we present the structural and compositional conditions for a statistical white-line analysis framework under which chemical state information can be directly extracted from the white-line peak position in voxel spectra without the need for voxel-wise background subtraction or normalization, under well-defined structural and compositional conditions. The method is validated on layered oxide cathode materials, where low-order polynomial fitting accurately reproduces white-line features, and the extracted energy distributions correlate strongly with VA-XANES results. This statistical approach enables high-throughput, dose-efficient, and noise-robust chemical state quantification in TXM-XANES, offering broad applicability to functional materials requiring nanoscale oxidation-state mapping.

TXM↗

Stochastic equilibrium Raman spectroscopy (STERS)

In this manuscript, we propose a new method for cavity- and surface-enhanced Raman spectroscopy (SERS) with improved temporal resolution in the measurement of stochastic Raman spectral fluctuations. Our approach combines Fourier spectroscopy and photon correlation to decouple the integration time from the temporal resolution. Using statistical optics Monte Carlo simulations, we establish the relationship between time resolution and Raman signal strength, revealing that typical Raman spectral fluctuations, commensurate with molecular conformational dynamics, can theoretically be resolved on micro- to millisecond timescales. The method can further extract average single-molecule dynamics from small sub-ensembles, thereby potentially mitigating challenges in achieving strictly single-molecule isolation on SERS substrates.

Cobb-Bruno, Colburn [University of California, Ber↗

Bipartite mutual information in classical many-body dynamics

Information theoretic measures have helped to sharpen our understanding of many-body quantum states. As perhaps the most well-known example, the entanglement entropy (or more generally, the bipartite mutual information) has become a powerful tool for characterizing the dynamical growth of quantum correlations. By contrast, although computable, the bipartite mutual information (MI) is almost never explored in classical many particle systems; this owes in part to the fact that computing the MI requires keeping track of the evolution of the full probability distribution, a feat which is rarely done (or thought to be needed) in classical many-body simulations. Here, we utilize the MI to analyze the spreading of information in 1D elementary cellular automata (CA). Broadly speaking, we find that the behavior of the MI in these dynamical systems exhibits a few different types of scaling that roughly correspond to known CA universality classes. Of particular note is that we observe a set of automata for which the MI converges parametrically slowly to its thermodynamic value. We develop a microscopic understanding of this behavior by analyzing a two-species model of annihilating particles moving in opposite directions. Furthermore, our work suggests the possibility that information theoretic tools such as the MI might enable a more fine-grained characterization of classical many-body states and dynamics.

Cellular automata↗

Geospatial analysis of preterm and small-for-gestational age births in Washington D.C.

Background: This study is based on the recognition that adverse pregnancy outcomes significantly affect maternal and infant health, leading to increased morbidity and mortality. These outcomes are shaped by a complex interplay of individual-level factors—like maternal age and education—and community-level influences, including socio-economic status and access to healthcare. Understanding these determinants is crucial for developing effective public health strategies, especially for marginalized populations, by identifying high-risk areas and informing targeted interventions that address both individual and structural barriers. Methods: We utilized geospatial analysis to explore the association between individual- and community-level factors and adverse pregnancy outcomes, specifically preterm birth (PTB) and small-for-gestational-age (SGA) birthweight in Washington, D.C. We used Empirical Bayes smoothing methods to calculate rates of adverse birth outcomes from 2010 to 2018 at the U.S. Census tract–level. Spatial scan statistics were used to investigate if adverse birth outcomes clustered in specific areas. ANOVA tests were conducted for individual- and community-level factors within identified clusters. Results: Spatial analysis identified significant high-risk clusters for PTB and SGA infants primarily in southeastern Washington, D.C., particularly in Wards 7 and 8. Individuals residing within these clusters experienced a 47% increased risk of PTB (RR = 1.467) and a 56% increased risk of SGA (RR = 1.560) compared to those outside clusters. Space–time analysis revealed temporal variation, with PTB clusters persisting from 2011 to 2014 and SGA clusters extending through 2017. Compared to low-risk clusters, high-risk clusters had younger birthing individuals (mean age ~26.5 vs. ~33 years), lower maternal college degree attainment (~20% vs. ~80%), higher rates of late or no prenatal care (~16% vs. 11%), and increased prevalence of smoking and hypertension (all P < 0.001). Community-level indicators showed lower median household incomes ($\$40,000$ vs. ~$\$105,000$), greater poverty (~16% vs. ~7% below $\$10,000$/year), higher public assistance use (~32% vs. ~5%), and reduced healthcare access (greater distances to emergency and specialty care) in high-risk areas (all P < 0.001). Neighborhood deprivation indices were significantly elevated, commutes were longer, and population density was lower in these clusters. These findings highlight that adverse birth outcomes cluster in neighborhoods with pronounced socioeconomic and health disparities. Conclusion: High-risk birth clusters highlight intertwined factors: individual, socio-economic, and geographic. Addressing these requires comprehensive interventions focusing on social and structural determinants of health.

Birth outcomes↗

The Atacama Cosmology Telescope: Mitigating the Impact of Extragalactic Foregrounds for the DR6 Cosmic Microwave Background Lensing Analysis

We investigate the impact and mitigation of extragalactic foregrounds for the cosmic microwave background (CMB) lensing power spectrum analysis of Atacama Cosmology Telescope (ACT) data release 6 (DR6) data. Two independent microwave sky simulations are used to test a range of mitigation strategies. We demonstrate that finding and then subtracting point sources, finding and then subtracting models of clusters, and using a profile bias-hardened lensing estimator together reduce the fractional biases to well below statistical uncertainties, with the inferred lensing amplitude, A lens , biased by less than 0.2σ. We also show that another method where a model for the cosmic infrared background (CIB) contribution is deprojected and high-frequency data from Planck is included has similar performance. Other frequency-cleaned options do not perform as well, either incurring a large noise cost or resulting in biased recovery of the lensing spectrum. In addition to these simulation-based tests, we also present null tests on the ACT DR6 data for sensitivity of our lensing spectrum estimation to differences in foreground levels between the two ACT frequencies used, while nulling the CMB lensing signal. These tests pass whether the nulling is performed at the map or bandpower level. The CIB-deprojected measurement performed on the DR6 data is consistent with our baseline measurement, implying that contamination from the CIB is unlikely to significantly bias the DR6 lensing spectrum. This collection of tests gives confidence that the ACT DR6 lensing measurements and cosmological constraints presented in companion papers to this work are robust to extragalactic foregrounds.

79 ASTRONOMY AND ASTROPHYSICS↗

Rapid measurement of soluble xylo-oligomers using near-infrared spectroscopy (NIRS) and multivariate statistics: calibration model development and practical approaches to model optimization

Rapid monitoring of biomass conversion processes using techniques such as near-infrared (NIR) spectroscopy can be substantially quicker and less labor-, resource-, and energy-intensive than conventional measurement techniques such as gas or liquid chromatography (GC or LC) due to the lack of solvents and preparation methods, as well as removing the need to transfer samples to an external lab for analytical evaluation. The purpose of this study was to determine the feasibility of rapid monitoring of a biomass conversion process using NIR spectroscopy combined with multivariate statistical modeling, and to examine the impact of (1) subsetting the samples in the original dataset by process location and (2) reducing the spectral range used in the calibration model on model performance. We develop multivariate calibration models for the concentrations of soluble xylo-oligosaccharides (XOS), monomeric xylose, and total solids at multiple points in a biomass conversion process which produces and then purifies XOS compounds from sugar cane bagasse. A single model using samples from multiple locations in the process stream showed acceptable performance as measured by standard statistical measures. However, compared to the single model, we show that separate models built by segregating the calibration samples according to process location show improved performance. We also show that combining an understanding of the sample spectra with simple multivariate analysis tools can result in a calibration model with a substantially smaller spectral range that provides essentially equal performance to the full-range model. We demonstrate that real-time monitoring of soluble xylo-oligosaccharides (XOS), monomeric xylose, and total solids concentration at multiple points in a process stream using NIR spectroscopy coupled with multivariate statistics is feasible. Segregation of sample populations by process location improves model performance. Models using a reduced spectral range containing the most relevant spectral signatures show very similar performance to the full-range model, reinforcing the importance of performing robust exploratory data analysis before beginning multivariate modeling.

09 BIOMASS FUELS↗

Disordered Rocksalts as High‐Energy and Earth‐Abundant Li‐Ion Cathodes

To address the growing demand for energy and support the shift toward transportation electrification and intermittent renewable energy, there is an urgent need for low‐cost, energy‐dense electrical storage. Research on Li‐ion electrode materials has predominantly focused on ordered materials with well‐defined lithium diffusion channels, limiting cathode design to resource‐constrained Ni‐ and Co‐based oxides and lower‐energy polyanion compounds. Recently, disordered rocksalts with lithium excess (DRX) have demonstrated high capacity and energy density when lithium excess and/or local ordering allow statistical percolation of lithium sites through the structure. This cation disorder can be induced by high temperature synthesis or mechanochemical synthesis methods for a broad range of compositions. DRX oxides and oxyfluorides containing Earth‐abundant transition metals have been prepared using various synthesis routes, including solid‐state, molten‐salt, and sol‐gel reactions. This review outlines DRX design principles and explains the effect of synthesis conditions on cation disorder and short‐range cation ordering (SRO), which determines the cycling stability and rate capability. In addition, strategies to enhance Li transport and capacity retention with Mn‐rich DRX possessing partial spinel‐like ordering are discussed. Finally, the review considers the optimization of carbon and electrolyte in DRX materials and addresses key challenges and opportunities for commercializing DRX cathodes.

Li-ion batteries↗

Decoding substance use disorder severity from clinical notes using a large language model

Substance use disorder (SUD) poses a major concern due to its detrimental effects on health and society. SUD identification and treatment depend on a variety of factors such as severity, co-determinants (e.g., withdrawal symptoms), and social determinants of health. Existing diagnostic coding systems used by insurance providers, like the International Classification of Diseases (ICD-10), lack granularity for certain diagnoses, but American clinicians will add this granularity (as that found within the Diagnostic and Statistical Manual of Mental Disorders classification or DSM-5) as supplemental unstructured text in clinical notes. Traditional natural language processing (NLP) methods face limitations in accurately parsing such diverse clinical language. Large language models (LLMs) offer promise in overcoming these challenges by adapting to diverse language patterns. This study investigates the application of LLMs for extracting severity-related information for various SUD diagnoses from clinical notes. We propose a workflow employing zero-shot learning of LLMs with carefully crafted prompts and post-processing techniques. Through experimentation with Flan-T5, an open-source LLM, we demonstrate its superior recall compared to the rule-based approach. Focusing on 11 categories of SUD diagnoses, we show the effectiveness of LLMs in extracting severity information, contributing to improved risk assessment and treatment planning for SUD patients.

60 APPLIED LIFE SCIENCES↗

The DESI One-Percent Survey: Modelling the clustering and halo occupation of all four DESI tracers with U CHUU

We present results from a set of mock lightcones for the DESI One-Percent Survey, created from the UCHUU simulation. This 8 h −3 Gpc 3 N-body simulation comprises 2.1 trillion particles and provides high-resolution dark matter (sub)haloes in the framework of the Planck-based ΛCDM cosmology. Employing the subhalo abundance matching (SHAM) technique, we populated the UCHUU (sub)haloes with all four DESI tracers – Bright Galaxy Survey (BGS), luminous red galaxies (LRGs), emission line galaxies (ELGs), and quasars (QSOs) – to z = 2.1. Our method accounts for redshift evolution as well as the clustering dependence on luminosity and stellar mass. The two-point clustering statistics of the DESI One-Percent Survey generally agree with predictions from UCHUU across scales ranging from 0.3 h −1 Mpc to 100 h −1 Mpc for the BGS and across scales ranging from 5 h −1 Mpc to 100 h −1 Mpc for the other tracers. We observed some differences in clustering statistics that can be attributed to incompleteness of the massive end of the stellar mass function of LRGs, our use of a simplified galaxy-halo connection model for ELGs and QSOs, and cosmic variance. We find that at the high precision of UCHUU, the shape of the halo occupation distribution (HOD) of the BGS and LRG samples is smaller bias values, likely due to cosmic variance. The bias dependence on absolute magnitude, stellar mass, and redshift aligns with that of previous surveys. These results provide DESI with tools to generate high-fidelity lightcones for the remainder of the survey and enhance our understanding of the galaxy-halo connection.

cosmology↗

A cross-dimensional analysis of data-driven short-term load forecasting methods with large-scale smart meter data

Electricity load forecasting is essential to utility operation and power grid stability. A wide spectrum of data-driven methods, ranging from linear regression models to more recent deep learning models have been adopted to forecast electric load over the years. However, there still lacks a holistic evaluation of the applicability of conventional statistical and machine learning based algorithms with respect to different temporal and spatial scopes, computational requirements, and sensitivity of model-tuning. Enabled by a large-scale electricity load profile dataset of over 40,000 residential customers in a utility region, we conducted a cross-dimensional analysis of data-driven load forecasting methods. Three regression-based and seven deep learning algorithms with different model configurations were evaluated in terms of their overall and peak load prediction accuracy, and training burdens, across spatial aggregation levels ranging from the transformer, feeder, substation, to neighborhood. We found, first, the load forecasting accuracy is constrained by a predictability boundary, influenced by the forecasting horizon and spatial aggregation level. Specifically, RandomForest, XGBoost, TFT, TSMixer, and TiDE models achieved less than 10 % prediction error for up to 96-h ahead forecasting for district, substation, and feeder levels, while other models struggle at long-horizon predictions; Second, for winter and summer peak load dates, most models were able to predict the peak demand timing within ± 1 h, but the prediction percentage error varied by models, with TFT and TiDE models being the top performers; Third, models with similar prediction accuracy can differ in training burden by an order of magnitude. Therefore, choosing model configurations that balance prediction performance and computational resource is an important practical consideration for large-scale deployment of the machine learning based load forecasting. The outcome of this study can guide researchers and practitioners to choose the proper load forecasting algorithms based on their problem scope, required accuracy, and available resources. The predictability boundary can serve as a benchmark for electricity load forecasting problems with new algorithms and datasets.

Li, Han↗

Adaptive Uncertainty Quantification for Stochastic Hyperbolic Conservation Laws

Here, we propose a predictor-corrector adaptive method for the study of hyperbolic partial differential equations (PDEs) under uncertainty. Constructed around the framework of stochastic finite volume (SFV) methods, our approach circumvents sampling schemes or simulation ensembles while also preserving fundamental properties, in particular hyperbolicity of the resulting systems and conservation of the discrete solutions. Furthermore, we augment the existing SFV theory with a priori convergence results for statistical quantities, in particular push-forward densities, which we demonstrate through numerical experiments. By linking refinement indicators to regions of the physical and stochastic spaces, we drive anisotropic refinements of the discretizations, introducing new degrees of freedom where deemed profitable. To illustrate our proposed method, we consider a series of numerical examples for nonlinear hyperbolic PDEs based on Burgers’ and Euler’s equations.

97 MATHEMATICS AND COMPUTING↗

Dark Energy Survey Year 6 Results: improved mitigation of spatially varying observational systematics with masking

As photometric surveys reach unprecedented statistical precision, systematic uncertainties increasingly dominate large-scale structure probes relying on galaxy number density. Defining the final survey footprint is critical, as it excludes regions affected by artefacts or suboptimal observing conditions. For galaxy clustering, spatially varying observational systematics, such as seeing, are a leading source of bias. Template maps of contaminants are used to derive spatially dependent corrections, but extreme values may fall outside the applicability range of mitigation methods, compromising correction reliability. The complexity and accuracy of systematics modelling depend on footprint conservativeness, with aggressive masking enabling simpler, robust mitigation. We present a unified approach to define the DES Year 6 joint footprint, integrating observational systematics templates and artefact indicators that degrade mitigation performance. This removes extreme values from an initial seed footprint, leading to the final joint footprint. By evaluating the DES Year 6 lens sample MagLim++ plus plus on this footprint, we enhance the Iterative Systematics Decontamination (ISD) method, detecting non-linear systematic contamination and improving correction accuracy. While the mask's impact on clustering is less significant than systematics decontamination, it remains non-negligible, comparable to statistical uncertainties in certain w(theta) scales and redshift bins. Supporting coherent analyses of galaxy clustering and cosmic shear, the final footprint spans 4031.04 deg2, setting the basis for DES Year 6 1x2pt, 2x2pt, and 3x2pt analyses. This work highlights how targeted masking strategies optimise the balance between statistical power and systematic control in Stage-III and -IV surveys.

Rodríguez-Monroy, M. [Madrid, IFT; IJCLab, Orsay]↗

Data Assimilation for Robust UQ Within Agent-Based Simulation on HPC Systems

Agent-based simulation provides a powerful tool for in silico system modeling. However, these simulations do not provide built-in methods for uncertainty quantification (UQ). Within these types of models a typical approach to UQ is to run multiple realizations of the model then compute aggregate statistics. This approach is limited due to the compute time required for a solution. When faced with an emerging biothreat, public health decisions need to be made quickly and solutions for integrating near real-time data with analytic tools are needed. We propose an integrated Bayesian UQ framework for agent-based models based on sequential Monte Carlo sampling. Given streaming or static data about the evolution of an emerging pathogen this Bayesian framework provides a distribution over the parameters governing the spread of a disease through a population. These estimates of the spread of a disease may be provided to public health agencies seeking to abate the spread. By coupling agent-based simulations with Bayesian modeling in a data assimilation, our proposed framework provides a powerful tool for modeling dynamical systems in silico. We propose a method which reduces model error and provides a range of realistic possible outcomes. Moreover, our method addresses two primary limitations of ABMs: the lack of UQ and an inability to assimilate data. Our proposed framework combines the flexibility of an agent-based model with UQ provided by the Bayesian paradigm in a workflow which scales well to HPC systems. We provide algorithmic details and results on a simulated outbreak with both static and streaming data.

Spannaus, Adam [ORNL] (ORCID:0000000225213657)↗