Search NASA⌕ Search

SEARCH · Search NASA

Results for “Models, Statistical”

Search indexed NASA NTRS and DOE OSTI research on propulsion, heat transfer, battery materials and energy systems. Follow report and document links to the original sources.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 379 records · Page 21

Producing two-dimensional dust clouds and clusters using a movable electrode for complex plasma and fundamental physics experiments

We report a Bidirectional Electrode Control Arm Assembly (BECAA) for precisely manipulating dust clouds levitated above the powered electrode in RF plasmas. The reported techniques allow the creation of perfectly 2D dust layers by eliminating off-plane particles by moving the electrode from outside the plasma chamber without altering the plasma conditions. Here, the tilting and moving of electrodes using BECAA also allows the precise and repeatable elimination of dust particles one by one to achieve any desired number of grains N without trial and error. Simultaneously acquired top and side view images of dust clusters show that they are perfectly planar or 2D. A demonstration of clusters with N = 1–28 without changing the plasma conditions is presented to show the utility of BECAA for complex plasma and statistical physics experimental design. Demonstration videos and 3D printable part files are available for easy reproduction and adaptation of this new method to repeatably produce 2D clusters in existing RF plasma chambers.

70 PLASMA PHYSICS AND FUSION TECHNOLOGY↗

Interpreting the spatial distribution of soil properties with a physically-based distributed hydrological model

Digital soil maps are commonly data-driven as the development of physically-based models for soil mapping is difficult due to the complexity of soils. However, physically-based hydrologic models have been successful in simulating water dynamics. Since water movement is a major driver of pedogenesis, the physical rules that govern water movement might help explain and predict the spatial variation of soil properties. Here, we demonstrate the novel use of a physically-based, distributed hydrologic model to inform the spatial distribution of soil properties. The Distributed Hydrology Soil Vegetation Model (DHSVM) was utilized to simulate soil moisture content (SM) and water table depth (WTD) in two hillslope catchments under pasture and forest management wherein hydrologic model outputs were then compared with soil properties measured in situ. SM sensors and wells were installed in both catchments to validate simulations of soil water movement via Nash-Sutcliffe Efficiency (E). In-situ observations were made at 87 sites within both catchments to study the connection between simulated water movement (SM and WTD) and observed soil properties, namely the depth and thickness of the argillic (Bt), fragic (Btx), and C horizons, and the depth of redoximorphic features. The simulated time series of SM and WTD were also clustered per season using Dynamic Time Warping (DTW), which identified similarity among time series at varying timescales. Model validation suggested that simulations of surficial SM (0–20 cm) were reasonable (E = 0.45), however, simulated subsurface SM (45–60 cm) and WTD were not sufficiently accurate. The thickness of Btx horizons were spatially grouped into different populations by SM clusters from every season except spring. For the other properties, only SM dynamics of specific seasons grouped into significantly different populations, suggesting that the explanatory power of simulated water movement varies seasonally and was greater during winter. Here, we show clusters of simulated SM separated soil properties into statistically different populations, showing that hydrologic models could inform areas that followed different water dynamics related to pedogenic trajectories and related biogeochemical processes not necessarily simulated by the model. As such, physically-based modeling of water dynamics can, therefore, inform and advance digital soil mapping by linking water movement patterns stemming from hydrologic model outputs to spatial patterns of soil properties and pedogenesis.

54 ENVIRONMENTAL SCIENCES↗

Two transitions in complex eigenvalue statistics: Hermiticity and integrability breaking

Open quantum systems have complex energy eigenvalues which are expected to follow non-Hermitian random matrix statistics, when chaotic, or two-dimensional (2d) Poisson statistics, when integrable. We investigate the spectral properties of a many-body quantum spin chain, i.e., the Hermitian Heisenberg model with imaginary disorder. Its rich complex eigenvalue statistics is found to separately break both Hermiticity and integrability at different scales of the disorder strength. With no disorder, the system is integrable and Hermitian, with spectral statistics corresponding to the 1d Poisson point process. At very small disorder, we find a transition from 1d Poisson statistics to an effective D -dimensional Poisson point process, showing Hermiticity breaking. At intermediate disorder, we find integrability breaking, as inferred from the statistics matching that of non-Hermitian complex symmetric random matrices in class AI † . For large disorder, as the spins align, we recover the expected integrability (now in the non-Hermitian setup), indicated by 2d Poisson statistics. These conclusions are based on fitting the spin-chain data of numerically generated nearest- and next-to-nearest-neighbor spacing distributions to an effective 2d Coulomb gas description at inverse temperature β . We confirm that such an effective description of random matrices also applies in classes AI † and AII † up to next-to-nearest-neighbor spacings. Published by the American Physical Society 2025

Akemann, Gernot (ORCID:0000000217104258)↗

Enhancing segmentation fairness through curriculum learning and progressive loss: a centralized and federated perspective on radiograph analysis

Bias in medical image segmentation can lead to unequal performance across demographic subgroups, raising concerns about fairness and reliability in clinical AI systems. While deep learning models have achieved high segmentation accuracy, ensuring equitable performance across race and gender remains a significant challenge, particularly in privacy-sensitive healthcare environments. This study investigates fairness-aware medical image segmentation for hip and knee radiographs using deep learning models evaluated in both centralized and Federated Learning (FL) settings. We introduce Curriculum Learning (CL) strategies and Progressive Loss (PL) functions to regulate sample difficulty during training. In addition, we propose two novel fairness-oriented federated learning algorithms, Federated Intersection over Union (FedIoU) and Federated Intersection over Union with Outlier Analysis (FedIoUoutlier). Experiments are conducted using multiple segmentation backbones and simulated multi-site data partitions derived from the Osteoarthritis Initiative dataset. Model performance is evaluated using Intersection over Union (IoU), IoU standard deviation, Skewed Error Ratio (SER), and Min-Max Disparity across race and gender subgroups. Statistical significance was verified using paired t-tests to compare per-sample IoU performance against baseline configurations. Across both hip and knee segmentation tasks, curriculum learning and progressive loss strategies consistently improved segmentation accuracy and reduced demographic performance disparities in centralized training. In federated settings, fairness-aware aggregation further enhanced performance. Notably, FedIoUoutlier combined with balanced curriculum learning and tiered progressive loss achieved the highest mean IoU while yielding the lowest SER and Min-Max Disparity, indicating improved fairness without sacrificing accuracy. In several configurations, federated models matched or exceeded the performance of optimized centralized models, with statistically significant improvements in per-sample IoU over baseline configurations. The results demonstrate that structured training strategies and fairness-aware federated aggregation can jointly improve accuracy, stability, and demographic fairness in medical image segmentation. By integrating curriculum learning, progressive loss, and novel FL algorithms, this work provides a practical pathway toward equitable and privacy-preserving AI systems for medical imaging.

97 MATHEMATICS AND COMPUTING↗

Automated vehicle microscopic energy consumption study (AV-Micro): Data collection and model development

While the Adaptive Cruise Control (ACC) system in automated vehicles (AVs) is expected to impact transportation energy significantly, existing AV energy consumption models only directly adopt those developed with Human-driven Vehicle (HV) data without even slight adaptation or calibration to accommodate unique AV energy consumption features. This study will investigate how accurately HV data-based models can predict the energy consumption of AVs. Empirical trajectory data and corresponding instantaneous energy consumption rates from both AVs and HVs were collected. We adopted two classical HV data-based models to fit these data. The calibration results indicated that these models yield around 20 30% prediction errors for AVs. To further improve the prediction accuracy, this study designed an AV-Micro model by incorporating components of multiple classic energy consumption models that better capture ACC energy consumption features, including piecewise driving behavior. With this, the AV-Micro model achieves lower than 10% prediction errors. The AV-Micro model’s high consistency across different test runs was verified with statistical significance tests, demonstrating its adaptability in different driving profiles. To confirm the discrepancies between the energy consumption features of AVs and HVs, more statistical significance tests were conducted to show that the AV-Micro model cannot be directly applied to HV data. The findings by calibrated AV-Micro models revealed that AVs consume approximately 80.5–146.4 J more energy than HVs for each meter traveled. Furthermore, the frequency analysis of energy consumption indicates that there is still some room for AVs to improve energy efficiency, particularly given their larger amplitude high-frequency fluctuations.

33 ADVANCED PROPULSION SYSTEMS↗

Stochastic Optimization and Uncertainty Quantification of Natrium-based Nuclear-Renewable Energy Systems for Flexible Power Applications in Deregulated Markets

Rapid integration of variable renewable energy sources (VRES) has made modeling and stochastic optimization of hybrid energy systems crucial for studying their long-term performance and viability. However, most studies have focused on just historical data, which may be unreliable for capturing short-term fluctuations, rare events, and long-term patterns of energy demand, price, and the variability of renewable energy sources. For this study, optimal synthetic time series models were developed using Wasserstein distance. The models were validated by comparing the key statistical measures against those of the historical data. They were then used to optimize the integrated Natrium-style advanced energy systems and their long-term (30 years) economics. The stochastic model performs bi-level optimization to find the optimal sizes for the balance of plant and thermal energy storage, while also optimizing energy dispatch to achieve the maximum net present value. In studies of two deregulated markets (California ISO and the Electric Reliability Council of Texas), the integrated Natrium-style system performed better in CAISO than in ERCOT, given higher and more consistent electricity prices during peak-demand periods. The potentially enlarged cost associated with the variable operation and maintenance of the TES system also plays a significant role in driving the system sizing, thus its impacts on the system are investigated in detail through comparison against a baseline case. The study also finds that the bi-level optimization results based on stochastic gradient descent closely match the grid search results. The uncertainty quantification of the stochastic signals provides further NPV-related insights and probability distributions for the case studies. The normal standard error of the mean of NPV for the case with and without TES VOM for CAISO were found to be 7.73M (plus-minus sign) 1.09M USD and 104.99M (plus-minus sign) 1.25M USD, respectively based on a 95% confidence. Given the relatively small NPV variance based on 150 samples, the analysis affords the most robust possible prediction of the techno-economic performance of the integrated Natrium-style energy systems.

25 ENERGY STORAGE↗

Deciphering baryonic feedback with galaxy clusters

Abstract Upcoming cosmic shear analyses will precisely measure the cosmic matter distribution at low redshifts. At these redshifts, the matter distribution is affected by galaxy formation physics, primarily baryonic feedback from star formation and active galactic nuclei. Employing measurements from theMagneticumandIllustrisTNGsimulations and a dark matter + baryon (DMB) halo model, this paper demonstrates that Sunyaev-Zel'dovich (SZ) effect observations of galaxy clusters, whose masses have been calibrated using weak gravitational lensing, can constrain the baryonic impact on cosmic shear with statistical and systematic errors subdominant to the measurement errors of DES-Y3 and LSST-Y1, with systematic errors on S 8 and Ω m reaching 10% and 50% of the statistical errors, respectively. For LSST-Y6 and Roman surveys, these systematic errors increase to 150% and 100% of the statistical errors, indicating the necessity for further model developments for future surveys. We further dissect the contributions from different scales and halos with different masses to cosmic shear, highlighting the dominant role of SZ clusters at scales critical for cosmic shear analyses. These findings suggest a promising avenue for future joint analyses of Cosmic Microwave Background (CMB) and lensing surveys.

Astronomy & Astrophysics↗

Fundamental microscopic properties as predictors of large-scale quantities of interest: Validation through grain boundary energy trends

Correlations between fundamental microscopic properties computable from first principles, which we term canonical properties, and complex large-scale quantities of interest (QoIs) provide an avenue to predictive materials discovery. Here, we propose that such correlations can be efficiently discovered through simulations utilizing approximate interatomic potentials (IPs), which serve as an ensemble of “synthetic materials”. As a proof of principle we build a regression model relating canonical properties to the symmetric tilt grain boundary (GB) energy curves in face-centered cubic crystals, characterized by the scaling factor in the universal lattice matching model of Runnels et al. (2016), which we take to be our QoI. Our analysis recovers known correlations of GB energy to other properties and discovers new ones. We also demonstrate, using available density functional theory (DFT) GB energy data, that the regression model constructed from IP data is consistent with DFT results, confirming the assumption that the IPs and DFT belong to same statistical pool and thereby validating the approach. Regression models constructed in this fashion can be used to predict large-scale QoIs based on first-principles data and provide a general method for training IPs for QoIs beyond the scope of first-principles calculations.

36 MATERIALS SCIENCE↗

Quantifying uncertainty in physics-based predictions of rare-isotope production cross sections via Bayesian-inspired model averaging across nuclear mass tables

Accurate prediction of fragmentation cross sections is essential for rare-isotope beam production, planning new-isotope searches, and designing experiments to study the most exotic regions of the nuclear chart. However, existing reaction models and phenomenological cross-section parametrizations often exhibit significant deviations over broad regions of mass and charge. In this work, a Bayesian-inspired model-averaging framework is developed to combine abrasion-ablation (AA) calculations based on multiple nuclear mass tables into a single statistically weighted estimate. For the calibrated systems, the model weights are assigned empirically according to the relative quality of fit to measured cross sections, thereby reducing systematic model bias while preserving the underlying physics content of the AA description. The weights are constrained using proton-rich fragmentation data for the 78 Kr and 124 Xe projectiles. The resulting parameter trends are then propagated to the 92 Mo and 144 Sm systems through a controlled scaling procedure. In the present implementation, the excitation-energy prescription is fixed, while the averaging is performed across nuclear-mass inputs; the framework provides both weighted cross sections and associated uncertainty estimates. Applied to proton-rich fragmentation, the present approach provides a practical basis for interpolation and limited extrapolation in regions relevant to rare-isotope production. The resulting predictions are used to assess the production of very proton-rich nuclei, and candidate new isotopes are discussed.

Bayesian methods↗

Analysis of differential scanning calorimetry data for aged plutonium

Differential scanning calorimetry data for samples of a 52 year old plutonium alloy with 3.3 at. % Ga that were heated beyond the melting point is analyzed using transition state theory to find activation energies for the δ to ε and ε to liquid phase transitions. A Bayesian statistical method involving a Gaussian process model is used to find mean values and confidence intervals for the activation energies. The activation energy for the δ to ε phase transition increases by 3.3 ± 3.8% per decade, relative to the case when all age related plutonium lattice point defects have been removed through annealing. The corresponding increase in activation energy for the ε to liquid transition is shown to be 7.1 ± 1.8% per decade. It is postulated that the change in activation energy with age for both phase transitions is caused, in part, by the accumulation of the same type of lattice point defects associated with the observed increase in elastic bulk modulus over time.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

A spatially-resolved model of neutron-irradiated tungsten coupling stochastic cluster dynamics and finite deformation plasticity

Structural materials used in nuclear reactors face severe degradation in mechanical properties, such as hardening and embrittlement. At the microscopic scale, this occurs due to creation and accumulation of irradiation-induced defects and their interaction with system dislocations. Although techniques exist which can model evolution of irradiation defects, for instance kinetic transport theory-based models, their interaction with mechanical deformation of the bulk material has not been investigated extensively. In this work, we demonstrate a novel spatially-resolved multiscale coupling between microscopic irradiation defect evolution, modeled using Stochastic Cluster Dynamics (SCD) and macroscopic mechanical deformation modeled using a finite-deformation plasticity model. SCD is used to determine the statistically averaged defect cluster spacing, dependent on operating conditions such as irradiation dose and temperature. This acts as an initial condition that governs the critical resolved shear stress of dislocation glide in the macroscopic plasticity model. This framework is used to predict mechanical behavior in post-mortem test of irradiated Tungsten samples, which has found its importance as structural material used in nuclear reactors. The results obtained using the coupled approach are in good agreement with experimental data of uniaxial tension tests. The model is able to capture the effect of temperature and irradiation dose on the material hardening. Two methods are proposed to estimate hardness – using Tabor's Law relating uniaxial yield stress to hardness and from flat-punch simulations. The results are in reasonable agreement with hardness data from micro-indentation experiments of irradiated Tungsten samples. Finally, the model is also able to reveal microstructural details such as spatial variation in defect density and local stress.

22 GENERAL STUDIES OF NUCLEAR REACTORS↗

Water under hydrophobic confinement: entropy and diffusion

The properties of liquid water are known to change drastically in confined geometries. A most interesting and intriguing phenomenon is that the diffusion of water is found to be strongly enhanced by the proximity of a hydrophobic confining wall relative to the bulk diffusion. We report a molecular dynamics simulation using a classical water model investigating the water diffusion near a non-interacting smooth confining wall, which is assumed to imitate a hydrophobic surface, revealing a pronounced diffusion enhancement within several water layers adjacent to the wall. We present evidence that the observed diffusion enhancement can be accounted for, with a quantitative accuracy, using the universal scaling law for liquid diffusion that relates the diffusion rate to the excess entropy. These results show that the scaling law, which has so far only been used for the description of the diffusion in simple liquids, can successfully describe the diffusion in water. It is shown that the law can be used for the analysis of water dynamics under nanoscale hydrophobic confinement, which is currently a subject of intense research activity.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Photometry of Outer Solar System Objects from the Dark Energy Survey. II. A Joint Analysis of Trans-Neptunian Absolute Magnitudes, Colors, Light Curves and Dynamics

For the 696 trans-Neptunian objects (TNOs) with absolute magnitudes 5.5 < H r < 8.2 detected in the Dark Energy Survey, we characterize the relationships between their dynamical state and physical properties—namely H r , indicating size; colors, indicating surface composition; and flux variation semiamplitude A, indicating asphericity and surface inhomogeneity. We seek “birth” physical distributions that can recreate these parameters in every dynamical class. We show that the observed colors of these TNOs are consistent with two Gaussian distributions in griz space, “near-infrared bright” (NIRB) and “near-infrared faint” (NIRF), presumably an inner and outer birth population, respectively. We find a model in which both the NIRB and NIRF H r and A distributions are independent of current dynamical states, supporting their assignment as birth populations. All objects are consistent with a common rolling p(H r ), but NIRF objects are significantly more variable. Cold classicals (CCs) are purely NIRF, while hot classical (HC), scattered, and detached TNOs are consistent with ≈ 70% NIRB and the resonance NIRB fractions show significant variation. The NIRB components of the HCs and of some resonances have broader inclination distributions than the NIRFs, i.e. their current dynamics retains information about birth location. We find evidence for radial stratification within the birth NIRB population, in that HC NIRBs are on average redder than detached or scattered NIRBs; a similar effect distinguishes CCs from other NIRFs. We estimate total object counts and masses of each class within our H r range. These results will strongly constrain models of the outer solar system.

79 ASTRONOMY AND ASTROPHYSICS↗

Effectiveness of nature-based solutions to reduce flooding in Quad Cities Metro Area (QCMA) using SWMM-HEC based flood model

Nature-based solutions (NbS) have gained significant attention as strategies for addressing urban environmental challenges, particularly since the establishment of the UN Sustainable Development Goals (SDGs) for 2030. However, the current research on NbS for urban flood management lacks comprehensive methodological approaches for identifying suitable areas and evaluating their effectiveness across different urban settings. Here, this study attempts to fill this gap by proposing a methodological framework integrating multi-criteria analysis with a SWMM-HEC-based hydrologic and hydraulic (HH) model to assess the suitability of NbS for the Quad Cities Metro Area (QCMA), consisting of Davenport, Bettendorf, Moline, and Rock Island. Eight NbS options-green roofs, rain gardens, infiltration trenches, permeable pavements, vegetative swales, dry detention basins, retention ponds, and rain barrels/cisterns - were considered based on volumetric efficiency and runoff reduction efficiency. The study reveals that implementing the proposed NbS could have substantially reduced flood depths in key historical flood events by 21% in 1993, 15% in 2008, 16% in 2011, 23% in 2014, 40% in 2019, and 10% in 2023. The findings highlight a critical trade-off between peak runoff and NbS implementation: while NbS effectively reduce flood impacts, they also enhance volumetric efficiency by approximately 43%. In high-density areas of the QCMA, flood depth reductions of around 20% suggest that NbS are a viable solution for dense urban environments with limited space. This shows the potential for integrating NbS into existing infrastructure, offering a promising approach for cities facing increasing flooding risks. The proposed methodology provides a practical framework for incorporating NbS into urban stormwater management, addressing gaps in optimizing NbS performance, and offering a pathway to scale their application in other urban areas with various environmental and social contexts.

CMIP6↗

Probing the Kitaev honeycomb model on a neutral-atom quantum computer

Quantum simulations of many-body systems are among the most promising applications of quantum computers. In particular, models based on strongly correlated fermions are central to our understanding of quantum chemistry and materials problems, and can lead to exotic, topological phases of matter. However, owing to the non-local nature of fermions, such models are challenging to simulate with qubit devices. Here we realize a digital quantum simulation architecture for two-dimensional fermionic systems based on reconfigurable atom arrays. We utilize a fermion-to-qubit mapping based on Kitaev’s model on a honeycomb lattice, in which fermionic statistics are encoded using long-range entangled states. We prepare these states efficiently using measurement and feedforward, realize subsequent fermionic evolution through Floquet engineering with tunable entangling gates interspersed with atom rearrangement, and improve results with built-in error detection. Leveraging this fermion description of the Kitaev spin model, we efficiently prepare topological states across its complex phase diagram and verify the non-Abelian spin-liquid phase by evaluating an odd Chern number. We further explore this two-dimensional fermion system by realizing tunable dynamics and directly probing fermion exchange statistics. Finally, we simulate strong interactions and study the dynamics of the Fermi–Hubbard model on a square lattice. These results pave the way for digital quantum simulations of complex fermionic systems for materials science, chemistry and high-energy physics.

atomic and molecular physics↗

Statistically-driven Experimental Design to Improve Reference-free Quantification of Small Molecules by Liquid Chromatography-Mass Spectrometry

Non-targeted analysis of small molecules and metabolites in unknown, complex samples using liquid chromatography-tandem mass spectrometry remains challenging. One of the main bottlenecks is the extensive unannotated regions of metabolomics mass spectrometry data, resulting in knowledge gaps. Small molecule annotation in mass spectrometry data has conventionally relied on reference standards and libraries for compound identification and confirmation, which can constrain compound identification to those molecules already known, thus limiting the ability to discover new knowledge and new markers. Retention time prediction can facilitate and expedite unknown compound identification in non-targeted analysis of complex metabolomics samples. Additionally, accurate retention time predictions can also inform sample mixture design for LC-MS/MS analyses. However, current machine learning-based methods for retention time prediction are typically developed for specific chromatographic platforms and are not generalizable across scales. And while technologies and methods to improve reference-free metabolite identification for more comprehensive annotation of unknowns has received much attention, development of the same for quantitation without reference standards has been much more limited, despite its importance in toxicological, environmental, food safety, forensics, and clinical applications. We believe that a reference-free quantitation strategy that exploits mass spectrometry data already collected for reference-free identification can provide much more insight on unknowns, and move the metabolomics field for more complete unknowns characterization. As such, we pursue two efforts to improve upon current state-of-the-art methods in non-targeted analysis: (1) machine learning-based retention time prediction and (2) statistical design of experiments framework for reference-free quantitation. In this work, we develop and demonstrate (1) a generalizable retention time prediction capability across chromatographic conditions and scales, and (2) a statistical design-based framework for response factor contribution elucidation and reference-free quantitation. Evaluation of our retention time prediction model, PrediToR, showed approximately 24% improvement over current models, and we observed approximately 10X improvement in concentration estimation accuracy from our statistical design-based response factor model over a primarily ionization efficiency-based model. We expect that future efforts to improve upon these new capabilities will further advance non-targeted analysis of small molecules towards truly reference-free metabolomics.

37 INORGANIC, ORGANIC, PHYSICAL, AND ANALYTICAL CH↗

Full-stack Quantification of Variability in Predicting Ion Transport Properties using Machine-learned Interatomic Potentials

Machine-learned interatomic potentials (MLIPs) have become the state-of-the-art for performing accurate, scalable molecular dynamics (MD) simulations. It is therefore crucial to understand and quantify the reliability of MLIPs for downstream property predictions. Uncertainty in predicted properties can arise from limitations in first-principles training data, intrinsic MLIP model errors in representing the data, and the statistical noise introduced during subsequent MD simulations. Using ion transport in Li7P3S11 as a case study, we systematically assess the impact of training set size and selection, neural network stochasticity, and MD sampling statistics on predicted diffusivity and activation energy. We find that when using equivariant MLIP architectures with standard MD protocols, uncertainty arising from MD sampling dominates over model-induced errors. In contrast, MLIP errors relative to the underlying first-principles data are consistently minor. Given this, there are two main routes to improving the accuracy of predictions based on MLIP potentials: adopting higher accuracy reference data generation methods, and improving the MD sampling statistics.

36 MATERIALS SCIENCE↗

Data-driven projection pursuit adaptation of polynomial chaos expansions for dependent high-dimensional parameters

Uncertainty quantification (UQ) and inference involving a large number of parameters are valuable tools for problems associated with heterogeneous and non-stationary behaviors. The difficulty with these problems is exacerbated when these parameters are statistically dependent requiring statistical characterization over joint measures. Probabilistic modeling methodologies stand as effective tools in the realms of UQ and inference. Among these, polynomial chaos expansions (PCE), when adapted to low-dimensional quantities of interest (QoI), provide effective yet accurate approximations for these QoI in terms of an adapted orthogonal basis. These adaptation techniques have been cast as projection pursuits in Gaussian Hilbert space in what has been referred to as a projection pursuit adaptation (PPA) by Xiaoshu Zeng and Roger Ghanem (2023). The PPA method efficiently identifies an optimal low-dimensional space for representing the QoI and simultaneously evaluates an optimal PCE within that space. The quality of this approximation clearly depends on the size of the training dataset, which is typically a function of the adapted reduced dimension. Here, the complexity of the problem is thus mediated by the complexity of the low-dimensional quantity of interest and not the complexity of the high-dimensional parameter space.

Data-driven↗