Search NASASearch

DOE OSTI · 3376987

Enhancing Gaussian Process Surrogates for Optimization and Posterior Approximation via Random Exploration

Abstract

This paper proposes novel noise-free Bayesian optimization strategies that rely on a random exploration step to enhance the accuracy of Gaussian process surrogate models. The new algorithms retain the ease of implementation of the classical GP-UCB algorithm, but the additional random exploration step accelerates their convergence, nearly achieving the optimal convergence rate. Furthermore, to facilitate Bayesian inference with intractable likelihoods, we propose to utilize optimization iterates for maximum a posteriori estimation to build a Gaussian process surrogate model for the unnormalized log-posterior density. We provide bounds for the Hellinger distance between the true and the approximate posterior distributions in terms of the number of design points. We demonstrate the effectiveness of our Bayesian optimization algorithms in nonconvex benchmark objective functions, in a machine learning hyperparameter tuning problem, and in a black-box engineering design problem. The effectiveness of our posterior approximation approach is demonstrated in two Bayesian inference problems for parameters of dynamical systems.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Kim, Hwanwoo [Duke University, Durham, NC (United States)], Sanz-Alonso, Daniel [University of Chicago, IL (United States)] (ORCID:000000025022864X). 2025-08-06. Enhancing Gaussian Process Surrogates for Optimization and Posterior Approximation via Random Exploration. https://doi.org/10.1137/24m1677009

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related reports

Scalable quantum computational science: A perspective from block-encodings and polynomial transformations

Significant developments made in quantum hardware and error correction recently have been driving quantum computing toward practical utility. However, gaps remain between abstract quantum algorithmic development and practical applications in computational sciences. In this perspective article, we propose several properties that scalable quantum computational science methods should possess. We further discuss how block-encodings and polynomial transformations can potentially serve as a unified framework with the desired properties. Recent advancements on these topics are presented, including the construction and assembly of block-encodings, and various generalizations of quantum signal processing (QSP) algorithms to perform polynomial transformations. The scalability of QSP methods on parallel and distributed quantum architectures is also highlighted. Promising applications in simulation and observable estimation in chemistry, physics, and optimization problems are presented. We hope this perspective serves as a gentle introduction to state-of-the-art quantum algorithms for the computational science community and inspires future development of scalable quantum computational science methodologies that bridge theory and practice.

Bayesian inference

Analysis of streaked images of x-ray self-emission in laser-driven spherical implosions

Imaging of x-ray self-emission provides a powerful in situ measurement of the spatial and temporal evolution of high-energy-density plasmas. However, interpretation of these measurements requires detailed understanding of the data-generating process. This work presents a case study in the interpretation of x-ray self-emission data for the specific application of streaked one-dimensional slit imaging of spherical laser-driven implosions. A comprehensive generative model of the streaked slit-imaging diagnostic is developed including detailed treatments of the radiation transfer, photometrics, and photostatistics associated with the measurement. The model is used to generate realistic synthetic streaked images and to analyze experimental streaked images to extract important physical quantities of interest. An example analysis of streaked images from implosion experiments on the OMEGA laser is presented, where the model developed in this work is used to constrain the trajectory and peak velocity of the implosion using Bayesian inference.

Bayesian inference

Accelerating Hamiltonian Monte Carlo for Bayesian inference in neural networks and neural operators

Hamiltonian Monte Carlo (HMC) is a powerful and accurate method to sample from the posterior distribution in Bayesian inference. However, HMC techniques are computationally demanding for Bayesian neural networks due to the high dimensionality of the network’s parameter space and the non-convexity of their posterior distributions. Therefore, various approximation techniques, such as variational inference (VI) or stochastic gradient MCMC, are often employed to infer the posterior distribution of the network parameters. Such approximations introduce inaccuracies in the inferred distributions, resulting in unreliable uncertainty estimates. In this work, we propose a hybrid approach that combines inexpensive VI and accurate HMC methods to efficiently and accurately quantify uncertainties in neural networks and neural operators. The proposed approach leverages an initial VI training on the full network. We examine the influence of individual parameters on the prediction uncertainty, which shows that a large proportion of the parameters do not contribute substantially to uncertainty in the network predictions. This information is then used to significantly reduce the dimension of the parameter space, and HMC is performed only for the subset of network parameters that strongly influence prediction uncertainties. This yields a framework for accelerating the full batch HMC for posterior inference in neural networks. We demonstrate the efficiency and accuracy of the proposed framework on deep neural networks and operator networks, showing that inference can be performed for large networks with tens to hundreds of thousands of parameters. Finally, we show that this method can effectively learn surrogates for complex physical systems by modeling the operator that maps from upstream conditions to wall-pressure data on a cone in hypersonic flow.

Bayesian inference