Publications Details

Publications / Journal Article

Cholesky-based experimental design for gaussian process and kernel-based emulation and calibration

Harbrecht, Helumt; Jakeman, John D.; Zaspel, Peter

Gaussian processes and other kernel-based methods are used extensively to construct approximations of multivariate data sets. The accuracy of these approximations is dependent on the data used. This paper presents a computationally efficient algorithm to greedily select training samples that minimize the weighted Lp error of kernel-based approximations for a given number of data. The method successively generates nested samples, with the goal of minimizing the error in high probability regions of densities specified by users. The algorithm presented is extremely simple and can be implemented using existing pivoted Cholesky factorization methods. Training samples are generated in batches which allows training data to be evaluated (labeled) in parallel. For smooth kernels, the algorithm performs comparably with the greedy integrated variance design but has significantly lower complexity. Numerical experiments demonstrate the efficacy of the approach for bounded, unbounded, multi-modal and non-tensor product densities. We also show how to use the proposed algorithm to efficiently generate surrogates for inferring unknown model parameters from data using Bayesian inference.

SAND2020-12052J

Communications in Computational Physics
Volume 29, Issue 4, Page 1152-1185

April 1, 2021

Global Science Press (United States)

Universitat Basel

Sandia National Laboratories, New Mexico

Sandia Laboratory Directed Research & Development (LDRD)

212952

Constructor University Bremen

19917120 18152406

10.4208/CICP.OA-2020-0060

Mathematics and Computing

Experimental design

Active learning

Gaussian Process

Radial basis function

Uncertainty quantification

Bayesian inference