We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

stat.ME

Change to browse by:

References & Citations

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo

Statistics > Methodology

Title: A Divide-and-Conquer Bayesian Approach to Large-Scale Kriging

Abstract: We propose a three-step divide-and-conquer strategy within the Bayesian paradigm that delivers massive scalability for any spatial process model. We partition the data into a large number of subsets, apply a readily available Bayesian spatial process model on every subset, in parallel, and optimally combine the posterior distributions estimated across all the subsets into a pseudo-posterior distribution that conditions on the entire data. The combined pseudo posterior distribution replaces the full data posterior distribution for predicting the responses at arbitrary locations and for inference on the model parameters and spatial surface. Based on distributed Bayesian inference, our approach is called "Distributed Kriging" (DISK) and offers significant advantages in massive data applications where the full data are stored across multiple machines. We show theoretically that the Bayes $L_2$-risk of the DISK posterior distribution achieves the near optimal convergence rate in estimating the true spatial surface with various types of covariance functions, and provide upper bounds for the number of subsets as a function of the full sample size. The model-free feature of DISK is demonstrated by scaling posterior computations in spatial process models with a stationary full-rank and a nonstationary low-rank Gaussian process (GP) prior. A variety of simulations and a geostatistical analysis of the Pacific Ocean sea surface temperature data validate our theoretical results.
Comments: 29 pages, including 4 figures and 5 tables
Subjects: Methodology (stat.ME)
Cite as: arXiv:1712.09767 [stat.ME]
  (or arXiv:1712.09767v3 [stat.ME] for this version)

Submission history

From: Sanvesh Srivastava [view email]
[v1] Thu, 28 Dec 2017 06:04:31 GMT (5073kb,D)
[v2] Mon, 10 Jun 2019 09:12:59 GMT (5133kb,D)
[v3] Wed, 12 Jun 2019 07:38:12 GMT (5129kb,D)

Link back to: arXiv, form interface, contact.