Gaussian predictive process models for large spatial data sets. Academic Article

Overview
Research
Identity
Additional Document Info
Other
View All

abstract

With scientific data available at geocoded locations, investigators are increasingly turning to spatial process models for carrying out statistical inference. Over the last decade, hierarchical models implemented through Markov chain Monte Carlo methods have become especially popular for spatial modelling, given their flexibility and power to fit models that would be infeasible with classical methods as well as their avoidance of possibly inappropriate asymptotics. However, fitting hierarchical spatial models often involves expensive matrix decompositions whose computational complexity increases in cubic order with the number of spatial locations, rendering such models infeasible for large spatial data sets. This computational burden is exacerbated in multivariate settings with several spatially dependent response variables. It is also aggravated when data are collected at frequent time points and spatiotemporal process models are used. With regard to this challenge, our contribution is to work with what we call predictive process models for spatial and spatiotemporal data. Every spatial (or spatiotemporal) process induces a predictive process model (in fact, arbitrarily many of them). The latter models project process realizations of the former to a lower dimensional subspace, thereby reducing the computational burden. Hence, we achieve the flexibility to accommodate non-stationary, non-Gaussian, possibly multivariate, possibly spatiotemporal processes in the context of large data sets. We discuss attractive theoretical properties of these predictive processes. We also provide a computational template encompassing these diverse settings. Finally, we illustrate the approach with simulated and real data sets.

authors

Sang, Huiyan

published proceedings

J R Stat Soc Series B Stat Methodol

altmetric score

15

author list (cited authors)

Banerjee, S., Gelfand, A. E., Finley, A. O., & Sang, H.

citation count

796

complete list of authors

Banerjee, Sudipto||Gelfand, Alan E||Finley, Andrew O||Sang, Huiyan

publication date

January 2008

publisher

Oxford University Press (OUP) Publisher

published in

Journal of the Royal Statistical Society Series B: Statistical Methodology Journal