AISTATS 2017poster18 citations

Hierarchically-partitioned Gaussian Process Approximation

Byung-Jun Lee, Jongmin Lee, Kee-Eung Kim

Abstract

The Gaussian process (GP) is a simple yet powerful probabilistic framework for various machine learning tasks. However, exact algorithms for learning and prediction are prohibitive to be applied to large datasets due to inherent computational complexity. To overcome this main limitation, various techniques have been proposed, and in particular, local GP algorithms that scales “truly linearly” with respect to the dataset size. In this paper, we introduce a hierarchical model based on local GP for large-scale datasets, which stacks inducing points over inducing points in layers. By using different kernels in each layer, the overall model becomes multi-scale and is able to capture both long- and short-range dependencies. We demonstrate the effectiveness of our model by speed-accuracy performance on challenging real-world datasets.

BibTeX
@InProceedings{pmlr-v54-lee17a,
  title = 	 {{Hierarchically-partitioned Gaussian Process Approximation}},
  author = 	 {Lee, Byung-Jun and Lee, Jongmin and Kim, Kee-Eung},
  booktitle = 	 {Proceedings of the 20th International Conference on Artificial Intelligence and Statistics},
  pages = 	 {822--831},
  year = 	 {2017},
  editor = 	 {Singh, Aarti and Zhu, Jerry},
  volume = 	 {54},
  series = 	 {Proceedings of Machine Learning Research},
  month = 	 {20--22 Apr},
  publisher =    {PMLR},
  pdf = 	 {http://proceedings.mlr.press/v54/lee17a/lee17a.pdf},
  url = 	 {https://proceedings.mlr.press/v54/lee17a.html},
  abstract = 	 {The Gaussian process (GP) is a simple yet powerful probabilistic framework for various machine learning tasks.  However, exact algorithms for learning and prediction are prohibitive to be applied to large datasets due to inherent computational complexity. To overcome this main limitation, various techniques have been proposed, and in particular, local GP algorithms that scales “truly linearly”  with respect to the dataset size. In this paper, we introduce a hierarchical model based on local GP for large-scale datasets, which stacks inducing points over inducing points in layers. By using different kernels in each layer, the overall model becomes multi-scale and is able to capture both long- and short-range dependencies.  We demonstrate the effectiveness of our model by speed-accuracy performance on challenging real-world datasets.}
}
Hierarchically-partitioned Gaussian Process Approximation · AISTATS 2017