IDEAS home Printed from https://ideas.repec.org/a/bla/jorssb/v84y2022i2p414-439.html
   My bibliography  Save this article

Graph based Gaussian processes on restricted domains

Author

Listed:
  • David B. Dunson
  • Hau‐Tieng Wu
  • Nan Wu

Abstract

In nonparametric regression, it is common for the inputs to fall in a restricted subset of Euclidean space. Typical kernel‐based methods that do not take into account the intrinsic geometry of the domain across which observations are collected may produce sub‐optimal results. In this article, we focus on solving this problem in the context of Gaussian process (GP) models, proposing a new class of Graph Laplacian based GPs (GL‐GPs), which learn a covariance that respects the geometry of the input domain. As the heat kernel is intractable computationally, we approximate the covariance using finitely‐many eigenpairs of the Graph Laplacian (GL). The GL is constructed from a kernel which depends only on the Euclidean coordinates of the inputs. Hence, we can benefit from the full knowledge about the kernel to extend the covariance structure to newly arriving samples by a Nyström type extension. We provide substantial theoretical support for the GL‐GP methodology, and illustrate performance gains in various applications.

Suggested Citation

  • David B. Dunson & Hau‐Tieng Wu & Nan Wu, 2022. "Graph based Gaussian processes on restricted domains," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 84(2), pages 414-439, April.
  • Handle: RePEc:bla:jorssb:v:84:y:2022:i:2:p:414-439
    DOI: 10.1111/rssb.12486
    as

    Download full text from publisher

    File URL: https://doi.org/10.1111/rssb.12486
    Download Restriction: no

    File URL: https://libkey.io/10.1111/rssb.12486?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    References listed on IDEAS

    as
    1. Abhirup Datta & Sudipto Banerjee & Andrew O. Finley & Alan E. Gelfand, 2016. "Hierarchical Nearest-Neighbor Gaussian Process Models for Large Geostatistical Datasets," Journal of the American Statistical Association, Taylor & Francis Journals, vol. 111(514), pages 800-812, April.
    2. Wenpin Tang & Lu Zhang & Sudipto Banerjee, 2021. "On identifiability and consistency of the nugget in Gaussian spatial process models," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 83(5), pages 1044-1070, November.
    3. Sudipto Banerjee & Alan E. Gelfand & Andrew O. Finley & Huiyan Sang, 2008. "Gaussian predictive process models for large spatial data sets," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 70(4), pages 825-848, September.
    4. Ming-yen Cheng & Hau-tieng Wu, 2013. "Local Linear Regression on Manifolds and Its Geometric Interpretation," Journal of the American Statistical Association, Taylor & Francis Journals, vol. 108(504), pages 1421-1434, December.
    5. Mu Niu & Pokman Cheung & Lizhen Lin & Zhenwen Dai & Neil Lawrence & David Dunson, 2019. "Intrinsic Gaussian processes on complex constrained domains," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 81(3), pages 603-627, July.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Matthias Katzfuss & Joseph Guinness & Wenlong Gong & Daniel Zilber, 2020. "Vecchia Approximations of Gaussian-Process Predictions," Journal of Agricultural, Biological and Environmental Statistics, Springer;The International Biometric Society;American Statistical Association, vol. 25(3), pages 383-414, September.
    2. Isabelle Grenier & Bruno Sansó & Jessica L. Matthews, 2024. "Multivariate nearest‐neighbors Gaussian processes with random covariance matrices," Environmetrics, John Wiley & Sons, Ltd., vol. 35(3), May.
    3. Marchetti, Yuliya & Nguyen, Hai & Braverman, Amy & Cressie, Noel, 2018. "Spatial data compression via adaptive dispersion clustering," Computational Statistics & Data Analysis, Elsevier, vol. 117(C), pages 138-153.
    4. Philip A. White & Alan E. Gelfand, 2021. "Multivariate functional data modeling with time-varying clustering," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 30(3), pages 586-602, September.
    5. Kelly R. Moran & Matthew W. Wheeler, 2022. "Fast increased fidelity samplers for approximate Bayesian Gaussian process regression," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 84(4), pages 1198-1228, September.
    6. Matthew J. Heaton & Abhirup Datta & Andrew O. Finley & Reinhard Furrer & Joseph Guinness & Rajarshi Guhaniyogi & Florian Gerber & Robert B. Gramacy & Dorit Hammerling & Matthias Katzfuss & Finn Lindgr, 2019. "A Case Study Competition Among Methods for Analyzing Large Spatial Data," Journal of Agricultural, Biological and Environmental Statistics, Springer;The International Biometric Society;American Statistical Association, vol. 24(3), pages 398-425, September.
    7. Jaewoo Park & Sangwan Lee, 2022. "A projection‐based Laplace approximation for spatial latent variable models," Environmetrics, John Wiley & Sons, Ltd., vol. 33(1), February.
    8. Sameh Abdulah & Yuxiao Li & Jian Cao & Hatem Ltaief & David E. Keyes & Marc G. Genton & Ying Sun, 2023. "Large‐scale environmental data science with ExaGeoStatR," Environmetrics, John Wiley & Sons, Ltd., vol. 34(1), February.
    9. Huang Huang & Sameh Abdulah & Ying Sun & Hatem Ltaief & David E. Keyes & Marc G. Genton, 2021. "Competition on Spatial Statistics for Large Datasets," Journal of Agricultural, Biological and Environmental Statistics, Springer;The International Biometric Society;American Statistical Association, vol. 26(4), pages 580-595, December.
    10. Xu Chen & Surya T. Tokdar, 2021. "Joint quantile regression for spatial data," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 83(4), pages 826-852, September.
    11. Si Cheng & Bledar A. Konomi & Georgios Karagiannis & Emily L. Kang, 2024. "Recursive nearest neighbor co‐kriging models for big multi‐fidelity spatial data sets," Environmetrics, John Wiley & Sons, Ltd., vol. 35(4), June.
    12. Jingjie Zhang & Matthias Katzfuss, 2022. "Multi-Scale Vecchia Approximations of Gaussian Processes," Journal of Agricultural, Biological and Environmental Statistics, Springer;The International Biometric Society;American Statistical Association, vol. 27(3), pages 440-460, September.
    13. Jialuo Liu & Tingjin Chu & Jun Zhu & Haonan Wang, 2022. "Large spatial data modeling and analysis: A Krylov subspace approach," Scandinavian Journal of Statistics, Danish Society for Theoretical Statistics;Finnish Statistical Society;Norwegian Statistical Association;Swedish Statistical Association, vol. 49(3), pages 1115-1143, September.
    14. Jorge Castillo-Mateo & Miguel Lafuente & Jesús Asín & Ana C. Cebrián & Alan E. Gelfand & Jesús Abaurrea, 2022. "Spatial Modeling of Day-Within-Year Temperature Time Series: An Examination of Daily Maximum Temperatures in Aragón, Spain," Journal of Agricultural, Biological and Environmental Statistics, Springer;The International Biometric Society;American Statistical Association, vol. 27(3), pages 487-505, September.
    15. Kirsner, Daniel & Sansó, Bruno, 2020. "Multi-scale shotgun stochastic search for large spatial datasets," Computational Statistics & Data Analysis, Elsevier, vol. 146(C).
    16. Chen, Yewen & Chang, Xiaohui & Luo, Fangzhi & Huang, Hui, 2023. "Additive dynamic models for correcting numerical model outputs," Computational Statistics & Data Analysis, Elsevier, vol. 187(C).
    17. Zilber, Daniel & Katzfuss, Matthias, 2021. "Vecchia–Laplace approximations of generalized Gaussian processes for big non-Gaussian spatial data," Computational Statistics & Data Analysis, Elsevier, vol. 153(C).
    18. Jafari Khaledi, Majid & Zareifard, Hamid & Boojari, Hossein, 2023. "A spatial skew-Gaussian process with a specified covariance function," Statistics & Probability Letters, Elsevier, vol. 192(C).
    19. Cole, D. Austin & Gramacy, Robert B. & Ludkovski, Mike, 2022. "Large-scale local surrogate modeling of stochastic simulation experiments," Computational Statistics & Data Analysis, Elsevier, vol. 174(C).
    20. Bledar A. Konomi & Emily L. Kang & Ayat Almomani & Jonathan Hobbs, 2023. "Bayesian Latent Variable Co-kriging Model in Remote Sensing for Quality Flagged Observations," Journal of Agricultural, Biological and Environmental Statistics, Springer;The International Biometric Society;American Statistical Association, vol. 28(3), pages 423-441, September.

    More about this item

    Statistics

    Access and download statistics

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:bla:jorssb:v:84:y:2022:i:2:p:414-439. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Wiley Content Delivery (email available below). General contact details of provider: https://edirc.repec.org/data/rssssea.html .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.