IDEAS home Printed from https://ideas.repec.org/a/eee/csdana/v134y2019icp1-16.html
   My bibliography  Save this article

Missing covariate data in generalized linear mixed models with distribution-free random effects

Author

Listed:
  • Liu, Li
  • Xiang, Liming

Abstract

We consider generalized linear mixed models in which random effects are free of parametric distributions and missing at random data are present in some covariates. To overcome the problem of missing data, we propose two novel methods relying on auxiliary variables: a penalized conditional likelihood method when covariates are independent of random effects, and a two-step procedure consisting of a pairwise likelihood for estimating fixed effects in the first step and a penalized conditional likelihood for estimating random effects in the second step while covariates can be related to random effects. Our methods allow a nonparametric structure for the missing covariate data and do not rely on distribution assumptions for random effects, which are not observed in the data, thus providing great flexibility in capturing a board range of the missingness mechanism and behaviors of random effects. We show that the proposed estimators enjoy desirable theoretical properties by relaxing the conditions for a finite number of clusters or finite cluster size imposed in the literature. The finite sample performance of the estimators is assessed through extensive simulations. We illustrate the application of the methods using a longitudinal data set on forest health monitoring.

Suggested Citation

  • Liu, Li & Xiang, Liming, 2019. "Missing covariate data in generalized linear mixed models with distribution-free random effects," Computational Statistics & Data Analysis, Elsevier, vol. 134(C), pages 1-16.
  • Handle: RePEc:eee:csdana:v:134:y:2019:i:c:p:1-16
    DOI: 10.1016/j.csda.2018.10.011
    as

    Download full text from publisher

    File URL: http://www.sciencedirect.com/science/article/pii/S0167947318302561
    Download Restriction: Full text for ScienceDirect subscribers only.

    File URL: https://libkey.io/10.1016/j.csda.2018.10.011?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    As the access to this document is restricted, you may want to search for a different version of it.

    References listed on IDEAS

    as
    1. Anthony Y. C. Kuk, 2007. "A Hybrid Pairwise Likelihood Method," Biometrika, Biometrika Trust, vol. 94(4), pages 939-952.
    2. Kuk, Anthony Y. C. & Nott, David J., 2000. "A pairwise likelihood approach to analyzing correlated binary data," Statistics & Probability Letters, Elsevier, vol. 47(4), pages 329-335, May.
    3. C.-Y. Huang & J. Qin & M.-C. Wang, 2010. "Semiparametric Analysis for Recurrent Event Data with Time-Dependent Covariates and Informative Censoring," Biometrics, The International Biometric Society, vol. 66(1), pages 39-49, March.
    4. Francis K. C. Hui & Samuel Müller & A. H. Welsh, 2017. "Joint Selection in Mixed Models using Regularized PQL," Journal of the American Statistical Association, Taylor & Francis Journals, vol. 112(519), pages 1323-1333, July.
    5. Thomas Kneib & Torsten Hothorn & Gerhard Tutz, 2009. "Variable Selection and Model Choice in Geoadditive Regression Models," Biometrics, The International Biometric Society, vol. 65(2), pages 626-634, June.
    6. Agresti, Alan & Caffo, Brian & Ohman-Strickland, Pamela, 2004. "Examples in which misspecification of a random effects distribution reduces efficiency, and possible remedies," Computational Statistics & Data Analysis, Elsevier, vol. 47(3), pages 639-653, October.
    7. Peng Wang & Guei-feng Tsai & Annie Qu, 2012. "Conditional Inference Functions for Mixed-Effects Models With Unspecified Random-Effects Distribution," Journal of the American Statistical Association, Taylor & Francis Journals, vol. 107(498), pages 725-736, June.
    8. Nicholas J. Horton & Nan M. Laird, 2001. "Maximum Likelihood Analysis of Logistic Regression Models with Incomplete Covariate Data and Auxiliary Information," Biometrics, The International Biometric Society, vol. 57(1), pages 34-42, March.
    9. Haibo Zhou & Jianwei Chen & Jianwen Cai, 2002. "Random Effects Logistic Regression Analysis with Auxiliary Covariates," Biometrics, The International Biometric Society, vol. 58(2), pages 352-360, June.
    10. Li Liu & Liming Xiang, 2014. "Semiparametric estimation in generalized linear mixed models with auxiliary covariates: A pairwise likelihood approach," Biometrics, The International Biometric Society, vol. 70(4), pages 910-919, December.
    11. J. G. Ibrahim & S. R. Lipsitz & M.‐H. Chen, 1999. "Missing covariates in generalized linear models when the missing data mechanism is non‐ignorable," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 61(1), pages 173-190.
    12. Kathryn M. Aloisio & Sonja A. Swanson & Nadia Micali & Alison Field & Nicholas J. Horton, 2014. "Analysis of partially observed clustered data using generalized estimating equations and multiple imputation," Stata Journal, StataCorp LP, vol. 14(4), pages 863-883, December.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Li Liu & Liming Xiang, 2014. "Semiparametric estimation in generalized linear mixed models with auxiliary covariates: A pairwise likelihood approach," Biometrics, The International Biometric Society, vol. 70(4), pages 910-919, December.
    2. Steele, Fiona & Clarke, Paul & Kuha, Jouni, 2019. "Modeling within-household associations in household panel studies," LSE Research Online Documents on Economics 88162, London School of Economics and Political Science, LSE Library.
    3. Paik, Jane & Ying, Zhiliang, 2012. "A composite likelihood approach for spatially correlated survival data," Computational Statistics & Data Analysis, Elsevier, vol. 56(1), pages 209-216, January.
    4. Francis K. C. Hui & Samuel Müller & Alan H. Welsh, 2021. "Random Effects Misspecification Can Have Severe Consequences for Random Effects Inference in Linear Mixed Models," International Statistical Review, International Statistical Institute, vol. 89(1), pages 186-206, April.
    5. Cristiano Varin, 2008. "On composite marginal likelihoods," AStA Advances in Statistical Analysis, Springer;German Statistical Society, vol. 92(1), pages 1-28, February.
    6. Cavit Pakel & Neil Shephard & Kevin Sheppard & Robert F. Engle, 2021. "Fitting Vast Dimensional Time-Varying Covariance Models," Journal of Business & Economic Statistics, Taylor & Francis Journals, vol. 39(3), pages 652-668, July.
    7. Benjamin Hofner & Andreas Mayr & Nikolay Robinzonov & Matthias Schmid, 2014. "Model-based boosting in R: a hands-on tutorial using the R package mboost," Computational Statistics, Springer, vol. 29(1), pages 3-35, February.
    8. M.-L. Feddag, 2016. "Pairwise likelihood estimation for the normal ogive model with binary data," AStA Advances in Statistical Analysis, Springer;German Statistical Society, vol. 100(2), pages 223-237, April.
    9. Renard, Didier & Molenberghs, Geert & Geys, Helena, 2004. "A pairwise likelihood approach to estimation in multilevel probit models," Computational Statistics & Data Analysis, Elsevier, vol. 44(4), pages 649-667, January.
    10. Zhengxin Zhang & Xiaosheng Si & Changhua Hu & Xiangyu Kong, 2015. "Degradation modeling–based remaining useful life estimation: A review on approaches for systems with heterogeneity," Journal of Risk and Reliability, , vol. 229(4), pages 343-355, August.
    11. Philip Kostov, 2010. "Do Buyers’ Characteristics and Personal Relationships Affect Agricultural Land Prices?," Land Economics, University of Wisconsin Press, vol. 86(1), pages 48-65.
    12. Zhang, Jing & Wang, Qihua & Kang, Jian, 2020. "Feature screening under missing indicator imputation with non-ignorable missing response," Computational Statistics & Data Analysis, Elsevier, vol. 149(C).
    13. Huang, Xianzheng, 2011. "Detecting random-effects model misspecification via coarsened data," Computational Statistics & Data Analysis, Elsevier, vol. 55(1), pages 703-714, January.
    14. Simona Buscemi & Antonella Plaia, 2020. "Model selection in linear mixed-effect models," AStA Advances in Statistical Analysis, Springer;German Statistical Society, vol. 104(4), pages 529-575, December.
    15. Hofner, Benjamin & Mayr, Andreas & Schmid, Matthias, 2016. "gamboostLSS: An R Package for Model Building and Variable Selection in the GAMLSS Framework," Journal of Statistical Software, Foundation for Open Access Statistics, vol. 74(i01).
    16. Gerda Claeskens & Fabrizio Consentino, 2008. "Variable Selection with Incomplete Covariate Data," Biometrics, The International Biometric Society, vol. 64(4), pages 1062-1069, December.
    17. Axel Munk & Tatyana Krivobokova, 2009. "Comments on: Goodness-of-fit tests in mixed models," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 18(2), pages 256-259, August.
    18. Mojtaba Ganjali & Taban Baghfalaki, 2018. "Application of Penalized Mixed Model in Identification of Genes in Yeast Cell-Cycle Gene Expression Data," Biostatistics and Biometrics Open Access Journal, Juniper Publishers Inc., vol. 6(2), pages 38-41, April.
    19. Juan Armando Torres Munguía, 2018. "What is behind homicide gender gaps in Mexico? A spatial semiparametric approach," Ibero America Institute for Econ. Research (IAI) Discussion Papers 236, Ibero-America Institute for Economic Research.
    20. Mohammad Rafiqul Islam & Masud Alam & Munshi Naser .Ibne Afzal & Sakila Alam, 2021. "Nighttime Light Intensity and Child Health Outcomes in Bangladesh," Papers 2108.00926, arXiv.org, revised Sep 2022.

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:csdana:v:134:y:2019:i:c:p:1-16. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: http://www.elsevier.com/locate/csda .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.