IDEAS home Printed from https://ideas.repec.org/a/eee/csdana/v134y2019icp1-16.html
   My bibliography  Save this article

Missing covariate data in generalized linear mixed models with distribution-free random effects

Author

Listed:
  • Liu, Li
  • Xiang, Liming

Abstract

We consider generalized linear mixed models in which random effects are free of parametric distributions and missing at random data are present in some covariates. To overcome the problem of missing data, we propose two novel methods relying on auxiliary variables: a penalized conditional likelihood method when covariates are independent of random effects, and a two-step procedure consisting of a pairwise likelihood for estimating fixed effects in the first step and a penalized conditional likelihood for estimating random effects in the second step while covariates can be related to random effects. Our methods allow a nonparametric structure for the missing covariate data and do not rely on distribution assumptions for random effects, which are not observed in the data, thus providing great flexibility in capturing a board range of the missingness mechanism and behaviors of random effects. We show that the proposed estimators enjoy desirable theoretical properties by relaxing the conditions for a finite number of clusters or finite cluster size imposed in the literature. The finite sample performance of the estimators is assessed through extensive simulations. We illustrate the application of the methods using a longitudinal data set on forest health monitoring.

Suggested Citation

  • Liu, Li & Xiang, Liming, 2019. "Missing covariate data in generalized linear mixed models with distribution-free random effects," Computational Statistics & Data Analysis, Elsevier, vol. 134(C), pages 1-16.
  • Handle: RePEc:eee:csdana:v:134:y:2019:i:c:p:1-16
    DOI: 10.1016/j.csda.2018.10.011
    as

    Download full text from publisher

    File URL: http://www.sciencedirect.com/science/article/pii/S0167947318302561
    Download Restriction: Full text for ScienceDirect subscribers only.

    File URL: https://libkey.io/10.1016/j.csda.2018.10.011?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    As the access to this document is restricted, you may want to search for a different version of it.

    References listed on IDEAS

    as
    1. Kuk, Anthony Y. C. & Nott, David J., 2000. "A pairwise likelihood approach to analyzing correlated binary data," Statistics & Probability Letters, Elsevier, vol. 47(4), pages 329-335, May.
    2. Francis K. C. Hui & Samuel Müller & A. H. Welsh, 2017. "Joint Selection in Mixed Models using Regularized PQL," Journal of the American Statistical Association, Taylor & Francis Journals, vol. 112(519), pages 1323-1333, July.
    3. Agresti, Alan & Caffo, Brian & Ohman-Strickland, Pamela, 2004. "Examples in which misspecification of a random effects distribution reduces efficiency, and possible remedies," Computational Statistics & Data Analysis, Elsevier, vol. 47(3), pages 639-653, October.
    4. Peng Wang & Guei-feng Tsai & Annie Qu, 2012. "Conditional Inference Functions for Mixed-Effects Models With Unspecified Random-Effects Distribution," Journal of the American Statistical Association, Taylor & Francis Journals, vol. 107(498), pages 725-736, June.
    5. Nicholas J. Horton & Nan M. Laird, 2001. "Maximum Likelihood Analysis of Logistic Regression Models with Incomplete Covariate Data and Auxiliary Information," Biometrics, The International Biometric Society, vol. 57(1), pages 34-42, March.
    6. Kathryn M. Aloisio & Sonja A. Swanson & Nadia Micali & Alison Field & Nicholas J. Horton, 2014. "Analysis of partially observed clustered data using generalized estimating equations and multiple imputation," Stata Journal, StataCorp LP, vol. 14(4), pages 863-883, December.
    7. Anthony Y. C. Kuk, 2007. "A Hybrid Pairwise Likelihood Method," Biometrika, Biometrika Trust, vol. 94(4), pages 939-952.
    8. C.-Y. Huang & J. Qin & M.-C. Wang, 2010. "Semiparametric Analysis for Recurrent Event Data with Time-Dependent Covariates and Informative Censoring," Biometrics, The International Biometric Society, vol. 66(1), pages 39-49, March.
    9. Thomas Kneib & Torsten Hothorn & Gerhard Tutz, 2009. "Variable Selection and Model Choice in Geoadditive Regression Models," Biometrics, The International Biometric Society, vol. 65(2), pages 626-634, June.
    10. Haibo Zhou & Jianwei Chen & Jianwen Cai, 2002. "Random Effects Logistic Regression Analysis with Auxiliary Covariates," Biometrics, The International Biometric Society, vol. 58(2), pages 352-360, June.
    11. Li Liu & Liming Xiang, 2014. "Semiparametric estimation in generalized linear mixed models with auxiliary covariates: A pairwise likelihood approach," Biometrics, The International Biometric Society, vol. 70(4), pages 910-919, December.
    12. J. G. Ibrahim & S. R. Lipsitz & M.‐H. Chen, 1999. "Missing covariates in generalized linear models when the missing data mechanism is non‐ignorable," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 61(1), pages 173-190.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Li Liu & Liming Xiang, 2014. "Semiparametric estimation in generalized linear mixed models with auxiliary covariates: A pairwise likelihood approach," Biometrics, The International Biometric Society, vol. 70(4), pages 910-919, December.
    2. Paik, Jane & Ying, Zhiliang, 2012. "A composite likelihood approach for spatially correlated survival data," Computational Statistics & Data Analysis, Elsevier, vol. 56(1), pages 209-216, January.
    3. Francis K. C. Hui & Samuel Müller & Alan H. Welsh, 2021. "Random Effects Misspecification Can Have Severe Consequences for Random Effects Inference in Linear Mixed Models," International Statistical Review, International Statistical Institute, vol. 89(1), pages 186-206, April.
    4. Steele, Fiona & Clarke, Paul & Kuha, Jouni, 2019. "Modeling within-household associations in household panel studies," LSE Research Online Documents on Economics 88162, London School of Economics and Political Science, LSE Library.
    5. Cristiano Varin, 2008. "On composite marginal likelihoods," AStA Advances in Statistical Analysis, Springer;German Statistical Society, vol. 92(1), pages 1-28, February.
    6. Shu Yang & Jae Kwang Kim, 2016. "Likelihood-based Inference with Missing Data Under Missing-at-Random," Scandinavian Journal of Statistics, Danish Society for Theoretical Statistics;Finnish Statistical Society;Norwegian Statistical Association;Swedish Statistical Association, vol. 43(2), pages 436-454, June.
    7. Cavit Pakel & Neil Shephard & Kevin Sheppard & Robert F. Engle, 2021. "Fitting Vast Dimensional Time-Varying Covariance Models," Journal of Business & Economic Statistics, Taylor & Francis Journals, vol. 39(3), pages 652-668, July.
    8. Benjamin Hofner & Andreas Mayr & Nikolay Robinzonov & Matthias Schmid, 2014. "Model-based boosting in R: a hands-on tutorial using the R package mboost," Computational Statistics, Springer, vol. 29(1), pages 3-35, February.
    9. Jorge I. Figueroa-Zúñiga & Cristian L. Bayes & Víctor Leiva & Shuangzhe Liu, 2022. "Robust beta regression modeling with errors-in-variables: a Bayesian approach and numerical applications," Statistical Papers, Springer, vol. 63(3), pages 919-942, June.
    10. Tatiyana V. Apanasovich & David Ruppert & Joanne R. Lupton & Natasa Popovic & Nancy D. Turner & Robert S. Chapkin & Raymond J. Carroll, 2008. "Aberrant Crypt Foci and Semiparametric Modeling of Correlated Binary Data," Biometrics, The International Biometric Society, vol. 64(2), pages 490-500, June.
    11. M.-L. Feddag, 2016. "Pairwise likelihood estimation for the normal ogive model with binary data," AStA Advances in Statistical Analysis, Springer;German Statistical Society, vol. 100(2), pages 223-237, April.
    12. Takahiro Hoshino & Yuya Shimizu, 2019. "Doubly Robust-type Estimation of Population Moments and Parameters in Biased Sampling," Keio-IES Discussion Paper Series 2019-006, Institute for Economics Studies, Keio University.
    13. Renard, Didier & Molenberghs, Geert & Geys, Helena, 2004. "A pairwise likelihood approach to estimation in multilevel probit models," Computational Statistics & Data Analysis, Elsevier, vol. 44(4), pages 649-667, January.
    14. Iddi Samuel & Nwoko Esther O., 2017. "Effect of covariate misspecifications in the marginalized zero-inflated Poisson model," Monte Carlo Methods and Applications, De Gruyter, vol. 23(2), pages 111-120, June.
    15. Zhengxin Zhang & Xiaosheng Si & Changhua Hu & Xiangyu Kong, 2015. "Degradation modeling–based remaining useful life estimation: A review on approaches for systems with heterogeneity," Journal of Risk and Reliability, , vol. 229(4), pages 343-355, August.
    16. Sinha, Sanjoy K. & Kaushal, Amit & Xiao, Wenzhong, 2014. "Inference for longitudinal data with nonignorable nonmonotone missing responses," Computational Statistics & Data Analysis, Elsevier, vol. 72(C), pages 77-91.
    17. Xiaoyu Wang & Liuquan Sun, 2023. "Joint modeling of generalized scale-change models for recurrent event and failure time data," Lifetime Data Analysis: An International Journal Devoted to Statistical Methods and Applications for Time-to-Event Data, Springer, vol. 29(1), pages 1-33, January.
    18. Francesco BARTOLUCCI & Silvia BACCI & Claudia PIGINI, 2015. "A Misspecification Test for Finite-Mixture Logistic Models for Clustered Binary and Ordered Responses," Working Papers 410, Universita' Politecnica delle Marche (I), Dipartimento di Scienze Economiche e Sociali.
    19. Philip Kostov, 2010. "Do Buyers’ Characteristics and Personal Relationships Affect Agricultural Land Prices?," Land Economics, University of Wisconsin Press, vol. 86(1), pages 48-65.
    20. Fei Jiang & Sebastien Haneuse, 2017. "A Semi-parametric Transformation Frailty Model for Semi-competing Risks Survival Data," Scandinavian Journal of Statistics, Danish Society for Theoretical Statistics;Finnish Statistical Society;Norwegian Statistical Association;Swedish Statistical Association, vol. 44(1), pages 112-129, March.

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:csdana:v:134:y:2019:i:c:p:1-16. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: http://www.elsevier.com/locate/csda .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.