首页 | 本学科首页   官方微博 | 高级检索  
     


Fast two-stage estimator for clustered count data with overdispersion
Authors:Alvaro J. Flórez  Geert Molenberghs  Geert Verbeke  Michael G. Kenward  Pavlos Mamouris  Bert Vaes
Affiliation:1. I-BioStat, Universiteit Hasselt, Diepenbeek, Belgium"ORCIDhttps://orcid.org/0000-0001-6127-8733;2. I-BioStat, Universiteit Hasselt, Diepenbeek, Belgium;3. I-BioStat, KU Leuven, Leuven, Belgium"ORCIDhttps://orcid.org/0000-0002-6453-5448;4. I-BioStat, KU Leuven, Leuven, Belgium;5. Emeritus from London School of Hygiene and Tropical Medicine, Ashkirk, London, UK;6. Academisch Centrum voor Huisartsgeneeskunde, KU Leuven, Leuven, Belgium
Abstract:
Clustered count data are commonly analysed by the generalized linear mixed model (GLMM). Here, the correlation due to clustering and some overdispersion is captured by the inclusion of cluster-specific normally distributed random effects. Often, the model does not capture the variability completely. Therefore, the GLMM can be extended by including a set of gamma random effects. Routinely, the GLMM is fitted by maximizing the marginal likelihood. However, this process is computationally intensive. Although feasible with medium to large data, it can be too time-consuming or computationally intractable with very large data. Therefore, a fast two-stage estimator for correlated, overdispersed count data is proposed. It is rooted in the split-sample methodology. Based on a simulation study, it shows good statistical properties. Furthermore, it is computationally much faster than the full maximum likelihood estimator. The approach is illustrated using a large dataset belonging to a network of Belgian general practices.
Keywords:Generalized linear mixed model  hierarchical data  negative binomial model  poisson model  random effects
设为首页 | 免责声明 | 关于勤云 | 加入收藏

Copyright©北京勤云科技发展有限公司  京ICP备09084417号