Fast two-stage estimator for clustered count data with overdispersion期刊界 All Journals 搜尽天下杂志传播学术成果专业期刊搜索期刊信息化学术搜索

按检索

Fast two-stage estimator for clustered count data with overdispersion

Authors:	Alvaro J Flórez Geert Molenberghs Geert Verbeke Michael G Kenward Pavlos Mamouris Bert Vaes

Institution:	1. I-BioStat, Universiteit Hasselt, Diepenbeek, Belgiumhttps://orcid.org/0000-0001-6127-8733;2. I-BioStat, Universiteit Hasselt, Diepenbeek, Belgium;3. I-BioStat, KU Leuven, Leuven, Belgiumhttps://orcid.org/0000-0002-6453-5448;4. I-BioStat, KU Leuven, Leuven, Belgium;5. Emeritus from London School of Hygiene and Tropical Medicine, Ashkirk, London, UK;6. Academisch Centrum voor Huisartsgeneeskunde, KU Leuven, Leuven, Belgium

Abstract:	Clustered count data are commonly analysed by the generalized linear mixed model (GLMM). Here, the correlation due to clustering and some overdispersion is captured by the inclusion of cluster-specific normally distributed random effects. Often, the model does not capture the variability completely. Therefore, the GLMM can be extended by including a set of gamma random effects. Routinely, the GLMM is fitted by maximizing the marginal likelihood. However, this process is computationally intensive. Although feasible with medium to large data, it can be too time-consuming or computationally intractable with very large data. Therefore, a fast two-stage estimator for correlated, overdispersed count data is proposed. It is rooted in the split-sample methodology. Based on a simulation study, it shows good statistical properties. Furthermore, it is computationally much faster than the full maximum likelihood estimator. The approach is illustrated using a large dataset belonging to a network of Belgian general practices.

Keywords:	Generalized linear mixed model hierarchical data negative binomial model poisson model random effects

设为首页 | 免责声明 | 关于勤云 | 加入收藏