Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Multivariate Generalized Linear Mixed Models for Count Data

View through CrossRef
Univariate regression models have rich literature for counting data. However, this is not the case for multivariate count data. Therefore, we present the Multivariate Generalized Linear Mixed Models framework that deals with a multivariate set of responses, measuring the correlation between them through random effects that follows a multivariate normal distribution. This model is based on a GLMM with a random intercept and the estimation process remains the same as a standard GLMM with random effects integrated out via Laplace approximation. We efficiently implemented this model through the TMB package available in R. We used Poisson, negative binomial (NB), and COM-Poisson distributions. To assess the estimator properties, we conducted a simulation study considering four different sample sizes and three different correlation values for each distribution. We achieved unbiased and consistent estimators for Poisson and NB distributions; for COM-Poisson estimators were consistent, but biased, especially for dispersion, variance, and correlation parameter estimators. These models were applied to two datasets. The first concerns a sample from 30 different sites collected in Australia where the number of times each one of the 41 different ant species was registered; which results in an impressive 820 variance-covariance and 41 dispersion parameters are estimated simultaneously, let alone the regression parameters. The second is from the Australia Health Survey with 5 response variables and 5190 respondents. These datasets can be considered overdispersed by the generalized dispersion index. The COM-Poisson model overcame the other two competitors considering three goodness-of-fit indexes, AIC, BIC, and maximized log-likelihood values. As a result, it estimated parameters with smaller standard errors and a greater number of significant correlation coefficients. Therefore, the proposed model is capable of dealing with multivariate count data, either under- equi- or overdispersed responses, and measuring any kind of correlation between them taking into account the effects of the covariates.
Title: Multivariate Generalized Linear Mixed Models for Count Data
Description:
Univariate regression models have rich literature for counting data.
However, this is not the case for multivariate count data.
Therefore, we present the Multivariate Generalized Linear Mixed Models framework that deals with a multivariate set of responses, measuring the correlation between them through random effects that follows a multivariate normal distribution.
This model is based on a GLMM with a random intercept and the estimation process remains the same as a standard GLMM with random effects integrated out via Laplace approximation.
We efficiently implemented this model through the TMB package available in R.
We used Poisson, negative binomial (NB), and COM-Poisson distributions.
To assess the estimator properties, we conducted a simulation study considering four different sample sizes and three different correlation values for each distribution.
We achieved unbiased and consistent estimators for Poisson and NB distributions; for COM-Poisson estimators were consistent, but biased, especially for dispersion, variance, and correlation parameter estimators.
These models were applied to two datasets.
The first concerns a sample from 30 different sites collected in Australia where the number of times each one of the 41 different ant species was registered; which results in an impressive 820 variance-covariance and 41 dispersion parameters are estimated simultaneously, let alone the regression parameters.
The second is from the Australia Health Survey with 5 response variables and 5190 respondents.
These datasets can be considered overdispersed by the generalized dispersion index.
The COM-Poisson model overcame the other two competitors considering three goodness-of-fit indexes, AIC, BIC, and maximized log-likelihood values.
As a result, it estimated parameters with smaller standard errors and a greater number of significant correlation coefficients.
Therefore, the proposed model is capable of dealing with multivariate count data, either under- equi- or overdispersed responses, and measuring any kind of correlation between them taking into account the effects of the covariates.

Related Results

Tracing Hematological Shifts in Pregnancy: How Anemia and Thrombocytopenia Evolve Across Trimesters
Tracing Hematological Shifts in Pregnancy: How Anemia and Thrombocytopenia Evolve Across Trimesters
Abstract Introduction Given pregnancy's significant impact on hematological parameters, monitoring these changes across trimesters is crucial. This study aims to evaluate hematolog...
Selection of Injectable Drug Product Composition using Machine Learning Models (Preprint)
Selection of Injectable Drug Product Composition using Machine Learning Models (Preprint)
BACKGROUND As of July 2020, a Web of Science search of “machine learning (ML)” nested within the search of “pharmacokinetics or pharmacodynamics” yielded over 100...
Association of T lymphocytes level and clinical severity in patients of COVID-19 in Shenzhen China
Association of T lymphocytes level and clinical severity in patients of COVID-19 in Shenzhen China
To explore the correlation between T lymphocytes and clinical severity in patients of COVID-19. A total of 183 COVID-19 patients were recruited in Shenzhen Third People’s Hospital ...
Semiparametric methods for regression analysis of panel count data and mixed panel count data
Semiparametric methods for regression analysis of panel count data and mixed panel count data
[ACCESS RESTRICTED TO THE UNIVERSITY OF MISSOURI AT AUTHOR'S REQUEST.] Recurrent event data and panel count data are two common types of data that have been studied extensively in ...
Novel/Old Generalized Multiplicative Zagreb Indices of Some Special Graphs
Novel/Old Generalized Multiplicative Zagreb Indices of Some Special Graphs
Topological descriptor is a fixed real number directly attached with the molecular graph to predict the physical and chemical properties of the chemical compound. Gutman and Trinaj...
Linear Dynamic Models
Linear Dynamic Models
It was emphasized in Chapter 1 that low-order, linear time-invariant models provide the foundation for much intuition about dynamic phenomena in the real world. This chapter provid...
A Multivariate Poisson Deep Learning Model for Genomic Prediction of Count Data
A Multivariate Poisson Deep Learning Model for Genomic Prediction of Count Data
Abstract The paradigm called genomic selection (GS) is a revolutionary way of developing new plants and animals. This is a predictive methodology, since it uses lear...
Testing for Interactions in Multivariate Data
Testing for Interactions in Multivariate Data
Abstract Factorial designs are a mainstay of the scientific paradigm, allowing the effects of multiple experimental factors and their interactions to be efficiently...

Back to Top