z-logo
Premium
Count transformation models
Author(s) -
Siegfried Sandra,
Hothorn Torsten
Publication year - 2020
Publication title -
methods in ecology and evolution
Language(s) - English
Resource type - Journals
SCImago Journal Rank - 3.425
H-Index - 105
ISSN - 2041-210X
DOI - 10.1111/2041-210x.13383
Subject(s) - count data , negative binomial distribution , quasi likelihood , transformation (genetics) , mathematics , generalized linear model , statistics , data transformation , linear model , overdispersion , poisson distribution , econometrics , computer science , data mining , gene , biochemistry , chemistry , data warehouse
The effect of explanatory environmental variables on a species' distribution is often assessed using a count regression model. Poisson generalized linear models or negative binomial models are common, but the traditional approach of modelling the mean after log or square root transformation remains popular and in some cases is even advocated. We propose a novel framework of linear models for count data. Similar to the traditional approach, the new models apply a transformation to count responses; however, this transformation is estimated from the data and not defined a priori. In contrast to simple least‐squares fitting and in line with Poisson or negative binomial models, the exact discrete likelihood is optimized for parameter estimation and inference. Simple interpretation of effects in the linear predictors is possible. Count transformation models provide a new approach to regressing count data in a distribution‐free yet fully parametric fashion, obviating the need to a priori commit to a specific parametric family of distributions or to a specific transformation. The models are a generalization of discrete Weibull models for counts and are thus able to handle over‐ and underdispersion. We demonstrate empirically that the models are more flexible than Poisson or negative binomial models but still maintain interpretability of multiplicative effects. A re‐analysis of deer–vehicle collisions and the results of artificial simulation experiments provide evidence of the practical applicability of the model framework. In ecology studies, uncertainties regarding whether and how to transform count data can be resolved in the framework of count transformation models, which were designed to simultaneously estimate an appropriate transformation and the linear effects of environmental variables by maximizing the exact count log‐likelihood. The application of data‐driven transformations allows over‐ and underdispersion to be addressed in a model‐based approach. Models in this class can be compared to Poisson or negative binomial models using the in‐ or out‐of‐sample log‐likelihood. Extensions to nonlinear additive or interaction effects, correlated observations, hurdle‐type models and other more complex situations are possible. A free software implementation is available in the cotram add‐on package to the R system for statistical computing.

This content is not available in your region!

Continue researching here.

Having issues? You can contact us here