A new alpha power type-1 family of distributions and modelling the overdispersed count outcome

Document Type : Original Scientific Paper

Authors

1 Department of Statistics, Yazd University, Yazd, Iran

2 Department of Statistics, University of Isfahan, Isfahan, Iran

Abstract

In this paper‎, ‎we introduce a novel family of statistical models called a new alpha power type-1 family of distributions‎. ‎Three sub-cases of the family are discussed‎. ‎Based on the novel family‎, ‎a special model‎, ‎explicitly‎, ‎a new alpha power type-1-Weibull distribution is studied in depth‎. ‎The new model has very interesting patterns of failure rates like increasing‎, ‎decreasing‎, ‎bathtub‎, ‎and parabola-down‎. ‎Hence‎, ‎it is so flexible‎. ‎Based on the comparison analysis‎, ‎among five well-known models‎, ‎it has an impact on health data analysis‎. ‎Furthermore‎, ‎the count data models capable of handling overdispersion and zero-inflation are discussed and applied the real health data‎. ‎The zero-inflated negative binomial model in the frequentist approach has shown its popularity in handling both overdispersion and zero-inflation simultaneously‎, ‎while the discrete Weibull model with the logit(q) link in the Bayesian approach outperformed its counterparts.

Keywords

Main Subjects


Abdullah, S.A. and Ahmad, W.M.A.W. (2013). Modelling the incident and risk factors of streptococcus pneumonia transmission among children in Malaysia. International Journal of Physical, Chemical and Mathematical Sciences, 2(2):106–116.
Adesina, O.S. (2021). Models for zero truncated count data in medicine and insurance. Theory and Practice of Mathematics and Computer Science, 6:129–141.
Adesina, O., Agunbiade, D. and Osundina, S. (2017). Bayesian regression model for counts in scholarship. Journal of Mathematical Theory and Modelling, 7(9):46–57.
Adesina, O.S., Agunbiade, D.A. and Oguntunde, P.E. (2021). Flexible Bayesian Dirichlet mixtures of generalized linear mixed models for count data. Scientific African, 13:e00963.
Ahmad, Z., Ampadu, C.B., Hamedani, G.G., Jamal, F. and Nasir, M.A. (2019). The new exponentiated TX class of distributions: properties, characterizations and application. Pakistan Journal of Statistics and Operation Research, XV(IV):941–962.
Akaike, H. (1974). A new look at the statistical model identification. IEEE Transactions on Automatic Control, 19(6):716–723.
Altun, E. (2018). A new zero-inflated regression model with application. İstatistikçiler Dergisi: İstatistik ve Aktüerya, 11(2):73–80.
Atienza, N., Garcia-Heras, J., Munoz-Pichardo, J.M. and Villa, R. (2008). An application of mixture distributions in modelization of length of hospital stay. Statistics in Medicine, 27(9):1403–1420.
Bao, Y., Vinciotti, V., Wit, E. and Hoen, P. (2014). Joint modelling of ChIP-seq data via a Markov random field model. Biostatistics, 15(2):296–310.
Bolstad, W.M. (2007). Introduction to Bayesian Statistics. New York: John Wiley & Sons. Inc.
Bozdogan, H. (1987). Model selection and Akaike's information criterion (AIC): The general theory and its analytical extensions. Psychometrika, 52(3):345–370.
Cameron, A.C. and Trivedi, P.K. (2013). Regression Analysis of Count Data. Vol. 53, Cambridge University Press.
Cameron, A.C. and Trivedi, P.K. (2005). Microeconometrics Methods and Application. Cambridge University Press.
Carter, E.M. and Potts, H.W. (2014). Predicting length of stay from an electronic patient record system: a primary total knee replacement example. BMC Medical Informatics and Decision Making, 14(1):1–13.
Chanialidis, C. (2015). Bayesian mixture models for count data. Doctoral dissertation, University of Glasgow.
Chesneau, C., Bakouch, H.S. and Hussain, T. (2019). A new class of probability distributions via cosine and sine functions with applications. Communications in Statistics-Simulation and Computation, 48(8):2287–2300.
Chesneau, C., Karakaya, K., Bakouch, H.S. and Kuş, C. (2022). An alternative to the Marshall-Olkin family of distributions: bootstrap, regression and applications. Communications on Applied Mathematics and Computation, 4:1–29.
Cordeiro, G.M., Ortega, E.M. and da Cunha, D.C. (2013). The exponentiated generalized class of distributions. Journal of Data Science, 11(1):1–27.
Demétrio, C.G., Hinde, J. and Moral, R.A. (2014). Models for overdispersed data in entomology. Ecological Modelling Applied to Entomology, 219–259.
El-Desouky, B.S., Mustafa, A. and Al-Garash, S. (2016). The exponential flexible Weibull extension distribution. arXiv preprint arXiv:1605.08152.
Famoye, F. and Singh, K.P. (2006). Zero-inflated generalized Poisson regression model with an application to domestic violence data. Journal of Data Science, 4(1):117–130.
Grunwald, G.K., Bruce, S.L., Jiang, L., Strand, M. and Rabinovitch, N. (2011). A statistical model for under-or overdispersed clustered and longitudinal count data. Biometrical Journal, 53(4):578–594.
Güneri, Ö.İ. and Durmuş, B. (2021). Models for overdispersion count data with generalized distribution: An application to parasites intensity. Journal of New Theory, 35:48–61.
Hadfield, J.D. (2010). MCMC methods for multi-response generalized linear mixed models: The MCMCglmm R package. Journal of Statistical Software, 33(2):1–22.
Haselimashhadi, H., Vinciotti, V. and Yu, K. (2016). A new Bayesian regression model for counts in medicine. arXiv:1601.02820.
Hannan, E.J. and Quinn, B.G. (1979). The determination of the order of an autoregression. Journal of the Royal Statistical Society: Series B (Methodological), 41(2):190–195.
Hassan, A.S. and Elgarhy, M. (2016). A new family of exponentiated Weibull-generated distributions. International Journal of Mathematics and its Applications, 4(1-D):135–148.
Hu, M.C., Pavlicova, M. and Nunes, E.V. (2011). Zero-inflated and hurdle models of count data with extra zeros: examples from an HIV-risk reduction intervention trial. The American Journal of Drug and Alcohol Abuse, 37(5):367–375.
José, A.F.M. and Santos Silva, J.M.C. (2005). Quantiles for counts. Journal of the American Statistical Association, 100(472):1226–1237.
Joshi, R.K. and Kumar, V. (2021). Poisson inverse Weibull distribution with theory and applications. International Journal of Statistics and Systems, 16(1):1–16.
Mahdavi, A. and Oliveira Silva, G. (2017). A method to expand family of continuous distributions based on truncated distributions. Journal of Statistical Research of Iran, 13(2):231–247.
Malehi, A.S., Pourmotahari, F. and Angali, K.A. (2015). Statistical models for the analysis of skewed healthcare cost data: a simulation study. Health Economics Review, 5(1):1–16.
Molenberghs, G., Verbeke, G., Demétrio, C.G. and Vieira, A.M. (2010). A family of generalized linear models for repeated measures with normal and conjugate random effects. Statistical Science, 25(3):325–347.
Molenberghs, G., Verbeke, G. and Demétrio, C.G. (2007). An extended random-effects approach to modelling repeated, overdispersed count data. Lifetime Data Analysis, 13(4):513–531.
Nakagawa, T. and Osaki, S. (1975). The discrete Weibull distribution. IEEE Transactions on Reliability, 24(5):300–301.
Nekoukhou, V. and Bidram, H. (2020). A new discrete distribution based on geometric odds ratio. Journal of Statistical Modelling: Theory and Applications, 1(2):153–166.
Oliveira, I.R.C.D. (2014). Modeling strategies for complex hierarchical and overdispersed data in the life sciences. Doctoral dissertation, Universidade de São Paulo, Uhasselt.
Ozsolak, F. and Milos, P.M. (2011). RNA sequencing: advances, challenges and opportunities. Nature Reviews Genetics, 12(2):87–98.
Remi, J.D., Olumide, S.A., Pelumi, E.O. and Olasunmbo, O.A. (2019). Adaptive regression model for highly skewed count data. International Journal of Mechanical Engineering and Technology, 10(1):1964–1972.
Robinson, M.D. and Smyth, G.K. (2008). Small-sample estimation of negative binomial dispersion, with applications to SAGE data. Biostatistics, 9(2):321–332.
Roozegar, R., Tekle, G. and Hamedani, G. (2022). A new generalized-X family of distributions: Applications, characterization and a mixture of random effect models. Pakistan Journal of Statistics and Operation Research, 18(2):483–504.
Schwarz, G. (1978). Estimating the dimension of a model. The Annals of Statistics, 6(2):461–464.
Sellers, K.F. and Shmueli, G. (2010). Predicting censored count data with COM-Poisson regression. Robert H. Smith School Research Paper No. RHS-06-129.
Shehata, W.A., Yousof, H. and Aboraya, M. (2021). A Novel generator of continuous probability distributions for the asymmetric left-skewed bimodal real-life data with properties and copulas. Pakistan Journal of Statistics and Operation Research, 17(4):943–961.
Workie, M.S. and Lakew, A.M. (2018). Bayesian count regression analysis for determinants of antenatal care service visits among pregnant women in Amhara regional state, Ethiopia. Journal of Big Data, 5(1):1–23.