IDEAS home Printed from https://ideas.repec.org/a/spr/advdac/v14y2020i2d10.1007_s11634-019-00369-4.html
   My bibliography  Save this article

Seemingly unrelated clusterwise linear regression

Author

Listed:
  • Giuliano Galimberti

    (University of Bologna)

  • Gabriele Soffritti

    (University of Bologna)

Abstract

Linear regression models based on finite Gaussian mixtures represent a flexible tool for the analysis of linear dependencies in multivariate data. They are suitable for dealing with correlated response variables when data come from a heterogeneous population composed of two or more sub-populations, each of which is characterised by a different linear regression model. Several types of finite mixtures of linear regression models have been specified by changing the assumptions on the parameters that differentiate the sub-populations and/or the vectors of regressors that affect the response variables. They are made more flexible in the class of models defined by mixtures of seemingly unrelated Gaussian linear regressions illustrated in this paper. With these models, the researcher is enabled to use a different vector of regressors for each dependent variable. The proposed class includes parsimonious models obtained by imposing suitable constraints on the variances and covariances of the response variables in the sub-populations. Details about the model identification and maximum likelihood estimation are given. The usefulness of these models is shown through the analysis of a real dataset. Regularity conditions for the model class are illustrated and a proof is provided that, when these conditions are met, the consistency of the maximum likelihood estimator under the examined models is ensured. In addition, the behaviour of this estimator in the presence of finite samples is numerically evaluated through the analysis of simulated datasets.

Suggested Citation

  • Giuliano Galimberti & Gabriele Soffritti, 2020. "Seemingly unrelated clusterwise linear regression," Advances in Data Analysis and Classification, Springer;German Classification Society - Gesellschaft für Klassifikation (GfKl);Japanese Classification Society (JCS);Classification and Data Analysis Group of the Italian Statistical Society (CLADAG);International Federation of Classification Societies (IFCS), vol. 14(2), pages 235-260, June.
  • Handle: RePEc:spr:advdac:v:14:y:2020:i:2:d:10.1007_s11634-019-00369-4
    DOI: 10.1007/s11634-019-00369-4
    as

    Download full text from publisher

    File URL: http://link.springer.com/10.1007/s11634-019-00369-4
    File Function: Abstract
    Download Restriction: Access to the full text of the articles in this series is restricted.

    File URL: https://libkey.io/10.1007/s11634-019-00369-4?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    As the access to this document is restricted, you may want to search for a different version of it.

    References listed on IDEAS

    as
    1. De Veaux, Richard D., 1989. "Mixtures of linear regressions," Computational Statistics & Data Analysis, Elsevier, vol. 8(3), pages 227-245, November.
    2. Bartolucci, F. & Scaccia, L., 2005. "The use of mixtures for dealing with non-normal regression errors," Computational Statistics & Data Analysis, Elsevier, vol. 48(4), pages 821-834, April.
    3. T. Rolf Turner, 2000. "Estimating the propagation rate of a viral infection of potato plants via mixtures of regressions," Journal of the Royal Statistical Society Series C, Royal Statistical Society, vol. 49(3), pages 371-384.
    4. Roberto Rocci & Stefano Antonio Gattone & Roberto Di Mari, 2018. "A data driven equivariant approach to constrained Gaussian mixture modeling," Advances in Data Analysis and Classification, Springer;German Classification Society - Gesellschaft für Klassifikation (GfKl);Japanese Classification Society (JCS);Classification and Data Analysis Group of the Italian Statistical Society (CLADAG);International Federation of Classification Societies (IFCS), vol. 12(2), pages 235-260, June.
    5. Grün, Bettina & Leisch, Friedrich, 2008. "FlexMix Version 2: Finite Mixtures with Concomitant Variables and Varying and Constant Parameters," Journal of Statistical Software, Foundation for Open Access Statistics, vol. 28(i04).
    6. Wayne DeSarbo & William Cron, 1988. "A maximum likelihood methodology for clusterwise linear regression," Journal of Classification, Springer;The Classification Society, vol. 5(2), pages 249-282, September.
    7. Henningsen, Arne & Hamann, Jeff D., 2007. "systemfit: A Package for Estimating Systems of Simultaneous Equations in R," Journal of Statistical Software, Foundation for Open Access Statistics, vol. 23(i04).
    8. Cathy Maugis & Gilles Celeux & Marie-Laure Martin-Magniette, 2009. "Variable Selection for Clustering with Gaussian Mixture Models," Biometrics, The International Biometric Society, vol. 65(3), pages 701-709, September.
    9. Judith A. Chevalier & Anil K. Kashyap & Peter E. Rossi, 2003. "Why Don't Prices Rise During Periods of Peak Demand? Evidence from Scanner Data," American Economic Review, American Economic Association, vol. 93(1), pages 15-37, March.
    10. W. A. Donnelly, 1982. "The Regional Demand for Petrol in Australia," The Economic Record, The Economic Society of Australia, vol. 58(4), pages 317-327, December.
    11. Adam Tashman & Robert Frey, 2009. "Modeling risk in arbitrage strategies using finite mixtures," Quantitative Finance, Taylor & Francis Journals, vol. 9(5), pages 495-503.
    12. Donnelly, W A, 1982. "The Regional Demand for Petrol in Australia," The Economic Record, The Economic Society of Australia, vol. 58(163), pages 317-327, December.
    13. Ingrassia, Salvatore & Rocci, Roberto, 2011. "Degeneracy of the EM algorithm for the MLE of multivariate Gaussian mixtures and dynamic constraints," Computational Statistics & Data Analysis, Elsevier, vol. 55(4), pages 1715-1725, April.
    Full references (including those not matched with items on IDEAS)

    Citations

    Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.
    as


    Cited by:

    1. Gabriele Perrone & Gabriele Soffritti, 2023. "Seemingly unrelated clusterwise linear regression for contaminated data," Statistical Papers, Springer, vol. 64(3), pages 883-921, June.
    2. Diani, Cecilia & Galimberti, Giuliano & Soffritti, Gabriele, 2022. "Multivariate cluster-weighted models based on seemingly unrelated linear regression," Computational Statistics & Data Analysis, Elsevier, vol. 171(C).

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Gabriele Perrone & Gabriele Soffritti, 2023. "Seemingly unrelated clusterwise linear regression for contaminated data," Statistical Papers, Springer, vol. 64(3), pages 883-921, June.
    2. Diani, Cecilia & Galimberti, Giuliano & Soffritti, Gabriele, 2022. "Multivariate cluster-weighted models based on seemingly unrelated linear regression," Computational Statistics & Data Analysis, Elsevier, vol. 171(C).
    3. Giuliano Galimberti & Lorenzo Nuzzi & Gabriele Soffritti, 2021. "Covariance matrix estimation of the maximum likelihood estimator in multivariate clusterwise linear regression," Statistical Methods & Applications, Springer;Società Italiana di Statistica, vol. 30(1), pages 235-268, March.
    4. Rainer Schlittgen, 2011. "A weighted least-squares approach to clusterwise regression," AStA Advances in Statistical Analysis, Springer;German Statistical Society, vol. 95(2), pages 205-217, June.
    5. Lloyd-Jones, Luke R. & Nguyen, Hien D. & McLachlan, Geoffrey J., 2018. "A globally convergent algorithm for lasso-penalized mixture of linear regression models," Computational Statistics & Data Analysis, Elsevier, vol. 119(C), pages 19-38.
    6. Gianfranco DI VAIO & Michele BATTISTI, 2010. "A Spatially-Filtered Mixture of Beta-Convergence Regression for EU Regions, 1980-2002," Regional and Urban Modeling 284100013, EcoMod.
    7. Sphiwe B. Skhosana & Salomon M. Millard & Frans H. J. Kanfer, 2023. "A Novel EM-Type Algorithm to Estimate Semi-Parametric Mixtures of Partially Linear Models," Mathematics, MDPI, vol. 11(5), pages 1-20, February.
    8. M. Battisti & F. Belloc & M. Del Gatto, 2017. "Technology-specific Production Functions," Working Paper CRENoS 201709, Centre for North South Economic Research, University of Cagliari and Sassari, Sardinia.
    9. Di Vaio, Gianfranco & Enflo, Kerstin, 2011. "Did globalization drive convergence? Identifying cross-country growth regimes in the long run," European Economic Review, Elsevier, vol. 55(6), pages 832-844, August.
    10. Atefeh Zarei & Zahra Khodadadi & Mohsen Maleki & Karim Zare, 2023. "Robust mixture regression modeling based on two-piece scale mixtures of normal distributions," Advances in Data Analysis and Classification, Springer;German Classification Society - Gesellschaft für Klassifikation (GfKl);Japanese Classification Society (JCS);Classification and Data Analysis Group of the Italian Statistical Society (CLADAG);International Federation of Classification Societies (IFCS), vol. 17(1), pages 181-210, March.
    11. Antonello Maruotti & Pierfrancesco Alaimo Di Loro, 2023. "CO2 emissions and growth: A bivariate bidimensional mean‐variance random effects model," Environmetrics, John Wiley & Sons, Ltd., vol. 34(5), August.
    12. Michele Battisti & Gianfranco Vaio, 2009. "A spatially filtered mixture of β-convergence regressions for EU regions, 1980–2002," Studies in Empirical Economics, in: Giuseppe Arbia & Badi H. Baltagi (ed.), Spatial Econometrics, pages 105-121, Springer.
    13. Michele Battisti & Filippo Belloc & Massimo Del Gatto, 2020. "Labor Productivity and Firm-Level TFP with Technology-Specific Production Function," Review of Economic Dynamics, Elsevier for the Society for Economic Dynamics, vol. 35, pages 283-300, January.
    14. Salvatore Ingrassia & Antonio Punzo, 2020. "Cluster Validation for Mixtures of Regressions via the Total Sum of Squares Decomposition," Journal of Classification, Springer;The Classification Society, vol. 37(2), pages 526-547, July.
    15. Galimberti, Giuliano & Soffritti, Gabriele, 2014. "A multivariate linear regression analysis using finite mixtures of t distributions," Computational Statistics & Data Analysis, Elsevier, vol. 71(C), pages 138-150.
    16. Shaw, Charles, 2020. "Econometric Analysis of Demand for Petrol in India, 1966-2019," MPRA Paper 104797, University Library of Munich, Germany.
    17. Gianfranco Di Vaio & Kerstin Enflo, 2009. "Did Globalization Lead to Segmentation? Identifying Cross-Country Growth Regimes in the Long-Run," Discussion Papers 09-08, University of Copenhagen. Department of Economics.
    18. Robert V. Breunig & Carol Gisz, 2009. "An Exploration of Australian Petrol Demand: Unobservable Habits, Irreversibility and Some Updated Estimates," The Economic Record, The Economic Society of Australia, vol. 85(268), pages 73-91, March.
    19. Sanjeena Subedi & Antonio Punzo & Salvatore Ingrassia & Paul McNicholas, 2013. "Clustering and classification via cluster-weighted factor analyzers," Advances in Data Analysis and Classification, Springer;German Classification Society - Gesellschaft für Klassifikation (GfKl);Japanese Classification Society (JCS);Classification and Data Analysis Group of the Italian Statistical Society (CLADAG);International Federation of Classification Societies (IFCS), vol. 7(1), pages 5-40, March.
    20. Brons, Martijn & Nijkamp, Peter & Pels, Eric & Rietveld, Piet, 2008. "A meta-analysis of the price elasticity of gasoline demand. A SUR approach," Energy Economics, Elsevier, vol. 30(5), pages 2105-2122, September.

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:spr:advdac:v:14:y:2020:i:2:d:10.1007_s11634-019-00369-4. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Sonal Shukla or Springer Nature Abstracting and Indexing (email available below). General contact details of provider: http://www.springer.com .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.