IDEAS home Printed from https://ideas.repec.org/a/eee/jmvana/v126y2014icp167-183.html
   My bibliography  Save this article

Strong consistency and robustness of the Forward Search estimator of multivariate location and scatter

Author

Listed:
  • Cerioli, Andrea
  • Farcomeni, Alessio
  • Riani, Marco

Abstract

The Forward Search is a powerful general method for detecting anomalies in structured data, whose diagnostic power has been shown in many statistical contexts. However, despite the wealth of empirical evidence in favor of the method, only few theoretical properties have been established regarding the resulting estimators. We show that the Forward Search estimators are strongly consistent at the multivariate normal model. We also obtain their finite sample breakdown point. Our results put the Forward Search approach for multivariate data on a solid statistical ground, which formally motivates its use in robust applied statistics. Furthermore, they allow us to compare the Forward Search estimators with other well known multivariate high-breakdown techniques.

Suggested Citation

  • Cerioli, Andrea & Farcomeni, Alessio & Riani, Marco, 2014. "Strong consistency and robustness of the Forward Search estimator of multivariate location and scatter," Journal of Multivariate Analysis, Elsevier, vol. 126(C), pages 167-183.
  • Handle: RePEc:eee:jmvana:v:126:y:2014:i:c:p:167-183
    DOI: 10.1016/j.jmva.2013.12.010
    as

    Download full text from publisher

    File URL: http://www.sciencedirect.com/science/article/pii/S0047259X14000037
    Download Restriction: Full text for ScienceDirect subscribers only

    File URL: https://libkey.io/10.1016/j.jmva.2013.12.010?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    As the access to this document is restricted, you may want to search for a different version of it.

    References listed on IDEAS

    as
    1. Marco Riani & Andrea Cerioli & Francesca Torti, 2014. "On consistency factors and efficiency of robust S-estimators," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 23(2), pages 356-387, June.
    2. Anthony C. Atkinson, 2002. "Forward search added-variable t-tests and the effect of masked outliers on model selection," Biometrika, Biometrika Trust, vol. 89(4), pages 939-946, December.
    3. Marco Riani & Anthony C. Atkinson & Andrea Cerioli, 2009. "Finding an unknown number of multivariate outliers," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 71(2), pages 447-466, April.
    4. Cerioli, Andrea & Farcomeni, Alessio & Riani, Marco, 2013. "Robust distances for outlier-free goodness-of-fit testing," Computational Statistics & Data Analysis, Elsevier, vol. 65(C), pages 29-45.
    5. Francesca De Battisti & Silvia Salini, 2013. "Robust analysis of bibliometric data," Statistical Methods & Applications, Springer;Società Italiana di Statistica, vol. 22(2), pages 269-283, June.
    6. Cator, Eric A. & Lopuhaä, Hendrik P., 2010. "Asymptotic expansion of the minimum covariance determinant estimators," Journal of Multivariate Analysis, Elsevier, vol. 101(10), pages 2372-2388, November.
    7. Søren Johansen & Bent Nielsen, 2013. "Asymptotic analysis of the Forward Search," Discussion Papers 13-01, University of Copenhagen. Department of Economics.
    8. Zani, Sergio & Riani, Marco & Corbellini, Aldo, 1998. "Robust bivariate boxplots and multiple outlier detection," Computational Statistics & Data Analysis, Elsevier, vol. 28(3), pages 257-270, September.
    9. Van Aelst, S. & Vandervieren, E. & Willems, G., 2012. "A Stahel–Donoho estimator based on huberized outlyingness," Computational Statistics & Data Analysis, Elsevier, vol. 56(3), pages 531-542.
    10. Atkinson, A.C. & Riani, M., 2007. "Exploratory tools for clustering multivariate data," Computational Statistics & Data Analysis, Elsevier, vol. 52(1), pages 272-285, September.
    11. Francesca DE BATTISTI & Silvia SALINI, 2011. "Robust analysis of bibliometric data," Departmental Working Papers 2011-36, Department of Economics, Management and Quantitative Methods at Università degli Studi di Milano.
    12. Garcia-Escudero, Luis Angel & Gordaliza, Alfonso, 2005. "Generalized Radius Processes for Elliptically Contoured Distributions," Journal of the American Statistical Association, American Statistical Association, vol. 100, pages 1036-1045, September.
    13. Van Aelst, Stefan & Willems, Gert, 2011. "Robust and Efficient One-Way MANOVA Tests," Journal of the American Statistical Association, American Statistical Association, vol. 106(494), pages 706-718.
    14. Croux, Christophe & Haesbroeck, Gentiane, 1999. "Influence Function and Efficiency of the Minimum Covariance Determinant Scatter Matrix Estimator," Journal of Multivariate Analysis, Elsevier, vol. 71(2), pages 161-190, November.
    15. Hossjer, O. & Croux, C. & Rousseeuw, P. J., 1994. "Asymptotics of Generalized S-Estimators," Journal of Multivariate Analysis, Elsevier, vol. 51(1), pages 148-177, October.
    16. Cerioli, Andrea, 2010. "Multivariate Outlier Detection With High-Breakdown Estimators," Journal of the American Statistical Association, American Statistical Association, vol. 105(489), pages 147-156.
    17. Arismendi, J.C., 2013. "Multivariate truncated moments," Journal of Multivariate Analysis, Elsevier, vol. 117(C), pages 41-75.
    18. Bent Nielsen & Soren Johansen, 2010. "Discussion of The Forward Search: Theory and Data Analysis," Economics Series Working Papers 2010-W02, University of Oxford, Department of Economics.
    Full references (including those not matched with items on IDEAS)

    Citations

    Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.
    as


    Cited by:

    1. Chakraborty, Chitradipa, 2019. "Testing multivariate scatter parameter in elliptical model based on forward search method," Statistics & Probability Letters, Elsevier, vol. 147(C), pages 66-72.
    2. Lisa Crosato & Luigi Grossi, 2019. "Correcting outliers in GARCH models: a weighted forward approach," Statistical Papers, Springer, vol. 60(6), pages 1939-1970, December.
    3. Riani, Marco & Atkinson, Anthony Curtis & Corbellini, Aldo & Farcomeni, Alessio & Laurini, Fabrizio, 2024. "Information Criteria for Outlier Detection Avoiding Arbitrary Significance Levels," Econometrics and Statistics, Elsevier, vol. 29(C), pages 189-205.
    4. Silvia Salini & Andrea Cerioli & Fabrizio Laurini & Marco Riani, 2016. "Reliable Robust Regression Diagnostics," International Statistical Review, International Statistical Institute, vol. 84(1), pages 99-127, April.
    5. Francesca Torti & Aldo Corbellini & Anthony C. Atkinson, 2021. "fsdaSAS: A Package for Robust Regression for Very Large Datasets Including the Batch Forward Search," Stats, MDPI, vol. 4(2), pages 1-21, April.
    6. Torti, Francesca & Corbellini, Aldo & Atkinson, Anthony C., 2021. "fsdaSAS: a package for robust regression for very large datasets including the batch forward search," LSE Research Online Documents on Economics 109895, London School of Economics and Political Science, LSE Library.
    7. Baishuai Zuo & Chuancun Yin & Jing Yao, 2023. "Multivariate range Value-at-Risk and covariance risk measures for elliptical and log-elliptical distributions," Papers 2305.09097, arXiv.org.
    8. Stephen Babos & Andreas Artemiou, 2021. "Cumulative Median Estimation for Sufficient Dimension Reduction," Stats, MDPI, vol. 4(1), pages 1-8, February.
    9. Arismendi, Juan C. & Broda, Simon, 2017. "Multivariate elliptical truncated moments," Journal of Multivariate Analysis, Elsevier, vol. 157(C), pages 29-44.
    10. Baishuai Zuo & Chuancun Yin, 2022. "Multivariate doubly truncated moments for generalized skew-elliptical distributions with application to multivariate tail conditional risk measures," Papers 2203.00839, arXiv.org.
    11. Andrea Cerioli & Marco Riani & Anthony C. Atkinson & Aldo Corbellini, 2018. "The power of monitoring: how to make the most of a contaminated multivariate sample," Statistical Methods & Applications, Springer;Società Italiana di Statistica, vol. 27(4), pages 559-587, December.
    12. Zuppiroli, Marco & Donati, Michele & Riani, Marco & Verga, Giovanni, 2015. "The Impact of Trading Activity in Agricultural Futures Markets," 2015 Fourth Congress, June 11-12, 2015, Ancona, Italy 207848, Italian Association of Agricultural and Applied Economics (AIEAA).
    13. Alessio Farcomeni & Francesco Dotto, 2018. "The power of (extended) monitoring in robust clustering," Statistical Methods & Applications, Springer;Società Italiana di Statistica, vol. 27(4), pages 651-660, December.
    14. Chitradipa Chakraborty & Subhra Sankar Dhar, 2020. "A Test for Multivariate Location Parameter in Elliptical Model Based on Forward Search Method," Sankhya A: The Indian Journal of Statistics, Springer;Indian Statistical Institute, vol. 82(1), pages 68-95, February.
    15. Brenton R. Clarke & Andrew Grose, 2023. "A further study comparing forward search multivariate outlier methods including ATLA with an application to clustering," Statistical Papers, Springer, vol. 64(2), pages 395-420, April.
    16. Atkinson, Anthony C. & Riani, Marco & Torti, Francesca, 2016. "Robust methods for heteroskedastic regression," Computational Statistics & Data Analysis, Elsevier, vol. 104(C), pages 209-222.

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Cerioli, Andrea & Farcomeni, Alessio & Riani, Marco, 2013. "Robust distances for outlier-free goodness-of-fit testing," Computational Statistics & Data Analysis, Elsevier, vol. 65(C), pages 29-45.
    2. Silvia Salini & Andrea Cerioli & Fabrizio Laurini & Marco Riani, 2016. "Reliable Robust Regression Diagnostics," International Statistical Review, International Statistical Institute, vol. 84(1), pages 99-127, April.
    3. Søren Johansen & Lukasz Gatarek, 2014. "Optimal hedging with the cointegrated vector autoregressive model," CREATES Research Papers 2014-40, Department of Economics and Business Economics, Aarhus University.
    4. Pokojovy, Michael & Jobe, J. Marcus, 2022. "A robust deterministic affine-equivariant algorithm for multivariate location and scatter," Computational Statistics & Data Analysis, Elsevier, vol. 172(C).
    5. Marco Riani & Andrea Cerioli & Francesca Torti, 2014. "On consistency factors and efficiency of robust S-estimators," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 23(2), pages 356-387, June.
    6. Søren Johansen & Bent Nielsen, 2014. "Outlier detection algorithms for least squares time series regression," Economics Papers 2014-W04, Economics Group, Nuffield College, University of Oxford.
    7. Andrea Cerioli & Marco Riani & Anthony C. Atkinson & Aldo Corbellini, 2018. "The power of monitoring: how to make the most of a contaminated multivariate sample," Statistical Methods & Applications, Springer;Società Italiana di Statistica, vol. 27(4), pages 559-587, December.
    8. Claudio Agostinelli & Luca Greco, 2019. "Weighted likelihood estimation of multivariate location and scatter," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 28(3), pages 756-784, September.
    9. Søren Johansen & Bent Nielsen, 2013. "Outlier Detection in Regression Using an Iterated One-Step Approximation to the Huber-Skip Estimator," Econometrics, MDPI, vol. 1(1), pages 1-18, May.
    10. Marco Riani & Anthony C. Atkinson & Andrea Cerioli, 2009. "Finding an unknown number of multivariate outliers," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 71(2), pages 447-466, April.
    11. Arismendi, Juan C. & Broda, Simon, 2017. "Multivariate elliptical truncated moments," Journal of Multivariate Analysis, Elsevier, vol. 157(C), pages 29-44.
    12. Luca Greco & Giovanni Saraceno & Claudio Agostinelli, 2021. "Robust Fitting of a Wrapped Normal Model to Multivariate Circular Data and Outlier Detection," Stats, MDPI, vol. 4(2), pages 1-18, June.
    13. Peter Filzmoser & Anne Ruiz-Gazen & Christine Thomas-Agnan, 2014. "Identification of local multivariate outliers," Statistical Papers, Springer, vol. 55(1), pages 29-47, February.
    14. Baishuai Zuo & Chuancun Yin, 2022. "Multivariate doubly truncated moments for generalized skew-elliptical distributions with application to multivariate tail conditional risk measures," Papers 2203.00839, arXiv.org.
    15. Meltem Ekiz & O.Ufuk Ekiz, 2017. "Outlier detection with Mahalanobis square distance: incorporating small sample correction factor," Journal of Applied Statistics, Taylor & Francis Journals, vol. 44(13), pages 2444-2457, October.
    16. Torti, Francesca & Corbellini, Aldo & Atkinson, Anthony C., 2021. "fsdaSAS: a package for robust regression for very large datasets including the batch forward search," LSE Research Online Documents on Economics 109895, London School of Economics and Political Science, LSE Library.
    17. Anthony C. Atkinson & Marco Riani & Aldo Corbellini, 2020. "The analysis of transformations for profit‐and‐loss data," Journal of the Royal Statistical Society Series C, Royal Statistical Society, vol. 69(2), pages 251-275, April.
    18. Salvatore Ingrassia & Simona Minotti & Giorgio Vittadini, 2012. "Local Statistical Modeling via a Cluster-Weighted Approach with Elliptical Distributions," Journal of Classification, Springer;The Classification Society, vol. 29(3), pages 363-401, October.
    19. Yunlu Jiang & Canhong Wen & Xueqin Wang, 2018. "Adaptive Exponential Power Depth with Application to Classification," Journal of Classification, Springer;The Classification Society, vol. 35(3), pages 466-480, October.
    20. Domenico Perrotta & Marco Riani & Francesca Torti, 2009. "New robust dynamic plots for regression mixture detection," Advances in Data Analysis and Classification, Springer;German Classification Society - Gesellschaft für Klassifikation (GfKl);Japanese Classification Society (JCS);Classification and Data Analysis Group of the Italian Statistical Society (CLADAG);International Federation of Classification Societies (IFCS), vol. 3(3), pages 263-279, December.

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:jmvana:v:126:y:2014:i:c:p:167-183. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: http://www.elsevier.com/wps/find/journaldescription.cws_home/622892/description#description .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.