IDEAS home Printed from https://ideas.repec.org/a/spr/compst/v39y2024i7d10.1007_s00180-024-01463-8.html
   My bibliography  Save this article

Some new invariant sum tests and MAD tests for the assessment of Benford’s law

Author

Listed:
  • Wolfgang Kössler

    (Humboldt Universität zu Berlin)

  • Hans-J. Lenz

    (Freie Universität Berlin)

  • Xing D. Wang

    (Humboldt Universität zu Berlin)

Abstract

The Benford law is used world-wide for detecting non-conformance or data fraud of numerical data. It says that the significand of a data set from the universe is not uniformly, but logarithmically distributed. Especially, the first non-zero digit is One with an approximate probability of 0.3. There are several tests available for testing Benford, the best known are Pearson’s $$\chi ^2$$ χ 2 -test, the Kolmogorov–Smirnov test and a modified version of the MAD-test. In the present paper we propose some tests, three of the four invariant sum tests are new and they are motivated by the sum invariance property of the Benford law. Two distance measures are investigated, Euclidean and Mahalanobis distance of the standardized sums to the orign. We use the significands corresponding to the first significant digit as well as the second significant digit, respectively. Moreover, we suggest inproved versions of the MAD-test and obtain critical values that are independent of the sample sizes. For illustration the tests are applied to specifically selected data sets where prior knowledge is available about being or not being Benford. Furthermore we discuss the role of truncation of distributions.

Suggested Citation

  • Wolfgang Kössler & Hans-J. Lenz & Xing D. Wang, 2024. "Some new invariant sum tests and MAD tests for the assessment of Benford’s law," Computational Statistics, Springer, vol. 39(7), pages 3779-3800, December.
  • Handle: RePEc:spr:compst:v:39:y:2024:i:7:d:10.1007_s00180-024-01463-8
    DOI: 10.1007/s00180-024-01463-8
    as

    Download full text from publisher

    File URL: http://link.springer.com/10.1007/s00180-024-01463-8
    File Function: Abstract
    Download Restriction: Access to the full text of the articles in this series is restricted.

    File URL: https://libkey.io/10.1007/s00180-024-01463-8?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    As the access to this document is restricted, you may want to search for a different version of it.

    References listed on IDEAS

    as
    1. Roy Cerqueti & Claudio Lupi, 2021. "Some New Tests of Conformity with Benford’s Law," Stats, MDPI, vol. 4(3), pages 1-17, September.
    2. Andreas Diekmann, 2007. "Not the First Digit! Using Benford's Law to Detect Fraudulent Scientif ic Data," Journal of Applied Statistics, Taylor & Francis Journals, vol. 34(3), pages 321-329.
    3. Liu, Huan & Tang, Yongqiang & Zhang, Hao Helen, 2009. "A new chi-square approximation to the distribution of non-negative definite quadratic forms in non-central normal variables," Computational Statistics & Data Analysis, Elsevier, vol. 53(4), pages 853-856, February.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Andreas Quatember, 2019. "A discussion of the two different aspects of privacy protection in indirect questioning designs," Quality & Quantity: International Journal of Methodology, Springer, vol. 53(1), pages 269-282, January.
    2. Roy Cerqueti & Claudio Lupi, 2023. "Severe testing of Benford’s law," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 32(2), pages 677-694, June.
    3. Dlugosz, Stephan & Müller-Funk, Ulrich, 2012. "Ziffernanalyse zur Betrugserkennung in Finanzverwaltungen: Prüfung von Kassenbelegen," Arbeitsberichte des Instituts für Wirtschaftsinformatik 133, University of Münster, Department of Information Systems.
    4. Zimmer, Zachary & Park, DoHwan & Mathew, Thomas, 2016. "Tolerance limits under normal mixtures: Application to the evaluation of nuclear power plant safety and to the assessment of circular error probable," Computational Statistics & Data Analysis, Elsevier, vol. 103(C), pages 304-315.
    5. Aineas Mallios & Taylan Mavruk, 2025. "Do ESG funds engage in portfolio pumping to gain higher flows? An application of Benford's Law," International Journal of Finance & Economics, John Wiley & Sons, Ltd., vol. 30(2), pages 1540-1563, April.
    6. Diekmann Andreas & Jann Ben, 2010. "Benford’s Law and Fraud Detection: Facts and Legends," German Economic Review, De Gruyter, vol. 11(3), pages 397-401, August.
    7. Bernhard Rauch & Max G�ttsche & Stephan Langenegger, 2014. "Detecting Problems in Military Expenditure Data Using Digital Analysis," Defence and Peace Economics, Taylor & Francis Journals, vol. 25(2), pages 97-111, April.
    8. Lee, Kang-Bok & Han, Sumin & Jeong, Yeasung, 2020. "COVID-19, flattening the curve, and Benford’s law," Physica A: Statistical Mechanics and its Applications, Elsevier, vol. 559(C).
    9. Debreceny, Roger S. & Gray, Glen L., 2010. "Data mining journal entries for fraud detection: An exploratory study," International Journal of Accounting Information Systems, Elsevier, vol. 11(3), pages 157-181.
    10. de Araújo Silva, Archibald & Aparecida Gouvêa, Maria, 2023. "Study on the effect of sample size on type I error, in the first, second and first-two digits excessmad tests," International Journal of Accounting Information Systems, Elsevier, vol. 48(C).
    11. Lin, Fengyi & Wu, Sheng-Fu, 2014. "Comparison of cosmetic earnings management for the developed markets and emerging markets: Some empirical evidence from the United States and Taiwan," Economic Modelling, Elsevier, vol. 36(C), pages 466-473.
    12. Brähler, Gernot & Bensmann, Markus & Emke, Anna-Lena, 2010. "Der Einsatz mathematisch-statistischer Methoden in der digitalen Betriebsprüfung," Ilmenauer Schriften zur Betriebswirtschaftslehre, Technische Universität Ilmenau, Institut für Betriebswirtschaftslehre, volume 4, number 42010, January.
    13. Sitsofe Tsagbey & Miguel de Carvalho & Garritt L. Page, 2017. "All Data are Wrong, but Some are Useful? Advocating the Need for Data Auditing," The American Statistician, Taylor & Francis Journals, vol. 71(3), pages 231-235, July.
    14. Philip E Hulme & Danish A Ahmed & Phillip J Haubrock & Brooks A Kaiser & Melina Kourantidou & Boris Leroy & Shana M Mcdermott, 2024. "Widespread imprecision in estimates of the economic costs of invasive alien species worldwide," Post-Print hal-04633043, HAL.
    15. Hui-Guo Zhang & Chang-Lin Mei, 2017. "Discussion," International Statistical Review, International Statistical Institute, vol. 85(1), pages 38-40, April.
    16. Sanae Rujivan & Athinan Sutchada & Kittisak Chumpong & Napat Rujeerapaiboon, 2023. "Analytically Computing the Moments of a Conic Combination of Independent Noncentral Chi-Square Random Variables and Its Application for the Extended Cox–Ingersoll–Ross Process with Time-Varying Dimens," Mathematics, MDPI, vol. 11(5), pages 1-29, March.
    17. Walter R. Schumm & Duane W. Crawford & Lorenza Lockett & Asma bin Ateeq & Abdullah AlRashed, 2023. "Can Retracted Social Science Articles Be Distinguished from Non-Retracted Articles by Some of the Same Authors, Using Benford’s Law or Other Statistical Methods?," Publications, MDPI, vol. 11(1), pages 1-13, March.
    18. Lasse Pröger & Paul Griesberger & Klaus Hackländer & Norbert Brunner & Manfred Kühleitner, 2021. "Benford’s Law for Telemetry Data of Wildlife," Stats, MDPI, vol. 4(4), pages 1-7, November.
    19. Bruno S. Frey, 2010. "Withering academia?," IEW - Working Papers 512, Institute for Empirical Research in Economics - University of Zurich.
    20. Teddy Lazebnik & Dan Gorlitsky, 2023. "Can We Mathematically Spot the Possible Manipulation of Results in Research Manuscripts Using Benford’s Law?," Data, MDPI, vol. 8(11), pages 1-11, October.

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:spr:compst:v:39:y:2024:i:7:d:10.1007_s00180-024-01463-8. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Sonal Shukla or Springer Nature Abstracting and Indexing (email available below). General contact details of provider: http://www.springer.com .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.