IDEAS home Printed from https://ideas.repec.org/a/spr/compst/v39y2024i7d10.1007_s00180-024-01463-8.html
   My bibliography  Save this article

Some new invariant sum tests and MAD tests for the assessment of Benford’s law

Author

Listed:
  • Wolfgang Kössler

    (Humboldt Universität zu Berlin)

  • Hans-J. Lenz

    (Freie Universität Berlin)

  • Xing D. Wang

    (Humboldt Universität zu Berlin)

Abstract

The Benford law is used world-wide for detecting non-conformance or data fraud of numerical data. It says that the significand of a data set from the universe is not uniformly, but logarithmically distributed. Especially, the first non-zero digit is One with an approximate probability of 0.3. There are several tests available for testing Benford, the best known are Pearson’s $$\chi ^2$$ χ 2 -test, the Kolmogorov–Smirnov test and a modified version of the MAD-test. In the present paper we propose some tests, three of the four invariant sum tests are new and they are motivated by the sum invariance property of the Benford law. Two distance measures are investigated, Euclidean and Mahalanobis distance of the standardized sums to the orign. We use the significands corresponding to the first significant digit as well as the second significant digit, respectively. Moreover, we suggest inproved versions of the MAD-test and obtain critical values that are independent of the sample sizes. For illustration the tests are applied to specifically selected data sets where prior knowledge is available about being or not being Benford. Furthermore we discuss the role of truncation of distributions.

Suggested Citation

  • Wolfgang Kössler & Hans-J. Lenz & Xing D. Wang, 2024. "Some new invariant sum tests and MAD tests for the assessment of Benford’s law," Computational Statistics, Springer, vol. 39(7), pages 3779-3800, December.
  • Handle: RePEc:spr:compst:v:39:y:2024:i:7:d:10.1007_s00180-024-01463-8
    DOI: 10.1007/s00180-024-01463-8
    as

    Download full text from publisher

    File URL: http://link.springer.com/10.1007/s00180-024-01463-8
    File Function: Abstract
    Download Restriction: Access to the full text of the articles in this series is restricted.

    File URL: https://libkey.io/10.1007/s00180-024-01463-8?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    As the access to this document is restricted, you may want to search for a different version of it.

    References listed on IDEAS

    as
    1. Roy Cerqueti & Claudio Lupi, 2021. "Some New Tests of Conformity with Benford’s Law," Stats, MDPI, vol. 4(3), pages 1-17, September.
    2. Andreas Diekmann, 2007. "Not the First Digit! Using Benford's Law to Detect Fraudulent Scientif ic Data," Journal of Applied Statistics, Taylor & Francis Journals, vol. 34(3), pages 321-329.
    3. Liu, Huan & Tang, Yongqiang & Zhang, Hao Helen, 2009. "A new chi-square approximation to the distribution of non-negative definite quadratic forms in non-central normal variables," Computational Statistics & Data Analysis, Elsevier, vol. 53(4), pages 853-856, February.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Roy Cerqueti & Claudio Lupi, 2023. "Severe testing of Benford’s law," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 32(2), pages 677-694, June.
    2. Zimmer, Zachary & Park, DoHwan & Mathew, Thomas, 2016. "Tolerance limits under normal mixtures: Application to the evaluation of nuclear power plant safety and to the assessment of circular error probable," Computational Statistics & Data Analysis, Elsevier, vol. 103(C), pages 304-315.
    3. Sitsofe Tsagbey & Miguel de Carvalho & Garritt L. Page, 2017. "All Data are Wrong, but Some are Useful? Advocating the Need for Data Auditing," The American Statistician, Taylor & Francis Journals, vol. 71(3), pages 231-235, July.
    4. Philip E Hulme & Danish A Ahmed & Phillip J Haubrock & Brooks A Kaiser & Melina Kourantidou & Boris Leroy & Shana M Mcdermott, 2024. "Widespread imprecision in estimates of the economic costs of invasive alien species worldwide," Post-Print hal-04633043, HAL.
    5. Sanae Rujivan & Athinan Sutchada & Kittisak Chumpong & Napat Rujeerapaiboon, 2023. "Analytically Computing the Moments of a Conic Combination of Independent Noncentral Chi-Square Random Variables and Its Application for the Extended Cox–Ingersoll–Ross Process with Time-Varying Dimens," Mathematics, MDPI, vol. 11(5), pages 1-29, March.
    6. Theoharry Grammatikos & Nikolaos I. Papanikolaou, 2021. "Applying Benford’s Law to Detect Accounting Data Manipulation in the Banking Industry," Journal of Financial Services Research, Springer;Western Finance Association, vol. 59(1), pages 115-142, April.
    7. Marcel Ausloos & Probowo Erawan Sastroredjo & Polina Khrennikova, 2025. "Note on Pre-Taxation Data Reported by UK FTSE-Listed Companies: Search for Compatibility with Benford’s Laws," Stats, MDPI, vol. 8(1), pages 1-17, February.
    8. Roeland de Kok & Giulia Rotundo, 2022. "Benford Networks," Stats, MDPI, vol. 5(4), pages 1-14, September.
    9. Xuemei Hu & Xiaohui Liu, 2013. "Empirical likelihood confidence regions for semi-varying coefficient models with linear process errors," Journal of Nonparametric Statistics, Taylor & Francis Journals, vol. 25(1), pages 161-180, March.
    10. Yamaguchi, Hikaru & Murakami, Hidetoshi, 2023. "The multi-aspect tests in the presence of ties," Computational Statistics & Data Analysis, Elsevier, vol. 180(C).
    11. Qianchuan He & Yang Liu & Ulrike Peters & Li Hsu, 2018. "Multivariate association analysis with somatic mutation data," Biometrics, The International Biometric Society, vol. 74(1), pages 176-184, March.
    12. Andreas Diekmann & Ben Jann, 2010. "Benford's Law and Fraud Detection: Facts and Legends," German Economic Review, Verein für Socialpolitik, vol. 11(3), pages 397-401, August.
    13. Chase Thiel & Zhanna Bagdasarov & Lauren Harkrider & James Johnson & Michael Mumford, 2012. "Leader Ethical Decision-Making in Organizations: Strategies for Sensemaking," Journal of Business Ethics, Springer, vol. 107(1), pages 49-64, April.
    14. Holz, Carsten A., 2014. "The quality of China's GDP statistics," China Economic Review, Elsevier, vol. 30(C), pages 309-338.
    15. Andreas Diekmann, 2012. "Making Use of “Benford’s Law†for the Randomized Response Technique," Sociological Methods & Research, , vol. 41(2), pages 325-334, May.
    16. Songhua Tan & Qianqian Zhu, 2022. "Asymmetric linear double autoregression," Journal of Time Series Analysis, Wiley Blackwell, vol. 43(3), pages 371-388, May.
    17. Montag, Josef, 2017. "Identifying odometer fraud in used car market data," Transport Policy, Elsevier, vol. 60(C), pages 10-23.
    18. Călin Vâlsan & Andreea-Ionela Puiu & Elena Druică, 2024. "From Whence Commeth Data Misreporting? A Survey of Benford’s Law and Digit Analysis in the Time of the COVID-19 Pandemic," Mathematics, MDPI, vol. 12(16), pages 1-20, August.
    19. Adriano Silva & Sergio Floquet & Ricardo Lima, 2023. "Newcomb–Benford’s Law in Neuromuscular Transmission: Validation in Hyperkalemic Conditions," Stats, MDPI, vol. 6(4), pages 1-19, October.
    20. Koch, Christoffer & Okamura, Ken, 2020. "Benford’s Law and COVID-19 reporting," Economics Letters, Elsevier, vol. 196(C).

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:spr:compst:v:39:y:2024:i:7:d:10.1007_s00180-024-01463-8. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Sonal Shukla or Springer Nature Abstracting and Indexing (email available below). General contact details of provider: http://www.springer.com .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.