The null hypothesis significance test in health sciences research (1995-2006): statistical analysis and interpretation.

Luis Carlos Silva-Ayçaguer,Patricio Suárez-Gil,Ana Fernández-Somoano

doi:10.1186/1471-2288-10-44

Abstract

BackgroundThe null hypothesis significance test (NHST) is the most frequently used statistical method, although its inferential validity has been widely criticized since its introduction. In 1988, the International Committee of Medical Journal Editors (ICMJE) warned against sole reliance on NHST to substantiate study conclusions and suggested supplementary use of confidence intervals (CI). Our objective was to evaluate the extent and quality in the use of NHST and CI, both in English and Spanish language biomedical publications between 1995 and 2006, taking into account the International Committee of Medical Journal Editors recommendations, with particular focus on the accuracy of the interpretation of statistical significance and the validity of conclusions.MethodsOriginal articles published in three English and three Spanish biomedical journals in three fields (General Medicine, Clinical Specialties and Epidemiology - Public Health) were considered for this study. Papers published in 1995-1996, 2000-2001, and 2005-2006 were selected through a systematic sampling method. After excluding the purely descriptive and theoretical articles, analytic studies were evaluated for their use of NHST with P-values and/or CI for interpretation of statistical "significance" and "relevance" in study conclusions.ResultsAmong 1,043 original papers, 874 were selected for detailed review. The exclusive use of P-values was less frequent in English language publications as well as in Public Health journals; overall such use decreased from 41% in 1995-1996 to 21% in 2005-2006. While the use of CI increased over time, the "significance fallacy" (to equate statistical and substantive significance) appeared very often, mainly in journals devoted to clinical specialties (81%). In papers originally written in English and Spanish, 15% and 10%, respectively, mentioned statistical significance in their conclusions.ConclusionsOverall, results of our review show some improvements in statistical management of statistical results, but further efforts by scholars and journal editors are clearly required to move the communication toward ICMJE advices, especially in the clinical setting, which seems to be imperative among publications in Spanish.

Highlights

The null hypothesis significance test (NHST) is the most frequently used statistical method, its inferential validity has been widely criticized since its introduction
Its origins dates back to 1279 [1] it was in the second decade of the twentieth century when the statistician Ronald Fisher formally introduced the concept of "null hypothesis" H0 which, generally speaking, establishes that certain parameters do not differ from each other
The general objective of the present study is to evaluate the extent and quality of use of NHST and confidence intervals (CI), both in English- and in Spanish-language biomedical publications, between 1995 and 2006 taking into account the International Committee of Medical Journal Editors recommendations, with particular focus on accuracy regarding interpretation of statistical significance and the validity of conclusions

Summary

Introduction

The null hypothesis significance test (NHST) is the most frequently used statistical method, its inferential validity has been widely criticized since its introduction. Our objective was to evaluate the extent and quality in the use of NHST and CI, both in English and Spanish language biomedical publications between 1995 and 2006, taking into account the International Committee of Medical Journal Editors recommendations, with particular focus on the accuracy of the interpretation of statistical significance and the validity of conclusions. Its origins dates back to 1279 [1] it was in the second decade of the twentieth century when the statistician Ronald Fisher formally introduced the concept of "null hypothesis" H0 which, generally speaking, establishes that certain parameters do not differ from each other. They established a rule to optimize the decision process, using the p-value introduced by Fisher, by setting the maximum frequency of errors that would be admissible

Objectives

Methods

Results

Discussion

Conclusion

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: BMC medical research methodology	Publication Date: May 19, 2010
Citations: 92	License type: cc-by

R Discovery Prime

R Discovery Prime

The null hypothesis significance test in health sciences research (1995-2006): statistical analysis and interpretation.

Abstract

Highlights

Summary

Talk to us

Similar Papers

More From: BMC medical research methodology

Lead the way for us

Similar Papers

Confidence Intervals Make a Difference
Rink Hoekstra ... Addie Johnson
Educational and Psychological Measurement | VOL. 72
Rink Hoekstra, et. al.Rink Hoekstra ... Addie Johnson
24 Jul 2012
Educational and Psychological Measurement | VOL. 72

Confidence Intervals of a Climatic Signal
Yoshikazu Hayashi
Journal of the Atmospheric Sciences | VOL. 39
Yoshikazu HayashiYoshikazu Hayashi
01 Sep 1982
Journal of the Atmospheric Sciences | VOL. 39

Journals like Acta Paediatrica should encourage authors to provide uncertainty estimates in manuscripts.
Ilari Kuitunen
Acta paediatrica (Oslo, Norway : 1992) | VOL. 112
Ilari KuitunenIlari Kuitunen
12 Apr 2023
Acta paediatrica (Oslo, Norway : 1992) | VOL. 112

In support of null hypothesis significance testing.
Michael Mogie
Proceedings. Biological sciences | VOL. Suppl 271 3
Michael MogieMichael Mogie
07 Feb 2004
Proceedings. Biological sciences | VOL. Suppl 271 3

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

The null hypothesis significance test in health sciences research (1995-2006): statistical analysis and interpretation.

Abstract

Highlights

Summary

Talk to us

Similar Papers

More From: BMC medical research methodology