Skip to main page content
U.S. flag

An official website of the United States government

Dot gov

The .gov means it’s official.
Federal government websites often end in .gov or .mil. Before sharing sensitive information, make sure you’re on a federal government site.

Https

The site is secure.
The https:// ensures that you are connecting to the official website and that any information you provide is encrypted and transmitted securely.

Access keys NCBI Homepage MyNCBI Homepage Main Content Main Navigation
. 2010 May 19:10:44.
doi: 10.1186/1471-2288-10-44.

The null hypothesis significance test in health sciences research (1995-2006): statistical analysis and interpretation

Affiliations

The null hypothesis significance test in health sciences research (1995-2006): statistical analysis and interpretation

Luis Carlos Silva-Ayçaguer et al. BMC Med Res Methodol. .

Abstract

Background: The null hypothesis significance test (NHST) is the most frequently used statistical method, although its inferential validity has been widely criticized since its introduction. In 1988, the International Committee of Medical Journal Editors (ICMJE) warned against sole reliance on NHST to substantiate study conclusions and suggested supplementary use of confidence intervals (CI). Our objective was to evaluate the extent and quality in the use of NHST and CI, both in English and Spanish language biomedical publications between 1995 and 2006, taking into account the International Committee of Medical Journal Editors recommendations, with particular focus on the accuracy of the interpretation of statistical significance and the validity of conclusions.

Methods: Original articles published in three English and three Spanish biomedical journals in three fields (General Medicine, Clinical Specialties and Epidemiology - Public Health) were considered for this study. Papers published in 1995-1996, 2000-2001, and 2005-2006 were selected through a systematic sampling method. After excluding the purely descriptive and theoretical articles, analytic studies were evaluated for their use of NHST with P-values and/or CI for interpretation of statistical "significance" and "relevance" in study conclusions.

Results: Among 1,043 original papers, 874 were selected for detailed review. The exclusive use of P-values was less frequent in English language publications as well as in Public Health journals; overall such use decreased from 41% in 1995-1996 to 21% in 2005-2006. While the use of CI increased over time, the "significance fallacy" (to equate statistical and substantive significance) appeared very often, mainly in journals devoted to clinical specialties (81%). In papers originally written in English and Spanish, 15% and 10%, respectively, mentioned statistical significance in their conclusions.

Conclusions: Overall, results of our review show some improvements in statistical management of statistical results, but further efforts by scholars and journal editors are clearly required to move the communication toward ICMJE advices, especially in the clinical setting, which seems to be imperative among publications in Spanish.

PubMed Disclaimer

Figures

Figure 1
Figure 1
Flow chart of the selection process for eligible papers.

Similar articles

Cited by

References

    1. Curran-Everett D. Explorations in statistics: hypothesis tests and P values. Adv Physiol Educ. 2009;33:81–86. doi: 10.1152/advan.90218.2008. - DOI - PubMed
    1. Fisher RA. Statistical Methods for Research Workers. Edinburgh: Oliver & Boyd; 1925.
    1. Neyman J, Pearson E. On the use and interpretation of certain test criteria for purposes of statistical inference. Biometrika. 1928;20:175–240.
    1. Silva LC. Los laberintos de la investigación biomédica. En defensa de la racionalidad para la ciencia del siglo XXI. Madrid: Díaz de Santos; 2009.
    1. Berkson J. Test of significance considered as evidence. J Am Stat Assoc. 1942;37:325–335. doi: 10.2307/2279000. - DOI

MeSH terms

LinkOut - more resources