Why and how we should join the shift from significance testing to estimation. (18th May 2022)
- Record Type:
- Journal Article
- Title:
- Why and how we should join the shift from significance testing to estimation. (18th May 2022)
- Main Title:
- Why and how we should join the shift from significance testing to estimation
- Authors:
- Berner, Daniel
Amrhein, Valentin - Abstract:
- Abstract: A paradigm shift away from null hypothesis significance testing seems in progress. Based on simulations, we illustrate some of the underlying motivations. First, p ‐values vary strongly from study to study, hence dichotomous inference using significance thresholds is usually unjustified. Second, 'statistically significant' results have overestimated effect sizes, a bias declining with increasing statistical power. Third, 'statistically non‐significant' results have underestimated effect sizes, and this bias gets stronger with higher statistical power. Fourth, the tested statistical hypotheses usually lack biological justification and are often uninformative. Despite these problems, a screen of 48 papers from the 2020 volume of the Journal of Evolutionary Biology exemplifies that significance testing is still used almost universally in evolutionary biology. All screened studies tested default null hypotheses of zero effect with the default significance threshold of p = 0.05, none presented a pre‐specified alternative hypothesis, pre‐study power calculation and the probability of 'false negatives' (beta error rate). The results sections of the papers presented 49 significance tests on average (median 23, range 0–390). Of 41 studies that contained verbal descriptions of a 'statistically non‐significant' result, 26 (63%) falsely claimed the absence of an effect. We conclude that studies in ecology and evolutionary biology are mostly exploratory and descriptive. WeAbstract: A paradigm shift away from null hypothesis significance testing seems in progress. Based on simulations, we illustrate some of the underlying motivations. First, p ‐values vary strongly from study to study, hence dichotomous inference using significance thresholds is usually unjustified. Second, 'statistically significant' results have overestimated effect sizes, a bias declining with increasing statistical power. Third, 'statistically non‐significant' results have underestimated effect sizes, and this bias gets stronger with higher statistical power. Fourth, the tested statistical hypotheses usually lack biological justification and are often uninformative. Despite these problems, a screen of 48 papers from the 2020 volume of the Journal of Evolutionary Biology exemplifies that significance testing is still used almost universally in evolutionary biology. All screened studies tested default null hypotheses of zero effect with the default significance threshold of p = 0.05, none presented a pre‐specified alternative hypothesis, pre‐study power calculation and the probability of 'false negatives' (beta error rate). The results sections of the papers presented 49 significance tests on average (median 23, range 0–390). Of 41 studies that contained verbal descriptions of a 'statistically non‐significant' result, 26 (63%) falsely claimed the absence of an effect. We conclude that studies in ecology and evolutionary biology are mostly exploratory and descriptive. We should thus shift from claiming to 'test' specific hypotheses statistically to describing and discussing many hypotheses (possible true effect sizes) that are most compatible with our data, given our statistical model. We already have the means for doing so, because we routinely present compatibility ('confidence') intervals covering these hypotheses. Abstract : Inference in ecology and evolution is still mostly based on significance testing. We summarize problems with this approach and argue that it should be replaced by the adequate description of effect size estimates. … (more)
- Is Part Of:
- Journal of evolutionary biology. Volume 35:Number 6(2022)
- Journal:
- Journal of evolutionary biology
- Issue:
- Volume 35:Number 6(2022)
- Issue Display:
- Volume 35, Issue 6 (2022)
- Year:
- 2022
- Volume:
- 35
- Issue:
- 6
- Issue Sort Value:
- 2022-0035-0006-0000
- Page Start:
- 777
- Page End:
- 787
- Publication Date:
- 2022-05-18
- Subjects:
- compatibility interval -- effect size -- null hypothesis -- p‐value -- scientific method -- statistical inference
Evolution (Biology) -- Periodicals
Biology -- Periodicals
576.8 - Journal URLs:
- http://onlinelibrary.wiley.com/journal/10.1111/(ISSN)1420-9101 ↗
http://www.blackwell-synergy.com/member/institutions/issuelist.asp?journal=jeb ↗
http://onlinelibrary.wiley.com/ ↗
http://firstsearch.oclc.org ↗
http://firstsearch.oclc.org/journal=1010-061x;screen=info;ECOIP ↗ - DOI:
- 10.1111/jeb.14009 ↗
- Languages:
- English
- ISSNs:
- 1010-061X
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - 4979.642100
British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 21807.xml