<?xml version="1.0" encoding="UTF-8"?>
<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Publishing DTD v1.1d1 20130915//EN" "http://jats.nlm.nih.gov/publishing/1.1d1/JATS-journalpublishing1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink" xmlns:mml="http://www.w3.org/1998/Math/MathML" article-type="research-article" xml:lang="en">
<front>
<journal-meta>
<journal-id journal-id-type="publisher-id">SAJIP</journal-id>
<journal-title-group>
<journal-title>SA Journal of Industrial Psychology</journal-title>
</journal-title-group>
<issn pub-type="ppub">0258-5200</issn>
<issn pub-type="epub">2071-0763</issn>
<publisher>
<publisher-name>AOSIS</publisher-name>
</publisher>
</journal-meta>
<article-meta>
<article-id pub-id-type="publisher-id">SAJIP-45-1717</article-id>
<article-id pub-id-type="doi">10.4102/sajip.v45i0.1717</article-id>
<article-categories>
<subj-group subj-group-type="heading">
<subject>Rebuttal</subject>
</subj-group>
</article-categories>
<title-group>
<article-title>Reducing our dependence on null hypothesis testing: A key to enhance the reproducibility and credibility of our science</article-title>
</title-group>
<contrib-group>
<contrib contrib-type="author" corresp="yes">
<contrib-id contrib-id-type="orcid">https://orcid.org/0000-0002-4886-9499</contrib-id>
<name>
<surname>Murphy</surname>
<given-names>Kevin R.</given-names>
</name>
<xref ref-type="aff" rid="AF0001">1</xref>
</contrib>
<aff id="AF0001"><label>1</label>Department of Work and Employment Studies, Kemmy Business School, University of Limerick, Limerick, Ireland</aff>
</contrib-group>
<author-notes>
<corresp id="cor1"><bold>Corresponding author:</bold> Kevin Murphy, <email xlink:href="Kevin.R.Murphy@ul.ie">Kevin.R.Murphy@ul.ie</email></corresp>
</author-notes>
<pub-date pub-type="epub"><day>05</day><month>11</month><year>2019</year></pub-date>
<pub-date pub-type="collection"><year>2019</year></pub-date>
<volume>45</volume>
<issue>0</issue>
<elocation-id>1717</elocation-id>
<history>
<date date-type="received"><day>16</day><month>07</month><year>2019</year></date>
<date date-type="accepted"><day>27</day><month>08</month><year>2019</year></date>
</history>
<permissions>
<copyright-statement>&#x00A9; 2019. The Authors</copyright-statement>
<copyright-year>2019</copyright-year>
<license license-type="open-access" xlink:href="https://creativecommons.org/licenses/by/4.0/">
<license-p>Licensee: AOSIS. This work is licensed under the Creative Commons Attribution License.</license-p>
</license>
</permissions>
<abstract>
<sec id="st1">
<title>Problemification</title>
<p>Over-reliance on null hypothesis significance testing (NHST) is one of the most important causes of the emerging crisis over the credibility and reproducibility of our science.</p>
</sec>
<sec id="st2">
<title>Implications</title>
<p>Most studies in the behavioural and social sciences have low levels of statistical power. Because &#x2018;significant&#x2019; results are often required, but often difficult to produce, the temptation to engage in questionable research practices that will produce these results is immense.</p>
</sec>
<sec id="st3">
<title>Purpose</title>
<p>Methodologists have been trying for decades to convince researchers, reviewers and editors that significance tests are neither informative nor useful. A recent set of articles published in top journals and endorsed by hundreds of scientists around the world seem to provide a fresh impetus for overturning the practice of using NHST as the primary, and sometimes sole basis for evaluating research results.</p>
</sec>
<sec id="st4">
<title>Recommendations</title>
<p>Authors, reviewers and journal editors are asked to change long-engrained habits and realise that &#x2018;statistically significant&#x2019; says more about the design of one&#x2019;s study than about the importance of one&#x2019;s results. They are urged to embrace the ATOM principle in evaluating research results, that is, <italic>accept</italic> that there will always be uncertainty, and be <italic>thoughtful, open</italic> and <italic>modest</italic> in evaluating what the data mean.</p>
</sec>
</abstract>
<kwd-group>
<kwd>Significance Testing</kwd>
<kwd>Confidence Intervals</kwd>
<kwd>Questionable Research Practices</kwd>
<kwd>Null Hypothesis</kwd>
</kwd-group>
</article-meta>
</front>
<body>
<sec id="s0001">
<title>Introduction</title>
<p>There are many indications that several sciences, including psychology, have a reproducibility crisis in their hands (Ioannidis, <xref ref-type="bibr" rid="CIT0009">2005</xref>; McNutt, <xref ref-type="bibr" rid="CIT0010">2014</xref>; Pashler &#x0026; Wagenmakers, <xref ref-type="bibr" rid="CIT0015">2012</xref>). Peer-reviewed research that is published in highly reputable journals has often failed to replicate; papers that attempt to replicate published research very often report smaller and non-significant effects (Open Science Collaboration, <xref ref-type="bibr" rid="CIT0014">2015</xref>). This persistent failure to replicate published findings calls the credibility and meaning of those findings and, by extension, of other published research into question. Efendic and Van Zyl (<xref ref-type="bibr" rid="CIT0007">2019</xref>) provide an excellent summary of the challenges this crisis poses to Industrial and Organizational Psychology, and they outline several thoughtful responses this journal might make to increase the robustness and credibility of the research published in the <italic>South African Journal of Industrial Psychology.</italic> The changes they propose will not be easy to implement, in part because they require authors, reviewers and editors to change their perspectives and their behaviours. However, there are reasons to be optimistic about one of the major changes the authors propose.</p>
<p>Many of the problems with reproducibility can be traced to our field&#x2019;s long reliance on Null Hypothesis Significance Testing (NHST). Efendic and Van Zyl (<xref ref-type="bibr" rid="CIT0007">2019</xref>) documented two of the most problematic aspects of the use of statistical tests of the null hypothesis in making decisions about study results. These were (1) the inadequate power of most studies and (2) the strong temptation to engage in a range of questionable research practices (ranging from <italic>p</italic> fishing [i.e. trying multiple statistical tests to find one in which <italic>p</italic> &#x003C; 0.05] to outright fabrication [Banks, Rogelberg, Woznyj, Landis, &#x0026; Rupp, <xref ref-type="bibr" rid="CIT0003">2016</xref>; Neuroskeptic, <xref ref-type="bibr" rid="CIT0013">2012</xref>]) in search of a &#x2018;significant&#x2019; (<italic>p</italic> &#x003C; 0.05) result. While many of the changes proposed by Efendic and Van Zyl (<xref ref-type="bibr" rid="CIT0007">2019</xref>) could help to address some of the problems caused by an over-reliance on NHST, I do not think they go far enough. As long as the pursuit of <italic>p</italic> &#x003C; 0.05 remains central to our evaluation of research results, I do not think we will make a meaningful dent in addressing the reproducibility crisis. Fortunately, there are reasons to believe that the long reign of NHST is coming to an end.</p>
</sec>
<sec id="s0002">
<title>The slow demise of null hypothesis significance testing</title>
<p>For decades, methodologists (e.g. Cohen, <xref ref-type="bibr" rid="CIT0004">1962</xref>, <xref ref-type="bibr" rid="CIT0005">1988</xref>, <xref ref-type="bibr" rid="CIT0006">1994</xref>; Fidler, Thomason, Cumming, Finch, &#x0026; Leeman, <xref ref-type="bibr" rid="CIT0008">2004</xref>; Meehl, <xref ref-type="bibr" rid="CIT0011">1978</xref>; Murphy, Myors, &#x0026; Wolach, <xref ref-type="bibr" rid="CIT0012">2014</xref>; Schmidt, <xref ref-type="bibr" rid="CIT0017">1996</xref>; Sterling, <xref ref-type="bibr" rid="CIT0020">1959</xref>) have been warning the research community about the perils of relying on NHST as a method for evaluating research results. Despite decades of criticism, null hypothesis tests are still widely used for evaluating the results of research. For example, Bakker, Van Dijk and Wicherst (<xref ref-type="bibr" rid="CIT0002">2012</xref>) cited several reviews that together suggest that over 95&#x0025; of papers in psychology use null hypothesis testing as one criterion for evaluating results. Authors who find that their results are not significant may decide to abandon that hypothesis or to not submit their work (self-censorship); reviewers and editors who see that the principal results of a study are not statistically significant may decide not to accept that study for publication. Several studies (e.g. Bakker et al., <xref ref-type="bibr" rid="CIT0002">2012</xref>; Ioannidis, <xref ref-type="bibr" rid="CIT0009">2005</xref>) have modelled the effects of reliance on null hypothesis testing to show how it biases the published literature and how it contributes to the reproducibility crisis in the social sciences.</p>
<p>There are many critiques of NHST, but in my view, two are paramount. Firstly, NHST tests a hypothesis that very few people believe to be credible &#x2013; the hypothesis that treatments have <italic>no</italic> effect whatsoever or that variables are <italic>completely</italic> uncorrelated (Cohen, <xref ref-type="bibr" rid="CIT0006">1994</xref>). It is indeed plausible that treatments might have a very small effect (perhaps so small that they can safely be ignored) or that the correlations between two variables might be quite small, but the hypothesis that they are <italic>precisely</italic> zero is not a credible one (Murphy et al., <xref ref-type="bibr" rid="CIT0012">2014</xref>). The null hypothesis is a point hypothesis that divides the set of possible outcomes into two zones, that is, either that a specific hypothesis about the value of some statistic is true (i.e. &#x03C1; is precisely equal to zero) or the entire range of alternatives (&#x03C1; is precisely equal to any value other than zero, to the millionth decimal place) contains the truth. It is well known that the likelihood of <italic>any</italic> point hypothesis being precisely true is exceedingly small, and that it does not matter whether a value of zero or some other specific value for a statistic is assumed. Because a point hypothesis is infinitely precise and the range of alternatives to a point hypothesis is infinitely large, the likelihood that any point hypothesis &#x2013; including the classic null that the effect is precisely zero &#x2013; is true will tend to approach zero. This undermines the whole architecture of NHST. For example, if H<sub>0</sub> is never, or essentially never true, it is impossible to make a type I error and all of the statistical tools designed to minimise these errors (e.g. stringent alpha levels, Bonferroni corrections) become meaningless. Several alternate approaches have been developed to test the more credible hypothesis, such as the hypothesis that the effects of treatments fall within a range of values that are all trivially small (Murphy et al., <xref ref-type="bibr" rid="CIT0012">2014</xref>; Rouanet, <xref ref-type="bibr" rid="CIT0016">1996</xref>; Serlin &#x0026; Lapsley, <xref ref-type="bibr" rid="CIT0018">1985</xref>, <xref ref-type="bibr" rid="CIT0019">1993</xref>), but their uptake has been limited.</p>
<p>Secondly, the outcomes of NHST are routinely misinterpreted. If you fail to reject H<sub>0</sub>, you are likely to conclude that your treatments did not work or that the variables you are interested in are not related to one another. That is, you are very likely to interpret NHST as telling you something about your results. This is wrong (Murphy et al., <xref ref-type="bibr" rid="CIT0012">2014</xref>). The failure to reject H<sub>0</sub> tells you something about the design of your study, in particular, that you did not build a study with sufficient statistical power. One of the lessons learnt by scanning power tables is that given a sufficiently large sample, you can reject virtually <italic>any</italic> null hypothesis, no matter what treatments or variables you are studying (Cohen, <xref ref-type="bibr" rid="CIT0005">1988</xref>; Murphy et al., <xref ref-type="bibr" rid="CIT0012">2014</xref>). Conversely, if your sample is small enough, you can be virtually certain that you will not reject H<sub>0</sub>, regardless of the treatments or variables being studied. Null hypothesis significance testing is essentially an assessment of whether or not your study was powerful enough to detect whatever effect you are studying, and it is very little else.</p>
<p>Recent developments in the scientific literature have finally given some reason for optimism that our over-reliance on NHST is coming to an end. A series of recent articles (Amrhein, Greenland, &#x0026; McShane, <xref ref-type="bibr" rid="CIT0001">2019</xref>; Wasserstein &#x0026; Lazar, <xref ref-type="bibr" rid="CIT0021">2016</xref>; Wasserstein, Schirm, &#x0026; Lazar, <xref ref-type="bibr" rid="CIT0022">2019</xref>) in high-profile journals (e.g. <italic>Nature</italic> and <italic>American Statistician</italic>) have accomplished three things that decades of research papers and chapters of previous critics of NHST had not been able to accomplish. Firstly, they have described in clear and largely non-technical language the deficiencies of NHST as a method for making decisions about the meaning of results. Secondly, they have gathered the support of hundreds of signatories (there were over 800 signatories from over 50 countries within a week of the distribution of Amrhein et al.&#x2019;s <xref ref-type="bibr" rid="CIT0001">2019</xref> draft) to statements calling for an end to mechanical reliance on significance testing for evaluating findings. Thirdly, and most importantly, they (particularly Wasserstein et al., <xref ref-type="bibr" rid="CIT0022">2019</xref>) have presented constructive alternatives. I believe this recent wave of papers presents an opportunity for researchers in virtually all disciplines to improve the methods they apply to make sense of data and to advance the cause of reproducible science.</p>
</sec>
<sec id="s0003">
<title>A call for action</title>
<p>I therefore urge the <italic>South African Journal of Industrial Psychology</italic> to adopt the core principles articulated by Wasserstein et al. (<xref ref-type="bibr" rid="CIT0022">2019</xref>). Firstly, it is not sufficient to simply say &#x2018;don&#x2019;t use significance testing&#x2019;; it is critical to help researchers and readers to understand and empower them to utilise the available alternatives. For example, Serlin and Lapsley (<xref ref-type="bibr" rid="CIT0018">1985</xref>, <xref ref-type="bibr" rid="CIT0019">1993</xref>) developed methods for assessing whether or not the hypothesis of no effect was a good enough approximation of reality to serve as a working description of one&#x2019;s result. Rather than testing whether the effect was <italic>precisely</italic> zero, their methods allowed one to evaluate the possibility that the effects were close enough to zero to be treated as such. Building on these concepts, Murphy et al. (<xref ref-type="bibr" rid="CIT0012">2014</xref>) showed how the entire body of methods subsumed under the general linear model (e.g. correlation, <italic>t</italic>-tests, Analysis of Variance (ANOVA), Analysis of Covariance (ANCOVA) and Multiple Regression) could be adapted to test the hypothesis that the effects of interventions fell within a range of values that were all so trivially small that we could conclude with confidence that whatever effect interventions might have, they were too small to care about. Alternatively, one might move away from the binary (significant vs. non-significant) classification systems these methods imply to simply describe the range of plausible valuables for key parameters, using tools such as confidence intervals (Fidler et al., <xref ref-type="bibr" rid="CIT0008">2004</xref>). I will return to this suggestion below.</p>
<p>Secondly, it is important to change the way in which we view and report our results. In essence, we need to change our collective &#x2018;scientific language&#x2019;. Readers often appear to interpret the phrase &#x2018;this result is statistically significant&#x2019; to mean &#x2018;this result is important&#x2019;. This has to stop. Thirdly, accept uncertainty. We sometimes compute confidence intervals or some similar measure, but we do this mainly to see whether or not our confidence interval includes zero (or whatever value is used to define a particular point hypothesis). It is much better to be aware of, and to understand, the implications of uncertainty in estimating population statistics from samples and to keep this uncertainty in mind when interpreting the results. Fourthly, be thoughtful in the design, analysis and interpretation of studies. In particular, we should use statistics and data analysis as a tool for helping us understand what the data mean. The use of NHST as a tool for making binary decisions (significant vs. non-significant) discourages thoughtful analysis; as we move away from a mechanical procedure for evaluating data, we will be forced to move towards methods of data analysis that compel us to think about what the data mean. Fifthly, be open. If we get rid of a simple widely known (but not widely understood) procedure for making sense of data, we are going to have to fall back on informed judgement for evaluating results. That is, we will have to present and defend criteria for evaluating data rather than falling back on a familiar but ultimately meaningless procedure such as NHST.</p>
<p>Finally, Wasserstein et al. (<xref ref-type="bibr" rid="CIT0022">2019</xref>) encourage scientists to be open and modest in the interpretation of their data. They advocate that we operate using the ATOM principle, that is, <italic>accept</italic> that there will always be uncertainty, and be <italic>thoughtful, open</italic> and <italic>modest</italic>. This strikes me as a very useful piece of advice for the <italic>South African Journal of Industrial Psychology</italic>, its contributing authors and incoming editorial board.</p>
</sec>
<sec id="s0004">
<title>Conclusion</title>
<p>Will reducing our reliance on NHST solve all of the problems that have come to light in recent research on the credibility and reproducibility of our results? Certainly not! However, it is a good start towards enhancing the credibility, reproducibility and meaning of the research published in our journals. Does this mean that we should abandon significance tests altogether? Probably not; this would be a step too far for many of the journal&#x2019;s readers, authors, editors and reviewers. However, I believe that it is a realistic step to require all studies to start with a realistic power analysis and to report the results of this analysis. This would have two salutatory effects. Firstly, it would allow authors to think concretely about what sort of effects they expect to observe and to justify their assumptions about effect size. Secondly, it would discourage authors from attempting to use small sample studies to address important problems. Unless the effect you expect to observe (and can articulate a realistic basis for this expectation) is quite large, power analyses will encourage you to collect larger sample than you might otherwise settle for. Larger samples would enhance the stability and reproducibility of the results reported in our journals. Serious attention to statistical power would also render most NHST tests essentially moot. The whole point of performing a power analysis before you collect data is to help you design a study where you will easily reject the null hypothesis H<sub>0</sub>.</p>
<p>Finally, a wholehearted adoption of power analysis will help to wean authors away from dependence on significance tests. If these tests become a foregone conclusion, their apparent value as evidence is likely to decline. This will not totally solve the problem of creating credible and reproducible science, but it strikes me as the best first step.</p>
</sec>
</body>
<back>
<ack>
<title>Acknowledgements</title>
<sec id="s20005" sec-type="COI-statement">
<title>Competing interests</title>
<p>The author declares that they have no financial or personal relationships which may have inappropriately influenced them in writing this article.</p>
</sec>
<sec id="s20006">
<title>Author&#x2019;s contributions</title>
<p>K.R.M. is the sole contributor to this article.</p>
</sec>
<sec id="s20007">
<title>Ethical considerations</title>
<p>I confirm that ethical clearance was not needed or required for the study.</p>
</sec>
<sec id="s20008">
<title>Funding information</title>
<p>This research received no specific grant from any funding agency in the public, commercial or not-for-profit sectors.</p>
</sec>
<sec id="s20009">
<title>Data availability statement</title>
<p>Data sharing is not applicable to this article as no new data were created or analysed in this study.</p>
</sec>
<sec id="s20010">
<title>Disclaimer</title>
<p>The views and opinions expressed in this article are those of the author and do not necessarily reflect the official policy or position of any affiliated agency of the author.</p>
</sec>
</ack>
<ref-list id="references">
<title>References</title>
<ref id="CIT0001"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Amrhein</surname>, <given-names>V</given-names></string-name>., <string-name><surname>Greenland</surname>, <given-names>S</given-names></string-name>., &#x0026; <string-name><surname>McShane</surname>, <given-names>B</given-names></string-name></person-group>. (<year>2019</year>). <article-title>Scientists rise up against statistical significance</article-title>. <source><italic>Nature</italic></source>, <volume>567</volume>, <fpage>305</fpage>&#x2013;<lpage>307</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1038/d41586-019-00857-9">https://doi.org/10.1038/d41586-019-00857-9</ext-link></comment></mixed-citation></ref>
<ref id="CIT0002"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Bakker</surname>, <given-names>M</given-names></string-name>., <string-name><surname>Van Dijk</surname>, <given-names>A</given-names></string-name>., &#x0026; <string-name><surname>Wicherts</surname>, <given-names>J.M</given-names></string-name></person-group>. (<year>2012</year>). <article-title>The rules of the game called psychological science</article-title>. <source><italic>Perspectives on Psychological Science</italic></source>, <volume>7</volume>, <fpage>534</fpage>&#x2013;<lpage>554</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1177/1745691612459060">https://doi.org/10.1177/1745691612459060</ext-link></comment></mixed-citation></ref>
<ref id="CIT0003"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Banks</surname>, <given-names>G.C</given-names></string-name>., <string-name><surname>Rogelberg</surname>, <given-names>S.G</given-names></string-name>., <string-name><surname>Woznyj</surname>, <given-names>H.M</given-names></string-name>., <string-name><surname>Landis</surname>, <given-names>R.S</given-names></string-name>., &#x0026; <string-name><surname>Rupp</surname>, <given-names>D.E</given-names></string-name></person-group>. (<year>2016</year>). <article-title>Editorial: Evidence on questionable research practices: The good, the bad and the ugly</article-title>. <source><italic>Journal of Business and Psychology</italic></source>, <volume>31</volume>, <fpage>323</fpage>&#x2013;<lpage>338</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1007/s10869-016-9456-7">https://doi.org/10.1007/s10869-016-9456-7</ext-link></comment></mixed-citation></ref>
<ref id="CIT0004"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Cohen</surname>, <given-names>J</given-names></string-name></person-group>. (<year>1962</year>). <article-title>The statistical power of abnormal-social psychological research: A review</article-title>. <source><italic>Journal of Abnormal Social Psychology</italic></source>, <volume>65</volume>, <fpage>145</fpage>&#x2013;<lpage>153</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1037/h0045186">https://doi.org/10.1037/h0045186</ext-link></comment></mixed-citation></ref>
<ref id="CIT0005"><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><surname>Cohen</surname>, <given-names>J</given-names></string-name></person-group>. (<year>1988</year>). <source><italic>Statistical power analysis for the behavioral sciences</italic></source> (<edition>2nd edn</edition>.). <publisher-loc>Hillsdale, NJ</publisher-loc>: <publisher-name>Erlbaum</publisher-name>.</mixed-citation></ref>
<ref id="CIT0006"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Cohen</surname>, <given-names>J</given-names></string-name></person-group>. (<year>1994</year>). <article-title>The earth is round (<italic>p</italic> &#x003C; .05)</article-title>. <source><italic>American Psychologist</italic></source>, <volume>49</volume>, <fpage>997</fpage>&#x2013;<lpage>1003</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1037/0003-066X.49.12.997">https://doi.org/10.1037/0003-066X.49.12.997</ext-link></comment></mixed-citation></ref>
<ref id="CIT0007"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Efendic</surname>, <given-names>E</given-names></string-name>., &#x0026; <string-name><surname>Van Zyl</surname>, <given-names>L.E</given-names></string-name></person-group>. (<year>2019</year>). <article-title>On reproducibility and replicability: Arguing for open science practices and methodological improvements at the South African Journal of Industrial Psychology</article-title>. <source><italic>SA Journal of Industrial Psychology</italic></source>, <volume>45</volume>(<issue>0</issue>), <fpage>a1607</fpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.4102/sajip.v45i0.1607">https://doi.org/10.4102/sajip.v45i0.1607</ext-link></comment></mixed-citation></ref>
<ref id="CIT0008"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Fidler</surname>, <given-names>F</given-names></string-name>., <string-name><surname>Thomason</surname>, <given-names>N</given-names></string-name>., <string-name><surname>Cumming</surname>, <given-names>G</given-names></string-name>., <string-name><surname>Finch</surname>, <given-names>S</given-names></string-name>., &#x0026; <string-name><surname>Leeman</surname>, <given-names>J</given-names></string-name></person-group>. (<year>2004</year>). <article-title>Editors can lead researchers to confidence intervals, but can&#x2019;t make them think</article-title>. <source><italic>Psychological Science</italic></source>, <volume>15</volume>, <fpage>119</fpage>&#x2013;<lpage>126</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1111/j.0963-7214.2004.01502008.x">https://doi.org/10.1111/j.0963-7214.2004.01502008.x</ext-link></comment></mixed-citation></ref>
<ref id="CIT0009"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Ioannidis</surname>, <given-names>J.P</given-names></string-name></person-group>. (<year>2005</year>). <article-title>Why most published research findings are false</article-title>. <source><italic>PLoS Medical</italic></source>, <volume>2</volume>, <fpage>e124</fpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1371/journal.pmed.0020124">https://doi.org/10.1371/journal.pmed.0020124</ext-link></comment></mixed-citation></ref>
<ref id="CIT0010"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>McNutt</surname>, <given-names>M</given-names></string-name></person-group>. (<year>2014</year>). <article-title>Reproducibility</article-title>. <source><italic>Science</italic></source>, <volume>343</volume>, <fpage>229</fpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1126/science.1250475">https://doi.org/10.1126/science.1250475</ext-link></comment></mixed-citation></ref>
<ref id="CIT0011"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Meehl</surname>, <given-names>P</given-names></string-name></person-group>. (<year>1978</year>). <article-title>Theoretical risks and tabular asterisks: Sir Karl, Sir Ronald, and the slow progress of psychology</article-title>. <source><italic>Journal of Consulting and Clinical Psychology</italic></source>, <volume>46</volume>, <fpage>806</fpage>&#x2013;<lpage>834</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1037/0022-006X.46.4.806">https://doi.org/10.1037/0022-006X.46.4.806</ext-link></comment></mixed-citation></ref>
<ref id="CIT0012"><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><surname>Murphy</surname>, <given-names>K</given-names></string-name>., <string-name><surname>Myors</surname>, <given-names>B</given-names></string-name>., &#x0026; <string-name><surname>Wolach</surname>, <given-names>A</given-names></string-name></person-group>. (<year>2014</year>). <source><italic>Statistical power analysis: A simple and general model for traditional and modern hypothesis tests</italic></source> (<edition>4th edn</edition>.). <publisher-loc>New York</publisher-loc>: <publisher-name>Tayor &#x0026; Francis</publisher-name>.</mixed-citation></ref>
<ref id="CIT0013"><mixed-citation publication-type="journal"><person-group person-group-type="author"><collab>Neuroskeptic</collab></person-group>. (<year>2012</year>). <article-title>The nine circles of scientific hell</article-title>. <source><italic>Perspectives on Psychological Science</italic></source>, <volume>7</volume>, <fpage>643</fpage>&#x2013;<lpage>644</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1177/1745691612459519">https://doi.org/10.1177/1745691612459519</ext-link></comment></mixed-citation></ref>
<ref id="CIT0014"><mixed-citation publication-type="journal"><person-group person-group-type="author"><collab>Open Science Collaboration</collab></person-group>. (<year>2015</year>). <article-title>Estimating the reproducibility of psychological science</article-title>. <source><italic>Science</italic></source>, <volume>349</volume>, <fpage>aac4716</fpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1126/science.aac4716">https://doi.org/10.1126/science.aac4716</ext-link></comment></mixed-citation></ref>
<ref id="CIT0015"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Pashler</surname>, <given-names>H</given-names></string-name>., &#x0026; <string-name><surname>Wagenmakers</surname>, <given-names>E.J</given-names></string-name></person-group>. (<year>2012</year>). <article-title>Editors&#x2019; introduction to the special section on replicability in psychological science: A crisis of confidence?</article-title> <source><italic>Perspectives on Psychological Science</italic></source>, <volume>7</volume>, <fpage>528</fpage>&#x2013;<lpage>530</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1177/1745691612465253">https://doi.org/10.1177/1745691612465253</ext-link></comment></mixed-citation></ref>
<ref id="CIT0016"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Rouanet</surname>, <given-names>H</given-names></string-name></person-group>. (<year>1996</year>). <article-title>Bayesian methods for assessing the importance of effects</article-title>. <source><italic>Psychological Bulletin</italic></source>, <volume>119</volume>, <fpage>149</fpage>&#x2013;<lpage>158</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1037/0033-2909.119.1.149">https://doi.org/10.1037/0033-2909.119.1.149</ext-link></comment></mixed-citation></ref>
<ref id="CIT0017"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Schmidt</surname>, <given-names>F.L</given-names></string-name></person-group>. (<year>1996</year>). <article-title>Statistical significance testing and cumulative knowledge in psychology: Implications for training of researchers</article-title>. <source><italic>Psychological Methods</italic></source>, <volume>1</volume>, <fpage>115</fpage>&#x2013;<lpage>129</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1037/1082-989X.1.2.115">https://doi.org/10.1037/1082-989X.1.2.115</ext-link></comment></mixed-citation></ref>
<ref id="CIT0018"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Serlin</surname>, <given-names>R.A</given-names></string-name>., &#x0026; <string-name><surname>Lapsley</surname>, <given-names>D.K</given-names></string-name></person-group>. (<year>1985</year>). <article-title>Rationality in psychological research: The good-enough principle</article-title>. <source><italic>American Psychologist</italic></source>, <volume>40</volume>, <fpage>73</fpage>&#x2013;<lpage>83</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1037/0003-066X.40.1.73">https://doi.org/10.1037/0003-066X.40.1.73</ext-link></comment></mixed-citation></ref>
<ref id="CIT0019"><mixed-citation publication-type="book"><person-group person-group-type="author"><string-name><surname>Serlin</surname>, <given-names>R.A</given-names></string-name>., &#x0026; <string-name><surname>Lapsley</surname>, <given-names>D.K</given-names></string-name></person-group>. (<year>1993</year>). <chapter-title>Rational appraisal of psychological research and the good-enough principle</chapter-title>. In <person-group person-group-type="editor"><string-name><given-names>G.</given-names> <surname>Keren</surname></string-name> &#x0026; <string-name><given-names>C.</given-names> <surname>Lewis</surname></string-name> (Eds.)</person-group>, <source><italic>A handbook for data analysis in the behavioral sciences: Methodological issues</italic></source> (pp. <fpage>199</fpage>&#x2013;<lpage>228</lpage>). <publisher-loc>Hillsdale, NJ</publisher-loc>: <publisher-name>Lawrence Erlbaum Associates</publisher-name>.</mixed-citation></ref>
<ref id="CIT0020"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Sterling</surname>, <given-names>T.D</given-names></string-name></person-group>. (<year>1959</year>). <article-title>Publication decisions and their possible effects on inferences drawn from tests of significance &#x2013; Or vice versa</article-title>. <source><italic>Journal of the American Statistical Association</italic></source>, <volume>54</volume>, <fpage>30</fpage>&#x2013;<lpage>34</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1080/01621459.1959.10501497">https://doi.org/10.1080/01621459.1959.10501497</ext-link></comment></mixed-citation></ref>
<ref id="CIT0021"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Wasserstein</surname>, <given-names>R.L</given-names></string-name>., &#x0026; <string-name><surname>Lazar</surname>, <given-names>N.A</given-names></string-name></person-group>. (<year>2016</year>). <article-title>The ASA&#x2019;s statement on p-values: Context, process, and purpose</article-title>, <source><italic>The American Statistician</italic></source>, <volume>70</volume>(<issue>2</issue>), <fpage>129</fpage>&#x2013;<lpage>133</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1080/00031305.2016.1154108">https://doi.org/10.1080/00031305.2016.1154108</ext-link></comment></mixed-citation></ref>
<ref id="CIT0022"><mixed-citation publication-type="journal"><person-group person-group-type="author"><string-name><surname>Wasserstein</surname>, <given-names>R.L</given-names></string-name>., <string-name><surname>Schirm</surname>, <given-names>A.L</given-names></string-name>., &#x0026; <string-name><surname>Lazar</surname>, <given-names>N.A</given-names></string-name></person-group>. (<year>2019</year>). <article-title>Moving to a world beyond &#x2018;<italic>p</italic> &#x003C; 0.05&#x2019;</article-title>. <source><italic>The American Statistician</italic></source>, <volume>73</volume>(<supplement>sup1</supplement>), <fpage>1</fpage>&#x2013;<lpage>19</lpage>. <comment><ext-link ext-link-type="uri" xlink:href="https://doi.org/10.1080/00031305.2019.1583913">https://doi.org/10.1080/00031305.2019.1583913</ext-link></comment></mixed-citation></ref>
</ref-list>
<fn-group>
<fn><p><bold>How to cite this article:</bold> Murphy, K.R. (2019). Reducing our dependence on null hypothesis testing: A key to enhance the reproducibility and credibility of our science. <italic>SA Journal of Industrial Psychology/SA Tydskrif vir Bedryfsielkunde, 45</italic>(0), a1717. <ext-link ext-link-type="uri" xlink:href="https://doi.org/10.4102/sajip.v45i0.1717">https://doi.org/10.4102/sajip.v45i0.1717</ext-link></p></fn>
</fn-group>
</back>
</article>