{"id":1601,"date":"2018-12-01T16:07:22","date_gmt":"2018-12-01T21:07:22","guid":{"rendered":"https:\/\/www.causeweb.org\/sbi\/?p=1601"},"modified":"2018-12-01T16:26:01","modified_gmt":"2018-12-01T21:26:01","slug":"thoughts-on-null-hypothesis-significance-testing-in-an-sbi-course","status":"publish","type":"post","link":"https:\/\/www.causeweb.org\/sbi\/?p=1601","title":{"rendered":"Thoughts on null hypothesis significance testing in an SBI course"},"content":{"rendered":"<p><a href=\"https:\/\/www.causeweb.org\/sbi\/wp-content\/uploads\/2018\/12\/Bruce-Evan-Blaine-1.jpg\"><img loading=\"lazy\" decoding=\"async\" class=\"alignleft size-full wp-image-1600\" src=\"https:\/\/www.causeweb.org\/sbi\/wp-content\/uploads\/2018\/12\/Bruce-Evan-Blaine-1.jpg\" alt=\"\" width=\"150\" height=\"150\" \/><\/a><strong>Bruce Evan Blaine, John Fisher College, Rochester NY<\/strong><\/p>\n<p>I teach simulation-based statistical inference methods (using R) in my 100-level Introduction to Data Science course. This course is the required first course for all Data Science minors, and a service course to numerous departments. I love teaching statistical inference this way because it reconnects me (and my students) with Fisher\u2019s original ideas and methods, and expresses Tukey\u2019s ideas that we learn about populations by being in dialogue with data. In the context of this welcome return to the empirical framework through which we understand and teach statistical inference, I wonder why we still teach students null hypothesis significance testing (NHST) in the same old way. I expect we\u2019re all aware of the vast literature accumulated over the past 40 years that is critical of NHST and its role in the reproducibility crises in many disciplines. I feel like an introductory statistics or data science course that embraces simulated-based inference should also move away from teaching students conventional NHST methods for learning about populations.<\/p>\n<p>[pullquote] I\u2019m just encouraging us to think about whether the formal, reflexive method of classical NHST fits within an SBI pedagogical framework. Cohen and many others have urged us to replace NHST with inferential tools such as parameter estimation, effect size estimation, replication, and meta-analysis\u2014tools that help us learn much more about our population of interest..[\/pullquote]<\/p>\n<p><!--more--><\/p>\n<p>Although I am still working out this larger project in my teaching, I offer these thoughts in the context of teaching inference in the 2 independent-sample design with the mean difference statistic (M<sub>1<\/sub> &#8211; M<sub>2<\/sub>) in an Introduction to Data Science class:<\/p>\n<p>First, I help my students think about all the possible values of the mean difference, one of which is the parameter (\u03bc<sub>1 <\/sub>&#8211; \u03bc<sub>2<\/sub>), and how ridiculously implausible it would be for that to be 0. If the true mean difference isn\u2019t 0, then the remaining possible values vary only in the sign and magnitude of the mean difference. This points our interest toward estimation and away from significance, and sets up the task of estimating the parameter rather than in establishing that it (probably) isn\u2019t 0. I\u2019ll paraphrase one of my idols, Jacob Cohen (from his famous <em>The Earth is Round (p&lt;.05)<\/em> paper): null hypotheses are rarely true, so rejecting them is hardly surprising.<\/p>\n<p>Second, I help the students think about what kind of probability distribution we need to estimate the parameter. We talk about the importance of having (or assuming we have) data from a random sample for this task, what M<sub>1<\/sub> &#8211; M<sub>2<\/sub> might be if we had a <em>different<\/em> random sample from this population, and how those random differences in M<sub>1<\/sub> &#8211; M<sub>2<\/sub> are important to estimation. Using R we generate the probability distribution under H<sub>A<\/sub>, not under the chance or null model used in NHST. Through resampling, students create a picture of the population of interest and its parameters. Once created, the distribution allows students to do inference through confidence interval estimates of the parameter using various methods (i.e., normal-theory, percentile).<\/p>\n<p>Third, students should see that the distribution described above is a probability distribution for testing all sorts of hypotheses including, if we\u2019re interested, the null (or Cohen would say, nil-null) hypothesis in which \u03bc<sub>1 <\/sub>&#8211; \u03bc<sub>2 <\/sub>= 0. Students can find the probability of any hypothesized \u03bc<sub>1 <\/sub>&#8211; \u03bc<sub>2<\/sub> occurring in <em>this<\/em> population. We can then point our students toward hypothesis testing that sets evidence thresholds with practically or clinically significant standards, rather than a standard of differing from 0. For example, if the data in our 2-sample study evaluated the effect of exercise on blood pressure, we could get students to consider questions like: a) whether this effect was large enough to merit changing one\u2019s behavior, b) what other behavioral interventions might be as, or more, effective, and c) how a replication of (or a failure to replicate) this sized outcome in another sample would change our parameter estimate and\/or our confidence in the estimate.<\/p>\n<p>I\u2019m not dumping on hypothesis testing\u2014inference by hypothesis testing is valuable way to learn about a population of interest and how it differs from other populations. I\u2019m just encouraging us to think about whether the formal, reflexive method of classical NHST fits within an SBI pedagogical framework. Cohen and many others have urged us to replace NHST with inferential tools such as parameter estimation, effect size estimation, replication, and meta-analysis\u2014tools that help us learn much more about our population of interest.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>Bruce Evan Blaine, John Fisher College, Rochester NY I teach simulation-based statistical inference methods (using R) in my 100-level Introduction to Data Science course. This course is the required first course for all Data Science minors, and a service course to numerous departments. I love teaching statistical inference this way because it reconnects me (and [&hellip;]<\/p>\n","protected":false},"author":72,"featured_media":0,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[3],"tags":[],"class_list":["post-1601","post","type-post","status-publish","format-standard","hentry","category-why"],"_links":{"self":[{"href":"https:\/\/www.causeweb.org\/sbi\/index.php?rest_route=\/wp\/v2\/posts\/1601","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.causeweb.org\/sbi\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.causeweb.org\/sbi\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.causeweb.org\/sbi\/index.php?rest_route=\/wp\/v2\/users\/72"}],"replies":[{"embeddable":true,"href":"https:\/\/www.causeweb.org\/sbi\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=1601"}],"version-history":[{"count":6,"href":"https:\/\/www.causeweb.org\/sbi\/index.php?rest_route=\/wp\/v2\/posts\/1601\/revisions"}],"predecessor-version":[{"id":1608,"href":"https:\/\/www.causeweb.org\/sbi\/index.php?rest_route=\/wp\/v2\/posts\/1601\/revisions\/1608"}],"wp:attachment":[{"href":"https:\/\/www.causeweb.org\/sbi\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=1601"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.causeweb.org\/sbi\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=1601"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.causeweb.org\/sbi\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=1601"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}