{"id":2384,"date":"2010-08-31T06:43:33","date_gmt":"2010-08-31T10:43:33","guid":{"rendered":"http:\/\/shermandorn.com\/wordpress\/?p=2384"},"modified":"2010-08-31T15:42:12","modified_gmt":"2010-08-31T19:42:12","slug":"please-leave-your-magic-numbers-on-the-magic-carpet-with-the-magic-wand","status":"publish","type":"post","link":"https:\/\/shermandorn.com\/?p=2384","title":{"rendered":"Please leave your magic numbers on the magic carpet with the magic wand"},"content":{"rendered":"<p>The shorter <a href=\"http:\/\/www.quickanded.com\/2010\/08\/the-magic-number-for-value-added.html\">Bill Tucker<\/a>: &quot;No, <em>you<\/em> pull a number out of the air.&quot; Well, buckaroos, if anyone&#39;s going to be all airy-fairy on this quantitative-component teacher-evaluation thing, <a href=\"http:\/\/papers.ssrn.com\/sol3\/papers.cfm?abstract_id=1461508\">it&#39;s going to be all Ivory Tower, historian me<\/a>. But I don&#39;t think you&#39;d like that.<\/p>\n<p>Maybe I should explain. Tucker&#39;s argument in a nutshell is, &quot;Okay, if the Economic Policy Institute paper authors don&#39;t think 50% of teacher evaluations should be based on test scores, what number <em>do<\/em> they think is the right one?&quot; The sly assumption here is that there <em>should be<\/em> an a priori percentage. Yes, yes, I get the point&#8211;if there isn&#39;t an a priori percentage, then why is 50% wrong?<\/p>\n<p>Since I wasn&#39;t involved in the development of the EPI white paper, I can step in here with my judgment: I don&#39;t think there&#39;s a magic number, but both my research and my teaching experience tells me that there is no pragmatic setting of a weight that should come close to 50% for any derivative of test scores. Maybe I&#39;ll be dead wrong, but I don&#39;t think so.<\/p>\n<p>First, the teaching experience, since that&#39;s more easily explained. I have taught hundreds of students, and I don&#39;t remember a single time when more than 40% of any term grade algorithm depended on a single assignment&#8211;not a paper nor an exam. Having multiple inputs to a term grade can be a headache, but I prefer having more information about student performance, and I suspect today&#39;s students hate classes where 50% or more of the grade depends on the final exam.<\/p>\n<p>In addition, my experience calculating grades for hundreds of students tells me that weighting has a nonlinear relationship with the influence of a single assignment on the final term grade. An assignment with 20% contribution to the final grade does not have twice as much influence as an assignment with 10% contribution. Part of the influence is in the effective range of scores (something I&#39;ve mentioned before). And part is a threshold effect: because three points near a grade threshold can be especially influential, components with larger effective ranges have a disproportionately larger influence on final grades. That teaching experience is consistent with my <a href=\"http:\/\/papers.ssrn.com\/sol3\/papers.cfm?abstract_id=1461508\">airy-fairy paper on the subject<\/a>, which you should take seriously because <del>it has lots of fancy equations<\/del> I think I&#39;ve been sufficiently careful to explain the consequences of a Bayesian approach to evaluation components: don&#39;t try to widgetize evaluation systems.<\/p>\n<p><em>So what should happen?<\/em> Glad you asked! Maybe we gather data for different components, as is happening now in a number of places, and run simulations for evaluations with different models and for different classes of teachers. I even have some bold predictions you can test:<\/p>\n<ul>\n<li>Any system for non-core academic teachers (e.g., teachers in the arts) is going to be clearly, obviously unstable.<\/li>\n<li>Within core academic areas, you will see less stability for teachers who share responsibility for some students (with disabilities, who are English language learners, etc.).<\/li>\n<li>Well-trained teams of peer evaluators will create more stable evaluation components than either test-score components or typical administrator-observation evaluations.<\/li>\n<\/ul>\n<p>As I&#39;ve said before, the 50% figure was pulled out of thin air in the same way that the &quot;65% solution&quot; figure was pulled out of thin air. Or maybe they were pulled out of pieces of someone&#39;s anatomy. We all do that on occasion, make sheer guesses that turn out later to be bold and utterly false. Most of us don&#39;t get to turn such braggadocio into policy, and I don&#39;t think it&#39;s wise for it to happen with teacher evaluation, either.<\/p>\n<p>Update: <a href=\"http:\/\/www.quickanded.com\/2010\/08\/my-value-added-number.html\">Tucker responds<\/a>.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>The shorter Bill Tucker: &quot;No, you pull a number out of the air.&quot; Well, buckaroos, if anyone&#39;s going to be all airy-fairy on this quantitative-component teacher-evaluation thing, it&#39;s going to be all Ivory Tower, historian me. But I don&#39;t think you&#39;d like that. Maybe I should explain. Tucker&#39;s argument in a nutshell is, &quot;Okay, if [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":0,"comment_status":"open","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"_jetpack_newsletter_access":"","_jetpack_dont_email_post_to_subs":false,"_jetpack_newsletter_tier_id":0,"_jetpack_memberships_contains_paywalled_content":false,"_jetpack_memberships_contains_paid_content":false,"footnotes":"","jetpack_publicize_message":"","jetpack_publicize_feature_enabled":true,"jetpack_social_post_already_shared":false,"jetpack_social_options":{"image_generator_settings":{"template":"highway","default_image_id":0,"font":"","enabled":false},"version":2},"jetpack_post_was_ever_published":false},"categories":[16,11],"tags":[],"class_list":["post-2384","post","type-post","status-publish","format-standard","hentry","category-accountability-frankenstein","category-education-policy"],"jetpack_publicize_connections":[],"jetpack_featured_media_url":"","jetpack_sharing_enabled":true,"jetpack_shortlink":"https:\/\/wp.me\/pag0MB-Cs","_links":{"self":[{"href":"https:\/\/shermandorn.com\/index.php?rest_route=\/wp\/v2\/posts\/2384","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/shermandorn.com\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/shermandorn.com\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/shermandorn.com\/index.php?rest_route=\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/shermandorn.com\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=2384"}],"version-history":[{"count":5,"href":"https:\/\/shermandorn.com\/index.php?rest_route=\/wp\/v2\/posts\/2384\/revisions"}],"predecessor-version":[{"id":2391,"href":"https:\/\/shermandorn.com\/index.php?rest_route=\/wp\/v2\/posts\/2384\/revisions\/2391"}],"wp:attachment":[{"href":"https:\/\/shermandorn.com\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=2384"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/shermandorn.com\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=2384"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/shermandorn.com\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=2384"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}