The Miscalculation of Interrater Reliability: A Case Study Involving the AAC&U VALUE Rubrics.

Q2 Social Sciences

Practical Assessment, Research and Evaluation Pub Date : 2017-12-01 DOI:10.7275/Y36W-HG55

R. F. Szafran

{"title":"The Miscalculation of Interrater Reliability: A Case Study Involving the AAC&U VALUE Rubrics.","authors":"R. F. Szafran","doi":"10.7275/Y36W-HG55","DOIUrl":null,"url":null,"abstract":"Institutional assessment of student learning objectives has become a fact-of-life in American higher education and the Association of American Colleges and Universities’ (AAC&U) VALUE Rubrics have become a widely adopted evaluation and scoring tool for student work. As faculty from a variety of disciplines, some less familiar with the psychometric literature, are drawn into assessment roles, it is important to point out two easily made but serious errors in what might appear to be one of the more straightforward assessments of measurement quality—interrater reliability. The first error which can occur when a third rater is brought in to adjudicate a discrepancy in the scores reported by an initial two raters has been well-documented in the literature but never before illustrated with AAC&U rubrics. The second error is to cease training before the raters have demonstrated a satisfactory level of interrater reliability. This research note describes an actual case study in which the interrater reliability of the AAC&U rubrics was incorrectly reported and when correctly reported found to be inadequate. The note concludes with recommendations for the correct measurement of interrater reliability.","PeriodicalId":20361,"journal":{"name":"Practical Assessment, Research and Evaluation","volume":"97 1","pages":"11"},"PeriodicalIF":0.0000,"publicationDate":"2017-12-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"","citationCount":"1","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Practical Assessment, Research and Evaluation","FirstCategoryId":"1085","ListUrlMain":"https://doi.org/10.7275/Y36W-HG55","RegionNum":0,"RegionCategory":null,"ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q2","JCRName":"Social Sciences","Score":null,"Total":0}

引用次数: 1

Abstract

Institutional assessment of student learning objectives has become a fact-of-life in American higher education and the Association of American Colleges and Universities’ (AAC&U) VALUE Rubrics have become a widely adopted evaluation and scoring tool for student work. As faculty from a variety of disciplines, some less familiar with the psychometric literature, are drawn into assessment roles, it is important to point out two easily made but serious errors in what might appear to be one of the more straightforward assessments of measurement quality—interrater reliability. The first error which can occur when a third rater is brought in to adjudicate a discrepancy in the scores reported by an initial two raters has been well-documented in the literature but never before illustrated with AAC&U rubrics. The second error is to cease training before the raters have demonstrated a satisfactory level of interrater reliability. This research note describes an actual case study in which the interrater reliability of the AAC&U rubrics was incorrectly reported and when correctly reported found to be inadequate. The note concludes with recommendations for the correct measurement of interrater reliability.

查看原文本刊更多论文

互译器可靠性的误算:以AAC&U值准则为例。

对学生学习目标的机构评估已经成为美国高等教育的一个现实，美国高校协会(AAC&U)的价值标准已经成为广泛采用的学生作业评估和评分工具。由于来自不同学科的教师，有些不太熟悉心理测量学文献，被吸引到评估角色中，重要的是要指出两个容易犯但严重的错误，这些错误可能是测量质量的一个更直接的评估-解释者可靠性。第一个错误可能发生在第三个评价者被引入来评判最初两个评价者所报告的分数的差异时，这在文献中有很好的记录，但从未在AAC&U标准中说明过。第二个错误是，在评价者的相互信度达到令人满意的水平之前，就停止了训练。本研究报告描述了一个实际的案例研究，其中AAC&U准则的互解释器可靠性被错误地报告，当正确报告时发现是不充分的。该说明最后提出了正确测量互传器可靠性的建议。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

Practical Assessment, Research and Evaluation Social Sciences-Education

CiteScore

2.60

自引率

0.00%

发文量