...
首页> 外文期刊>Computational linguistics >What Determines Inter-Coder Agreement in Manual Annotations? A Meta-Analytic Investigation
【24h】

What Determines Inter-Coder Agreement in Manual Annotations? A Meta-Analytic Investigation

机译:是什么决定手册注释中的编码间协议?荟萃分析

获取原文
           

摘要

Recent discussions of annotator agreement have mostly centered around its calculation and interpretation, and the correct choice of indices. Although these discussions are important, they only consider the “back-end” of the story, namely, what to do once the data are collected. Just as important in our opinion is to know how agreement is reached in the first place and what factors influence coder agreement as part of the annotation process or setting, as this knowledge can provide concrete guidelines for the planning and set-up of annotation projects. To investigate whether there are factors that consistently impact annotator agreement we conducted a meta-analytic investigation of annotation studies reporting agreement percentages. Our meta-analysis synthesized factors reported in 96 annotation studies from three domains (word-sense disambiguation, prosodic transcriptions, and phonetic transcriptions) and was based on a total of 346 agreement indices. Our analysis identified seven factors that influence reported agreement values: annotation domain, number of categories in a coding scheme, number of annotators in a project, whether annotators received training, the intensity of annotator training, the annotation purpose, and the method used for the calculation of percentage agreements. Based on our results we develop practical recommendations for the assessment, interpretation, calculation, and reporting of coder agreement. We also briefly discuss theoretical implications for the concept of annotation quality.
机译:最近对注释者协议的讨论主要集中在其计算和解释以及正确选择索引上。尽管这些讨论很重要,但它们仅考虑了故事的“后端”,即一旦收集到数据该怎么办。在我们看来,重要的是要首先了解如何达成协议以及在注释过程或设置中会影响编码器协议的因素,因为该知识可以为注释项目的规划和设置提供具体指导。为了调查是否存在持续影响注释者协议的因素,我们对注释研究报告协议百分比的情况进行了荟萃分析。我们的荟萃分析综合了来自三个领域(词义歧义消除,韵律转录和语音转录)的96个注释研究中报告的因素,并且基于总共346个协议索引。我们的分析确定了影响报告的协议值的七个因素:注释域,编码方案中的类别数量,项目中注释者的数量,注释者是否接受培训,注释者培训的强度,注释目的以及用于注释的方法。百分比协议的计算。根据我们的结果,我们针对编码器协议的评估,解释,计算和报告提出了实用的建议。我们还简要讨论了注释质量概念的理论含义。

著录项

相似文献

  • 外文文献
  • 中文文献
  • 专利
获取原文

客服邮箱:kefu@zhangqiaokeyan.com

京公网安备:11010802029741号 ICP备案号:京ICP备15016152号-6 六维联合信息科技 (北京) 有限公司©版权所有
  • 客服微信

  • 服务号