Article detail · 2020
Annotator rationales for labeling tasks in crowdsourcing
Journal
The Journal of Artificial Intelligence ResearchISSN 1076 - 9757
- Year
- 2020
- Type
- article
Data source split
- YÖKSİS YÖKSİS article record
- YÖKSİS venue The Journal of Artificial Intelligence Research
- OpenAlex OpenAlex enrichment (abstract, citations, topics)
Abstract
OpenAlex · English
When collecting item ratings from human judges, it can be difficult to measure and enforce data quality due to task subjectivity and lack of transparency into how judges make each rating decision. To address this, we investigate asking judges to provide a specific form of rationale supporting each rating decision. We evaluate this approach on an information retrieval task in which human judges rate the relevance of Web pages for different search topics. Cost-benefit analysis over 10,000 judgments collected on Amazon’s Mechanical Turk suggests a win-win. Firstly, rationales yield a multitude of benefits: more reliable judgments, greater transparency for evaluating both human raters and their judgments, reduced need for expert gold, the opportunity for dual-supervision from ratings and rationales, and added value from the rationales themselves. Secondly, once experienced in the task, crowd workers provide rationales with almost no increase in task completion time. Consequently, we can realize the above benefits with minimal additional cost.
Topics
Citations
OpenAlex cited_by_count. Not a WoS or Scopus citation count; those sources have no separate column here.
32 citations
OpenAlex cited_by_count (cache / database)