2022 TheCrowdisMadeofPeopleObservati

(Thomas et al., 2022) ⇒ Paul Thomas, Gabriella Kazai, Ryen White, and Nick Craswell. (2022). “The Crowd is Made of People: Observations from Large-scale Crowd Labelling.” In: Proceedings of the 2022 Conference on Human Information Interaction and Retrieval. doi:10.1145/3498366.3505815

Subject Headings: Data Annotation Task, Information Retrieval Evaluation.

Notes

Cited By

http://scholar.google.com/scholar?q=%222022%22+The+Crowd+is+Made+of+People%3A+Observations+from+Large-scale+Crowd+Labelling

Quotes

Abstract

Like many other researchers, at Microsoft Bing we use external “crowd” judges to label results from a search engine—especially, although not exclusively, to obtain relevance labels for offline evaluation in the Cranfield tradition. Crowdsourced labels are relatively cheap, and hence very popular, but are prone to disagreements, spam, and various biases which appear to be unexplained “noise” or “error”. In this paper, we provide examples of problems we have encountered running crowd labelling at large scale and around the globe, for search evaluation in particular. We demonstrate effects due to the time of day and day of week that a label is given; fatigue; anchoring; exposure; left-side bias; task switching; and simple disagreement between judges. Rather than simple “error”, these effects are consistent with well-known physiological and cognitive factors. “The crowd” is not some abstract machinery, but is made of people. Human factors that affect people’s judgement behaviour must be considered when designing research evaluations and in interpreting evaluation metrics.

References

;

	Author	volume	Date Value	title	type	journal	titleUrl	doi	note	year
2022 TheCrowdisMadeofPeopleObservati	Paul Thomas Gabriella Kazai Ryen White Nick Craswell			The Crowd is Made of People: Observations from Large-scale Crowd Labelling				10.1145/3498366.3505815		2022