Improving crowdsourcing for AI through cognitive-inspired data engineering.
Authors
Affiliations (9)
Affiliations (9)
- Cognitive Science Program, Indiana University, Indiana, USA. [email protected].
- Department of Psychological and Brain Sciences, Indiana University, Indiana, USA. [email protected].
- Centaur Labs, Massachusetts, USA. [email protected].
- Department of Economics, New York University, New York, USA.
- Centaur Labs, Massachusetts, USA.
- Cognitive Science Program, Indiana University, Indiana, USA.
- Department of Mathematics, Indiana University, Indiana, USA.
- Department of Economics, University of California, Santa Barbara, USA.
- Department of Psychological and Brain Sciences, Indiana University, Indiana, USA.
Abstract
Crowdsourcing offers a fast and cost-efficient approach to obtaining human-labeled datasets. However, crowdsourced datasets and the models trained on them can inherit the cognitive constraints and biases of their annotators. In a process we refer to as cognitive-inspired data engineering, we investigate whether ideas from cognitive science can be applied to mitigate the presence of cognitive constraints and cognitive biases in crowdsourced datasets and, as a result, improve the performance of models trained on these datasets. We evaluate our approach by crowdsourcing labels for medical image diagnostic tasks using two different crowdsourcing platforms across two experiments. In Experiment 1, we collect subjective probability judgments from novice annotators through Amazon Mechanical Turk and, in Experiment 2, we collect subjective probability judgments and binary classifications from skilled annotators through DiagnosUs, a crowdsourcing platform specializing in medical and scientific data annotation. In both experiments, we find that recalibrating subjective probability judgments reduces bias, yielding more accurate crowdsourced datasets and more accurate models trained on them. Our results suggest that cognitive-inspired data engineering offers a promising avenue to improve the quality of crowdsourced datasets, with consistent downstream benefits for machine learning models.