Crowdsourced comparative judgement for evaluating learner texts: How reliable are judges recruited from an online crowdsourcing platform?

Thwaites, Peter; Vandeweerd, Nathan; Paquot, Magali

DIAL.pr - BOREAL

Accès à distance ? S'identifier sur le proxy UCLouvain

Crowdsourced comparative judgement for evaluating learner texts: How reliable are judges recruited from an online crowdsourcing platform?

Primary tabs

download

THWAITES VANDEWEERD PAQUOT AL 2024.pdf

Open access
PDF
630.21 K

Thwaites, Peter [UCL]

Vandeweerd, Nathan Paquot, Magali [UCL]

Recent studies of proficiency measurement and reporting practices in applied linguists have revealed widespread use of unsatisfactory practices such as the use of proxy measures of proficiency in place of explicit tests. Learner corpus research is one specific area affected by this problem: few learner corpora contain reliable, valid evaluations of text proficiency. This has led to calls for the development of new L2 writing proficiency measures for use in research contexts. Answering this call, a recent study by Paquot et al. (2022) generated assessments of learner corpus texts using a community-driven approach in which judges, recruited from the linguistic community, conducted assessments using comparative judgement. Although the approach generated reliable assessments, its practical use is limited because linguists are not always available to contribute to data collections. This paper therefore explores an alternative approach, in which judges are recruited through a crowdsourcing platform. We find that assessments generated in this way can reach near identical levels of reliability and concurrent validity to those produced by members of the linguistic community.

metadata

Document type	Article de périodique (Journal article) – Article de recherche
Access type	Accès libre
Publication date	2024
Language	Anglais
Journal information	"Applied Linguistics" - Vol. forthcoming, no.forth, p. forth (2024)
Peer reviewed	yes
Publication status	Publié
Affiliation	UCL - SSH/ILC/PLIN - Pôle de recherche en linguistique
Keywords	comparative judgement ; crowdsourcing ; learner corpus ; learner corpus research ; prolific ; proficiency
Links	http://hdl.handle.net/2078.1/290245[Handle] https://doi.org/10.1093/applin/amae048[DOI]

Bibliographic reference	Thwaites, Peter ; Vandeweerd, Nathan ; Paquot, Magali. Crowdsourced comparative judgement for evaluating learner texts: How reliable are judges recruited from an online crowdsourcing platform?. In: Applied Linguistics, Vol. forthcoming, no.forth, p. forth (2024)
Permanent URL	http://hdl.handle.net/2078.1/290245

User menu

Crowdsourced comparative judgement for evaluating learner texts: How reliable are judges recruited from an online crowdsourcing platform?

Primary tabs

Footer Help

Languages

Footer menu

User menu

Search form

You are here

Crowdsourced comparative judgement for evaluating learner texts: How reliable are judges recruited from an online crowdsourcing platform?

Primary tabs

Footer Help

Languages

Footer menu