Skip to main content

ETHNOS_APP

Home • Search • Journals • List 0

Crowdsourced Comparative Judgement for Evaluating Learner Texts

How Reliable are Judges Recruited from an Online Crowdsourcing Platform?

Bibliographic Data

ID22871285
AuthorsPeter Thwaites (0000-0003-2176-9620, Centre for English Corpus Linguistics, Institut Langage et Communication , Collège Érasme, Cardinal Mercier, 1, Louvain-la-Neuve 1348 UCLouvain), Nathan Vandeweerd (0000-0002-6498-9474, Radboud University Nijmegen), Magali Paquot (0000-0001-5687-5074, Centre for English Corpus Linguistics, Institut Langage et Communication , Collège Érasme, Cardinal Mercier, 1, Louvain-la-Neuve 1348 UCLouvain, corresponding author)
Year2025
Volume46
Issue4
Pages611-628
Publication date2025-08-28
Peer ReviewedYes
Open AccessYes
TypeARTICLE
VenueApplied Linguistics (JOURNAL)
Journal identifiersISSN: 0142-6001 • E-ISSN: 1477-450X
PublisherOxford University Press (OUP) (PUBLISHER)
DOI10.1093/applin/amae048
OpenAlexW4401158115
LanguageEN
Citations received3
References cited46

Recent studies of proficiency measurement and reporting practices in applied linguists have revealed widespread use of unsatisfactory practices such as the use of proxy measures of proficiency in place of explicit tests. Learner corpus research is one specific area affected by this problem: few learner corpora contain reliable, valid evaluations of text proficiency. This has led to calls for the development of new L2 writing proficiency measures for use in research contexts. Answering this call, a recent study by Paquot et al. (2022) generated assessments of learner corpus texts using a community-driven approach in which judges, recruited from the linguistic community, conducted assessments using comparative judgement. Although the approach generated reliable assessments, its practical use is limited because linguists are not always available to contribute to data collections. This paper, therefore, explores an alternative approach, in which judges are recruited through a crowdsourcing platform. We find that assessments generated in this way can reach near identical levels of reliability and concurrent validity to those produced by members of the linguistic community.

Crowdsourcing · Inter-rater reliability · Judgement · Language proficiency · Linguistics · Mathematics education · Natural language processing · Proxy (statistics) · Rating scale · Reliability (semiconductor) · World Wide Web · Artificial Intelligence · Computer Science · Discourse Analysis in Language Studies · Interpreting and Communication in Healthcare · Natural Language Processing Techniques · Psychology

  • Contribution of linguistic complexity to L2 learners’ perception on Chinese text comprehensibility and reading speed

    Open Access•Xiaopeng Zhang•Applied Linguistics•2025

  • Testing crowdsourcing as a means of recruitment for the comparative judgement of L2 argumentative essays

    Open Access•Peter Thwaites, Magali Paquot•Journal of Second Language Writing•2025

  • Comparative Judgment

    Open Access•Qian Wu•SAGE Open•2025

  • Measuring L2 Proficiency

    Pascale Leclercq, Amanda Edmonds et al.•Measuring L2 Proficiency•2014

  • Assessing Writing

    Open Access•Sara Cushing Weigle•Assessing Writing•2002

  • Modern Applied Statistics with S

    Open Access•W N Venables, Brian D Ripley•Modern Applied Statistics with S…•2002

  • Emmeans

    Russell V Lenth, Julia Piaskowski•CRAN: Contributed Packages•2017

  • Proficiency Assessment Standards in Second Language Acquisition Research

    Open Access•Angelo Tremblay, Annie Tremblay•Studies in Second Language…•2011

  • Language Proficiency in Native and Nonnative Speakers

    Jan H Hulstijn•Language Assessment Quarterly•2011

  • Data quality in online human-subjects research

    Open Access•Benjamin D Douglas, Patrick J Ewell et al.•PLoS ONE•2023

  • Rank Analysis of Incomplete Block Designs

    Ralph Allan Bradley, Milton E Terry•Biometrika•1952

  • Proficiency Level--a Fuzzy Variable in Computer Learner Corpora

    Carl Carlsen•Applied Linguistics•2012

  • Exploring the Validity of Comparative Judgement

    Open Access•Lucy Chambers, Euan Cunningham•Frontiers in Education•2022

  • Validity of Comparative Judgment Scores

    Open Access•Marije Lesterhuis, Renske Bouwer et al.•Frontiers in Education•2022

  • Research synthesis and historiography

    Margaret Thomas•Synthesizing Research on Language…•2006

  • The assessment of writing ability

    Open Access•Rob Schoonen, Margaretha Vergeer et al.•Language Testing•1997

  • Replication in Second Language Research

    Open Access•Emma Marsden, Kara Morgan‐short et al.•Language Learning•2018

  • Coming of age

    Open Access•Susan M Ga, Sasha Loewen et al.•Language Teaching•2021

  • Instructional manipulation checks

    Open Access•Daniel M Oppenheimer, Tom Meyvis et al.•Journal of Experimental Social…•2009

  • (Why) Are Open Research Practices the Future for the Study of Language Learning

    Open Access•Emma Marsden, Kara Morgan‐short•Language Learning•2023

  • Crowdsourced Adaptive Comparative Judgment

    Open Access•Magali Paquot, Ruth Rubin et al.•Language Learning•2022

  • Assessment of L2 Proficiency in Second Language Acquisition Research

    Open Access•Margaret Thomas•Language Learning•1994

  • Guidelines for Reporting Quantitative Methods and Results in Primary Research

    Open Access•John M Norris, Luke Plonsky et al.•Language Learning•2015

  • Proficiency Reporting Practices in Research on Second Language Acquisition

    Open Access•Hae In Park, Megan Solon et al.•Language Learning•2022

  • Language Proficiency in Native and Non-native Speakers

    Jan H Hulstijn•Language proficiency in native…•2015

Unique citing works3
Citations per year3
Citation span2025 - 2025 (1)
Citation velocityrecent
Highly citedNo
Citation typesNeutral: 3

Tools

Open DOIOpen Access
Ethnos_APP • Open Source Project • MIT License • Frontend v2.0.0 • Privacy and Cookies • API Documentation: api.ethnos.app/docs • API Source Code: GitHub • DOI: 10.5281/zenodo.17049435 • Frontend Source Code: GitHub • DOI: 10.5281/zenodo.17050053 • cruz.rio.br • Expectantes Misericordiae