Developing pedagogically appropriate language corpora through crowdsourcing and gamification
Abstract
Despite the unquestionable academic interest on corpus-based approaches to language education, the use of corpora by teachers in their everyday practice is still not very widespread. One way to promote usage of corpora in language teaching is by making pedagogically appropriate corpora, labelled with different types of problems (for instance, sensitive content, offensive language, structural problems), so that teachers can select authentic examples according to their needs. Because manually labelling corpora is extremely time-consuming, we propose to use crowdsourcing for this task. After a first exploratory phase, we are currently developing a multimode, multilanguage game in which players first identify problematic sentences and then classify them.
Authors 8
-
Affiliation as printed
Ruppin Academic Center, Emek Hefer, Israel
-
Centro de Estudos de Linguística Geral e Aplicada · University of Coimbra
Affiliation as printed
CELGA-ILTEC/University of Coimbra, Coimbra, Portugal
-
Centro de Estudos de Linguística Geral e Aplicada · University of Coimbra
Affiliation as printed
CELGA-ILTEC/University of Coimbra, Coimbra, Portugal
-
Institute of the Estonian Language
Affiliation as printed
Institute of the Estonian Language, Tallinn, Estonia
-
Affiliation as printed
University of Belgrade, Belgrade, Serbia
-
Affiliation as printed
University of Ljubljana, Ljubljana, Slovenia
-
Instituut voor de Nederlandse Taal
Affiliation as printed
Dutch Language Institute, Leiden, Netherlands
-
University of Ljubljana · Jožef Stefan Institute
Affiliation as printed
University of Ljubljana & Jožef Stefan Institute, Ljubljana, Slovenia
Cited by 2 stored of 2
2 results
No patents citing this paper on Lens.org (checked 2026-10-11).
References 5
5 results