Gablasova, Dana and Harding, Luke and Bottini, Raffaella and Brezina, Vaclav and Ren, Haoshan (Sally) and Iamartino, Giovanni and Li, Yingyu and Liu, Tanjun and Pogessi, Laura and Savski, Kristof and Toomaneejinda, Anuchit and Zottola, Angela (2024) Building a corpus of student academic writing in EMI contexts : Challenges in corpus design and data collection across international higher education settings. Research Methods in Applied Linguistics, 3 (3): 100140. ISSN 2772-7661
Building_a_corpus_of_academic_writing_final.pdf - Accepted Version
Available under License Creative Commons Attribution.
Download (520kB)
Abstract
The article discusses methodological procedures and challenges in a project requiring multi-site, transnational data collection for the construction of a corpus of academic writing in EMI higher education contexts. Drawing on our decision-making experiences as a research team, together with empirical data generated through data collection logs recorded by a network of researchers involved in the project, we reflect on key issues in conducting the project and the solutions we found to address specific challenges. After describing the background to the project and the current status of the corpus, we focus on four broad challenges: (1) selecting partners and managing a multi-site project; (2) defining a working construct of academic writing; (3) categorising data according to disciplinary areas; and (4) managing data collection “on the ground”. Throughout, we provide descriptions of our solutions to the challenges identified, and we conclude with a call for further publication of corpus construction records to provide greater transparency and detail around decisions and judgements made at all stages of a corpus construction project.