Constructing CEFR-Aligned English Reading Assessments with AI: A Practical Guide for Thai English Teachers in Higher Education

Main Article Content

Sun-Young Shin
Suchada Sanonguthai

Abstract

This paper provides practical guidance for developing English reading assessments aligned with the Common European Framework of Reference for Languages (CEFR) with the support of generative AI. Although the CEFR is widely used in curriculum design, instruction, and assessment, many instructors face challenges when translating reading descriptors into assessment tasks that appropriately reflect specific proficiency levels, particularly transitional levels such as B1+ and B2+. AI tools offer new opportunities for generating reading materials and assessment tasks, but recent studies show that CEFR level labels alone do not ensure appropriate alignment. Drawing on language assessment theory and recent research on AI-generated CEFR-aligned texts, this paper proposes a specification-driven approach in which reading test specifications are derived from CEFR-aligned instructional materials and target CEFR descriptors to guide AI-assisted assessment development through a three-stage prompting process consisting of reading test specification development, reading assessment generation, and optional review and revision. For illustrative purposes, a baseline prompting approach based on CEFR-aligned textbook tasks and target CEFR descriptors is also presented to represent a realistic classroom assessment development process. Illustrative applications at the B1, B1+, and B2 levels show how reading test specifications can be incorporated into AI prompts to guide the development of CEFR-aligned reading assessments. The paper concludes with practical guidance for adapting AI-assisted assessment development to classroom contexts while maintaining instructor judgment and alignment with the intended reading construct.

Article Details

How to Cite
Shin, S.-Y., & Sanonguthai, S. (2026). Constructing CEFR-Aligned English Reading Assessments with AI: A Practical Guide for Thai English Teachers in Higher Education. LEARN Journal: Language Education and Acquisition Research Network, 19(2), 71–155. https://doi.org/10.70730/FDFL1823
Section
Academic Articles
Author Biographies

Sun-Young Shin, Department of Linguistics, Second Language Studies, Indiana University Bloomington, USA

A Professor in the Department of Linguistics at Indiana University. His research interests include GenAI-mediated language assessment, L2 listening assessment, L2 pragmatics assessment, and standard setting. His work has been published in numerous journals and edited volumes, and he has been invited to deliver lectures and workshops on L2 assessment in several countries. He currently serves on the editorial boards of Language Testing and Language Assessment Quarterly.

Suchada Sanonguthai, Language Institute, Thammasat University, Thailand

A Lecturer at the Language Institute, Thammasat University. She holds a Ph.D. in Second Language Studies from Indiana University, with a specialization in Second Language Assessment and a minor in Inquiry Methodology. Her research focuses on test validation, CEFR-aligned test development, standard setting, and differential item functioning.

References

Aryadoust, V., & Wong, J. (2026). How to train your dragon: Evaluating prompting and fine-tuning for GPT-based item generation in L2 listening assessment. Computers and Education: Artificial Intelligence, 10, 100623. https://doi.org/10.1016/j.caeai.2026.100623

Bax, S. (n.d.). Text Inspector. Retrieved July 5, 2026, from

https://textinspector.com/

Charttrakul, K., & Damnet, A. (2021). Role of the CEFR and English teaching in Thailand: A case study of Rajabhat Universities. Advances in Language and Literary Studies, 12(2), 82–89. https://journals.aiac.org.au/index.php/alls/article/view/6666

Council of Europe. (2001). Common European Framework of Reference for Languages: Learning, teaching, assessment. Cambridge University Press. https://rm.coe.int/1680459f97

Council of Europe. (2020). Common European Framework of Reference for Languages: Learning, teaching, assessment – Companion volume. Council of Europe Publishing. https://rm.coe.int/common-european-framework-of-reference-for-languages-learning-teaching/16809ea0d4

Roberts, R., Buchanan, H., & Pathare, E. (2015). Navigate intermediate B1+ coursebook with video and Oxford online skills. Oxford University Press.

Nation, I. S. P. (2012). The BNC/COCA word family lists.

https://www.wgtn.ac.nz/lals/resources/paul-nations-resources/paul-nations-publications/publications/documents/Information-on-the-BNC_COCA-word-family-lists.pdf

Nguyen, N. T., Huang, X., & Dang, T. N. Y. (2026). Lexical profile of

ChatGPT-generated reading materials targeting EFL learners across the CEFR levels. International Journal of TESOL Studies, 8(3), 113– 133. https://doi.org/10.58304/ijts.260211

OpenAI. (n.d.). ChatGPT (GPT-5.5) [Large language model]. Retrieved June

, 2026, from https://chatgpt.com

Phoolaikao, W., & Sukying, A. (2021). Insights into CEFR and its

implementation through the lens of preservice English teachers in Thailand. English Language Teaching, 14(6), 25–35.

https://doi.org/10.5539/elt.v14n6p25

Piamsai, C. (2023). Development and use of CEFR-based self-assessment in

English language learning and assessment in Thailand. PASAA, 66, 45–76. https://doi.org/10.58837/CHULA.PASAA.66.1.3

Shin, S.-Y. (2026). Proficiency scales. In C. A. Chapelle (Ed.), The Encyclopedia of applied linguistics (2nd ed.). Wiley. https://doi.org/10.1002/9781405198431.wbeal1423.pub2

Siripol, P., Rhee, S., Thirakunkovit, S., & Liang-Itsara, A. (2025). Evaluating the consistency of automated CEFR analyzers: a study of English language text classification. International Journal of Evaluation and Research in Education, 14(4), 3283–3294.

https://doi.org/10.11591/ijere.v14i4.33528

Styring, J., & Tims, N. (2019). Prepare level 4 student’s book (2nd ed.) Cambridge University Press.

Styring, J., Tims, N., & Chilton, H. (2019). Prepare level 7 student’s book (2nd ed.). Cambridge University Press.

Uchida, S. (2025). Generative AI and CEFR levels: Evaluating the accuracy of text generation with ChatGPT-4o through textual features. Vocabulary Learning and Instruction, 14(1), 1–13. https://doi.org/10.29140/vli.v14n1.2078

Waluyo, B. (2019). Thai first-year university students’ English proficiency on CEFR levels: A case study of Walailak University, Thailand. The New English Teacher, 13(2), 51–71. http://www.assumptionjournal.au.edu/index.php/newEnglishTeac her/article/view/3651

Wudthayagorn, J. (2022). An exploration of the English exit examination policy in Thai public universities. Language Assessment Quarterly, 19(2), 107–123. https://doi.org/10.1080/15434303.2021.1937174