1 LEXICAL IDENTITY AND DISCIPLINARY DIFFERENTIATION IN ENGLISH PROGRAM CURRICULA: A CORPUS-BASED KEYNESS ANALYSIS ACROSS 38 THAI RAJABHAT UNIVERSITIES
Main Article Content
Abstract
Curriculum design in higher education has traditionally relied on expert judgment rather than systematic linguistic analysis, leaving the lexical coherence and disciplinary identity of course descriptions largely unexamined. This study addresses that gap through a macro level corpus analysis of English course descriptions drawn from 38 Rajabhat universities across Thailand, spanning three professional tracks, namely General English (GE), Business English (BE), and English Education (ED). A corpus of 177,065 words compiled from 4,353 course description files was analyzed using keyness statistics, specifically Log-Likelihood, Log Ratio, and Juilland's D, with the British National Corpus and the Leipzig General Written English corpus serving as reference corpora. The findings show that 95 percent of the top 20 keywords differed significantly across curriculum types (p < .001). The BE curriculum was grounded in commercial discourse, the ED curriculum in pedagogical and institutional language, and the GE curriculum in humanities and translation related vocabulary, reflecting distinct English for Specific Purposes (ESP) orientations across the three tracks. A subsequent MANOVA confirmed a multivariate difference with a large effect size for the word cluster comprising "principles," "educational," and "management" (Pillai's Trace = 0.175, p < .001). Taken together, these results characterize the lexical differentiation underlying English curriculum discourse and offer a corpus-based framework for curriculum evaluation that can be applied in comparable English as a Foreign Language (EFL) context.
Article Details
References
Anthony, L. (2024). AntConc (Version 4.2.4) [Computer software]. Waseda University. https://www.laurenceanthony.net/software/antconc
Baker, P. (2004). Querying keywords: Questions of difference, frequency, and sense in keyword analysis. Journal of English Linguistics, 32(4), 346–359.
Biber, D., Conrad, S., & Reppen, R. (1998). Corpus linguistics: Investigating language structure and use. Cambridge University Press.
Biber, D., Reppen, R., Schnur, E., & Ghanem, R. (2016). On the (non)utility of Juilland's D to measure lexical dispersion in large corpora. International Journal of Corpus Linguistics, 21(4), 439–464. https://doi.org/10.1075/ijcl.21.4.01bib
BNC Consortium. (2007). The British National Corpus (Version 3, BNC XML edition). Oxford University Computing Services. http://www.natcorp.ox.ac.uk/
Charttrakul, K., & Damnet, A. (2021). Role of the CEFR and English teaching in Thailand: A case study of Rajabhat universities. Advances in Language and Literary Studies, 12(2), 82–89.
Coxhead, A. (2000). A new academic word list. TESOL Quarterly, 34(2), 213–238.
Curry, N., & McEnery, T. (2025). Corpus linguistics for language teaching and learning: A research agenda. Language Teaching, 58(2), 232–251. https://doi.org/10.1017/S0261444824000430
Dunning, T. (1993). Accurate methods for the statistics of surprise and coincidence. Computational Linguistics, 19(1), 61–74.
EF Education First. (2025). EF English Proficiency Index 2025. https://www.ef.com/wwen/epi/
Gabrielatos, C., & Marchi, A. (2012, September). Keyness: Appropriate metrics and practical issues [Paper presentation]. CADS International Conference, University of Bologna, Bologna, Italy.
Goldhahn, D., Eckart, T., & Quasthoff, U. (2012). Building large monolingual dictionaries at the Leipzig Corpora Collection: From 100 to 200 languages. In Proceedings of the Eighth International Conference on Language Resources and Evaluation (LREC’12) (pp. 759–765). European Language Resources Association.
Hardie, A. (2014). Log ratio: An informal introduction [Unpublished working paper]. ESRC Centre for Corpus Approaches to Social Science, Lancaster University. http://cass.lancs.ac.uk/log-ratio-an-informal-introduction/
Hyland, K. (2008). Academic discourse: English in a global context. Continuum.
Jantanukul, W. (2024). Rajabhat University concept for local development: Strategies, implementation, and community impact. Journal of Education and Learning Reviews, 1(6), 47–60.
Juilland, A., Brodin, D., & Davidovitch, D. (1970). Frequency dictionary of French words. Mouton.
Metang, P., & Narathakoon, A. (2025). A corpus-based study of lexical bundles of Keywords Found in Online News Articles. THAITESOL Journal, 38(1), 1–21.
Narkprom, N. (2021). A comparative corpus study on lexical bundles and keywords [Master's thesis, Thammasat University]. TU Digital Collections.
Olson, C. L. (1976). On choosing a test statistic in multivariate analysis of variance. Psychological Bulletin, 83(4), 579–586.
Rayson, P., & Garside, R. (2000). Comparing corpora using frequency profiling. In Proceedings of the ACL Workshop on Comparing Corpora (pp. 1–6). Association for Computational Linguistics.
Scott, M. (1997). PC analysis of key words – and key key words. System, 25(2), 233–245.
Scott, M., & Tribble, C. (2006). Textual patterns: Key words and corpus analysis in language education. John Benjamins.
Sriwichai, C., Duangfai, C., Paicharoen, N., & Yuensak, S. (2026). Insights into English proficiency: The analysis of test results among undergraduates in Thailand. Indonesian Journal of Applied Linguistics, 15(3), 630–644. https://doi.org/10.17509/hv44am08
Suriyapee, P., & Pongpairoj, N. (2022). Corpus to enhance verbal complements among low English proficiency Thai learners. PASAA: Journal of Language Teaching and Learning in Thailand, 63, 205–231. https://doi.org/
58837/CHULA.PASAA.63.1.8