generalizability-theory
Filtering by topic generalizability-theory(2)Clear all filters
- PaperLanguage Teaching Research16 Jul 2026
Comparing Teacher and Artificial Intelligence Scoring in Writing Assessment: A Generalizability Theory Analysis
Burak Asma
Compared teacher and AI scoring of middle school essays with and without rubrics using generalizability theory. AI tools showed higher consistency and better differentiation of individual differences, while teachers were influenced by biases and mood. Teachers acknowledged the potential of rubrics and AI feedback for more consistent results.
Original abstract
This study examined the use of artificial intelligence tools, which have garnered significant attention in recent years, in the assessment and evaluation processes of language education. For this purpose, student essays were scored by Turkish middle school teachers and artificial intelligence tools both with and without the use of a rubric, and the findings were evaluated based on generalizability theory. Additionally, the research findings were shared with participants to gather qualitative data, which were analysed using the inductive thematic analysis method to support the research results. The findings revealed that in evaluations conducted without a rubric, teachers were limited in their ability to distinguish individual differences and demonstrated low scoring consistency. In contrast, artificial intelligence tools were more effective in distinguishing individual differences and exhibited high consistency. In evaluations conducted using a rubric, scoring consistency increased in both groups, although, as in the first evaluation, artificial intelligence tools demonstrated a higher level of consistency. Regarding the research findings, teachers expressed that individual biases, mood, and professional experiences influenced their scoring processes and emphasized the potential of rubrics and artificial intelligence-supported feedback systems for achieving more consistent results. Artificial intelligence tools, on the other hand, highlighted their independence from subjective factors but stressed the need for more diverse and generalizable datasets to further enhance their evaluation capacities.
- PaperJournal of Second Language Writing14 May 2026
Evaluating ChatGPT-4o as an AI assessor in Chinese as a Second Language writing: Reliability through generalizability theory, feedback actionability, and teacher-student perceptions
Xiaosheng Zhou, Hanwei Wu, Ying Soon Goh
Evaluates ChatGPT-4o's reliability as an AI assessor for Chinese L2 writing using generalizability theory, examines feedback actionability, and explores teacher-student perceptions.