Home > Research > Publications & Outputs > Application of cluster analysis to identify dif...

Associated organisational unit

Electronic data

  • MaCainUshakovaC_E2024

    Accepted author manuscript, 1 MB, PDF document

    Available under license: CC BY: Creative Commons Attribution 4.0 International License

Links

Text available via DOI:

View graph of relations

Application of cluster analysis to identify different reader groups through their engagement with a digital reading supplement

Research output: Contribution to Journal/MagazineJournal articlepeer-review

Published
Article number105025
<mark>Journal publication date</mark>1/06/2024
<mark>Journal</mark>Computers and Education
Volume214
Publication StatusPublished
Early online date29/02/24
<mark>Original language</mark>English

Abstract

The focus of this study is the identification of reader profiles that differ in performance and progression in an educational literacy app. A total of 19,830 students in Grade 2 from 347 Elementary schools located in 30 different districts in the United States played the app from 2020 to 2021. Our aim was to identify unique groups of readers using an unsupervised statistical learning technique - cluster analysis. Six indicators generated from the students’ log files were included to provide insights into engagement and learning across four different reading-related skills: phonological awareness, early decoding, vocabulary, and comprehension processes. A key aim was to evaluate the implementation and performance of Gaussian mixture models, k-means, k-medoids, clustering large applications and hierarchical clustering, alongside provision of detailed guidance that can benefit researchers in the field. K-means algorithm performed the best and identified nine groups of readers. Children with low initial reading ability showed greater engagement with code-related games (phonological awareness, early decoding) and took longer to master these games, whereas children with higher initial ability showed more engagement with meaning-related games (vocabulary, comprehension processes). Our findings can inform further research that aims to understand individual differences in learning behaviour within digital environments both over time and across various cohorts of children.