Exploring an effective automated grading model with reliability detection for large‐scale online peer assessment-Reference-Cited by-同舟云学术

Exploring an effective automated grading model with reliability detection for large‐scale online peer assessment

Published:2024-03-12 Issue:4 Volume:40 Page:1535-1551
ISSN:0266-4909
Container-title:Journal of Computer Assisted Learning
language:en
Short-container-title:Computer Assisted Learning

Author:

Lin Zirou¹,Yan Hanbing²,Zhao Li³^ORCID

Affiliation:

1. Department of Educational Information Technology East China Normal University Shanghai China

2. School of Teacher Development East China Normal University Shanghai China

3. School of Education Science Nanjing Normal University Nanjing China

Abstract

AbstractBackgroundPeer assessment has played an important role in large‐scale online learning, as it helps promote the effectiveness of learners' online learning. However, with the emergence of numerical grades and textual feedback generated by peers, it is necessary to detect the reliability of the large amount of peer assessment data, and then develop an effective automated grading model to analyse the data and predict learners' learning results.ObjectivesThe present study aimed to propose an automated grading model with reliability detection.MethodsA total of 109,327 instances of peer assessment from a large‐scale teacher online learning program were tested in the experiments. The reliability detection approach included three steps: recurrent convolutional neural networks (RCNN) was used to detect grade consistency, bidirectional encoder representations from transformers (BERT) was used to detect text originality, and long short‐term memory (LSTM) was used to detect grade‐text consistency. Furthermore, the automated grading was designed with the BERT‐RCNN model.Results and ConclusionsThe effectiveness of the automated grading model with reliability detection was shown. For reliability detection, RCNN performed best in detecting grade consistency with an accuracy rate of 0.889, BERT performed best in detecting text originality with an improvement of 4.47% compared to the benchmark model, and LSTM performed best with an accuracy rate of 0.883. Moreover, the automated grading model with reliability detection achieved good performance, with an accuracy rate of 0.89. Compared to the absence of reliability detection, it increased by 12.1%.ImplicationsThe results strongly suggest that the automated grading model with reliability detection for large‐scale peer assessment is effective, with the following implications: (1) The introduction of reliability detection is necessary to help filter out low reliability data in peer assessment, thus promoting effective automated grading results. (2) This solution could assist assessors in adjusting the exclusion threshold of peer assessment reliability, providing a controllable automated grading tool to reducing manual workload with high quality. (3) This solution could shift educational institutions from labour‐intensive grading procedures to a more efficient educational assessment pattern, allowing for more investment in supporting instructors and learners to improve the quality of peer feedback.

Funder

National Social Science Fund of China

Publisher

Wiley

Link

https://onlinelibrary.wiley.com/doi/pdf/10.1111/jcal.12970

Reference98 articles.

1. An adaptable and personalised E-learning system applied to computer science Programmes design

2. How automated feedback through text mining changes plagiaristic behavior in online assignments

3. The role of quality factors in supporting self-regulated learning (SRL) skills in MOOC environment

4. The emergent role of artificial intelligence, natural learning processing, and large language models in higher education and research

5. Peer-assessment in higher education – twenty-first century practices, challenges and the way forward