Facebook pixel tracking

The Science and Information (SAI) Organization publishes open-access peer-reviewed journals in computer science and artificial intelligence.

Contact Info
Website thesai.org
Follow Us
Contact Info
Follow Us
Research Article | Open Access |

Evaluating ChatGPT for Grading Programming Assignments: Effectiveness, Fairness, and Student Perceptions

Author 1: Abedallah Zaid Abualkishik Author 2: Sherzod Turaev Author 3: Ali A. Alwan Author 4: Mohamed Elhoseny Author 5: Mohsin Murtaza
International Journal of Advanced Computer Science and Applications (IJACSA) · Vol. 17, No. 3 · Published 2026

DOI: https://doi.org/10.14569/IJACSA.2026.0170328

Abstract

This study investigates ChatGPT as an automated grading tool for programming assignments in higher education. Three datasets comprising Python, C++, and Java assignments were graded three times by ChatGPT and compared with faculty evaluations. Results show that ChatGPT achieves high grading accuracy, closely aligning with faculty scores and demonstrating statistically significant correlations. Statistical analyses using the Kolmogorov–Smirnov test, paired t-test, and Wilcoxon signed-rank test confirm overall agreement, although ChatGPT tends to apply stricter grading criteria. High intraclass correlation coefficients further indicate strong reliability and consistency across repeated grading attempts. The study highlights the critical role of well-defined rubrics in improving grading alignment and proposes an Instructor–AI Collaborative Rubric Development framework to support effective AI integration in assessment. A survey of 158 students indicates increased satisfaction and trust following disclosure of AI-assisted grading, although some still prefer human evaluation. Overall, the findings provide strong evidence that ChatGPT is a reliable and consistent grading tool, demonstrating close alignment with faculty evaluations and high reproducibility across attempts. However, its effectiveness is critically dependent on well-defined rubrics and requires human oversight to mitigate strictness, ensure fairness, and account for contextual nuances. These results strongly support a hybrid AI–human grading approach, grounded in transparent rubric design and reinforced by appropriate ethical safeguards.

Keywords

How to Cite this Article

Abualkishik, A. Z., Turaev, S., Alwan, A. A., Elhoseny, M., & Murtaza, M. (2026). Evaluating ChatGPT for Grading Programming Assignments: Effectiveness, Fairness, and Student Perceptions. International Journal of Advanced Computer Science and Applications, 17(3). https://doi.org/10.14569/IJACSA.2026.0170328

Abualkishik, Abedallah Zaid, et al.. "Evaluating ChatGPT for Grading Programming Assignments: Effectiveness, Fairness, and Student Perceptions." International Journal of Advanced Computer Science and Applications, vol. 17, no. 3, 2026, https://doi.org/10.14569/IJACSA.2026.0170328.

@article{Abualkishik2026,
  title     = {Evaluating ChatGPT for Grading Programming Assignments: Effectiveness, Fairness, and Student Perceptions},
  journal   = {International Journal of Advanced Computer Science and Applications},
  volume    = {17},
  number    = {3},
  year      = {2026},
  publisher = {The Science and Information Organization},
  author    = {Abedallah Zaid Abualkishik and Sherzod Turaev and Ali A. Alwan and Mohamed Elhoseny and Mohsin Murtaza},
  doi       = {10.14569/IJACSA.2026.0170328},
  url       = {https://doi.org/10.14569/IJACSA.2026.0170328}
}

Open Access — licensed under a Creative Commons Attribution 4.0 International License. Unrestricted use, distribution, and reproduction in any medium, even commercially, as long as the original work is properly cited.