Beyond Accuracy: How EFL Students Navigate GPT-Generated Feedback in Argumentative Writing across Proficiency Levels

Authors

  • Moh. Shofi Zuhri Universitas Tarbiyatut Tholabah Lamongan, Indonesia
  • Moh. Kavin Lidinillah

DOI:

https://doi.org/10.58518/jelp.v5i2.5578

Keywords:

Automated Grammatical Error, GPT, EFL Writing

Abstract

Essay-writing proficiency is an important indicator of English language competence among English as a Foreign Language (EFL) learners. However, providing written corrective feedback manually can place considerable demands on instructors’ time and workload, particularly in large classes. This study evaluates the performance of a Generative Pre-trained Transformer (GPT) model in detecting and correcting grammatical and lexical errors in EFL students’ argumentative essays, while also exploring how students perceive and respond to the feedback it provides. An exploratory descriptive qualitative design was employed with 15 English Language Education students from three semester levels (1, 3, and 5). Essay drafts and semi-structured interview transcripts were analyzed using the Error Analysis framework (Corder, 1981; Ellis, 2008) and Thematic Analysis (Braun & Clarke, 2006). The findings show that GPT achieved its highest correction accuracy in Semester 1 students’ essays (78%), whereas the rate of over-correction increased to 20% in Semester 5 as sentence structures became more complex. The qualitative findings reveal a clear difference in students’ metacognitive responses. Semester 1 students tended to exhibit automation bias by accepting GPT-generated corrections with little critical evaluation, while Semester 3 and 5 students demonstrated greater learner agency by evaluating, filtering, and negotiating the suggested corrections. Overall, the findings suggest that GPT can serve as a useful form of scaffolding for formative feedback, but its integration into EFL writing pedagogy should be accompanied by critical AI literacy to help students maintain their metalinguistic awareness and authorial voice

References

Braun, V., & Clarke, V. (2006). Using thematic analysis in psychology. Qualitative Research in Psychology, 3(2), 77–101. https://doi.org/10.1191/1478088706qp063oa

Bryant, C., Felice, M., Andersen, O. E., & Briscoe, T. (2019). The BEA-2019 shared task on grammatical error correction. In Proceedings of the Fourteenth Workshop on Innovative Use of NLP for Building Educational Applications (pp. 52–75). Association for Computational Linguistics. https://doi.org/10.18653/v1/W19-4406

Corder, S. P. (1981). Error analysis and interlanguage. Oxford University Press.

Creswell, J. W., & Poth, C. N. (2018). Qualitative inquiry and research design: Choosing among five approaches (4th ed.). SAGE Publications.

Ellis, R. (2008). The study of second language acquisition (2nd ed.). Oxford University Press.

Ferris, D. R. (2002). Treatment of error in second language student writing. University of Michigan Press. https://doi.org/10.3998/mpub.9059

Hyland, K. (2003). Second language writing. Cambridge University Press. https://doi.org/10.1017/CBO9780511667251

Long, H. S. (2024). Exploring the use of ChatGPT as a tool for written corrective feedback in an EFL classroom. The Journal of AsiaTEFL, 21(2), 397–412. https://doi.org/10.18823/asiatefl.2024.21.2.8.397

Long, M. H. (1996). The role of the linguistic environment in second language acquisition. In W. C. Ritchie & T. K. Bhatia (Eds.), Handbook of second language acquisition (pp. 413–468). Academic Press. https://doi.org/10.1016/B978-012588825-7/50015-3

Mercer, S. (2011). Understanding learner agency as a complex dynamic system. System, 39(4), 427–436. https://doi.org/10.1016/j.system.2011.08.001

Nagata, N., & Nakatani, K. (2010). The effectiveness of intelligent computer-assisted language learning in learning Japanese grammar. Computer Assisted Language Learning, 23(4), 291–302. https://doi.org/10.1080/09588221.2010.502018

Nurjati, N., Syahria, N., & Khan, A. K. B. S. (2025). ChatGPT-based reflective feedback to improve EFL graduate students’ academic writing for publication. JOLLT Journal of Languages and Language Teaching, 13(4), 1802–1816. https://doi.org/10.33394/jollt.v13i4.14858

Parasuraman, R., & Manzey, D. H. (2010). Complacency and bias in human interaction with automated systems: An integrated review. Human Factors, 52(3), 381–410. https://doi.org/10.1177/0018720810376055

Prasetya, R., Syarif, A., & Mayof, A. (2026). Comparative analysis of ChatGPT and Grammarly in supporting grammar correction and writing development in EFL students. Journal of English and Education, 12(1), 83–101. https://doi.org/10.20885/jee.v12i1.44160

Puteri, C. G., Cahyono, B. Y., Widiati, U., & Suryati, N. (2026). Using ChatGPT as a writing assistant: A study on essay quality development among Indonesian EFL university students. Indonesian Journal of Applied Linguistics, 15(2), 332–346. https://doi.org/10.17509/3ebmy158

Romeo, G., & Conti, D. (2025). Exploring automation bias in human–AI collaboration: A review and implications for explainable AI. AI & SOCIETY. https://doi.org/10.1007/s00146-025-02422-7

Rozovskaya, A., & Roth, D. (2014). Building a state-of-the-art grammatical error correction system. Transactions of the Association for Computational Linguistics, 2, 419–434. https://doi.org/10.1162/tacl_a_00192

Selinker, L. (1972). Interlanguage. International Review of Applied Linguistics in Language Teaching, 10(1–4), 209–232. https://doi.org/10.1515/iral.1972.10.1-4.209

Sidoti, O., Park, E., & Gottfried, J. (2025). About a quarter of U.S. teens have used ChatGPT for schoolwork – double the share in 2023. Pew Research Center. https://www.pewresearch.org/short-reads/2025/01/15/about-a-quarter-of-us-teens-have-used-chatgpt-for-schoolwork-double-the-share-in-2023/

Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., & Polosukhin, I. (2017). Attention is all you need. In Advances in Neural Information Processing Systems 30 (NIPS 2017) (pp. 5998–6008). Curran Associates, Inc.

Vygotsky, L. S. (1978). Mind in society: The development of higher psychological processes. Harvard University Press.

Yan, D., & Zhang, S. (2024). L2 writer engagement with automated written corrective feedback provided by ChatGPT: A mixed-method multiple case study. Humanities and Social Sciences Communications, 11, 1086. https://doi.org/10.1057/s41599-024-03543-y

Zhang, Y., Zong, C., & Li, J. (2020). Neural grammatical error correction: A survey. IEEE/ACM Transactions on Audio, Speech, and Language Processing, 28, 2744–2760. https://doi.org/10.1109/TASLP.2020.3023023

Downloads

Published

2026-09-24

How to Cite

Beyond Accuracy: How EFL Students Navigate GPT-Generated Feedback in Argumentative Writing across Proficiency Levels. (2026). JELP Journal of English Language and Pedagogy, 5(2), 141-154. https://doi.org/10.58518/jelp.v5i2.5578