Generative AI in language and literature: impact on argumentative writing and process‑based academic integrity in urban public general basic education (Guayaquil)

Authors

DOI:

https://doi.org/10.64747/sfe6qg28

Keywords:

generative AI, argumentative writing, general basic education, academic integrity, Guayaquil

Abstract

This study investigates the impact of a short, classroom‑embedded generative AI (GAI) intervention—strictly limited to metacognitive functions of ideation, diagnosis, and revision—on the quality of argumentative writing in Ecuador’s urban public General Basic Education (EGB). We implemented a cluster quasi‑experimental pretest–posttest design with an active control. The final sample comprised 8 classrooms (N = 236). The treatment group completed three GAI‑assisted sessions without machine‑generated final prose; the control followed comparable traditional practices. Products were double‑blind rated with a four‑dimension analytic rubric (thesis/focus, evidence/counterargumentation, cohesion‑coherence, conventions/voice). We also evaluated a process‑based academic integrity protocol requiring disclosure of use, traceable drafts, and documented prompts. ANCOVA with cluster‑robust errors and linear mixed models indicated a moderate advantage for the intervention on adjusted total scores (d ≈ 0.48; adjusted mean difference ≈ 1.26 points, 95% CI [0.68, 1.84]; p < .001). By dimension, the largest gains appeared in cohesion/coherence and evidence/counterargumentation. The group × baseline performance interaction was significant, with stronger benefits for the lowest tercile (d ≈ 0.62). Perceptually, the treatment reported lower cognitive load and higher perceived usefulness of feedback. Regarding integrity, the protocol correlated with a decrease in unattributed textual overlap at posttest (Δdiff ≈ −2.2 percentage points; p = .002) relative to control. We conclude that positioning GAI as a Socratic mentor that externalizes metacognition—without authoring students’ final text—enhances argumentative quality while supporting responsible authorship. For school systems with large classes, this approach is feasible and scalable when coupled with explicit rubrics and transparent process protocols.

References

Alnemrat, A., Aldamen, H., Almashour, M., Al-Deaibes, M., & AlSharefeen, R. (2025). AI vs. teacher feedback on EFL argumentative writing: A quantitative study. Frontiers in Education, 10, 1614673. https://doi.org/10.3389/feduc.2025.1614673

Arce, C. M., Ureña, R., & Játiva, J. (2025). Attitudes toward AI and dependency among Ecuadorian university students: A predictive model. Sustainability, 17(17), 7741. https://doi.org/10.3390/su17177741

Ardito, C. G., Gopal, T., Ceppini, E., & Burnett, P. C. (2024). Generative AI detection in higher education assessments. New Directions for Teaching and Learning, 2024(180), 39–54. https://doi.org/10.1002/tl.20624

Baldeón Medina, P. J., González Yagual, T. B., Izquierdo Corozo, C. D., & Villalba Vélez, K. M. (2025). Aplicación de IA en Educación Básica para fortalecer el aprendizaje y la motivación escolar. Horizonte Científico International Journal, 3(2), 1–9. https://doi.org/10.64747/dbkg2w76

Banihashem, S. K., Nejati, R., Sadeghi, M. R., & Babaii, E. (2024). Feedback sources in essay writing: Peer-generated or AI-generated? International Journal of Educational Technology in Higher Education, 21, 26. https://doi.org/10.1186/s41239-024-00455-4

Bland, J. M., & Altman, D. G. (1986). Statistical methods for assessing agreement between two methods of clinical measurement. The Lancet, 327(8476), 307–310. https://doi.org/10.1016/S0140-6736(86)90837-8

Buele, J., Sabando-García, Á. R., Sabando-García, B. J., & Yánez-Rueda, H. (2025). Ethical use of generative artificial intelligence among Ecuadorian university students. Sustainability, 17(10), 4435. https://doi.org/10.3390/su17104435

Castro Macías, N. M. (2025). Aplicación de la inteligencia artificial como recurso para desarrollar pensamiento crítico en clases de Física. Horizonte Científico International Journal, 3(2), 1–19. https://doi.org/10.64747/7wvch719

Doshi, A. R., Yamauchi, T., Salehi, N., & Bernstein, M. S. (2024). Generative AI enhances individual creativity but reduces collaboration quality. Science Advances, 10(44), eadn5290. https://doi.org/10.1126/sciadv.adn5290

Faul, F., Erdfelder, E., Lang, A.-G., & Buchner, A. (2007). GPower 3: A flexible statistical power analysis program for the social, behavioral, and biomedical sciences. Behavior Research Methods, 39*(2), 175–191. https://doi.org/10.3758/BF03193146

Fleckenstein, J., Liebenow, L. W., & Meyer, J. (2023). Automated feedback and writing: A multi‑level meta‑analysis of effects on students’ performance. Frontiers in Artificial Intelligence, 6, 1162454. https://doi.org/10.3389/frai.2023.1162454

Galeas Gaibor, M. L., Meza Bravo, D. M., Ocampo Andrade, W. F., & Romero Arias, J. B. (2025). Aplicación de inteligencia artificial y gamificación para fortalecer el aprendizaje en estudiantes de básica en contextos rurales. Horizonte Científico International Journal, 3(2), 1–12. https://doi.org/10.64747/f718q395

Hattie, J., & Timperley, H. (2007). The power of feedback. Review of Educational Research, 77(1), 81–112. https://doi.org/10.3102/003465430298487

Huang, Y., Wu, C., & Lin, S. (2025). Exploring the effectiveness of large‑scale automated writing evaluation in education: A systematic review. Journal of Computer Assisted Learning, 41(3), 987–1005. https://doi.org/10.1111/jcal.70009

Liang, W., Yuksekgonul, M., Mao, Y., Wu, E., & Zou, J. (2023). GPT detectors are biased against non‑native English writers. Patterns, 4(7), 100779. https://doi.org/10.1016/j.patter.2023.100779

Nature Machine Intelligence Editorial. (2023). The AI writing on the wall. Nature Machine Intelligence, 5(1), 1. https://doi.org/10.1038/s42256-023-00613-9

Nguyen, A., & Wang, X. (2024). Human–AI collaboration patterns in AI‑assisted academic writing. Studies in Higher Education, 49(10), 1991–2012. https://doi.org/10.1080/03075079.2024.2323593

Nordstokke, D. W., & Zumbo, B. D. (2010). A new nonparametric Levene test for equal variances. Psicologica, 31(2), 401–430. https://doi.org/10.2478/v10053-008-0081-1

Shrout, P. E., & Fleiss, J. L. (1979). Intraclass correlations: Uses in assessing rater reliability. Psychological Bulletin, 86(2), 420–428. https://doi.org/10.1037/0033-2909.86.2.420

The jamovi project. (2022). jamovi (Version 2.3) [Computer software]. https://doi.org/10.17605/OSF.IO/GXRQ9

Wisniewski, B., Zierer, K., & Hattie, J. (2020). The power of feedback revisited: A meta‑analysis of the effects of feedback in educational contexts. Educational Research Review, 30, 100331. https://doi.org/10.1016/j.edurev.2020.100331

Wilson, J., Zhang, F., Palermo, C., Cruz Cordero, T., Myers, M. C., Eacker, H., Potter, A., & Coles, J. (2024). Predictors of middle school students’ perceptions of automated writing evaluation. Computers & Education, 211, 104985. https://doi.org/10.1016/j.compedu.2023.104985

Xue, Y., Wang, Z., & Li, J. (2024). Towards automated writing evaluation: A comprehensive bibliometric review. Education and Information Technologies, 29(7), 8895–8919. https://doi.org/10.1007/s10639-024-12596-0

Zhai, N., Xie, F., Zou, D., & Wang, F. L. (2023). The effectiveness of automated writing evaluation on students’ writing: A meta‑analysis. Journal of Educational Computing Research, 61(2), 363–394. https://doi.org/10.1177/07356331221127300

Zhang, K. (2025). Enhancing critical writing through AI feedback: A randomized control study. Behavioral Sciences, 15(5), 600. https://doi.org/10.3390/bs15050600

Downloads

Published

2025-03-10

How to Cite

Tenecela Calderón, M. E. (2025). Generative AI in language and literature: impact on argumentative writing and process‑based academic integrity in urban public general basic education (Guayaquil). Horizonte Cientifico Educativo International Journal, 1(1), 20-36. https://doi.org/10.64747/sfe6qg28

Most read articles by the same author(s)