AI Detection Tools: A Systematic Review of Empirical Evidence and Their Implications for Education.
The rapid advancement of generative artificial intelligence has significantly transformed academic writing practices, prompting institutions to implement tools designed to verify authorship and uphold academic integrity. Artificial intelligence detection systems have emerged as a prominent, albeit i...
| Publicado en: | Khazar Journal of Humanities & Social Sciences Vol. 28; no. 4; pp. 234 - 247 |
|---|---|
| Autor principal: | |
| Formato: | Artículo |
| Publicado: |
Khazar University Press
2025
|
| Materias: | |
| Acceso en línea: | Ver este registro en EBSCOhost |
| fields | @attributes: recordID: 1 pdfLink: plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=191829331&site=ehost-live header: @attributes: shortDbName: hlh uiTerm: 191829331 longDbName: Humanities International Complete uiTag: AN controlInfo: bkinfo: jinfo: jid: 22232613 DRXR jtl: Khazar Journal of Humanities & Social Sciences issn: 22232613 maglogo: N pubinfo: dt: 2025 vid: 28 iid: 4 pid: 73448 pub: Khazar University Press artinfo: ui: 191829331 10.5782/2223-2621.1580 ppf: 234 ppct: 13 formats: tig: atl: AI Detection Tools: A Systematic Review of Empirical Evidence and Their Implications for Education. aug: au: CANER, Halime Nuran affil: Akdeniz University, Antalya, TURKIYE su: Generative artificial intelligence Algorithmic bias Educational evaluation Academia Software development tools Education ethics sug: subj: Generative artificial intelligence Algorithmic bias Educational evaluation Academia Software development tools Education ethics keyword: Academic integrity AI detection tools Authorship verification Generative AI PRISMA Systematic review ab: The rapid advancement of generative artificial intelligence has significantly transformed academic writing practices, prompting institutions to implement tools designed to verify authorship and uphold academic integrity. Artificial intelligence detection systems have emerged as a prominent, albeit increasingly debated response to these challenges. This systematic review synthesizes empirical evidence to assess the reliability, fairness, and pedagogical implications of artificial intelligence text-detection tools in educational settings. Adhering to PRISMA 2020 standards, this review identified twenty-five peer-reviewed empirical studies, 18 of which were conducted directly within educational settings. This review synthesizes empirical studies published between 2022 and 2025, encompassing quantitative, qualitative, and mixed-methods designs across diverse disciplinary and linguistic contexts. The findings indicate that AI text detection tools are unsuitable for high-stakes academic integrity decisions in their current form. Furthermore, there is substantial variability and instability in detection accuracy across tools, genres, and linguistic backgrounds; a noticeable weakness in paraphrasing, translation, and other adversarial techniques; and systemic biases that disproportionately aect non-native English writers. Human judgment was also found to be inconsistent, reinforcing the difficulty in reliably distinguishing AI-generated text from human-authored text. Collectively, these results raise significant ethical, pedagogical and institutional concerns. This review underscores the need for integrity strategies that prioritize transparency, AI literacy, fairness-aware design, and process-based assessment rather than relying on detection-centered approaches. The findings suggest the necessity of hybrid approaches that combine watermarking and fairness-aware detection algorithms with process-oriented assessment, AI literacy initiatives, and cross-linguistic benchmarking, alongside interpretability-focused and longitudinal research on students' perceptions of AI detection. pubtype: Academic Journal doctype: Article src: R language: English refInfo: copyright: @attributes: flag: Y dt: @attributes: year: 2025 holdings: @attributes: islocal: N |
|---|