AI Detection Tools: A Systematic Review of Empirical Evidence and Their Implications for Education.

The rapid advancement of generative artificial intelligence has significantly transformed academic writing practices, prompting institutions to implement tools designed to verify authorship and uphold academic integrity. Artificial intelligence detection systems have emerged as a prominent, albeit i...

Descripción completa

Detalles Bibliográficos
Publicado en:Khazar Journal of Humanities & Social Sciences Vol. 28; no. 4; pp. 234 - 247
Autor principal: CANER, Halime Nuran
Formato: Artículo
Publicado: Khazar University Press 2025
Materias:
Acceso en línea:Ver este registro en EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=191829331&site=ehost-live
header:
  @attributes:
    shortDbName: hlh
    uiTerm: 191829331
    longDbName: Humanities International Complete
    uiTag: AN
  controlInfo:
    bkinfo:
    jinfo:
      jid:
        22232613
        DRXR
      jtl: Khazar Journal of Humanities & Social Sciences
      issn: 22232613
      maglogo: N
    pubinfo:
      dt: 2025
      vid: 28
      iid: 4
      pid: 73448
      pub: Khazar University Press
    artinfo:
      ui:
        191829331
        10.5782/2223-2621.1580
      ppf: 234
      ppct: 13
      formats:
      tig:
        atl: AI Detection Tools: A Systematic Review of Empirical Evidence and Their Implications for Education.
      aug:
        au: CANER, Halime Nuran
        affil: Akdeniz University, Antalya, TURKIYE
      su:
        Generative artificial intelligence
        Algorithmic bias
        Educational evaluation
        Academia
        Software development tools
        Education ethics
      sug:
        subj:
          Generative artificial intelligence
          Algorithmic bias
          Educational evaluation
          Academia
          Software development tools
          Education ethics
      keyword:
        Academic integrity
        AI detection tools
        Authorship verification
        Generative AI
        PRISMA
        Systematic review
      ab: The rapid advancement of generative artificial intelligence has significantly transformed academic writing practices, prompting institutions to implement tools designed to verify authorship and uphold academic integrity. Artificial intelligence detection systems have emerged as a prominent, albeit increasingly debated response to these challenges. This systematic review synthesizes empirical evidence to assess the reliability, fairness, and pedagogical implications of artificial intelligence text-detection tools in educational settings. Adhering to PRISMA 2020 standards, this review identified twenty-five peer-reviewed empirical studies, 18 of which were conducted directly within educational settings. This review synthesizes empirical studies published between 2022 and 2025, encompassing quantitative, qualitative, and mixed-methods designs across diverse disciplinary and linguistic contexts. The findings indicate that AI text detection tools are unsuitable for high-stakes academic integrity decisions in their current form. Furthermore, there is substantial variability and instability in detection accuracy across tools, genres, and linguistic backgrounds; a noticeable weakness in paraphrasing, translation, and other adversarial techniques; and systemic biases that disproportionately aect non-native English writers. Human judgment was also found to be inconsistent, reinforcing the difficulty in reliably distinguishing AI-generated text from human-authored text. Collectively, these results raise significant ethical, pedagogical and institutional concerns. This review underscores the need for integrity strategies that prioritize transparency, AI literacy, fairness-aware design, and process-based assessment rather than relying on detection-centered approaches. The findings suggest the necessity of hybrid approaches that combine watermarking and fairness-aware detection algorithms with process-oriented assessment, AI literacy initiatives, and cross-linguistic benchmarking, alongside interpretability-focused and longitudinal research on students' perceptions of AI detection.
      pubtype: Academic Journal
      doctype: Article
      src: R
    language: English
    refInfo:
    copyright:
      @attributes:
        flag: Y
      dt:
        @attributes:
          year: 2025
    holdings:
      @attributes:
        islocal: N