AI Chatbots as Sources of STD Information: A Study on Reliability and Readability.

Background: Artificial intelligence (AI) chatbots are increasingly used for medical inquiries, including sensitive topics like sexually transmitted diseases (STDs). However, concerns remain regarding the reliability and readability of the information they provide. This study aimed to assess the reli...

Descripción completa

Detalles Bibliográficos
Publicado en:Journal of Medical Systems Vol. 49; no. 1; pp. 1 - 11
Autores principales: Yıldız, Hüseyin Alperen, Söğütdelen, Emrullah
Formato: equations & formulas research tables/charts Journal Article
Publicado: Springer Nature 4/3/2025
Acceso en línea:Ver este registro en EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=ccm&AN=185070844&site=ehost-live
header:
  @attributes:
    shortDbName: ccm
    uiTerm: 185070844
    longDbName: CINAHL Complete
    uiTag: AN
  controlInfo:
    bkinfo:
    dissinfo:
    jinfo:
      jid:
        01485598
        4N0
      jtl: Journal of Medical Systems
      issn: 01485598
      maglogo: N
    pubinfo:
      dt: 4/3/2025
      vid: 49
      iid: 1
      pid: 237
      pub: Springer Nature
      place: New York, New York
    artinfo:
      ui:
        185070844
        185070844
        185070844
        10.1007/s10916-025-02178-z
        185070844
      ppf: 1
      ppct: 10
      formats:
        fmt:
          – @attributes:
              type: T
          – @attributes:
              type: P
      tig:
        atl: AI Chatbots as Sources of STD Information: A Study on Reliability and Readability.
      aug:
        au:
          Yıldız, Hüseyin Alperen
          Söğütdelen, Emrullah
        affil: https://ror.org/01x1kqx83 Faculty of Medicine, Bolu Abant İzzet Baysal University, 14030, Bolu, Türkiye
      sug:
        subj:
          Sexually Transmitted Diseases
          Health Information
          Information Resources
          Chatbot Evaluation
          Reliability Evaluation
          Readability Evaluation
          Patient Education Standards
          Teaching Materials Standards
          Human
          Health Literacy
          Public Health
          Communication
          Language
          Descriptive Statistics
          Benchmarking
          Scales
      ab: Background: Artificial intelligence (AI) chatbots are increasingly used for medical inquiries, including sensitive topics like sexually transmitted diseases (STDs). However, concerns remain regarding the reliability and readability of the information they provide. This study aimed to assess the reliability and readability of AI chatbots in providing information on STDs. The key objectives were to determine (1) the reliability of STD-related information provided by AI chatbots, and (2) whether the readability of this information meets the recommended standarts for patient education materials. Methods: Eleven relevant STD-related search queries were identified using Google Trends and entered into four AI chatbots: ChatGPT, Gemini, Perplexity, and Copilot. The reliability of the responses was evaluated using established tools, including DISCERN, EQIP, JAMA, and GQS. Readability was assessed using six widely recognized metrics, such as the Flesch-Kincaid Grade Level and the Gunning Fog Index. The performance of chatbots was statistically compared in terms of reliability and readability. Results: The analysis revealed significant differences in reliability across the AI chatbots. Perplexity and Copilot consistently outperformed ChatGPT and Gemini in DISCERN and EQIP scores, suggesting that these two chatbots provided more reliable information. However, results showed that none of the chatbots achieved the 6th-grade readability standard. All the chatbots generated information that was too complex for the general public, especially for individuals with lower health literacy levels. Conclusion: While Perplexity and Copilot showed better reliability in providing STD-related information, none of the chatbots met the recommended readability benchmarks. These findings highlight the need for future improvements in both the accuracy and accessibility of AI-generated health information, ensuring it can be easily understood by a broader audience.
      pubtype: Academic Journal
      doctype:
        equations & formulas
        research
        tables/charts
        Journal Article
      ougenre: Article
    language: English
    refInfo:
    holdings:
      @attributes:
        islocal: N