Maintenance Notice

Due to necessary scheduled maintenance, the JMIR Publications website will be unavailable from Wednesday, July 01, 2020 at 8:00 PM to 10:00 PM EST. We apologize in advance for any inconvenience this may cause you.

Who will be affected?

Previously submitted to: JMIR AI (no longer under consideration since Sep 17, 2025)

Date Submitted: Jun 3, 2024

Warning: This is an author submission that is not peer-reviewed or edited. Preprints - unless they show as "accepted" - should not be relied on to guide clinical practice or health-related behavior and should not be reported in news media as established information.

Use of Conversational AI Chatbots for Retinitis Pigmentosa Patients: a Semantic Textual Analysis

  • Milan Del Buono; 
  • Gloria Wu; 
  • Adrial Wong

ABSTRACT

Background:

Retinitis Pigmentosa (RP) is a genetic disease that causes progression vision loss, which almost invariably leads to blindness. Being a rare disease, RP is costly to diagnose and takes a long time to diagnose. Patients may inevitably turn to ChatGPT or similar large language AI models (LLMs) to answer some of their questions, however, given the amount of incorrect information on the internet combined with small training data available for RP, the answers may not be correct.

Objective:

We sought to investigate how correct 4 LLMs’ (ChatGPT 3.5 and 4.0, Claude, and Copilot) answers were for RP by comparing them semantically via cosine score to the American Academy of Opthalmology’s (AAO) webpage on RP.

Methods:

Embeddings for the LLM's outputs were computed using the MiniLM-v2 model from HuggingFace and the cosine between the AAO's sentence and the LLM's sentence were taken. To summarize the data, New Dale-Chall readability scores were calculated.

Results:

We find that the LLMs answer the questions reasonably similarly to the AAO, and there was no significant difference between ChatGPT 3.5, 4.0, and Claude; Copilot had lower cosine scores. However, every LLM had a significantly harder readability level, indicating that LLM outputs may be difficult for the lay public to comprehend.

Conclusions:

This study demonstrates that AI models can produce conversational text that is highly accurate. Their high semantic similarity to the AAO's official website indicates that they provide good background information to curious patients. LLM's future use in patient education cannot be ignored, as AI technology is further adopted.


 Citation

Please cite as:

Del Buono M, Wu G, Wong A

Use of Conversational AI Chatbots for Retinitis Pigmentosa Patients: a Semantic Textual Analysis

JMIR Preprints. 03/06/2024:62873

DOI: 10.2196/preprints.62873

URL: https://preprints.jmir.org/preprint/62873

PDF not available

The author of this paper has made a PDF available, but requires the user to login, or create an account.

© The authors. All rights reserved. This is a privileged document currently under peer-review/community review (or an accepted/rejected manuscript). Authors have provided JMIR Publications with an exclusive license to publish this preprint on it's website for review and ahead-of-print citation purposes only. While the final peer-reviewed paper may be licensed under a cc-by license on publication, at this stage authors and publisher expressively prohibit redistribution of this draft paper other than for review purposes.