Currently submitted to: JMIR Medical Informatics
Date Submitted: May 25, 2026
Open Peer Review Period: Jun 5, 2026 - Jul 31, 2026
(closed for review but you can still tweet)
NOTE: This is an unreviewed Preprint
Warning: This is a unreviewed preprint (What is a preprint?). Readers are warned that the document has not been peer-reviewed by expert/patient reviewers or an academic editor, may contain misleading claims, and is likely to undergo changes before final publication, if accepted, or may have been rejected/withdrawn (a note "no longer under consideration" will appear above).
Peer review me: Readers with interest and expertise are encouraged to sign up as peer-reviewer, if the paper is within an open peer-review period (in this case, a "Peer Review Me" button to sign up as reviewer is displayed above). All preprints currently open for review are listed here. Outside of the formal open peer-review period we encourage you to tweet about the preprint.
Citation: Please cite this preprint only for review purposes or for grant applications and CVs (if you are the author).
Final version: If our system detects a final peer-reviewed "version of record" (VoR) published in any journal, a link to that VoR will appear below. Readers are then encourage to cite the VoR instead of this preprint.
Settings: If you are the author, you can login and change the preprint display settings, but the preprint URL/DOI is supposed to be stable and citable, so it should not be removed once posted.
Submit: To post your own preprint, simply submit to any JMIR journal, and choose the appropriate settings to expose your submitted version as preprint.
Warning: This is an author submission that is not peer-reviewed or edited. Preprints - unless they show as "accepted" - should not be relied on to guide clinical practice or health-related behavior and should not be reported in news media as established information.
Challenges of Generative Artificial Intelligence for Patients With Hypertension and Diabetes: A Cross-Platform Network Analysis and Latent Profile Analysis of the Quality, Readability, and Actionability of Texts Generated by Large Language Models
ABSTRACT
Background:
Patients with comorbid hypertension and diabetes require sustained self-management support and multidisciplinary care. Although generative artificial intelligence may support chronic disease education, quantitative evidence on the clinical quality and actionability of its outputs remains limited. This cross-sectional study compared the quality, readability, and actionability of patient education materials generated by 16 large language models.
Objective:
This study aimed to compare the quality, readability, understandability, and actionability of patient education materials generated by 16 large language models for patients with comorbid hypertension and diabetes.
Methods:
We developed an 80-item guideline-based question bank and submitted each item to 16 large language models, yielding 1,280 texts. Under blinded conditions, the texts were independently evaluated using the Patient Education Materials Assessment Tool for Printable Materials, the Ensuring Quality Information for Patients, 36-item version, the Global Quality Score, and seven readability formulas. Kruskal–Wallis tests, extended Bayesian information criterion graphical least absolute shrinkage and selection operator partial correlation network analysis, and latent profile analysis were performed.
Results:
Median understandability on the Patient Education Materials Assessment Tool for Printable Materials was moderate to high (75.0), whereas actionability remained low (20.0). Network analysis revealed a trade-off between information quality and accessibility: longer texts tended to score higher on the Ensuring Quality Information for Patients, 36-item version but imposed a greater reading burden, and higher Simple Measure of Gobbledygook scores were associated with lower actionability. Latent profile analysis identified five latent profiles; only 19.2% of texts were classified as “high-quality, comprehensive-guidance type.” Profile distribution differed significantly across platforms (p < 0.001), with actionability showing the clearest between-profile separation.
Conclusions:
Mainstream large language models can explain core medical knowledge for this comorbidity, but they remain limited in generating explicit, actionable behavioral guidance. More comprehensive content often comes at the cost of readability. Future digital nursing tools should incorporate multidisciplinary collaboration and behavior-change theory to move generative artificial intelligence from information delivery toward clinically actionable intervention.
Citation
Request queued. Please wait while the file is being generated. It may take some time.
Copyright
© The authors. All rights reserved. This is a privileged document currently under peer-review/community review (or an accepted/rejected manuscript). Authors have provided JMIR Publications with an exclusive license to publish this preprint on it's website for review and ahead-of-print citation purposes only. While the final peer-reviewed paper may be licensed under a cc-by license on publication, at this stage authors and publisher expressively prohibit redistribution of this draft paper other than for review purposes.