Maintenance Notice

Due to necessary scheduled maintenance, the JMIR Publications website will be unavailable from Wednesday, July 01, 2020 at 8:00 PM to 10:00 PM EST. We apologize in advance for any inconvenience this may cause you.

Who will be affected?

Accepted for/Published in: Journal of Medical Internet Research

Date Submitted: Dec 22, 2025
Date Accepted: Jun 18, 2026

The final, peer-reviewed published version of this preprint can be found here:

Effectiveness of ChatGPT and DeepSeek in Urology Medical Education: Randomized Controlled Trial

Niu H, Yang W, Xu T, Wei J, Zheng W, Yan W, Wang J

Effectiveness of ChatGPT and DeepSeek in Urology Medical Education: Randomized Controlled Trial

J Med Internet Res 2026;28:e89315

DOI: 10.2196/89315

PMID: 42766393

PMCID: 13592387

Effectiveness of Generative AI Tools in Urology Medical Education: A Randomized Controlled Trial Comparing ChatGPT and DeepSeek

  • Haitao Niu; 
  • Wentong Yang; 
  • Ting Xu; 
  • Junjie Wei; 
  • Wentao Zheng; 
  • Wenbo Yan; 
  • Jingkai Wang

ABSTRACT

Background:

Background:

Since its release in November 2022, generative artificial intelligence (GenAI) tools, including ChatGPT, have gained widespread attention across various sectors, including medical education.

Objective:

Objective:

This study seeks to examine the effectiveness and feasibility of generative AI tools (ChatGPT and DeepSeek) in enhancing urology teaching outcomes for medical undergraduates.

Methods:

Methods:

We assessed the accuracy of responses from ChatGPT and DeepSeek to authoritative urology multiple-choice questions. Then, a randomized controlled trial was performed to compare the learning outcomes of students using ChatGPT and DeepSeek with those using traditional learning methods. Additionally, a questionnaire was designed to survey medical undergraduates' perspectives on the application of AI in urology education.

Results:

Results:

DeepSeek demonstrated higher accuracy than ChatGPT in answering urology-related multiple-choice questions. In the testfollowing the self-study period, the DeepSeek group surpassed both the control and ChatGPT groups in total scores across various question types. Despite the superior scores in the ChatGPT group, statistical significance was not achieved relative to the control group. Survey results revealed that most students had a positive attitude toward AI-assisted learning, believing it could effectively enhance medical education.

Conclusions:

Conclusion: The study suggests that GenAI like ChatGPT and DeepSeek can effectively support medical education.These findings offer evidence-based insights into the embedding of generative AI within medical educational frameworks, providing guidance for educators in developing teaching strategies and for institutions in formulating relevant policies.


 Citation

Please cite as:

Niu H, Yang W, Xu T, Wei J, Zheng W, Yan W, Wang J

Effectiveness of ChatGPT and DeepSeek in Urology Medical Education: Randomized Controlled Trial

J Med Internet Res 2026;28:e89315

DOI: 10.2196/89315

PMID: 42766393

PMCID: 13592387

Download PDF


Request queued. Please wait while the file is being generated. It may take some time.

© The authors. All rights reserved. This is a privileged document currently under peer-review/community review (or an accepted/rejected manuscript). Authors have provided JMIR Publications with an exclusive license to publish this preprint on it's website for review and ahead-of-print citation purposes only. While the final peer-reviewed paper may be licensed under a cc-by license on publication, at this stage authors and publisher expressively prohibit redistribution of this draft paper other than for review purposes.