Maintenance Notice

Due to necessary scheduled maintenance, the JMIR Publications website will be unavailable from Wednesday, July 01, 2020 at 8:00 PM to 10:00 PM EST. We apologize in advance for any inconvenience this may cause you.

Who will be affected?

Accepted for/Published in: Journal of Medical Internet Research

Date Submitted: Jan 26, 2026
Date Accepted: Jul 1, 2026

The final, peer-reviewed published version of this preprint can be found here:

A Bilingual Benchmark for Evaluating Diagnostic Performance of Multimodal Large Language Models in Radiology (RadM-Bench): Evaluation Development and Validation

Wu Q, Wu Q, Zhang P, Yi Z, Shen Y, Bai Y, Tan H, Dong P, Xue Z, Roberts N, Li Y, Wang M

A Bilingual Benchmark for Evaluating Diagnostic Performance of Multimodal Large Language Models in Radiology (RadM-Bench): Evaluation Development and Validation

J Med Internet Res 2026;28:e92183

DOI: 10.2196/92183

PMID: 42566748

PDF not available

Per the author's request this version is not available.