Maintenance Notice

Due to necessary scheduled maintenance, the JMIR Publications website will be unavailable from Wednesday, July 01, 2020 at 8:00 PM to 10:00 PM EST. We apologize in advance for any inconvenience this may cause you.

Who will be affected?

Currently submitted to: JMIR AI

Date Submitted: Dec 19, 2025

Warning: This is an author submission that is not peer-reviewed or edited. Preprints - unless they show as "accepted" - should not be relied on to guide clinical practice or health-related behavior and should not be reported in news media as established information.

Beyond the First Generated Summarization: Comparing Human and AI-Generated Reports in Geriatric Outpatient Consultations on Content, Linguistic Form, and Healthcare Professionals’ Reporting Preferences

  • Lourens Kraft van Ermel; 
  • Janneke Noordman; 
  • Diede Merkelbach; 
  • Tom Huibers; 
  • Sjaak Brinkkemper; 
  • Sandra van Dulmen; 
  • Yvonne Schoon

ABSTRACT

Background:

Administrative reporting is a major contributor to healthcare professionals’ (HCPs) workload, available patient time and HCP burnout rates, and large staffing costs. Automated medical reporting (AMR) systems, systems using large language models to automatically generate medical reports, have been proposed as a solution, but their clinical validity remains uncertain, despite their promised benefits. While some studies address content accuracy, linguistic form in AMR reports and their alignment with HCP preferences are understudied.

Objective:

The goal is 1) to evaluate the first development iteration in terms of content (what is reported), linguistic form (how it is reported), and HCPs’ reporting preferences; and 2) to generate insights for the AMR system’s further development.

Methods:

We evaluated the first development iteration of an AMR system for geriatric outpatient care by focusing on the history taking section (or anamnesis) of the multifaceted and time-consuming Comprehensive Geriatric Assessment (CGA). This first iteration of the CGA reporting model was developed based on the Dutch CGA guidelines and CGA blueprints used in the hospital this study was conducted at. A mixed-methods design was used, employing: (1) content comparison of ten audio-recorded consultations and their corresponding conventional and AMR reports, (2) a focus group in which HCPs discuss reporting differences and HCPs’ reporting preferences were elicited, (3) a linguistic sentence complexity analysis, as a proxy to study linguistic form.

Results:

Compared to the conventional reports, AMR reports were shorter, contained less and different information (including hallucinations), repeated more content, had higher sentence complexity, and adhered to a rigid structure. HCPs evaluated AMR as too concise and conclusive, and lacking specific textual elements they deemed important for medical and legal accountability, HCPs appreciated AMR’s structural organization.

Conclusions:

This study yielded multiple insights for the further development of the AMR system used, specifically on information selection, how it is linguistically realized, and HCPs’ reporting preferences. It showcases how not only the content of reports should be evaluated, but that linguistic form and HCP reporting preferences are important to evaluate for AMR implementations as well. This methodology serves as an initial proof-of-concept validation framework that future research can use to evaluate AI technology beyond content-based measures.


 Citation

Please cite as:

Kraft van Ermel L, Noordman J, Merkelbach D, Huibers T, Brinkkemper S, van Dulmen S, Schoon Y

Beyond the First Generated Summarization: Comparing Human and AI-Generated Reports in Geriatric Outpatient Consultations on Content, Linguistic Form, and Healthcare Professionals’ Reporting Preferences

JMIR Preprints. 19/12/2025:89971

DOI: 10.2196/preprints.89971

URL: https://preprints.jmir.org/preprint/89971

Download PDF


Request queued. Please wait while the file is being generated. It may take some time.

© The authors. All rights reserved. This is a privileged document currently under peer-review/community review (or an accepted/rejected manuscript). Authors have provided JMIR Publications with an exclusive license to publish this preprint on it's website for review and ahead-of-print citation purposes only. While the final peer-reviewed paper may be licensed under a cc-by license on publication, at this stage authors and publisher expressively prohibit redistribution of this draft paper other than for review purposes.