Hiring in life sciences? Share your open positions with our professional community. Read more Close

Advertisement

Large Language Model Simplification of Open Access Pediatric Strabismus Literature: Cross-Sectional Validation of Readability and Clinical Fidelity.

Created on 15 Aug 2026

Authors

Mingming Jiang, Mingming Zhou, Xiaomei Wan, Jing Zhang

Published in

JMIR formative research. Volume 10. Pages e91572. Aug 14, 2026. Epub Aug 14, 2026.

Abstract

Peer-reviewed medical literature consistently violates established health literacy readability targets, creating a gap that effectively excludes patients and caregivers from accessing evidence-based information.
This study aimed to evaluate whether a large language model (LLM) can generate plain-language summaries of pediatric strabismus literature while preserving clinical fidelity and meeting established health literacy readability targets.
This cross-sectional study analyzed 85 open access, peer-reviewed pediatric strabismus articles published between 2022 and 2025, stratified by strabismus subtype, surgical relevance, and publication type. Full-text articles were processed using DeepSeek-V3 (DeepSeek) via a structured prompt, which instructed the model to provide a simplified summary meeting the following requirements for each article: a seventh-grade or lower reading level, a maximum length of 800 words, and strict preservation of medically significant data. Primary outcomes were readability scores measured by the Flesch-Kincaid Grade Level (FKGL) and Simple Measure of Gobbledygook (SMOG) indices. Secondary outcomes included clinical fidelity, which was independently assessed by 2 fellowship-trained pediatric strabismus specialists.
Baseline articles demonstrated a mean FKGL score of 15.79 (SD 1.53) and a mean SMOG score of 14.41 (SD 1.09). Following LLM simplification, the mean FKGL score significantly decreased from 15.79 (SD 1.53) to 7.84 (SD 1.30), representing a mean difference of 7.95 (95% CI 7.52-8.38; P<.001). Similarly, the mean SMOG score decreased from 14.41 (SD 1.09) to 7.68 (SD 0.94), representing a mean difference of 6.73 (95% CI 6.42-7.04; P<.001). Postsimplification readability did not differ significantly by strabismus subtype or surgical relevance (all adjusted P>.05). However, case reports retained slightly higher FKGL scores (mean 8.35, SD 0.89) compared to reviews (mean 7.89, SD 1.37) and original research (mean 7.48, SD 1.40) (adjusted P=.003). Out of the 85 summaries, clinical fidelity was rated good in 81 (95.29%), moderate in 4 (4.71%; these were exclusively summaries of review articles), and poor in 0 (0%).
DeepSeek-V3 effectively reduced the reading level of complex pediatric strabismus literature by approximately 8 grade levels, achieving National Institutes of Health-recommended eighth-grade or lower targets without compromising clinical accuracy. When integrated with clinician oversight, LLM-generated summaries offer a scalable, equitable tool to enhance health literacy and support shared decision-making for patients and caregivers.

PMID:
42600075
Bibliographic data and abstract were imported from PubMed on 15 Aug 2026.

Read full publication at:
Please sign in to see all details.

Advertisement

Stats

  • Community rating n/a 0 votes
  • Reviewers' rating n/a 0 votes
  • Your rating

1-terrible, 9-excellent. How would you rate this publication? Sign in in to submit your rating.

  • Recommendations n/a n/a positive of 0 vote(s)
  • Views 6
  • Comments 0

Recommended by

  • No recommendations yet.

Post a comment

You need to be signed in to post comments. You can sign in here.

Comments

There are no comments yet.

Advertisement