Hiring in life sciences? Share your open positions with our professional community. Read more Close

Advertisement

Artificial Intelligence in Medical Writing: Subtle Errors and Their Complex Consequences.

Created on 11 Aug 2026

Authors

Shigeki Matsubara, Daisuke Matsubara, Kazuhiko Kotani

Published in

JMA journal. Volume 9. Issue 4. Pages 992-996. Jul 15, 2026. Epub May 08, 2026.

Abstract

Although artificial intelligence (AI; such as ChatGPT) can improve manuscript readability, AI-related false descriptions (so-called hallucinations) and incorrect reference retrieval have been repeatedly reported. We tested (1) whether ChatGPT, when provided with medically correct inputs, generates a manuscript containing medically incorrect statements-particularly whether it misunderstands or overlooks "subtle but important" medical issues-and (2) whether ChatGPT cites inappropriate references. We input bullet points on placenta percreta, tasked ChatGPT-5 with generating a mini-review, and asked it to confirm whether the output was medically correct. We introduced a small, deliberate trap. In percreta, a recent conceptual change has gained increased attention: this condition is considered to result from uterine abnormality rather than abnormal placental invasion. This etiopathological shift could be misinterpreted as implying "weaker adherence" and therefore "less difficult surgery," leading to the notion that percreta could be managed at secondary-level institutions. The ChatGPT-generated manuscript cited appropriate references and was almost medically correct, except for one critical issue: it stated that percreta could be managed in secondary-level institutions, which is incorrect and potentially dangerous. During the "confirmation" stage, ChatGPT raised a caution regarding this issue, but not in a definitive manner. An additional experiment was conducted on hypercholesterolemia. ChatGPT again failed to address an important issue, familial hypercholesterolemia, which requires a management strategy different from that for non-familial hypercholesterolemia. Overall, ChatGPT generated a linguistically appealing manuscript with largely correct context, but it produced incorrect and potentially dangerous statements in clinically critical areas. When incorrect statements are subtle rather than obvious, authors, journals, and readers may fail to recognize them. This is paradoxical: advances in AI may reduce "apparent" errors while generating less recognizable ones. When evaluating AI-assisted manuscripts, careful review by individuals with deep domain knowledge is mandatory.

PMID:
42577064
Bibliographic data and abstract were imported from PubMed on 11 Aug 2026.

Read full publication at:
Please sign in to see all details.

Advertisement

Stats

  • Community rating n/a 0 votes
  • Reviewers' rating n/a 0 votes
  • Your rating

1-terrible, 9-excellent. How would you rate this publication? Sign in in to submit your rating.

  • Recommendations n/a n/a positive of 0 vote(s)
  • Views 7
  • Comments 0

Recommended by

  • No recommendations yet.

Post a comment

You need to be signed in to post comments. You can sign in here.

Comments

There are no comments yet.

Advertisement