Hiring in life sciences? Share your open positions with our professional community. Read more Close

Advertisement

Comparing Radiologists' and Artificial Intelligence Performance Detecting Suspicious Microcalcifications on Screening Mammograms: A Pilot Cross-Sectional Study.

Created on 28 Jul 2026

Authors

Sarah J Lewis, Jayden B Wells, Zhengqiang Jiang, Dania Abu Awwad, Melissa L Barron, Phuong Dung Yun Trieu

Published in

Cancer control : journal of the Moffitt Cancer Center. Volume 33. Pages 10732748261470916. Epub Jul 27, 2026.

Abstract

IntroductionEarly breast cancer detection through periodic screening is crucial for reducing mortality due to later stage detection. Microcalcifications are common mammographic findings, present in both malignant lesions, benign pathologies, and normal tissues. This study aimed to assess radiologists' observer performance in determining suspicious calcifications requiring recall, and to compare reader performance to an in-house AI model trained and tested on Australian screening mammograms.MethodsIn this pilot proof-of-concept cross-sectional study, radiologists (n=27), breast physicians (n=2), and final year radiology trainees (n=6), completed the same mammographic test set consisting of 30 mammographic cases displaying different types of calcifications (10 breast cancer, 20 normal/benign). An in-house trained artificial intelligence model (Sydney-GMIC) was also applied to the same test set. Performance between readers was compared to AI via Spearman Rank-Order Correlation test. Work experience and caseload trends were compared using independent T tests and Mann-Whitney-U.ResultsSensitivity was significantly higher in radiologists with ≤ 10 years of experience compared to radiology trainees (72.2% vs 53.3%, p=0.042). Furthermore, readers with higher cases read per week (CPW) (i.e. 101-200 CPW compared to 0-20 CPW) had a decreased specificity (58.8% vs 74.6%, p=0.041), but higher sensitivity (68.8% vs 59.2%, p=0.12). In this pilot dataset, the Sydney-GMIC AI model demonstrated higher sensitivity and specificity than the radiologist means. The cases perceived as difficult by AI differed substantially from those challenging for human readers.ConclusionsThis study highlights the challenging nature of recalling calcifications from screening mammograms only, with variable performance among readers. The findings of this small pilot study are exploratory in nature, but the AI performance signals the potential utility of AI models in mammographic analysis of screening cases to progress to recall.

PMID:
42507819
Bibliographic data and abstract were imported from PubMed on 28 Jul 2026.

Read full publication at:
Please sign in to see all details.

Advertisement

Stats

  • Community rating n/a 0 votes
  • Reviewers' rating n/a 0 votes
  • Your rating

1-terrible, 9-excellent. How would you rate this publication? Sign in in to submit your rating.

  • Recommendations n/a n/a positive of 0 vote(s)
  • Views 8
  • Comments 0

Recommended by

  • No recommendations yet.

Post a comment

You need to be signed in to post comments. You can sign in here.

Comments

There are no comments yet.

Advertisement