Hiring in life sciences? Share your open positions with our professional community. Read more Close

Advertisement

Large Language Models in Spine Surgery: A Scoping Review of Clinical Efficacy, Technical Integration, and Ethical Paradigms.

Created on 20 Aug 2026

Authors

Samer G Salman, Rohan A Phadke, Anne E Tatooles, Adithya Nair, Alireza Tavakkoli, James Rizkalla

Published in

Global spine journal. Pages 21925682261481450. Aug 20, 2026. Epub Aug 20, 2026.

Abstract

Study DesignScoping review.ObjectivesTo map spine literature on large language models, characterize reported use cases, and identify evidence gaps limiting implementation.MethodsA scoping review was conducted according to Joanna Briggs Institute methodology and PRISMA-ScR guidance. PubMed, Embase, Scopus, Web of Science, and Cochrane were searched for English-language, peer-reviewed studies published from January 2023 through May 2026 that evaluated large language models in spinal disease, spine surgery, or spine-related care. Eligible studies were synthesized across clinical decision support, triage, patient communication, automation, surgical education, and implementation barriers.ResultsFifteen studies met inclusion criteria. Most evidence involved early evaluation of commercially available or general-purpose models rather than prospectively validated spine-specific systems. Reported applications included patient education, report simplification, coding support, emergency consultation simulation, spinal cord stimulation referral screening, conservative triage, and surgical education. Performance was strongest for structured text-based tasks, patient communication, documentation support, and simplified decision pathways. Performance was weaker for image interpretation, quantitative radiographic assessment, individualized operative planning, and granular procedure selection. Recurrent limitations included hallucinated or unsupported outputs, unreliable citation generation, limited multimodal capability, privacy and data-governance concerns, bias, unclear medicolegal accountability, and minimal validation.ConclusionsLarge language models are an adjunct in spine surgery, with the near-term role in clinician-supervised, text-centered workflows including patient communication, education, documentation, coding, guideline retrieval, and preliminary triage. Current evidence does not support autonomous diagnostic, radiographic, or operative decision-making. Future studies should prioritize spine-specific retrieval-augmented systems, validated multimodal workflows, privacy-preserving deployment, fairness assessment, and prospective evaluation using clinically meaningful outcomes.

PMID:
42622533
Bibliographic data and abstract were imported from PubMed on 20 Aug 2026.

Read full publication at:
Please sign in to see all details.

Advertisement

Stats

  • Community rating n/a 0 votes
  • Reviewers' rating n/a 0 votes
  • Your rating

1-terrible, 9-excellent. How would you rate this publication? Sign in in to submit your rating.

  • Recommendations n/a n/a positive of 0 vote(s)
  • Views 1
  • Comments 0

Recommended by

  • No recommendations yet.

Post a comment

You need to be signed in to post comments. You can sign in here.

Comments

There are no comments yet.

Advertisement