Hiring in life sciences? Share your open positions with our professional community. Read more Close

Advertisement

[Construction of a genotype-phenotype database and machine learning models for pseudohypertrophic muscular dystrophy].

Created on 18 Aug 2026

Authors

Yi-Xuan Chu, Ci-Liu Zhang, Jing Peng

Published in

Zhongguo dang dai er ke za zhi = Chinese journal of contemporary pediatrics. Volume 28. Issue 8. Pages 991-997. Aug 15, 2026.

Abstract

To establish a genotype-phenotype database and machine learning models for pseudohypertrophic muscular dystrophy (PMD), and to explore genotype-phenotype correlations of the disease.
Clinical data of children with PMD admitted to Xiangya Hospital, Central South University from January 2010 to December 2024, together with cases retrieved from the PubMed database between January 1987 and December 2024, were retrospectively collected to construct a genotype-phenotype database, with an online query function via a WeChat mini-program. Based on this database, Random Forest, Extreme Gradient Boosting, and Light Gradient Boosting Machine algorithms were integrated using a soft voting ensemble strategy to build a machine learning model predicting clinical phenotypes associated with small variants. The predictive performance of the model was compared with that of the reading-frame rule. The model was deployed online via the Streamlit platform.
The database included 17 053 PMD cases, comprising 472 patients in the local cohort and 16 581 literature-derived cases. Modeling and validation were performed on a filtered dataset comprising small variants. In the internal test set, the machine learning model achieved an area under the receiver operating characteristic curve (AUC) of 0.924 (95%CI: 0.881-0.963), significantly higher than the reading-frame rule AUC of 0.652 (95%CI: 0.591-0.717) (P<0.001). In the external test set, the machine learning model achieved an AUC of 0.854 (95%CI: 0.736-1.000), compared to 0.667 (95%CI: 0.500-1.000) for the reading-frame rule, with no statistically significant difference (P>0.05).
The constructed PMD genotype-phenotype database and machine learning prediction model provide an efficient and reliable novel tool for phenotype prediction in PMD.

PMID:
42608308
Bibliographic data and abstract were imported from PubMed on 18 Aug 2026.

Read full publication at:
Please sign in to see all details.

Advertisement

Stats

  • Community rating n/a 0 votes
  • Reviewers' rating n/a 0 votes
  • Your rating

1-terrible, 9-excellent. How would you rate this publication? Sign in in to submit your rating.

  • Recommendations n/a n/a positive of 0 vote(s)
  • Views 7
  • Comments 0

Recommended by

  • No recommendations yet.

Post a comment

You need to be signed in to post comments. You can sign in here.

Comments

There are no comments yet.

Advertisement