Authors
Hailin Pan, Fengqin Luo, Yishuo Zhang, Shizhe Jiang, Wenzhi Wang, Huixian Zeng, Yunfeng Zhang, Ou Wang, Wenbing Qin, Lingxuan Zhang, Dapeng Wang, Junyi Chen, Ji Wang
Published in
Small methods. Pages e70949. Aug 11, 2026. Epub Aug 11, 2026.
Abstract
The recent repurposing of nanopore arrays for proteomics generates extensive high-throughput datasets, but stochastic molecular translocations and sensor heterogeneity inevitably introduce massive non-informative signal artifacts. Therefore, an automated and unbiased data curation method is essential. Here, we present NanoCurator, an iterative, self-supervised CNN-LSTM autoencoder that automatically extracts high-fidelity translocation fingerprints from complex raw data. By evaluating reconstruction error and Lempel-Ziv complexity via adaptive thresholding, NanoCurator effectively segregates genuine signals from diverse noise profiles in both simulated and empirical datasets. Crucially, this curation significantly elevates classification accuracy by 1.8% for a 15-peptide panel, 2.3% for post-translational modifications, and 1.5% for heterogeneous mixtures, while reducing data volume requirements, thereby establishing a robust, highly generalizable paradigm for automated signal quality control. Ultimately, this versatile framework accelerates precise peptide identification and enables the reliable, large-scale application of massively parallel nanopore sensing in proteomics.
PMID:
42581496
Bibliographic data and abstract were imported from PubMed on 12 Aug 2026.
Read full publication at:
Please sign in
to see all details.
Advertisement
Stats
- Recommendations n/a n/a positive of 0 vote(s)
- Views 10
- Comments 0