Hiring in life sciences? Share your open positions with our professional community. Read more Close

Advertisement

Hidden assumptions in nascent RNA sequencing pipelines define reproducibility states

Created on 18 Jul 2026

Authors

Zhou, X., Feng, C., Zhao, Y.

Abstract

Reproducibility of sequencing analyses is often assumed when identical data are processed with established pipelines, yet outcomes can depend on library assumptions that are not explicit to users. Here we examined commonly used pipelines for nascent RNA sequencing. Across public human PRO-seq datasets, identical inputs generated structured divergence in transcriptional profiles. Diagnostic processing combinations traced this divergence to interactions between paired-end library design, UMI organization and pipeline-embedded assumptions for read trimming, alignment and signal generation. This pattern persisted in independent human and pig PRO-seq libraries sharing a dual-end UMI design, reflecting pipeline-defined assumptions not fully accessible through user-specified parameters. Beyond PRO-seq, GRO-seq analyses showed that assay-specific library architecture can distort positional signal profiles without UMI processing, whereas PRO-cap and reannotated PRO-seq datasets showed that incomplete metadata can prevent pipeline execution or cause silent signal loss. Together, these results define reproducibility states shaped by library design, pipeline assumptions and metadata availability.

Preprint server: bioRxiv
The authors list and abstract were imported from bioRxiv on 18 Jul 2026.

Advertisement

Stats

  • Community rating n/a 0 votes
  • Your rating

1-terrible, 9-excellent. How would you rate this preprint? Sign in in to submit your rating.

  • Recommendations n/a n/a positive of 0 vote(s)
  • Views 56
  • Comments 0

Recommended by

  • No recommendations yet.

Post a comment

You need to be signed in to post comments. You can sign in here.

Comments

There are no comments yet.

Advertisement