Authors
Saikat Barua, Humayra Shemushee
Published in
Data in brief. Volume 68. Pages 113149. Epub Aug 06, 2026.
Abstract
This article presents a harmonized corpus of scholarly publications and patents concerning quantum error correction, the encoding techniques that protect quantum information against noise and that underpin the pursuit of fault-tolerant quantum computation. The corpus spans 1 January 1995 to 31 December 2025 and contains 16,274 records: 14,385 scholarly publications harvested from OpenAlex and enriched from Crossref, and 1889 United States patent publications retrieved from Google Patents Public Data. Publications were collected by matching twenty scope phrases against titles and abstracts, intersected with a quantum-computing topic to suppress unrelated matches; patents were collected under the Cooperative Patent Classification class for quantum error correction. Both record types were mapped onto a common eighteen-field schema, deduplicated within type, and assigned to one of three retrieval tiers according to the strength of the topical match. Each record carries subfield tags and flags marking machine-learning decoder records and adjacent-topic mentions. The precision of the core tier was assessed by manual annotation of a stratified sample, and inter-annotator agreement was measured. The corpus, its variable dictionary, provenance metadata, validation samples, and the complete construction notebook are openly available, supporting reuse in scientometrics, patent analysis, and text mining of the field.
PMID:
42676606
Bibliographic data and abstract were imported from PubMed on 01 Sep 2026.
Read full publication at:
Please sign in
to see all details.
Advertisement
Stats
- Recommendations n/a n/a positive of 0 vote(s)
- Views 9
- Comments 0