Provenance — who produced it, who reused it
Linked to 1 papers in the literature. Roles are inferred factual signals (who deposited the data vs who reused it), with counts — never a judgement about any author.
1 further paper cites this accession but reuse could not be confirmed.
Deep data QC
87/100 · BStandardized, field-standard QC computed by touching the data — every metric states how it was obtained · evidence: measured
This NovaSeq 6000 SARS-CoV-2 amplicon dataset contains 4.5M reads across 2.2B bases with 94.4% Q20 and 86% Q30, lower than ideal but sufficient for consensus genome assembly and major variant identification. The amplicon framework concentrates reads on target regions, enabling deep coverage despite lower per-read accuracy. Reuse caveats include sensitivity to primer-binding polymorphisms and diminished power to detect subclonal variants.
The B grade is a transparent weighted average. Each metric below scored from 0–100% against the published amplicon thresholds, weighted by its importance; nothing is hidden or subjective.
measured = computed from the data · extrapolated/reported = derived or from the repository · dq-1.0