Provenance — who produced it, who reused it
Linked to 1 papers in the literature. Roles are inferred factual signals (who deposited the data vs who reused it), with counts — never a judgement about any author.
1 further paper cites this accession but reuse could not be confirmed.
Deep data QC
70/100 · CStandardized, field-standard QC computed by touching the data — every metric states how it was obtained · evidence: measured
This NovaSeq 6000 amplicon dataset for SARS-CoV-2 contains 2.1 million short reads (617.8 million bases) with excellent quality (99.1% Q20, 96.6% Q30) and zero N-content. The 38.2% GC content is slightly lower than most SARS-CoV-2 amplicon runs, though still within expected bounds. Good read depth and quality support confident consensus calling and variant discovery; suitable for moderately-sized surveillance studies.
The C grade is a transparent weighted average. Each metric below scored from 0–100% against the published amplicon thresholds, weighted by its importance; nothing is hidden or subjective.
measured = computed from the data · extrapolated/reported = derived or from the repository · dq-1.0