Provenance — who produced it, who reused it
Linked to 1 papers in the literature. Roles are inferred factual signals (who deposited the data vs who reused it), with counts — never a judgement about any author.
1 further paper cites this accession but reuse could not be confirmed.
Deep data QC
66/100 · DStandardized, field-standard QC computed by touching the data — every metric states how it was obtained · evidence: measured
This NovaSeq 6000 amplicon run for SARS-CoV-2 yielded 4.2 million short reads (1.25 billion bases) with good Q20 (96.7%) but notably lower Q30 (89.8%), the lowest among NovaSeq runs in this set. The 39% GC content remains stable. Very high read count partially compensates for lower per-base quality; aggressive quality filtering is recommended. Best suited for large multiplexed studies where throughput outweighs per-base quality concerns.
The D grade is a transparent weighted average. Each metric below scored from 0–100% against the published amplicon thresholds, weighted by its importance; nothing is hidden or subjective.
measured = computed from the data · extrapolated/reported = derived or from the repository · dq-1.0