Provenance — who produced it, who reused it
Linked to 1 papers in the literature. Roles are inferred factual signals (who deposited the data vs who reused it), with counts — never a judgement about any author.
1 further paper cites this accession but reuse could not be confirmed.
Deep data QC
100/100 · AStandardized, field-standard QC computed by touching the data — every metric states how it was obtained · evidence: measured
This Illumina HiSeq 2000 amplicon dataset from Homo sapiens contains 187 million reads and 34.6 billion bases with excellent quality (98.4% Q20, 93.6% Q30), enabling deep targeted profiling of specific genomic regions. The 41.4% GC and massive scale support single-read-level variant enumeration and rare-allele detection across amplified loci (e.g., immune repertoire, pathogen detection, or mutation hotspots). This dataset is valuable for large-scale clinical screening and population-frequency estimation at targeted genomic coordinates.
The A grade is a transparent weighted average. Each metric below scored from 0–100% against the published amplicon thresholds, weighted by its importance; nothing is hidden or subjective.
measured = computed from the data · extrapolated/reported = derived or from the repository · dq-1.0