Provenance — who produced it, who reused it
Linked to 1 papers in the literature. Roles are inferred factual signals (who deposited the data vs who reused it), with counts — never a judgement about any author.
1 further paper cites this accession but reuse could not be confirmed.
Deep data QC
66/100 · DStandardized, field-standard QC computed by touching the data — every metric states how it was obtained · evidence: measured
This PacBio RS II long-read WGS dataset provides ~89.5 Gb of sequence from Homo sapiens at 28.9× coverage, enabling structural variant discovery and complex region resolution. The dataset is suitable for identifying large insertions, deletions, and repetitive-element variants that short-read sequencing misses. Researchers searching for long-read whole-genome sequencing or PacBio SMRT genotyping data will find this valuable for haplotype phasing and breakpoint mapping. Note: Q-score metrics reported as zero, suggesting older-generation base-quality reporting.
The D grade is a transparent weighted average. Each metric below scored from 0–100% against the published WGS thresholds, weighted by its importance; nothing is hidden or subjective.
measured = computed from the data · extrapolated/reported = derived or from the repository · dq-1.0