Provenance — who produced it, who reused it
Linked to 1 papers in the literature. Roles are inferred factual signals (who deposited the data vs who reused it), with counts — never a judgement about any author.
1 further paper cites this accession but reuse could not be confirmed.
Deep data QC
21/100 · FStandardized, field-standard QC computed by touching the data — every metric states how it was obtained · evidence: measured
This AB 310 Genetic Analyzer capillary sequencing dataset from Entamoeba invadens IP1 contains 100,495 reads and 96.8 million bases with moderate quality (80% Q20, 65.4% Q30), representing an older single-pass Sanger-like protocol for eukaryotic genome surveying. The unusually low 31.4% GC warrants verification for potential sequencing bias or contamination. Modern researchers should treat this as a legacy resource requiring integration with contemporary sequencing technologies for structural accuracy and completeness.
The F grade is a transparent weighted average. Each metric below scored from 0–100% against the published WGS thresholds, weighted by its importance; nothing is hidden or subjective.
measured = computed from the data · extrapolated/reported = derived or from the repository · dq-1.0