Corpus 1,272 assessed · 1,173 scored · 643 reproduced ≥75 · 168 flagged ·∅ 74.1/100
← Dataset search

GSE115404

GEO first seen 2018

Single cell transcriptome profiling of retinal ganglion cells identifies cellular subtypes

Organism
Mus musculus
Samples
2
Type
Expression profiling by high...
Submitted
2018-06-06

Retinal ganglion cells (RGCs) convey the major output of information collected from the eye to the brain. Thirty subtypes of RGCs have been identified to date. Here, we analyze 6,225 RGCs (average of 5,000 genes per cell) from right and left eyes by single cell RNA-seq and classify them into 40 subtypes using clustering algorithms. We identify additional subtypes and markers, as well as transcription factors predicted to cooperate in specifying RGC subtypes. Zic1, a marker of the right eye-enric...

Provenance — who produced it, who reused it

Linked to 1 papers in the literature. Roles are inferred factual signals (who deposited the data vs who reused it), with counts — never a judgement about any author.

Deposited / produced by
Ephraim TrakhtenbergBruce Rheaume

Deep data QC

73/100 · C

Standardized, field-standard QC computed by touching the data — every metric states how it was obtained · evidence: measured

What this means
claude:haiku

Bulk RNA-seq (mouse). Grade C: problematic duplication (71.61%) combines with below-optimal base quality (Q30=84.3%, mean Q=36) to create a lower-tier dataset. Large sample size (731M reads, 71B bases) does not compensate for these technical issues, particularly the severe amplification bias. Reuse is not recommended without deep investigation and likely requires extensive filtering or filtering-aware quantitation.

Data type / assay
bulk-RNA-seq
Organism
Mus musculus
Instrument
Illumina HiSeq 4000
Platform
ILLUMINA
Read type
short-read
Files available
FASTQ (raw reads)
N numbers (samples, groups)
2 / 2 runs
Completeness
100%
Metrics (value · how obtained)
checksum ok yes reported
total bases 71691710272 reported
total reads 731548064 reported
n content pct 0.02 measured
pct q20 bases 94.2 measured
pct q30 bases 84.3 measured
gc content pct 41.2 measured
mean read length 98 measured
mean base quality 36 measured
adapter content pct 0 measured
duplication rate pct 71.61 measured
supplementary file types CSV reported
How this grade was computed
Weighted mean of 4 scored metric(s) → 73/100

The C grade is a transparent weighted average. Each metric below scored from 0–100% against the published bulk-RNA-seq thresholds, weighted by its importance; nothing is hidden or subjective.

pct q30 bases 84.3 measured ×1 72%
mean base quality 36 measured ×0.6 100%
adapter content pct 0 measured ×0.4 100%
duplication rate pct 71.61 measured ×0.4 8%
QC cost 18 s compute

measured = computed from the data · extrapolated/reported = derived or from the repository · dq-1.0