Corpus 1,286 assessed · 1,187 scored · 648 reproduced ≥75 · 174 flagged ·∅ 73.9/100
← Dataset search

GSE39756

GEO first seen 2016

Provenance — who produced it, who reused it

Linked to 9 papers in the literature. Roles are inferred factual signals (who deposited the data vs who reused it), with counts — never a judgement about any author.

Reused by

7 further papers cite this accession but reuse could not be confirmed.

Deep data QC

insufficient data to score

Standardized, field-standard QC computed by touching the data — every metric states how it was obtained

Data type / assay
ChIP-seq
Organism
Mus musculus
Instrument
Illumina Genome Analyzer
Platform
ILLUMINA
Read type
short-read
Files available
FASTQ (raw reads)
N numbers (samples, groups)
38 / 38 runs
Completeness
100%
Metrics (value · how obtained)
gc sd 13.33 measured
checksum ok yes reported
total bases 39559014450 reported
total reads 936923017 reported
n content pct 0.014 measured
pct q20 bases 93 measured
sampled bases 24000000 measured
sampled reads 1000000 measured
gc content pct 52.2 measured
polyg tail pct 0 measured
read length sd 0 measured
quality dropoff -2.7 measured
read length max 24 measured
read length min 24 measured
read length n50 24 measured
max base quality 27 measured
mean read length 24 measured
max n pct per pos 0.061 measured
mean base quality 25.5 measured
pct reads lt 100bp 100 measured
read length median 24 measured
adapter content pct 0.25 measured
median read quality 26.3 measured
duplication rate pct 34.98 measured
overrepresented top pct 9.38 measured
supplementary file types BED, TXT reported
QC cost 3.6 min compute

measured = computed from the data · extrapolated/reported = derived or from the repository · dq-1.0

Scientific quality

Based on hands-on reproduction of the papers that use this dataset. A reproducible paper that stands on this data is positive evidence; a flagged one is a prompt to look closer — never a verdict on the dataset itself without the evidence.

1 studies use it 1 reproduced mean score 100