← Dataset search
ERR1248472
ENAProvenance — who produced it, who reused it
Linked to 0 papers in the literature. Roles are inferred factual signals (who deposited the data vs who reused it), with counts — never a judgement about any author.
No linked papers found in the corpus yet.
Deep data QC
62/100 · DStandardized, field-standard QC computed by touching the data — every metric states how it was obtained · evidence: measured
Data type / assay
bulk-RNA-seq
Organism
Listeria monocytogenes EGD-e
Instrument
Illumina HiSeq 2000
Platform
ILLUMINA
Read type
short-read
Files available
FASTQ (raw reads)
N numbers (samples, groups)
1 runs
Metrics (value · how obtained)
gc sd
8.8
measured
checksum ok
yes
reported
total bases
427014281
reported
total reads
6765359
reported
n content pct
0.094
measured
pct q20 bases
89.5
measured
pct q30 bases
79.7
measured
pct reads q30
80
measured
sampled bases
63141329
measured
sampled reads
1000000
measured
gc content pct
45.4
measured
polyg tail pct
0.06
measured
read length sd
10.9
measured
quality dropoff
1.1
measured
read length max
70
measured
read length min
32
measured
read length n50
69
measured
max base quality
37
measured
mean read length
63.1
measured
max n pct per pos
1.081
measured
mean base quality
32.3
measured
pct reads lt 100bp
100
measured
read length median
69
measured
adapter content pct
0.14
measured
median read quality
33.4
measured
duplication rate pct
55.73
measured
overrepresented top pct
23.34
measured
How this grade was computed
Weighted mean of 4 scored metric(s) → 62/100
The D grade is a transparent weighted average. Each metric below scored from 0–100% against the published bulk-RNA-seq thresholds, weighted by its importance; nothing is hidden or subjective.
pct q30 bases
79.7
measured
×1
49%
mean base quality
32.3
measured
×0.6
72%
adapter content pct
0.14
measured
×0.4
100%
duplication rate pct
55.73
measured
×0.4
43%
QC cost
18 s compute
measured = computed from the data · extrapolated/reported = derived or from the repository · dq-1.0