Corpus 1,272 assessed · 1,173 scored · 643 reproduced ≥75 · 168 flagged ·∅ 74.1/100
← Dataset search

GSE20846

GEO first seen 2015

Transcript assembly and abundance estimation from RNA-Seq reveals thousands of new transcripts and switching among isoforms

Organism
Mus musculus
Samples
4
Type
Expression profiling by high...
Submitted
2010-03-11

We introduce an approach to transcript discovery coupled with a statistical model for RNA-Seq experiments that produces estimates of transcript abundances. Our algorithms are implemented in an open source software program called Cufflinks. To test Cufflinks, we sequenced and analyzed more than 430 million paired 75bp RNA-Seq reads from a mouse myoblast cell line representing a differentiation timeseries. We detected 13,689 known transcripts and 3,724 previously unannotated ones, 62% of which ar...

Provenance — who produced it, who reused it

Linked to 5 papers in the literature. Roles are inferred factual signals (who deposited the data vs who reused it), with counts — never a judgement about any author.

Deposited / produced by
Cole TrapnellBrian A WilliamsGeo PerteaAli MortazaviGordon KwanMarijke J van BarenSteven L SalzbergBarbara J WoldLior Pachter
Reused by

3 further papers cite this accession but reuse could not be confirmed.

Deep data QC

metadata only · no data-level QC for this type

Standardized, field-standard QC computed by touching the data — every metric states how it was obtained

No quantitative QC rubric exists for this data type yet, so it is deliberately left unscored — this is an honest "not applicable", not a poor rating.

QC cost 22 s compute

measured = computed from the data · extrapolated/reported = derived or from the repository · dq-1.0 · provisional — verify independently