Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 78
Fungal metabarcoding data integration framework for the MycoDiversity DataBase (MDDB).
PMID 32463383 · PMC7734503 · Journal of integrative bioinformatics · 2020 · 7 claims · 4 setups
Public fungal metabarcoding raw DNA data and their associated environmental metadata in sequence archives are heterogeneously annotated and lack a uniform processing pipeline, preventing large-scale biodiversity/distribution assessments.
-
Full-text index only
ChimerDB 2.0--a knowledgebase for fusion genes updated.
PMID 19906715 · PMC2808913 · Nucleic acids research · 2010 · 8 claims · 4 setups
ChimerDB 2.0 is an updated knowledgebase integrating fusion transcripts from GenBank transcriptome analysis with Sanger CGP, OMIM, PubMed, and Mitelman's database data.
-
Has reproduction
All of gene expression (AOE): An integrated index for public gene expression databases.
PMID 31978081 · PMC6980531 · PloS one · 2020 · 6 claims · 4 setups
ArrayExpress (AE) discontinued importing data from GEO in 2017, causing new GEO entries to become unavailable from AE
-
Has reproduction · 74
Wide-Open: Accelerating public data release by automating detection of overdue datasets.
PMID 28594819 · PMC5464523 · PLoS biology · 2017 · 6 claims · 5 setups
A general text-mining + API-query approach (Wide-Open) can automatically identify datasets that are overdue for public release in a repository
-
Has reproduction · 80
Curation of over 10 000 transcriptomic studies to enable data reuse.
PMID 33599246 · PMC7904053 · Database : the journal of biological databases and curation · 2021 · 8 claims · 6 setups
Gemma is a curated database and bioinformatics system that addresses metadata, probe annotation, and expression data inconsistencies in GEO to enable transcriptomic data reuse