Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 82
Reusable building blocks in biological systems.
PMID 30958230 · PMC6303794 · Journal of the Royal Society, Interface · 2018 · 8 claims · 5 setups
Biological systems can be decomposed into phenotypic building blocks (PBBs) whose reusability ranges from single-use (condition-specific) to constitutive
-
Has reproduction · 92
Similarities and Differences in Gene Expression Networks Between the Breast Cancer Cell Line Michigan Cancer Foundation-7 and Invasive Human Breast Cancer Tissues.
PMID 34056582 · PMC8155268 · Frontiers in artificial intelligence · 2021 · 8 claims · 8 setups
MCF-7 cell lines and human breast cancer tissues share only minimal similarity in biological processes, though fundamental functions such as cell cycle are conserved
-
Has reproduction · 63
RummaGEO: Automatic mining of human and mouse gene sets from GEO.
PMID 39569206 · PMC11573963 · Patterns (New York, N.Y.) · 2024 · 8 claims · 7 setups
RummaGEO is a gene expression signature search engine built from automatically mined human and mouse RNA-seq perturbation studies in GEO
-
Has reproduction
D2H2: diabetes data and hypothesis hub.
PMID 38107655 · PMC10723036 · Bioinformatics advances · 2023 · 6 claims · 5 setups
D2H2 is a web portal hosting hundreds of curated, uniformly reprocessed diabetes-relevant transcriptomics datasets from GEO with per-study visualization, differential expression, and single-gene queries.
-
Has reproduction · 68
Systematic identification of ACE2 expression modulators reveals cardiomyopathy as a risk factor for mortality in COVID-19 patients.
PMID 35012625 · PMC8743438 · Genome biology · 2022 · 7 claims · 8 setups
GENEVA is a semi-automated, study-design-agnostic framework that mines large-scale public RNA-seq data to identify conditions modulating a gene of interest's expression
-
Has reproduction · 96
Scalable Prediction of Acute Myeloid Leukemia Using High-Dimensional Machine Learning and Blood Transcriptomics.
PMID 31918046 · PMC6992905 · iScience · 2020 · 8 claims · 8 setups
Assembled the largest reference blood gene expression profiling (GEP) dataset for AML to date: 12,029 samples from 105 studies across three platforms.
-
Has reproduction · 88
AuPairWise: A Method to Estimate RNA-Seq Replicability through Co-expression.
PMID 27082953 · PMC4833304 · PLoS computational biology · 2016 · 7 claims · 6 setups
Sample-sample correlation of transcript abundances is a misleading measure of replicability for assessing differential expression, because it is dominated by gene-specific dynamic ranges rather than condition-dependent variation.
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Full-text index only
miRGator: an integrated system for functional annotation of microRNAs.
PMID 17942429 · PMC2238850 · Nucleic acids research · 2008 · 8 claims · 8 setups
miRGator integrates target prediction, functional enrichment analysis (GO/pathway/disease), and expression data (miRNA/mRNA/protein) into one system for functional annotation of miRNAs