Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Phenotypic categorization of genetic skin diseases reveals new relations between phenotypes, genes and pathways.
PMID 19744994 · PMC2773259 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 5 setups
560 genetic skin diseases can be decomposed into 71 elementary phenotypic features (42 dermatologic, 29 systemic) that combine to represent each disease as a point in a multidimensional phenotype space
-
Full-text index only
Information-based methods for predicting gene function from systematic gene knock-downs.
PMID 18959798 · PMC2596148 · BMC bioinformatics · 2008 · 8 claims · 4 setups
Information-based metrics, which incorporate a phenotype's genomic frequency, outperform non-information-based metrics for detecting gene-gene functional similarity from phenotypic knock-down profiles.
-
Full-text index only
Cell atlases and the developmental foundations of the phenotype.
PMID 41662466 · PMC12904592 · PLoS computational biology · 2026 · 8 claims · 6 setups
There is a proportional relationship between average developmental similarity (⟨simD⟩) and average phenotypic similarity (⟨simP⟩) across genes, supporting the D–P rule on average
-
Full-text index only
Commonality of functional annotation: a method for prioritization of candidate genes from genome-wide linkage studies.
PMID 18263617 · PMC2275105 · Nucleic acids research · 2008 · 8 claims · 7 setups
Genes correlated with a common complex trait are more likely to share GO functional annotations than genes not correlated with that trait
-
Has reproduction · 70
GAUGE-Annotated Microbial Transcriptomic Data Facilitate Parallel Mining and High-Throughput Reanalysis To Form Data-Driven Hypotheses.
PMID 33758032 · PMC8547006 · mSystems · 2021 · 8 claims · 7 setups
GAUGE automatically annotates GEO microbial transcriptomic data sets (microarray and RNA-seq), increasing the proportion of annotatable studies from 4% to 33%
-
Has reproduction · 80
Comprehensive analysis of transcriptomics and radiomics revealed the potential of TEDC2 as a diagnostic marker for lung adenocarcinoma.
PMID 39553728 · PMC11569783 · PeerJ · 2024 · 8 claims · 8 setups
WGCNA identified 214 key genes in the blue module most correlated with LUAD
-
Has reproduction · 80
GenomeChronicler: The Personal Genome Project UK Genomic Report Generator Pipeline.
PMID 33193602 · PMC7541957 · Frontiers in genetics · 2020 · 7 claims · 4 setups
GenomeChronicler is the first pipeline able to run offline or in the cloud to generate personal genomics reports (not limited to disease) from whole genome or whole exome sequencing data.
-
Full-text index only
Identification of gene interactions associated with disease from gene expression data using synergy networks.
PMID 18234101 · PMC2258206 · BMC systems biology · 2008 · 8 claims · 4 setups
Synergy of a gene pair with respect to disease, defined as I(G1,G2;C) - [I(G1;C)+I(G2;C)], identifies gene pairs that interact cooperatively with respect to a phenotype rather than independently.
-
Full-text index only
SpaPheno: linking spatial transcriptomics to clinical phenotypes with interpretable machine learning.
PMID 41975540 · PMC13185361 · Genome medicine · 2026 · 8 claims · 8 setups
SpaPheno integrates spatial transcriptomics with clinically annotated bulk RNA-seq to identify spatially resolved biomarkers predictive of patient outcomes including survival, tumor stage, and immunotherapy response
-
Has reproduction · 90
Systematic clustering algorithm for chromatin accessibility data and its application to hematopoietic cells.
PMID 33253153 · PMC7728210 · PLoS computational biology · 2020 · 7 claims · 5 setups
A systematic clustering algorithm for ATAC-seq data can be built by binarizing the genome into open/closed chromatin (1/0) strings and computing Hamming distances between samples for hierarchical clustering.
-
Has reproduction · 80
Colorectal Cancer Prediction Based on Weighted Gene Co-Expression Network Analysis and Variational Auto-Encoder.
PMID 32825264 · PMC7563725 · Biomolecules · 2020 · 6 claims · 7 setups
Combining WGCNA hub genes and VAE 10-dimensional representation as features for an SVM classifier achieves high accuracy in predicting CRC
-
Has reproduction · 76
GeneSetCart: assembling, augmenting, combining, visualizing, and analyzing gene sets.
PMID 40208796 · PMC11984350 · GigaScience · 2025 · 8 claims · 8 setups
GeneSetCart is a web-based platform that lets users assemble, augment, combine, visualize, and analyze gene sets from multiple sources in one place
-
Full-text index only
BaGPipe: an automated, reproducible, and flexible pipeline for bacterial genome-wide association studies.
PMID 41896736 · PMC13147680 · BMC microbiology · 2026 · 7 claims · 8 setups
BaGPipe is an automated, reproducible Nextflow pipeline that integrates pre-processing, Pyseer-based association analysis, and downstream visualisation into a unified bacterial GWAS workflow
-
Has reproduction · 54
EpiDiverse Toolkit: a pipeline suite for the analysis of bisulfite sequencing data in ecological plant epigenetics.
PMID 34805989 · PMC8598301 · NAR genomics and bioinformatics · 2021 · 8 claims · 5 setups
EpiDiverse Toolkit provides Nextflow-based pipelines for WGBS mapping, methylation calling, variant calling, differential methylation, and EWAS tailored to non-model plant ecology
-
Full-text index only
Genomic expression during human myelopoiesis.
PMID 17683550 · PMC2045681 · BMC genomics · 2007 · 8 claims · 5 setups
An integrated myelopoiesis expression dataset of 9,425 genes, each mapped to a unique genomic position, was generated from 24 microarray experiments across 8 myeloid cell types.