Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Ontological Discovery Environment: a system for integrating gene-phenotype associations.
PMID 19733230 · PMC2783409 · Genomics · 2009 · 8 claims · 8 setups
ODE is a web-based system for storing, sharing, retrieving and analyzing phenotype-centered genomic data sets across species and experimental systems
-
Full-text index only
Filtering high-throughput protein-protein interaction data using a combination of genomic features.
PMID 15833142 · PMC1127019 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A combination of three genomic features (interacting Pfam domains, GO annotations, sequence homology) using naive Bayesian networks predicts true protein-protein interactions with high sensitivity and good specificity.
-
Has reproduction · 63
Target identification for repurposed drugs active against SARS-CoV-2 via high-throughput inverse docking.
PMID 34825285 · PMC8616721 · Journal of computer-aided molecular design · 2022 · 8 claims · 6 setups
Combining Vinardo, Ledock, and Korp-PL scoring functions (via averaged Z-scores) improves correct target identification over any single scoring function.
-
Full-text index only
A hierarchical and modular approach to the discovery of robust associations in genome-wide association studies from pooled DNA samples.
PMID 18194558 · PMC2248205 · BMC genetics · 2008 · 8 claims · 5 setups
A hierarchical/modular approach integrating quality control, LD, physical distance, and gene ontology identifies authentic associations among those found by statistical tests in pooled DNA GWAS
-
Full-text index only
Leveraging two-way probe-level block design for identifying differential gene expression with high-density oligonucleotide arrays.
PMID 15099405 · PMC411067 · BMC bioinformatics · 2004 · 7 claims · 2 setups
Two-way ANOVA and Mack-Skillings tests on probe-level data with FDR control are substantially more powerful than t-test/Wilcoxon on probe-set level data for detecting differential expression
-
Has reproduction · 100
Proneural-mesenchymal antagonism dominates the patterns of phenotypic heterogeneity in glioblastoma.
PMID 38433919 · PMC10905000 · iScience · 2024 · 8 claims · 8 setups
The four proposed GBM molecular subtypes (Proneural, Neural, Classical, Mesenchymal / NPC-like, OPC-like, AC-like, MES-like) are not mutually exclusive or independent of one another
-
Has reproduction · 71
Assessment tool based on fatty acid metabolic signatures for predicting the prognosis and treatment response in bladder cancer.
PMID 38076064 · PMC10703629 · Heliyon · 2023 · 8 claims · 8 setups
Consensus clustering of prognosis-related fatty acid metabolism genes (FAMGs) identifies three molecular subtypes of BLCA (FAMC1, FAMC2, FAMC3) with distinct prognoses and tumor microenvironments
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.
-
Full-text index only
Cancer-specific high-throughput annotation of somatic mutations: computational prediction of driver missense mutations.
PMID 19654296 · PMC2763410 · Cancer research · 2009 · 7 claims · 7 setups
CHASM, a Random Forest-based computational method, was developed to identify and prioritize missense mutations likely to be functional drivers of tumor cell proliferation.
-
Has reproduction · 76
GeneSetCart: assembling, augmenting, combining, visualizing, and analyzing gene sets.
PMID 40208796 · PMC11984350 · GigaScience · 2025 · 8 claims · 8 setups
GeneSetCart is a web-based platform that lets users assemble, augment, combine, visualize, and analyze gene sets from multiple sources in one place
-
Has reproduction · 100
DeepRNA-Reg: a deep-learning based approach for comparative analysis of CLIP experiments.
PMID 41055236 · PMC12505516 · RNA biology · 2025 · 7 claims · 5 setups
DeepRNA-Reg, a recurrent neural network-based algorithm, predicts differentially enriched sites in paired HITS-CLIP datasets and outperforms dCLIP 1.7.
-
Has reproduction · 95
Increased prevalence of hybrid epithelial/mesenchymal state and enhanced phenotypic heterogeneity in basal breast cancer.
PMID 38974967 · PMC11225361 · iScience · 2024 · 7 claims · 7 setups
Luminal breast cancer gene expression signature is closely/positively associated with an epithelial signature
-
Has reproduction · 67
binny: an automated binning algorithm to recover high-quality genomes from complex metagenomic datasets.
PMID 36239393 · PMC9677464 · Briefings in bioinformatics · 2022 · 8 claims · 8 setups
binny outperforms or is highly competitive with commonly used and state-of-the-art binning methods (MetaBAT2, MaxBin2, CONCOCT, VAMB, SemiBin, MetaDecoder)
-
Has reproduction · 83
Gene-expression patterns in peripheral blood classify familial breast cancer susceptibility.
PMID 26538066 · PMC4634735 · BMC medical genomics · 2015 · 8 claims · 5 setups
A multigene peripheral-blood gene-expression biomarker accurately classifies which women from high-risk families develop familial breast cancer.
-
Has reproduction · 79
Enriched domain detector: a program for detection of wide genomic enrichment domains robust against local variations.
PMID 24782521 · PMC4066758 · Nucleic acids research · 2014 · 8 claims · 5 setups
EDD is a new algorithm that detects broad (megabase-size) enrichment domains from ChIP-seq data of widely distributed chromatin proteins such as A- and B-type lamins.
-
Full-text index only
Babelomics: advanced functional profiling of transcriptomics, proteomics and genomics experiments.
PMID 18515841 · PMC2447758 · Nucleic acids research · 2008 · 8 claims · 5 setups
Babelomics is a web suite offering both conventional functional enrichment methods and more advanced gene set analysis (GSA) methods, a combination offered by only one other tool (FuncAssociate) among competitors.
-
Full-text index only
A systematic comparative and structural analysis of protein phosphorylation sites based on the mtcPTM database.
PMID 17521420 · PMC1929158 · Genome biology · 2007 · 7 claims · 6 setups
mtcPTM is a hierarchically organized database of human and mouse phosphosites that preserves experimental context, enabling comparison of phosphorylation patterns across conditions
-
Full-text index only
Prediction of missed cleavage sites in tryptic peptides aids protein identification in proteomics.
PMID 17203985 · PMC2664920 · Journal of proteome research · 2007 · 8 claims · 4 setups
An information-theoretic log-likelihood scoring method can predict experimentally observed missed cleavage sites from amino acid sequence alone with up to 90% accuracy.
-
Full-text index only
PA-GOSUB: a searchable database of model organism protein sequences with their predicted Gene Ontology molecular function and subcellular localization.
PMID 15608166 · PMC540074 · Nucleic acids research · 2005 · 7 claims · 4 setups
PA-GOSUB significantly extends the coverage of GO molecular function and subcellular localization annotations for 10 model organism proteomes compared with existing databases (GOA, Swiss-Prot).
-
Has reproduction · 59
Comparison between short-term stress and long-term adaptive responses reveal common paths to molecular adaptation.
PMID 35243257 · PMC8873613 · iScience · 2022 · 8 claims · 7 setups
Short-term stress and long-term adaptations share common metabolic pathways