Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The use of edge-betweenness clustering to investigate biological function in protein interaction networks.
PMID 15740614 · PMC555937 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Edge-Betweenness clustering separates protein interaction graphs into subgraphs whose GO term distributions show significant correlations, revealing biologically meaningful functional modules.
-
Full-text index only
Network properties of complex human disease genes identified through genome-wide association studies.
PMID 19956617 · PMC2779513 · PloS one · 2009 · 7 claims · 6 setups
Complex disease genes are significantly less central (lower degree/closeness, higher eccentricity) in the human interactome than essential and monogenic disease genes, occupying an intermediate niche between monogenic disease genes and non-disease genes
-
Full-text index only
Discovering cancer genes by integrating network and functional properties.
PMID 19765316 · PMC2758898 · BMC medical genomics · 2009 · 8 claims · 6 setups
Cancer genes have distinct PPI network topology (higher connectivity, higher clustering coefficient, shorter path length to known cancer genes) compared to non-cancer genes
-
Has reproduction · 75
Identification of Key Differentially Expressed Genes in Arabidopsis thaliana Under Short- and Long-Term High Light Stress.
PMID 40869111 · PMC12386182 · International journal of molecular sciences · 2025 · 7 claims · 5 setups
Short- and long-term HL responses in Arabidopsis leaves are driven by distinct transcriptional programs, with duration of HL treatment as the primary factor separating transcriptomic clusters.
-
Full-text index only
Human synthetic lethal inference as potential anti-cancer target gene detection.
PMID 20015360 · PMC2804737 · BMC systems biology · 2009 · 7 claims · 8 setups
Targeting the synthetic lethal partner of a gene mutated in cancer selectively damages tumor cells while sparing healthy cells, offering a rationale for anti-cancer drug design
-
Full-text index only
Building disease-specific drug-protein connectivity maps from molecular interaction networks and PubMed abstracts.
PMID 19649302 · PMC2709445 · PLoS computational biology · 2009 · 7 claims · 4 setups
A computational framework can build disease-specific drug-protein connectivity maps by integrating protein interaction networks and PubMed literature mining, without gene expression profiles from drug perturbation experiments
-
Has reproduction · 75
PHA4GE quality control contextual data tags: standardized annotations for sharing public health sequence datasets with known quality issues to facilitate testing and training.
PMID 38860884 · PMC11261899 · Microbial genomics · 2024 · 7 claims · 3 setups
PHA4GE developed a set of standardized contextual data tags (five fields plus controlled-vocabulary terms) for annotating pathogen sequence datasets with known quality issues.
-
Has reproduction · 51
SGCP: a spectral self-learning method for clustering genes in co-expression networks.
PMID 38956463 · PMC11221046 · BMC bioinformatics · 2024 · 7 claims · 4 setups
SGCP, a spectral self-learning method, yields gene co-expression modules with higher GO enrichment than WGCNA, CoExpNets, and CEMiTool across 12 real gene expression datasets.
-
Has reproduction · 85
A mechanistic model captures the emergence and implications of non-genetic heterogeneity and reversible drug resistance in ER+ breast cancer cells.
PMID 34316714 · PMC8271219 · NAR cancer · 2021 · 7 claims · 8 setups
EMT and tamoxifen-resistance (TamR) regulatory axes can drive one another, enabling non-genetic heterogeneity via six co-existing phenotypes (ES, ER, HS, HR, MS, MR)
-
Full-text index only
Reconstruction of human protein interolog network using evolutionary conserved network.
PMID 17493278 · PMC1885812 · BMC bioinformatics · 2007 · 8 claims · 7 setups
A relative conservation score derived from maximal quasi-cliques in protein interaction networks, combined with other interaction features, can score and rank predicted human interologs for confidence.
-
Full-text index only
Improvements to cardiovascular gene ontology.
PMID 19046747 · PMC2706316 · Atherosclerosis · 2009 · 8 claims · 8 setups
Gene Ontology (GO) provides a controlled vocabulary that links current functional knowledge of genes to high-throughput genomic and proteomic datasets, aiding data interpretation.
-
Has reproduction · 71
Artificial intelligence-guided discovery of gastric cancer continuum.
PMID 36692601 · PMC9871434 · Gastric cancer : official journal of the International Gastric Cancer Association and the Japanese Gastric Cancer Association · 2023 · 8 claims · 8 setups
A Boolean implication network built from GSE66229 yields a GC-BoNE gene signature (Boolean paths C#11-2-4-14 and C#7-13-14) that classifies tumor vs normal/adjacent-normal gastric samples
-
Has reproduction · 94
Systematic assessment of pathway databases, based on a diverse collection of user-submitted experiments.
PMID 36088548 · PMC9487593 · Briefings in bioinformatics · 2022 · 8 claims · 6 setups
Well-established, hierarchically organized pathway annotation systems (e.g. GO, Reactome, KEGG) yield the best overall enrichment performance despite covering much of the human genome only in general terms.
-
Full-text index only
Molecular markers of preterm labor in the choriodecidua.
PMID 20009011 · PMC2852874 · Reproductive sciences (Thousand Oaks, Calif.) · 2010 · 8 claims · 4 setups
Preterm choriodecidua displays distinct gene and protein expression patterns associated with preterm labor
-
Has reproduction · 86
Plasmid transmission dynamics and evolution of partner quality in a natural population of Rhizobium leguminosarum.
PMID 41212030 · PMC12691615 · mBio · 2025 · 8 claims · 8 setups
Of the four most frequent plasmid types, types II and III have more stable size, larger core genomes, and track the chromosomal phylogeny (more vertical transmission), while types I and IV (pSym) vary in size and gene content with phylogenies consistent with frequent horizontal transmission.
-
Has reproduction · 83
Functional module detection through integration of single-cell RNA sequencing data with protein-protein interaction networks.
PMID 33138772 · PMC7607865 · BMC genomics · 2020 · 8 claims · 6 setups
scPPIN integrates scRNA-seq-derived p-values with PPINs to detect maximum-weight connected subgraphs (active/functional modules) via an exact Steiner-tree approach
-
Full-text index only
SysPIMP: the web-based systematical platform for identifying human disease-related mutated sequences from mass spectrometry.
PMID 19036792 · PMC2686442 · Nucleic acids research · 2009 · 8 claims · 7 setups
SysPIMP is a web-based platform integrating disease mutation databases with X!Tandem and BLAST to identify disease-related mutated proteins from MS results
-
Full-text index only
Comparative Toxicogenomics Database: a knowledgebase and discovery tool for chemical-gene-disease networks.
PMID 18782832 · PMC2686584 · Nucleic acids research · 2009 · 8 claims · 5 setups
CTD is a manually curated knowledgebase that integrates chemical-gene interactions, chemical-disease relationships, and gene-disease relationships into a chemical-gene-disease triad
-
Has reproduction · 50
Comparative analysis of circular RNAs between soybean cytoplasmic male-sterile line NJCMS1A and its maintainer NJCMS1B by high-throughput sequencing.
PMID 30208848 · PMC6134632 · BMC genomics · 2018 · 8 claims · 7 setups
2867 circRNAs were identified in soybean flower buds via high-throughput sequencing with RNase R enrichment, of which 1009 were differentially expressed between NJCMS1A and NJCMS1B
-
Has reproduction · 57
Diapause vs. reproductive programs: transcriptional phenotypes in a keystone copepod.
PMID 33782539 · PMC8007741 · Communications biology · 2021 · 8 claims · 7 setups
t-SNE clustering of all-gene expression data groups field-collected (diapause program) samples into one cluster while early and late culture (reproductive program) samples separate into two distinct phenotypes