Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 71
Parsimonious Gene Correlation Network Analysis (PGCNA): a tool to define modular gene co-expression for refined molecular stratification in cancer.
PMID 30993001 · PMC6459838 · NPJ systems biology and applications · 2019 · 8 claims · 7 setups
Retaining only the top ~3 most correlated edges per gene (EPG3) combined with FastUnfold clustering (termed PGCNA) produces gene co-expression modules with significantly better separation and enrichment of known biology than using all edges or other clustering methods.
-
Full-text index only
Identification of gene interactions associated with disease from gene expression data using synergy networks.
PMID 18234101 · PMC2258206 · BMC systems biology · 2008 · 8 claims · 4 setups
Synergy of a gene pair with respect to disease, defined as I(G1,G2;C) - [I(G1;C)+I(G2;C)], identifies gene pairs that interact cooperatively with respect to a phenotype rather than independently.
-
Full-text index only
Needles in the haystack: identifying individuals present in pooled genomic data.
PMID 19798441 · PMC2747273 · PLoS genetics · 2009 · 8 claims · 7 setups
The distribution of T for null samples (individuals not in F or G) deviates strongly from the assumed standard normal, in both location and width.
-
Full-text index only
CLEAN: CLustering Enrichment ANalysis.
PMID 19640299 · PMC2734555 · BMC bioinformatics · 2009 · 8 claims · 4 setups
The gene-specific CLEAN score improves reproducibility of cluster analysis conclusions across independent datasets compared to the traditional cluster-wide score (cwCLEAN).
-
Full-text index only
PubMatrix: a tool for multiplex literature mining.
PMID 14667255 · PMC317283 · BMC bioinformatics · 2003 · 8 claims · 3 setups
PubMatrix is a web-based CGI tool that queries PubMed with two lists of terms (search terms vs modifier terms) and returns a matrix of pairwise co-occurrence frequency counts
-
Full-text index only
Mutation analysis of 24 known cancer genes in the NCI-60 cell line set.
PMID 17088437 · PMC2705832 · Molecular cancer therapeutics · 2006 · 8 claims · 3 setups
137 oncogenic mutations were identified across 14 of 24 screened cancer genes (APC, BRAF, CDKN2A, CTNNB1, HRAS, KRAS, NRAS, SMAD4, PIK3CA, PTEN, RB1, STK11, TP53, VHL) in the NCI-60 panel
-
Full-text index only
Constitutional genetic variation at the human aromatase gene (Cyp19) and breast cancer risk.
PMID 10027313 · PMC2362434 · British journal of cancer · 1999 · 7 claims · 5 setups
Allelic distribution of the Cyp19 intron 4 STRP differs significantly between breast cancer cases and controls
-
Full-text index only
Strong signature of natural selection within an FHIT intron implicated in prostate cancer risk.
PMID 18953408 · PMC2568805 · PloS one · 2008 · 8 claims · 8 setups
Re-sequencing and genotyping across a 28.5 kb region delineates the prostate cancer risk association within FHIT intron 5 to a 15 kb LD block in European-Americans.
-
Full-text index only
HapMap-based study of the 17q21 ERBB2 amplicon in susceptibility to breast cancer.
PMID 17117180 · PMC2360759 · British journal of cancer · 2006 · 6 claims · 5 setups
Common genetic variation (tSNPs and haplotypes) across the 400-kb 17q21 ERBB2 amplicon is not associated with breast cancer risk in British women.
-
Full-text index only
Finding disease candidate genes by liquid association.
PMID 17915034 · PMC2246280 · Genome biology · 2007 · 7 claims · 6 setups
LA can detect functionally associated genes that are not directly co-expressed by identifying a mediating gene Z whose expression level changes the correlation between X and Y.
-
Full-text index only
Comprehensive resequence analysis of a 136 kb region of human chromosome 8q24 associated with prostate and colon cancers.
PMID 18704501 · PMC2525844 · Human genetics · 2008 · 6 claims · 5 setups
Next-generation (Roche/454) resequencing of 136 kb at 8q24 in 39 prostate cancer cases and 40 controls generated a comprehensive catalog of common SNPs (MAF>1%), including 442 novel SNPs
-
Full-text index only
Linkage disequilibrium mapping of a breast cancer susceptibility locus near RAI/PPP1R13L/iASPP.
PMID 18588689 · PMC2474586 · BMC medical genetics · 2008 · 8 claims · 6 setups
A region spanning the gene RAI and the 5' portion of XPD is associated with postmenopausal breast cancer
-
Full-text index only
Comprehensive resequence analysis of a 97 kb region of chromosome 10q11.2 containing the MSMB gene associated with prostate cancer.
PMID 19644707 · PMC2778717 · Human genetics · 2009 · 7 claims · 5 setups
Resequencing of the 97-kb 10q11.2 region identified 241 novel polymorphisms not previously reported in dbSNP
-
Has reproduction · 78
Emergent dynamics of underlying regulatory network links EMT and androgen receptor-dependent resistance in prostate cancer.
PMID 36851919 · PMC9957767 · Computational and structural biotechnology journal · 2023 · 8 claims · 7 setups
Simulations of the EMT-AR crosstalk network reveal four possible phenotypes: epithelial-sensitive (ES), epithelial-resistant (ER), mesenchymal-resistant (MR), and mesenchymal-sensitive (MS), with MS occurring rarely
-
Has reproduction
Using random walks to identify cancer-associated modules in expression data.
PMID 24128261 · PMC4015830 · BioData mining · 2013 · 8 claims · 8 setups
Walktrap-GM, a random-walk community detection algorithm adapted with stopping criteria (maximum modularity, maximum size, maximum module score), identifies modules significantly enriched with cancer genes in expression-weighted interaction networks.
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Has reproduction · 79
TSUNAMI: Translational Bioinformatics Tool Suite for Network Analysis and Mining.
PMID 33705981 · PMC9403021 · Genomics, proteomics & bioinformatics · 2021 · 8 claims · 6 setups
TSUNAMI is a freely accessible web-based tool suite that mines gene co-expression network (GCN) modules from public (GEO, TCGA) or user-uploaded numerical omics data and performs downstream gene set enrichment analysis.
-
Has reproduction · 94
Hierarchical cell-type identifier accurately distinguishes immune-cell subtypes enabling precise profiling of tissue microenvironment with single-cell RNA-sequencing.
PMID 36681937 · PMC10025442 · Briefings in bioinformatics · 2023 · 8 claims · 8 setups
HiCAT is a hierarchical, marker-based cell-type identifier that uses gene set analysis (GSA) scoring with markers structured in a three-level taxonomy tree (major-type, minor-type, subset)
-
Has reproduction · 65
FusionQ: a novel approach for gene fusion detection and quantification from paired-end RNA-Seq.
PMID 23768108 · PMC3691734 · BMC bioinformatics · 2013 · 8 claims · 8 setups
FusionQ is a novel tool that detects gene fusions, constructs chimerical transcript structures, and estimates their abundances from paired-end RNA-Seq data.
-
Full-text index only
InSite: a computational method for identifying protein-protein interaction binding sites on a proteome-wide scale.
PMID 17868464 · PMC2375030 · Genome biology · 2007 · 8 claims · 8 setups
InSite predicts protein-pair-specific binding motifs ('Motif M on protein A binds to protein B') by integrating heterogeneous PPI and motif-motif interaction evidence within a Bayesian network trained by EM