Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Patterns of somatic mutation in human cancer genomes.
PMID 17344846 · PMC2712719 · Nature · 2007 · 8 claims · 5 setups
Systematic resequencing of a large gene family (protein kinases) across diverse cancers reveals a larger repertoire of cancer genes than previously anticipated
-
Has reproduction · 59
Refining breast cancer biomarker discovery and drug targeting through an advanced data-driven approach.
PMID 38253993 · PMC10810249 · BMC bioinformatics · 2024 · 8 claims · 8 setups
The BGWO_SA_Ens algorithm (hybrid BGWO + simulated annealing with an ensemble classifier objective function) selects predictive breast cancer biomarker genes with high classification performance
-
Full-text index only
BRCA1 mutations in southern England.
PMID 9649133 · PMC2150412 · British journal of cancer · 1998 · 7 claims · 4 setups
Early age of onset combined with a strong family history is the most effective selection criterion for detecting BRCA1 mutations, more so than bilaterality alone
-
Full-text index only
Haplotype analysis of common variants in the BRCA1 gene and risk of sporadic breast cancer.
PMID 15743496 · PMC1064127 · Breast cancer research : BCR · 2005 · 7 claims · 5 setups
A common BRCA1 haplotype (haplotype 2, C A G G) is associated with a modest increase in sporadic breast cancer risk
-
Full-text index only
Positive selection for the male functionality of a co-retroposed gene in the hominoids.
PMID 19832993 · PMC2773790 · BMC evolutionary biology · 2009 · 8 claims · 8 setups
PIPSL is an extraordinary co-retroposed protein-coding gene that may participate in male-specific functions of humans and close relatives
-
Full-text index only
Proteome analysis enables separate clustering of normal breast, benign breast and breast cancer tissues.
PMID 12865921 · PMC2394238 · British journal of cancer · 2003 · 6 claims · 3 setups
Hierarchical cluster analysis of 2-DE proteome data can distinguish normal breast, benign breast, and breast cancer tissues based on protein expression profiles
-
Full-text index only
Incidence of mutation and deletion in topoisomerase II alpha mRNA of etoposide and mAMSA-resistant cell lines.
PMID 11676865 · PMC5926608 · Japanese journal of cancer research : Gann · 2001 · 7 claims · 6 setups
Acquired mutations of the topoisomerase IIα gene are an important and frequent mechanism of resistance to topoisomerase II inhibitors, independent of the degree of resistance.
-
Full-text index only
HapMap-based study of the 17q21 ERBB2 amplicon in susceptibility to breast cancer.
PMID 17117180 · PMC2360759 · British journal of cancer · 2006 · 6 claims · 5 setups
Common genetic variation (tSNPs and haplotypes) across the 400-kb 17q21 ERBB2 amplicon is not associated with breast cancer risk in British women.
-
Full-text index only
CLEAN: CLustering Enrichment ANalysis.
PMID 19640299 · PMC2734555 · BMC bioinformatics · 2009 · 8 claims · 4 setups
The gene-specific CLEAN score improves reproducibility of cluster analysis conclusions across independent datasets compared to the traditional cluster-wide score (cwCLEAN).
-
Full-text index only
Integrating complex genomic datasets and tumour cell sensitivity profiles to address a 'simple' question: which patients should get this drug?
PMID 20003409 · PMC2799438 · BMC medicine · 2009 · 8 claims · 5 setups
A panel of 48 genomically characterized breast cancer cell lines can model patient tumour heterogeneity to identify biomarkers predicting response to PG-11047
-
Has reproduction · 69
Automatic discovery of 100-miRNA signature for cancer classification using ensemble feature selection.
PMID 31533612 · PMC6751684 · BMC bioinformatics · 2019 · 7 claims · 8 setups
An ensemble feature selection method based on classifier consensus identifies a robust 100-miRNA signature from TCGA data.
-
Full-text index only
Interpretation of genomic data: questions and answers.
PMID 18582627 · PMC2528831 · Seminars in hematology · 2008 · 8 claims · 6 setups
The main challenge in using genomic technology in cancer research is not managing the volume of data but the proper design, analysis, and reporting of studies.
-
Has reproduction · 83
Gene-expression patterns in peripheral blood classify familial breast cancer susceptibility.
PMID 26538066 · PMC4634735 · BMC medical genomics · 2015 · 8 claims · 5 setups
A multigene peripheral-blood gene-expression biomarker accurately classifies which women from high-risk families develop familial breast cancer.
-
Has reproduction · 48
Comparative analysis of molecular signatures reveals a hybrid approach in breast cancer: Combining the Nottingham Prognostic Index with gene expressions into a hybrid signature.
PMID 35143511 · PMC8830616 · PloS one · 2022 · 8 claims · 6 setups
A hybrid signature combining the Nottingham Prognostic Index with SIS-selected gene expressions can be built in a data-driven fashion (NPI treated as a gene expression during feature selection).
-
Has reproduction · 100
Smart spatial omics (S2-omics) optimizes region of interest selection to capture molecular heterogeneity in diverse tissues.
PMID 41298871 · PMC12662399 · Nature cell biology · 2025 · 7 claims · 6 setups
S2-omics is an end-to-end workflow that automatically selects ROIs from H&E histology images to maximize molecular information content for spatial omics profiling.
-
Has reproduction · 92
A network-guided protocol to discover susceptibility genes in genome-wide association studies using stability selection.
PMID 36609152 · PMC9850185 · STAR protocols · 2023 · 5 claims · 5 setups
The protocol identifies genes that are both statistically associated with a phenotype and functionally interconnected in a biological network
-
Has reproduction · 67
Generative and integrative modeling for transcriptomics with formalin fixed paraffin embedded material.
PMID 41029822 · PMC12486589 · Journal of translational medicine · 2025 · 8 claims · 5 setups
fRNA-seq transcript counts are best fit by the negative binomial distribution, with little evidence supporting zero-inflated extensions
-
Full-text index only
Microarray analysis: genome-scale hypothesis scanning.
PMID 14551912 · PMC212694 · PLoS biology · 2003 · 8 claims · 5 setups
Microarrays can be used to both test and generate hypotheses, not merely to fish for candidate genes.
-
Full-text index only
Prioritization of candidate cancer genes--an aid to oncogenomic studies.
PMID 18710882 · PMC2566894 · Nucleic acids research · 2008 · 8 claims · 8 setups
Computational classifiers using combinations of protein conservation, gene structure, protein domains, protein interactions, and regulatory data can distinguish known cancer genes (CD/CR) from unlabelled human genes