Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 86
Molecular Classification Models for Triple Negative Breast Cancer Subtype Using Machine Learning.
PMID 34575658 · PMC8472680 · Journal of personalized medicine · 2021 · 6 claims · 4 setups
A training gene set of 719 unique upregulated DEGs (subtype-specific) can be used to build ML models that classify TNBC into BLIA, BLIS, MES, and LAR subtypes.
-
Has reproduction · 68
Molecular subtype of recurrent implantation failure reveals distinct endometrial etiology of female infertility.
PMID 40660214 · PMC12257665 · Journal of translational medicine · 2025 · 8 claims · 8 setups
RIF endometrial samples segregate into two reproducible molecular subtypes: an immune-driven subtype (RIF-I) and a metabolic-driven subtype (RIF-M)
-
Has reproduction · 59
Identifying Molecular Subtypes and 6-Gene Prognostic Signature Based on Hypoxia for Optimizing Targeted Therapies in Non-Small Cell Lung Cancer.
PMID 35509605 · PMC9058021 · International journal of general medicine · 2022 · 8 claims · 8 setups
NSCLC samples can be classified into two molecular subtypes (C1 and C2) based on hypoxia-related gene expression via consensus clustering
-
Full-text index only
High-resolution aCGH and expression profiling identifies a novel genomic subtype of ER negative breast cancer.
PMID 17925008 · PMC2246289 · Genome biology · 2007 · 7 claims · 8 setups
A novel subtype of high-grade ER-negative breast cancer exists, characterized by a low genomic instability index (GII)
-
Full-text index only
Molecular tumor profiling: translating genomic insights into clinical advances.
PMID 15287965 · PMC507868 · Genome biology · 2004 · 8 claims · 8 setups
Gene-expression profiling can distinguish BRCA1- and BRCA2-linked breast tumors from sporadic breast tumors with similar hormone-receptor status
-
Full-text index only
Towards precise classification of cancers based on robust gene functional expression profiles.
PMID 15774002 · PMC1274255 · BMC bioinformatics · 2005 · 6 claims · 7 setups
Functional expression profiles (FEPs) achieve comparable or better classification performance than conventional gene expression profiles (GEPs) across four public microarray datasets
-
Full-text index only
Predicting survival within the lung cancer histopathological hierarchy using a multi-scale genomic model of development.
PMID 16800721 · PMC1483910 · PLoS medicine · 2006 · 8 claims · 8 setups
Multi-scale genomic similarities exist between four human lung cancer subtypes and the developing mouse lung, and these similarities are prognostically meaningful.
-
Full-text index only
Expression genomics in breast cancer research: microarrays at the crossroads of biology and medicine.
PMID 17397520 · PMC1868923 · Breast cancer research : BCR · 2007 · 8 claims · 8 setups
Genome-wide expression microarray studies reveal transcriptional networks/signatures that explain breast cancer biological and clinical heterogeneity
-
Full-text index only
Integrated genomics identifies five medulloblastoma subtypes with distinct genetic profiles, pathway signatures and clinicopathological features.
PMID 18769486 · PMC2518524 · PloS one · 2008 · 8 claims · 4 setups
Unsupervised clustering of expression profiles from 62 medulloblastomas identifies 5 molecular subtypes (A–E)
-
Full-text index only
Seeded Bayesian Networks: constructing genetic networks from microarray data.
PMID 18601736 · PMC2474592 · BMC systems biology · 2008 · 8 claims · 4 setups
Seeding Bayesian Network analysis with prior networks derived from literature and/or PPI data improves recovery of known gene-gene interactions compared to BN analysis without a seed
-
Full-text index only
Proteomics as a method for early detection of cancer: a review of proteomics, exhaled breath condensate, and lung cancer screening.
PMID 18095050 · PMC2150625 · Journal of general internal medicine · 2008 · 8 claims · 7 setups
Protein expression is closely aligned with cellular activity, unlike genomic changes which may have no functional significance
-
Full-text index only
Integrating complex genomic datasets and tumour cell sensitivity profiles to address a 'simple' question: which patients should get this drug?
PMID 20003409 · PMC2799438 · BMC medicine · 2009 · 8 claims · 5 setups
A panel of 48 genomically characterized breast cancer cell lines can model patient tumour heterogeneity to identify biomarkers predicting response to PG-11047
-
Full-text index only
Emerging translational bioinformatics: knowledge-guided biomarker identification for cancer diagnostics.
PMID 19964620 · PMC5003034 · Annual International Conference of the IEEE Engineering in Medicine and Biology Society. IEEE Engineering in Medicine and Biology Society. Annual International Conference · 2009 · 8 claims · 3 setups
omniBiomarker, a web-based application, uses prior biological knowledge to identify the most biologically relevant gene ranking metric for a given clinical problem
-
Full-text index only
Overview of microarray analysis of gene expression and its applications to cervical cancer investigation.
PMID 18182341 · PMC7129792 · Taiwanese journal of obstetrics & gynecology · 2007 · 6 claims · 5 setups
Oligonucleotide microarray and cDNA microarray are the two main microarray platforms used to study gene expression genome-wide.
-
Has reproduction · 87
CoINcIDE: A framework for discovery of patient subtypes across multiple datasets.
PMID 26961683 · PMC4784276 · Genome medicine · 2016 · 8 claims · 6 setups
CoINcIDE is a methodological framework that discovers replicable patient subtypes (meta-clusters) across multiple datasets by finding consensus across dataset-specific clusterings, requiring no between-dataset transformations.
-
Has reproduction · 96
Scalable Prediction of Acute Myeloid Leukemia Using High-Dimensional Machine Learning and Blood Transcriptomics.
PMID 31918046 · PMC6992905 · iScience · 2020 · 8 claims · 8 setups
Data-driven, high-dimensional ML approaches that learn multivariate signatures directly from genome-wide transcriptomic data (no prior gene selection) yield accurate and robust AML classifiers.
-
Full-text index only
Microarray analysis: genome-scale hypothesis scanning.
PMID 14551912 · PMC212694 · PLoS biology · 2003 · 8 claims · 5 setups
Microarrays can be used to both test and generate hypotheses, not merely to fish for candidate genes.
-
Full-text index only
Expoldb: expression linked polymorphism database with inbuilt tools for analysis of expression and simple repeats.
PMID 17038195 · PMC1618849 · BMC genomics · 2006 · 8 claims · 6 setups
EXPOLDB is a novel database integrating human gene expression variability data (including monozygotic twin comparisons) with (TG/CA)n repeat polymorphism information
-
Full-text index only
Independent component analysis reveals new and biologically significant structures in micro array data.
PMID 16762055 · PMC1557674 · BMC bioinformatics · 2006 · 7 claims · 8 setups
ICA applied to three microarray datasets reveals many biologically significant components, including low-ranking ones not obvious by rank alone
-
Has reproduction · 77
Comparison of RNA-Seq by poly (A) capture, ribosomal RNA depletion, and DNA microarray for expression profiling.
PMID 24888378 · PMC4070569 · BMC genomics · 2014 · 8 claims · 8 setups
Ribo-Zero-Seq removes rRNA with efficiency comparable to poly(A)-based mRNA-Seq in both FF and FFPE RNA, whereas DSN-Seq leaves significantly more rRNA and shows greater variation.