Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.
-
Full-text index only
CLEAN: CLustering Enrichment ANalysis.
PMID 19640299 · PMC2734555 · BMC bioinformatics · 2009 · 8 claims · 4 setups
The gene-specific CLEAN score improves reproducibility of cluster analysis conclusions across independent datasets compared to the traditional cluster-wide score (cwCLEAN).
-
Full-text index only
Recurring genomic breaks in independent lineages support genomic fragility.
PMID 17090315 · PMC1636669 · BMC evolutionary biology · 2006 · 6 claims · 6 setups
The propensity of a chromosomal region to break is significantly correlated among independent lineages, even after accounting for covariates like region length and functional class.
-
Full-text index only
Swarm intelligence based wavelet coefficient feature selection for mass spectral classification: an application to proteomics data.
PMID 19733729 · PMC2748225 · Analytica chimica acta · 2009 · 8 claims · 4 setups
ACA-based wavelet coefficient feature selection can achieve up to 100% classification accuracy on training, validating, and independent testing sets using only 5 selected features.
-
Full-text index only
Patterns of somatic mutation in human cancer genomes.
PMID 17344846 · PMC2712719 · Nature · 2007 · 8 claims · 5 setups
Systematic resequencing of a large gene family (protein kinases) across diverse cancers reveals a larger repertoire of cancer genes than previously anticipated
-
Has reproduction · 86
Molecular Classification Models for Triple Negative Breast Cancer Subtype Using Machine Learning.
PMID 34575658 · PMC8472680 · Journal of personalized medicine · 2021 · 6 claims · 4 setups
A training gene set of 719 unique upregulated DEGs (subtype-specific) can be used to build ML models that classify TNBC into BLIA, BLIS, MES, and LAR subtypes.
-
Has reproduction · 50
STAT1 and IL-7 as potential diagnostic biomarkers for distinguishing high-grade from low-grade serous ovarian cancer: a multi-cohort analysis.
PMID 42058211 · PMC13120972 · Frontiers in immunology · 2026 · 7 claims · 6 setups
STAT1 and IL-7 are differentially expressed immune-related genes that can distinguish HGSOC from LGSOC and may serve as ancillary diagnostic biomarkers.
-
Full-text index only
Identification of deleterious non-synonymous single nucleotide polymorphisms using sequence-derived information.
PMID 18588693 · PMC2446391 · BMC bioinformatics · 2008 · 8 claims · 5 setups
A decision tree built on 10 selected sequence-derived features classifies SAPs as Disease or Polymorphism with 82.6% accuracy and 0.607 MCC in cross-validation.
-
Has reproduction · 49
Integration of Transcriptomics With Interpretable Artificial Intelligence for Identifying Molecular Signatures of Physiological Stress in Sleep Deprivation.
PMID 42216239 · PMC13240488 · Journal of cellular and molecular medicine · 2026 · 8 claims · 8 setups
S100A3 is a robust candidate biomarker showing consistent discriminatory performance across the acute sleep deprivation training cohort, an independent sleep deprivation cohort, and a chronic insomnia cohort.
-
Full-text index only
Integration of text- and data-mining using ontologies successfully selects disease gene candidates.
PMID 15767279 · PMC1065256 · Nucleic acids research · 2005 · 7 claims · 6 setups
Integrating eVOC anatomical ontology-based text-mining of PubMed abstracts with data-mining of gene expression annotation successfully selects and prioritizes candidate disease genes
-
Has reproduction · 90
A Decentralized Kidney Transplant Biopsy Classifier for Transplant Rejection Developed Using Genes of the Banff-Human Organ Transplant Panel.
PMID 35619722 · PMC9128066 · Frontiers in immunology · 2022 · 6 claims · 6 setups
A random forest model trained solely on B-HOT panel genes (B-HOT Model) accurately classifies kidney transplant biopsies as NR, ABMR, or TCMR.
-
Full-text index only
Meta-analysis of inter-species liver co-expression networks elucidates traits associated with common human diseases.
PMID 20019805 · PMC2787626 · PLoS computational biology · 2009 · 8 claims · 8 setups
A novel semi-parametric meta-analysis method (based on a gene-centric Glass's d effect size) outperforms existing parametric and non-parametric meta-analysis methods at identifying functionally coherent gene pairs across species.
-
Full-text index only
Genetic variability of the P120' surface protein gene of Mycoplasma hominis isolates recovered from Tunisian patients with uro-genital and infertility disorders.
PMID 18053243 · PMC2225410 · BMC infectious diseases · 2007 · 7 claims · 5 setups
The P120' surface-exposed N-terminal region undergoes substantial genetic variability among Tunisian M. hominis clinical isolates
-
Has reproduction · 85
PDGFRA defines the mesenchymal stem cell Kaposi's sarcoma progenitors by enabling KSHV oncogenesis in an angiogenic environment.
PMID 31881074 · PMC6980685 · PLoS pathogens · 2019 · 8 claims · 8 setups
PDGFRA(+)/SCA-1(+) bone marrow-derived MSCs (Pα(+)S MSCs) are KS spindle-cell progenitors