Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Prioritization of candidate cancer genes--an aid to oncogenomic studies.
PMID 18710882 · PMC2566894 · Nucleic acids research · 2008 · 8 claims · 8 setups
Computational classifiers using combinations of protein conservation, gene structure, protein domains, protein interactions, and regulatory data can distinguish known cancer genes (CD/CR) from unlabelled human genes
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.
-
Full-text index only
Genome-wide prioritization of disease genes and identification of disease-disease associations from an integrated human functional linkage network.
PMID 19728866 · PMC2768980 · Genome biology · 2009 · 6 claims · 6 setups
Integrating 16 genomic features (32 sub-features) via a naïve Bayes classifier produces a genome-scale FLN of 21,657 human genes and 22,388,609 weighted links that outperforms any individual data source for inferring functional linkages.
-
Full-text index only
A comparison of classification methods for predicting Chronic Fatigue Syndrome based on genetic data.
PMID 19772600 · PMC2765429 · Journal of translational medicine · 2009 · 7 claims · 3 setups
The naive Bayes model with the wrapper-based feature selection approach performed best among all predictive models tested for distinguishing CFS from controls.
-
Full-text index only
SePaCS--a web-based application for classification of seroreactivity profiles.
PMID 17478503 · PMC1933220 · Nucleic acids research · 2007 · 8 claims · 4 setups
SePaCS is a freely available web-based tool that trains and applies multiple classification methods (4 Naive Bayes variants, SVM with RBF kernel, LDA, DLDA) to seroreactivity profiles and outputs results as a summary table plus a detailed PDF report
-
Full-text index only
Comprehensive analysis of the causal risk factor from hypertension associated with prognosis and therapeutic response in renal cell carcinoma by multi-omics analysis and validation.
PMID 41680825 · PMC12998095 · Biology direct · 2026 · 8 claims · 8 setups
A 48-gene cross-species hypertension (HTN) gene module identified from human and SHR rat scRNA-seq can classify ccRCC patients into two molecular subgroups with distinct survival and targeted therapy response
-
Has reproduction · 59
Integrative network modeling reveals mechanisms underlying T cell exhaustion.
PMID 32024856 · PMC7002445 · Scientific reports · 2020 · 8 claims · 7 setups
TCE arises from changes in diverse gene regulatory interactions across a shared network rather than dysregulation of a single gene
-
Full-text index only
Candidate vaccine sequences to represent intra- and inter-clade HIV-1 variation.
PMID 19812689 · PMC2753653 · PloS one · 2009 · 7 claims · 5 setups
Natural CTL immunodominance toward variable proteome regions increases epitope mismatch with challenge strains and recapitulates the escape-driven CTL failure seen in natural infection, contributing to HIV vaccine failure
-
Full-text index only
The specificity and polymorphism of the MHC class I prevents the global adaptation of HIV-1 to the monomorphic proteasome and TAP.
PMID 18949050 · PMC2569417 · PloS one · 2008 · 6 claims · 5 setups
Within individual hosts, proteasome and TAP escape mutations in HIV-1 occur frequently
-
Full-text index only
Critical evaluation of drug response prediction models with DrEval.
PMID 42120410 · PMC13168506 · Nature communications · 2026 · 8 claims · 6 setups
DrEval is a living open-source benchmarking pipeline for unbiased, biologically meaningful evaluation of cancer drug response prediction models, integrating standardized preprocessing, hyperparameter tuning, statistically rigorous evaluation, cross-study benchmarks, and ablation studies.
-
Full-text index only
Filtering high-throughput protein-protein interaction data using a combination of genomic features.
PMID 15833142 · PMC1127019 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A combination of three genomic features (interacting Pfam domains, GO annotations, sequence homology) using naive Bayesian networks predicts true protein-protein interactions with high sensitivity and good specificity.
-
Has reproduction · 86
Molecular Classification Models for Triple Negative Breast Cancer Subtype Using Machine Learning.
PMID 34575658 · PMC8472680 · Journal of personalized medicine · 2021 · 7 claims · 4 setups
TNBC can be divided into four gene-expression-defined subtypes: BLIA, BLIS, MES, and LAR
-
Full-text index only
HIT: a versatile proteomics platform for multianalyte phenotyping of cytokines, intracellular proteins and surface molecules.
PMID 18849997 · PMC3334282 · Nature medicine · 2008 · 8 claims · 8 setups
HIT uses oligonucleotide-tagged (Fab- or mSA-conjugated) antibodies, T7 polymerase amplification, and DNA microarray hybridization to indirectly measure multiple analytes in a fluid-phase multiplex format
-
Has reproduction · 83
Integrative transcriptomic and machine learning framework reveals candidate genes and potential mechanisms of aflatoxin B1 exposure in breast cancer.
PMID 41688730 · PMC12982753 · Scientific reports · 2026 · 7 claims · 8 setups
170 unique human AFB1 targets were identified by merging ChEMBL, SwissTargetPrediction, and PharmMapper predictions