Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Gut microbiome signatures associated with depression and obesity.
PMID 41615149 · PMC13011431 · mSystems · 2026 · 8 claims · 4 setups
Taxonomic gut microbiome profiles classify depressed vs non-depressed subjects with a balanced accuracy of 0.90 using machine learning
-
Full-text index only
Robust and efficient annotation of cell states through gene signature scoring.
PMID 41708334 · PMC12951948 · Genome research · 2026 · 8 claims · 8 setups
Established scoring methods (Seurat, SCANPY, UCell, JASMINE) fail to provide robust and comparable score distributions across diverse signatures and experimental conditions, precluding accurate unsupervised cell-state annotation.
-
Full-text index only
A comprehensive toolkit for analyzing cell-free DNA genomic sequencing data in liquid biopsy.
PMID 42111187 · PMC13157187 · iScience · 2026 · 8 claims · 8 setups
cfDNAanalyzer integrates feature extraction, feature processing/selection, and machine learning model building into a single one-command-line toolkit for cfDNA genomic sequencing data
-
Full-text index only
EXPLANA: a user-friendly workflow for EXPLoratory ANAlysis and feature selection in cross-sectional and longitudinal microbiome studies.
PMID 41416890 · PMC12766912 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 3 setups
EXPLANA is a feature selection workflow for longitudinal microbiome studies (LMS) that supports numerical and categorical data and also accommodates cross-sectional studies.
-
Has reproduction
WDR11, a WD protein that interacts with transcription factor EMX1, is mutated in idiopathic hypogonadotropic hypogonadism and Kallmann syndrome.
PMID 20887964 · PMC2948809 · American journal of human genetics · 2010 · 5 claims · 3 setups
WDR11 is a gene involved in human puberty, identified via the chromosomal breakpoint of a balanced t(10;12) translocation in a Kallmann syndrome subject.
-
Has reproduction · 73
treeclimbR pinpoints the data-dependent resolution of hierarchical hypotheses.
PMID 34001188 · PMC8127214 · Genome biology · 2021 · 7 claims · 6 setups
treeclimbR proposes multiple candidate resolutions on a tree and selects the optimal one in a data-driven manner using three criteria (FDR-controlling range of t, number of rejected leaves, fewest internal nodes)
-
Has reproduction · 100
Prediction of Antimicrobial Resistance in Gram-Negative Bacteria From Whole-Genome Sequencing Data.
PMID 32528441 · PMC7262952 · Frontiers in microbiology · 2020 · 8 claims · 5 setups
WGS-derived antibiotic resistance gene (ARG) coverage can be used to predict antimicrobial resistance in Gram-negative bacteria via machine learning
-
Has reproduction · 44
An OMICs-based meta-analysis to support infection state stratification.
PMID 33560295 · PMC8388022 · Bioinformatics (Oxford, England) · 2021 · 7 claims · 6 setups
Multi-class Random Forest models built from meta-analyzed blood gene expression data can predict infection state (bacterial/viral/none) with high accuracy, correctly classifying 93% of bacterial and 89% of viral samples in the best model.
-
Has reproduction
Fast, accurate, and racially unbiased pan-cancer tumor-only variant calling with tabular machine learning.
PMID 36611079 · PMC9825621 · NPJ precision oncology · 2023 · 8 claims · 8 setups
Tree-based (XGBoost, LightGBM) and deep-learning (TabNet) tabular ML classifiers achieve state-of-the-art somatic vs germline classification in tumor-only WES samples, outperforming PureCN.
-
Full-text index only
CONTRAST: a discriminative, phylogeny-free approach to multiple informant de novo gene prediction.
PMID 18096039 · PMC2246271 · Genome biology · 2007 · 8 claims · 5 setups
CONTRAST predicts exact coding region structures for 65% more human genes than the previous state-of-the-art de novo predictor (N-SCAN)
-
Has reproduction · 58
A comparative study of techniques for differential expression analysis on RNA-Seq data.
PMID 25119138 · PMC4132098 · PloS one · 2014 · 8 claims · 8 setups
edgeR performs slightly better than DESeq and Cuffdiff2 in terms of the ability to uncover true positives.
-
Full-text index only
DNA methylation biomarkers-based pan-cancer classifier: predictive modeling for cancer classification.
PMID 42152108 · PMC13185202 · Genome medicine · 2026 · 8 claims · 5 setups
Relatively simple ML models (logistic regression) outperform complex algorithms such as deep neural networks for methylation-based cancer classification
-
Full-text index only
The 32nd Annual Congress of the Society of Critical Care Medicine, 28 January - 2 February 2003, San Antonio, USA.
PMID 12720569 · PMC270663 · Critical care (London, England) · 2003 · 8 claims · 8 setups
Proteomics is more useful than genomics for identifying regulatory pathways and druggable targets in disease because transcriptional responses to different stimuli often converge while protein interaction networks reveal distinct regulatory nodes
-
Full-text index only
High-resolution array comparative genomic hybridization of single micrometastatic tumor cells.
PMID 18344524 · PMC2367728 · Nucleic acids research · 2008 · 7 claims · 8 setups
A protocol combining PCR-based whole genome amplification with arrays of highly purified BAC clones enables detection of DNA copy number changes in single cells
-
Full-text index only
A comprehensive sensitivity analysis of microarray breast cancer classification under feature variability.
PMID 19941644 · PMC2789744 · BMC bioinformatics · 2009 · 7 claims · 4 setups
Feature variability strongly influences breast cancer signature composition even when array platform and patient stratification are identical.
-
Full-text index only
Early feature extraction drives model performance in high-resolution chromatin accessibility prediction.
PMID 41526189 · PMC12951969 · Genome research · 2026 · 8 claims · 6 setups
Early feature extraction (via ConvNeXt V2 blocks), rather than downstream architecture type, is the primary determinant of prediction accuracy in high-resolution chromatin accessibility prediction.
-
Full-text index only
Transcriptome-based high-frequency recurrence index predicts frequent recurrence in non-muscle-invasive bladder cancer after Bacillus Calmette-Guérin therapy.
PMID 41749284 · PMC13040970 · BMC medicine · 2026 · 8 claims · 7 setups
A 75-gene HfRI signature predicts high-frequency recurrence (≥2 recurrences) in NMIBC patients