Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 90
LoRA-TV: read depth profile-based clustering of tumor cells in single-cell sequencing.
PMID 38877886 · PMC11179121 · Briefings in bioinformatics · 2024 · 6 claims · 2 setups
LoRA-TV jointly processes read-depth profiles of all cells by stacking them into a matrix and applying low-rank approximation plus total-variation smoothing to capture shared genomic signatures for clustering.
-
Has reproduction · 10
RADAR: differential analysis of MeRIP-seq data with a random effect model.
PMID 31870409 · PMC6927177 · Genome biology · 2019 · 8 claims · 6 setups
RADAR is a novel analytical tool for differential methylation analysis of MeRIP-seq data combining gene-level INPUT normalization with a Poisson random effect model.
-
Full-text index only
Iterative class discovery and feature selection using Minimal Spanning Trees.
PMID 15355552 · PMC520744 · BMC bioinformatics · 2004 · 7 claims · 5 setups
Iterating between MST-based clustering and t-statistic feature selection removes noise genes step-wise while sharpening the sample clustering
-
Full-text index only
A statistical change point model approach for the detection of DNA copy number variations in array CGH data.
PMID 19875853 · PMC4154476 · IEEE/ACM transactions on computational biology and bioinformatics · 2009 · 7 claims · 4 setups
A novel mean and variance change point model (MVCM) is proposed to detect CNVs/breakpoints in aCGH data.
-
Full-text index only
Design and analysis issues in genome-wide somatic mutation studies of cancer.
PMID 18692126 · PMC2820387 · Genomics · 2009 · 6 claims · 4 setups
Two-stage (discovery + validation) sequencing designs efficiently allocate resources and can produce highly informative candidate driver gene lists even with relatively small sample sizes.
-
Has reproduction · 87
CoINcIDE: A framework for discovery of patient subtypes across multiple datasets.
PMID 26961683 · PMC4784276 · Genome medicine · 2016 · 8 claims · 6 setups
CoINcIDE is a methodological framework that discovers replicable patient subtypes (meta-clusters) across multiple datasets by finding consensus across dataset-specific clusterings, requiring no between-dataset transformations.
-
Full-text index only
Statistical challenges in preprocessing in microarray experiments in cancer.
PMID 18829474 · PMC3529914 · Clinical cancer research : an official journal of the American Association for Cancer Research · 2008 · 8 claims · 7 setups
Choice of pre-processing method materially changes which features are found significantly associated with survival in the Beer et al. lung cancer microarray dataset
-
Has reproduction · 63
Community assessment of methods to deconvolve cellular composition from bulk gene expression.
PMID 39191725 · PMC11350143 · Nature communications · 2024 · 8 claims · 4 setups
Most deconvolution methods accurately predict coarse-grained immune/stromal cell populations from bulk expression.
-
Has reproduction · 100
Intratumoral heterogeneity in microsatellite instability status at single-cell resolution.
PMID 41767255 · PMC12936829 · iScience · 2026 · 8 claims · 8 setups
MSI status can be heterogeneous at the single-cell level within a tumor, challenging its use as a binary biomarker
-
Has reproduction · 59
Integrative network modeling reveals mechanisms underlying T cell exhaustion.
PMID 32024856 · PMC7002445 · Scientific reports · 2020 · 8 claims · 6 setups
An integrative, literature-curated and data-driven gene regulatory network underlies CD8+ T cell exhaustion and accurately captures expression states in chronic infection and tumor settings.
-
Has reproduction · 85
Digital sorting of complex tissues for cell type-specific gene expression profiles.
PMID 23497278 · PMC3626856 · BMC bioinformatics · 2013 · 8 claims · 8 setups
The Digital Sorting Algorithm (DSA) deconvolves mixed tissue expression into cell type-specific profiles using only marker genes, without requiring prior knowledge of cell type frequencies or in vitro pure-cell profiles.
-
Full-text index only
Estimation of relevant variables on high-dimensional biological patterns using iterated weighted kernel functions.
PMID 18509521 · PMC2396875 · PloS one · 2008 · 7 claims · 6 setups
wKIERA combines a weighted-kernel discriminant (kernel perceptron) with an iterative stochastic probability estimation-of-distribution algorithm to estimate a relevance distribution over variables