Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A novel wavelet-based thresholding method for the pre-processing of mass spectrometry data that accounts for heterogeneous noise.
PMID 18615428 · PMC2855839 · Proteomics · 2008 · 6 claims · 4 setups
Noise in SELDI-TOF/MALDI-TOF mass spectrometry data is heteroscedastic across the m/z range, with larger variance at lower m/z values, contrary to the homogeneous noise assumption of existing wavelet denoising methods.
-
Has reproduction · 68
Enhancing cell subpopulation discovery in cancer by integrating single-cell transcriptome and expressed variants.
PMID 41647537 · PMC12869734 · Fundamental research · 2026 · 6 claims · 3 setups
scCluster, an end-to-end deep clustering model integrating gene expression and expressed variant (eSNP) features, stratifies cell subpopulations in cancer scRNA-seq data.
-
Has reproduction · 76
Correcting scale distortion in RNA sequencing data.
PMID 39875825 · PMC11776150 · BMC bioinformatics · 2025 · 8 claims · 8 setups
Local averaging reveals expression-level-dependent biases that differ from sample to sample across all RNA-seq datasets studied, and are not corrected by conventional normalization (TPM/FPKM)
-
Full-text index only
Swarm intelligence based wavelet coefficient feature selection for mass spectral classification: an application to proteomics data.
PMID 19733729 · PMC2748225 · Analytica chimica acta · 2009 · 8 claims · 4 setups
ACA-based wavelet coefficient feature selection can achieve up to 100% classification accuracy on training, validating, and independent testing sets using only 5 selected features.
-
Has reproduction · 95
nf-rnaSeqCount: A Nextflow pipeline for obtaining raw read counts from RNA-seq data.
PMID 35574063 · PMC9097006 · South African computer journal = Suid-Afrikaanse rekenaartydskrif · 2021 · 7 claims · 5 setups
nf-rnaSeqCount is a portable, reproducible Nextflow pipeline that maps RNA-seq reads to a reference genome and quantifies gene abundance for differential expression analysis
-
Has reproduction · 87
Genetic demultiplexing of pooled single-cell RNA-sequencing samples in cancer facilitates effective experimental design.
PMID 34553212 · PMC8458035 · GigaScience · 2021 · 8 claims · 7 setups
Genetic variation–based demultiplexing tools can be effectively deployed on cancer scRNA-seq tissue using a pooled experimental design, achieving high recall at acceptable precision-recall tradeoffs in both high-CNV (HGSOC) and high-SNV (lung adenocarcinoma) cancers, even with extremely high doublet proportions.
-
Has reproduction · 88
Human methylome variation across Infinium 450K data on the Gene Expression Omnibus.
PMID 33937763 · PMC8061458 · NAR genomics and bioinformatics · 2021 · 8 claims · 8 setups
Among annotated HM450K GEO samples, about two-thirds were from blood, one-quarter from brain, and about one-third were from cancer patients.
-
Full-text index only
On the analysis of glycomics mass spectrometry data via the regularized area under the ROC curve.
PMID 18076765 · PMC2211327 · BMC bioinformatics · 2007 · 8 claims · 4 setups
The TGDR-AUC algorithm regularizes the empirical AUC by replacing the non-differentiable 0-1 loss with a smooth sigmoid surrogate function and applies constrained threshold gradient descent regularization
-
Has reproduction · 67
Generative and integrative modeling for transcriptomics with formalin fixed paraffin embedded material.
PMID 41029822 · PMC12486589 · Journal of translational medicine · 2025 · 8 claims · 6 setups
The negative binomial distribution best fits fRNA-seq transcript counts, with little evidence supporting zero-inflated extensions