Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Benchmarking component choices for unpaired single cell RNA and epigenomic integration.
PMID 41987329 · PMC13192178 · Genome biology · 2026 · 7 claims · 8 setups
Gene activity scores (GAS) show limited correlation with actual gene expression but effectively preserve cellular neighborhood structure and support clustering.
-
Has reproduction · 95
Mouse-Geneformer: A deep learning model for mouse single-cell transcriptome and its cross-species utility.
PMID 40106407 · PMC11964219 · PLoS genetics · 2025 · 7 claims · 6 setups
Mouse-Geneformer, a Transformer Encoder model pre-trained via masked-token self-supervised learning on mouse-Genecorpus-20M, was successfully constructed following the original human Geneformer architecture.
-
Has reproduction · 91
Whole genome and transcriptome maps of the entirely black native Korean chicken breed Yeonsan Ogye.
PMID 30010758 · PMC6065499 · GigaScience · 2018 · 8 claims · 6 setups
A draft genome (Ogye_1.1) was assembled using a hybrid de novo method combining high-depth Illumina short reads (376.6X) and low-depth PacBio long reads (9.7X)
-
Full-text index only
GENCODE: producing a reference annotation for ENCODE.
PMID 16925838 · PMC1810553 · Genome biology · 2006 · 8 claims · 8 setups
GENCODE annotation combines initial manual annotation by HAVANA, experimental validation, and refinement based on results to identify protein-coding genes in ENCODE regions
-
Full-text index only
FLASH-MM: fast and scalable single-cell differential expression analysis using linear mixed-effects models.
PMID 41644528 · PMC12982622 · Nature communications · 2026 · 8 claims · 6 setups
FLASH-MM produces LMM parameter estimates identical to lmer (lme4) up to the sixth decimal place while being 50- to 140-fold faster as sample size increases from 20,000 to 120,000 cells
-
Full-text index only
Integrative Learning of Disentangled Representations from Single-Cell RNA-Sequencing Datasets.
PMID 41971949 · PMC13068006 · Computational and structural biotechnology journal · 2026 · 8 claims · 6 setups
spVIPES decomposes unpaired scRNA-seq datasets with nonmatching features into shared and private latent representations using a Product of Experts framework
-
Has reproduction · 100
Smart spatial omics (S2-omics) optimizes region of interest selection to capture molecular heterogeneity in diverse tissues.
PMID 41298871 · PMC12662399 · Nature cell biology · 2025 · 7 claims · 6 setups
S2-omics is an end-to-end workflow that automatically selects ROIs from H&E histology images to maximize molecular information content for spatial omics profiling.