Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
PolyAseqTrap: a universal tool for genome-wide identification and quantification of polyadenylation sites from different 3' end sequencing data.
PMID 41620776 · PMC12947541 · Genome biology · 2026 · 6 claims · 7 setups
PolyAseqTrap is a universal R package for identifying and quantifying polyA sites from diverse 3' end sequencing data
-
Full-text index only
ChromBERT: A foundation model for learning interpretable representations for context-specific transcriptional regulatory networks.
PMID 41592570 · PMC13069865 · Cell genomics · 2026 · 8 claims · 7 setups
ChromBERT is pre-trained via masked reconstruction on the Cistrome-Human-6K dataset (6,391 cistromes, 991 transcription regulators) to learn genome-wide interaction syntax of transcription regulators
-
Has reproduction · 74
SpaGene: A Deep Adversarial Framework for Spatial Gene Imputation.
PMID 42146899 · PMC13176606 · Computational and structural biotechnology journal · 2026 · 8 claims · 6 setups
SpaGene improves average PCC and SSIM and reduces RMSE compared to 6 baseline methods (SpaGE, gimVI, Tangram, VISTA, spRefine, stDiff) across 8 diverse ST-SC dataset pairs under gene-holdout evaluation.
-
Has reproduction · 53
Combining evidence of preferential gene-tissue relationships from multiple sources.
PMID 23950964 · PMC3741196 · PloS one · 2013 · 8 claims · 8 setups
A high-level integration approach combining three methods across four human microarray datasets, merged by consensus voting and a rule-based inner/total score, predicts preferentially expressed genes while reducing method- and study-specific bias.
-
Full-text index only
Hi-Compass: a depth-aware deep learning framework for predicting cell-type-specific 3D genome organization from single-cell to spatial resolution.
PMID 41980945 · PMC13250166 · Nature communications · 2026 · 8 claims · 8 setups
Hi-Compass predicts cell-type-specific Hi-C contact maps using only ATAC-seq as cell-type-specific input, plus DNA sequence and a generalized CTCF binding profile
-
Full-text index only
Separating selection from mutation in antibody language models.
PMID 41944291 · PMC13056363 · eLife · 2026 · 8 claims · 6 setups
Masked antibody language models such as AbLang2 are biased by nucleotide-level mutation processes (germline memorization, codon table, SHM rate variation)
-
Full-text index only
Evaluating the Utilities of Foundation Models in Single-Cell Data Analysis.
PMID 41869863 · PMC13170260 · Advanced science (Weinheim, Baden-Wurttemberg, Germany) · 2026 · 8 claims · 8 setups
Among ten/eleven evaluated single-cell FMs, scGPT, Geneformer, and CellFM are the top models considering both performance and user accessibility
-
Has reproduction · 89
Graph Random Forest: A Graph Embedded Algorithm for Identifying Highly Connected Important Features.
PMID 37509188 · PMC10377046 · Biomolecules · 2023 · 8 claims · 6 setups
GRF identifies effective features that form highly connected sub-graphs on the underlying biological network
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Full-text index only
Deep-learning prediction of gene expression from personal genomes.
PMID 41495833 · PMC12869966 · Genome biology · 2026 · 8 claims · 8 setups
Fine-tuning Enformer on paired personal WGS and RNA-seq data (Variformer) corrects Enformer's failure to predict inter-individual gene expression differences across held-out people.