Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The gene expression landscape of disease genes.
PMID 41664080 · PMC12983648 · Genome biology · 2026 · 8 claims · 8 setups
In tissues and cell types with established disease relevance, disease genes show higher and more specific gene expression than control genes
-
Full-text index only
TE-SCALE: a comprehensive database for exploring transposable element expression across human cancers at single-cell resolution.
PMID 41296555 · PMC12807651 · Nucleic acids research · 2026 · 8 claims · 8 setups
TE-SCALE is a comprehensive single-cell database integrating TE expression across 20 human cancer types and 12 tissue origins
-
Full-text index only
Variant-resolved prediction of context-specific isoform variation with a graph-based attention model.
PMID 41547351 · PMC13069856 · Cell genomics · 2026 · 8 claims · 8 setups
Otari, an attention-based graph neural network trained on long-read transcriptomes across 30 tissues/brain regions, predicts tissue-specific differential isoform abundance
-
Has reproduction · 63
Community assessment of methods to deconvolve cellular composition from bulk gene expression.
PMID 39191725 · PMC11350143 · Nature communications · 2024 · 8 claims · 4 setups
Most deconvolution methods accurately predict coarse-grained immune/stromal cell populations from bulk expression.
-
Full-text index only
CLAMP: predicting specific protein-mediated chromatin loops in diverse species with a chromatin accessibility language model.
PMID 41555433 · PMC12903630 · Genome biology · 2026 · 8 claims · 8 setups
CLAMP, a chromatin-accessibility language model, predicts protein-mediated chromatin loops across 10 species, 18 proteins, and 24 cell types with superior performance versus existing methods.
-
Full-text index only
Retentive Network promotes efficient RNA language modeling of long sequences.
PMID 41814064 · PMC13111708 · Communications biology · 2026 · 8 claims · 6 setups
RNAret, a RetNet-based RNA language model with O(n) complexity, achieves training parallelism and low computational overhead while processing long RNA sequences
-
Full-text index only
ChromBERT: A foundation model for learning interpretable representations for context-specific transcriptional regulatory networks.
PMID 41592570 · PMC13069865 · Cell genomics · 2026 · 8 claims · 7 setups
ChromBERT is pre-trained via masked reconstruction on the Cistrome-Human-6K dataset (6,391 cistromes, 991 transcription regulators) to learn genome-wide interaction syntax of transcription regulators
-
Full-text index only
Cross-ancestry genome-wide association studies of liver function biomarkers uncover pleiotropic variants, systemic disease links and therapeutic targets.
PMID 41689074 · PMC13005531 · Genome medicine · 2026 · 8 claims · 8 setups
5,507 lead signals (P<5x10^-9) were identified for seven LFQBs across ancestries, including 210 novel loci