Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
SNAP: predict effect of non-synonymous polymorphisms on function.
PMID 17526529 · PMC1920242 · Nucleic acids research · 2007 · 7 claims · 8 setups
SNAP, a neural network-based method using sequence-derived information, predicts whether a non-synonymous SNP is neutral or non-neutral for protein function
-
Has reproduction · 87
Genetic demultiplexing of pooled single-cell RNA-sequencing samples in cancer facilitates effective experimental design.
PMID 34553212 · PMC8458035 · GigaScience · 2021 · 8 claims · 6 setups
Genetic variation-based demultiplexing tools can be effectively deployed on pooled scRNA-seq experimental designs in cancer tissue (HGSOC and lung adenocarcinoma) despite somatic variation.
-
Has reproduction · 50
MEDUSA: A Pipeline for Sensitive Taxonomic Classification and Flexible Functional Annotation of Metagenomic Shotgun Sequences.
PMID 35330728 · PMC8940201 · Frontiers in genetics · 2022 · 7 claims · 6 setups
MEDUSA correctly identifies more species than MEGAN 6 CE, especially less abundant species.
-
Full-text index only
Improved reconstruction of transcripts and coding sequences from RNA-seq data.
PMID 41700087 · PMC12910111 · Nucleic acids research · 2026 · 7 claims · 3 setups
GeMoSeq combines combinatorial enumeration of candidate transcripts, splitting heuristics, and likelihood-based (EM) quantification for transcript reconstruction from RNA-seq data
-
Has reproduction · 61
TEMP: a computational method for analyzing transposable element polymorphism in populations.
PMID 24753423 · PMC4066757 · Nucleic acids research · 2014 · 8 claims · 8 setups
TEMP combines pair-end (discordant) read and split (soft-clipped) read information to identify both presence and absence of TE insertions in genomic DNA from heterogeneous/pooled samples.
-
Has reproduction · 37
A Bayesian approach to accurate and robust signature detection on LINCS L1000 data.
PMID 32003771 · PMC7203754 · Bioinformatics (Oxford, England) · 2020 · 7 claims · 4 setups
A novel Bayesian peak deconvolution algorithm gives unbiased likelihood estimations for peak locations and derives probability-based z-scores.
-
Full-text index only
sCellST predicts single-cell gene expression from H& E images.
PMID 41513659 · PMC12858858 · Nature communications · 2026 · 7 claims · 6 setups
sCellST is a weakly supervised (Multiple Instance Learning) deep learning framework that predicts single-cell gene expression from H&E images alone, trained using paired spatial transcriptomics (Visium) and H&E slides
-
Full-text index only
An end-to-end generalizable deep learning framework to comprehensively analyze transcriptional regulation.
PMID 41922356 · PMC13212934 · Nature communications · 2026 · 8 claims · 7 setups
BioSeq2Seq is a transformer-based deep learning framework that predicts genome-wide transcriptional regulatory profiles at 128-bp resolution by jointly using RO-seq data and DNA sequence as tri-modal input