Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 42
KAGE: fast alignment-free graph-based genotyping of SNPs and short indels.
PMID 36195962 · PMC9531401 · Genome biology · 2022 · 7 claims · 7 setups
KAGE combines population-based kmer count modeling with single-variant prior adjustment into an alignment-free genotyper that matches the accuracy of the best existing alignment-free genotypers while being an order of magnitude faster.
-
Has reproduction · 43
Compression of structured high-throughput sequencing data.
PMID 24260313 · PMC3832420 · PloS one · 2013 · 8 claims · 7 setups
Leveraging an explicit data schema (separate field encoding, field modeling, template compression, domain modeling) enables stronger compression of HTS alignment data than general-purpose compression of serialized bytes.
-
Full-text index only
nf-core/proteinfamilies: a scalable pipeline for the generation of protein families.
PMID 41563008 · PMC12950615 · GigaScience · 2026 · 8 claims · 3 setups
nf-core/proteinfamilies is a scalable, parametrizable, open-source Nextflow pipeline that generates new protein families or assigns sequences to existing families using profile HMMs and MSAs
-
Full-text index only
Beyond blacklists: a critical assessment of exclusion set generation strategies and alternative approaches.
PMID 41826793 · PMC13020910 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 8 setups
Pre-generated Blacklist exclusion sets were difficult to reproduce due to sensitivity to input BAM data, aligner choice, and read length
-
Full-text index only
CIRCE: a scalable Python package to predict cis-regulatory DNA interactions from single-cell chromatin accessibility data.
PMID 41734268 · PMC12987762 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 5 setups
CIRCE re-implements the Cicero co-accessibility algorithm in Python, producing near-identical results while running much faster and using far less memory
-
Has reproduction · 68
Bayesian transcriptome assembly.
PMID 25367074 · PMC4397945 · Genome biology · 2014 · 8 claims · 8 setups
Bayesembler, a probabilistic transcriptome assembler built on a Bayesian model of the RNA sequencing process with Gibbs sampling over expressed candidates, abundances and read assignments, is introduced.