Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 84
Pharokka: a fast scalable bacteriophage annotation tool.
PMID 36453861 · PMC9805569 · Bioinformatics (Oxford, England) · 2023 · 8 claims · 5 setups
Pharokka is a one-line, fast, scalable bacteriophage annotation tool producing standards-compliant outputs, installable via a two-line bioconda command
-
Full-text index only
Frag'n'Flow: automated workflow for large-scale quantitative proteomics in high performance computing environments.
PMID 41486154 · PMC12828970 · BMC bioinformatics · 2026 · 8 claims · 8 setups
Frag'n'Flow is a Nextflow-based pipeline that encapsulates FragPipe, automating manifest/workflow generation, tool dependency management, and downstream analysis for HPC/cloud/cluster environments.
-
Full-text index only
DupyliCate: mining, classifying, and characterizing gene duplications.
PMID 42209743 · PMC13219399 · Scientific reports · 2026 · 8 claims · 8 setups
DupyliCate is a Python tool for identifying and classifying gene duplication arrays, using BUSCO-based species-specific thresholds and offering integrated expression divergence and Ka/Ks analysis.
-
Has reproduction · 42
KAGE: fast alignment-free graph-based genotyping of SNPs and short indels.
PMID 36195962 · PMC9531401 · Genome biology · 2022 · 7 claims · 7 setups
KAGE combines population-based kmer count modeling with single-variant prior adjustment into an alignment-free genotyper that matches the accuracy of the best existing alignment-free genotypers while being an order of magnitude faster.
-
Has reproduction · 89
HTSQualC is a flexible and one-step quality control software for high-throughput sequencing data analysis.
PMID 34548573 · PMC8455540 · Scientific reports · 2021 · 8 claims · 5 setups
HTSQualC is a standalone, one-step QC software that performs filtering and trimming of raw HTS data in a single run
-
Full-text index only
Multi-context seeds enable fast and high-accuracy read mapping.
PMID 41764549 · PMC13059148 · Genome biology · 2026 · 7 claims · 5 setups
Multi-context seeds (MCS) allow storage of seeds with different lengths in the same index structure by splitting hash bits among strobes, enabling full and partial matches
-
Has reproduction · 78
QuasiFlow: a Nextflow pipeline for analysis of NGS-based HIV-1 drug resistance data.
PMID 36699347 · PMC9722223 · Bioinformatics advances · 2022 · 6 claims · 8 setups
QuasiFlow is a Nextflow pipeline that runs entirely locally via command-line tools and a local HIVdb database copy to analyze NGS-based HIV-1 drug resistance testing data.
-
Full-text index only
FEDRANN: effective long-read overlap detection based on dimensionality reduction and approximate nearest neighbors.
PMID 42102720 · PMC13201080 · GigaScience · 2026 · 8 claims · 6 setups
A pipeline combining IDF transformation, sparse random projection (SRP), and NNDescent (the FEDRANN strategy) enables accurate overlap detection across diverse long-read datasets
-
Full-text index only
pmid-42277027
PMID 42277027 · PMC13260335 · 8 claims · 5 setups
TOFU-MAaPO yields significantly more high-quality MAGs than metaFun, nf-core/mag, and ATLAS due to integration of multiple complementary binning tools with unified MAGScoT refinement
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.