Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 67
Comparison of Metagenomics and Metatranscriptomics Tools: A Guide to Making the Right Choice.
PMID 36553546 · PMC9777648 · Genes · 2022 · 8 claims · 1 setups
16S rRNA gene sequencing enables taxonomic identification of bacteria/archaea via hypervariable regions without amplifying human DNA, but is limited by short-read biases (GC bias, sequencing errors) and poor species-level resolution
-
Full-text index only
Frag'n'Flow: automated workflow for large-scale quantitative proteomics in high performance computing environments.
PMID 41486154 · PMC12828970 · BMC bioinformatics · 2026 · 8 claims · 8 setups
Frag'n'Flow is a Nextflow-based pipeline that encapsulates FragPipe, automating manifest/workflow generation, tool dependency management, and downstream analysis for HPC/cloud/cluster environments.
-
Has reproduction · 58
iCOMIC: a graphical interface-driven bioinformatics pipeline for analyzing cancer omics data.
PMID 35899080 · PMC9310080 · NAR genomics and bioinformatics · 2022 · 8 claims · 4 setups
iCOMIC provides a GUI-driven, Snakemake-based pipeline integrating multiple tools for DNA-Seq and RNA-Seq analysis with minimal command-line interaction.
-
Full-text index only
MetaPepticon: automated prediction of anticancer peptides from microbial genomes and metagenomes.
PMID 41918857 · PMC13034871 · PeerJ · 2026 · 7 claims · 6 setups
MetaPepticon is a modular, end-to-end Snakemake pipeline that predicts ACP candidates directly from raw genomic, metagenomic, transcriptomic, metatranscriptomic reads, assembled contigs, or peptide sequences.
-
Full-text index only
umite: fast quantification of Smart-seq3 libraries with improved UMI retrieval.
PMID 41692984 · PMC12989134 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 6 setups
umite offers efficient mismatch-tolerant (fuzzy) UMI detection that boosts UMI retrieval by 5%-15% compared to standard position-based matching
-
Full-text index only
DoBSeqWF: a framework for sensitive detection of individual genetic variation in pooled sequencing data.
PMID 41704565 · PMC12907731 · NAR genomics and bioinformatics · 2026 · 7 claims · 5 setups
DoBSeqWF, a Nextflow-based pipeline, processes pooled DoBSeq sequencing data through alignment, variant calling, machine-learning-based filtering, and variant pinpointing/assignment to individuals.
-
Full-text index only
Metapipeline-DNA: A comprehensive germline and somatic genomics Nextflow pipeline.
PMID 41850291 · PMC13030954 · Cell reports methods · 2026 · 8 claims · 7 setups
Metapipeline-DNA automates germline and somatic DNA sequencing analysis end-to-end, from raw reads through preprocessing, feature detection, QC, and visualization.
-
Has reproduction · 83
Macrel: antimicrobial peptide screening in genomes and metagenomes.
PMID 33384902 · PMC7751412 · PeerJ · 2020 · 8 claims · 8 setups
Macrel introduces a novel set of 22 peptide features (6 local, 16 global), including a new Free Energy Transition (FET) feature group, for AMP and hemolytic activity classification
-
Full-text index only
Alpseq: an open-source workflow to turbocharge nanobody discovery with high-throughput sequencing.
PMID 41631412 · PMC12885427 · mAbs · 2026 · 8 claims · 8 setups
alpseq is an open-source, end-to-end workflow combining a PCR-free sequencing library prep protocol with a Nextflow pre-processing pipeline and an R-based analysis/reporting module for nanobody NGS data.
-
Full-text index only
Eduomics: a Nextflow pipeline to simulate -omics data for education.
PMID 41816779 · PMC12972896 · NAR genomics and bioinformatics · 2026 · 8 claims · 4 setups
Eduomics is a Nextflow DSL2 pipeline that automates generation of validated variant-calling and RNA-seq datasets for education while abstracting away technical requirements
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.
-
Full-text index only
MobiCT: a UMI-based circulating tumor DNA analysis pipeline.
PMID 41503160 · PMC12770973 · NAR genomics and bioinformatics · 2026 · 7 claims · 7 setups
MobiCT is a Nextflow/nf-core UMI-based ctDNA pipeline (deduplication, alignment, variant calling with VarDict, annotation with VEP) achieving sensitivity, precision, and F1-score around 90% after comprehensive filtering.
-
Full-text index only
nf-core/crisprseq: a versatile pipeline for comprehensive analysis of CRISPR gene editing and screening assays.
PMID 41551929 · PMC12805889 · NAR genomics and bioinformatics · 2026 · 8 claims · 5 setups
nf-core/crisprseq is the first generic pipeline enabling analysis of the broad spectrum of CRISPR designs, from targeted gene edits (KO, KI, BE, PE) to large-scale functional screens
-
Has reproduction · 95
nf-rnaSeqCount: A Nextflow pipeline for obtaining raw read counts from RNA-seq data.
PMID 35574063 · PMC9097006 · South African computer journal = Suid-Afrikaanse rekenaartydskrif · 2021 · 7 claims · 5 setups
nf-rnaSeqCount is a portable, reproducible Nextflow pipeline that maps RNA-seq reads to a reference genome and quantifies gene abundance for differential expression analysis
-
Full-text index only
nf-core/viralmetagenome: A novel pipeline for untargeted viral genome reconstruction.
PMID 42057295 · PMC13141149 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 5 setups
nf-core/viralmetagenome is a Nextflow pipeline that automates untargeted reconstruction and variant analysis of eukaryotic DNA and RNA viruses from short-read metagenomic or hybridisation-capture data.
-
Full-text index only
GeneKeyDB: a lightweight, gene-centric, relational database to support data mining environments.
PMID 15790402 · PMC1274265 · BMC bioinformatics · 2005 · 8 claims · 6 setups
GeneKeyDB is a lightweight, gene-centric relational database that supports data mining and integration with computational analysis tools.
-
Full-text index only
ANOMALY: a Snakemake pipeline for identifying NuMTs from long-read sequencing data.
PMID 41647924 · PMC12869244 · NAR genomics and bioinformatics · 2026 · 8 claims · 8 setups
ANOMALY is a novel Snakemake pipeline for detecting NuMTs from long-read sequencing data
-
Has reproduction · 79
RetroSnake: A modular pipeline to detect human endogenous retroviruses in genome sequencing data.
PMID 36339261 · PMC9626663 · iScience · 2022 · 8 claims · 4 setups
RetroSnake is an end-to-end, modular, computationally efficient Snakemake pipeline for detecting HERV-K insertions in short-read NGS data, from raw alignment files to an annotated interactive HTML report
-
Full-text index only
Duplex-Indel: a Snakemake pipeline for somatic Indel calling in Tn5 transposase-based duplex sequencing data.
PMID 42046229 · PMC13171174 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 8 setups
Duplex-Indel is a Snakemake pipeline for somatic Indel calling from Tn5 transposase-based duplex sequencing data that requires consensus support from both DNA strands to minimize technical artifacts.
-
Has reproduction · 100
poreCov-An Easy to Use, Fast, and Robust Workflow for SARS-CoV-2 Genome Reconstruction via Nanopore Sequencing.
PMID 34394197 · PMC8355734 · Frontiers in genetics · 2021 · 8 claims · 8 setups
poreCov is an easy-to-use, fast, and robust Nextflow-based workflow for reference-based SARS-CoV-2 genome reconstruction and lineage determination from nanopore sequencing data