Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 67
binny: an automated binning algorithm to recover high-quality genomes from complex metagenomic datasets.
PMID 36239393 · PMC9677464 · Briefings in bioinformatics · 2022 · 8 claims · 8 setups
binny outperforms or is highly competitive with commonly used and state-of-the-art binning methods (MetaBAT2, MaxBin2, CONCOCT, VAMB, SemiBin, MetaDecoder)
-
Has reproduction · 78
Taxonomic analysis of metagenomic data with kASA.
PMID 33784400 · PMC8266618 · Nucleic acids research · 2021 · 8 claims · 3 setups
kASA achieves high sensitivity and precision by using an amino acid-like encoding of k-mers together with a range of multiple k's
-
Has reproduction · 89
miRge 2.0 for comprehensive analysis of microRNA sequencing data.
PMID 30153801 · PMC6112139 · BMC bioinformatics · 2018 · 8 claims · 6 setups
An SVM-based novel miRNA detection model achieves an average MCC of 0.939 across 32 human cell datasets and outperforms miRDeep2 and miRAnalyzer on phylogenetic conservation of predicted miRNAs
-
Full-text index only
Trimmomatic: a decade of feature-rich, high-performance NGS read preprocessing.
PMID 42178219 · PMC13242794 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 4 setups
A high-performance multithreading architecture allows batches of read pairs to be processed independently by a pool of worker threads, scaling efficiently with available hardware.
-
Has reproduction · 50
MEDUSA: A Pipeline for Sensitive Taxonomic Classification and Flexible Functional Annotation of Metagenomic Shotgun Sequences.
PMID 35330728 · PMC8940201 · Frontiers in genetics · 2022 · 7 claims · 6 setups
MEDUSA correctly identifies more species than MEGAN 6 CE, especially less abundant species.
-
Has reproduction · 83
MetaGT: A pipeline for de novo assembly of metatranscriptomes with the aid of metagenomic data.
PMID 36386613 · PMC9651917 · Frontiers in microbiology · 2022 · 7 claims · 4 setups
MetaGT is a pipeline that combines metatranscriptomic and metagenomic data from the same sample to assemble complete transcript sequences
-
Has reproduction · 89
HTSQualC is a flexible and one-step quality control software for high-throughput sequencing data analysis.
PMID 34548573 · PMC8455540 · Scientific reports · 2021 · 8 claims · 5 setups
HTSQualC is a standalone, one-step QC software that performs filtering and trimming of raw HTS data in a single run
-
Has reproduction · 95
nf-rnaSeqCount: A Nextflow pipeline for obtaining raw read counts from RNA-seq data.
PMID 35574063 · PMC9097006 · South African computer journal = Suid-Afrikaanse rekenaartydskrif · 2021 · 7 claims · 5 setups
nf-rnaSeqCount is a portable, reproducible Nextflow pipeline that maps RNA-seq reads to a reference genome and quantifies gene abundance for differential expression analysis
-
Has reproduction · 42
KAGE: fast alignment-free graph-based genotyping of SNPs and short indels.
PMID 36195962 · PMC9531401 · Genome biology · 2022 · 7 claims · 7 setups
KAGE combines population-based kmer count modeling with single-variant prior adjustment into an alignment-free genotyper that matches the accuracy of the best existing alignment-free genotypers while being an order of magnitude faster.
-
Has reproduction · 88
Wochenende - modular and flexible alignment-based shotgun metagenome analysis.
PMID 36368923 · PMC9650795 · BMC genomics · 2022 · 8 claims · 6 setups
Wochenende is a modular, transparent alignment-based pipeline for shotgun metagenome analysis supporting short and long reads across all kingdoms of life
-
Has reproduction · 89
Improved eukaryotic detection compatible with large-scale automated analysis of metagenomes.
PMID 37032329 · PMC10084625 · Microbiome · 2023 · 8 claims · 7 setups
MAPQ ≥30 filtering improves precision but substantially reduces recall, especially for unrepresented/divergent eukaryotic taxa
-
Has reproduction · 90
Optimal Dual RNA-Seq Mapping for Accurate Pathogen Detection in Complex Eukaryotic Hosts.
PMID 39959292 · PMC11825298 · Bio-protocol · 2025 · 7 claims · 6 setups
Mapping adapter-trimmed reads first to the pathogen genome recovers more pathogen reads than the traditional host-first mapping approach.
-
Full-text index only
Benchmarking methods for genome annotation using nanopore direct RNA in a non-model crop plant.
PMID 41800382 · PMC12967217 · Bioinformatics advances · 2026 · 6 claims · 8 setups
Annotation tools show substantial variation in isoform detection, structural completeness, splicing classification, and handling of 5' read truncation when applied to plant dRNA-seq data.
-
Has reproduction · 93
aCLImatise: automated generation of tool definitions for bioinformatics workflows.
PMID 33325479 · PMC8016486 · Bioinformatics (Oxford, England) · 2021 · 6 claims · 3 setups
aCLImatise automatically generates workflow-language tool definitions by parsing a command-line tool's help output
-
Full-text index only
L2L: a simple tool for discovering the hidden significance in microarray expression data.
PMID 16168088 · PMC1242216 · Genome biology · 2005 · 8 claims · 4 setups
L2L systematically compares a user's differentially expressed gene list against a database of published differentially expressed gene lists to find statistically significant overlaps and generate hypotheses about shared mechanisms
-
Full-text index only
Phylogenomic approaches to common problems encountered in the analysis of low copy repeats: the sulfotransferase 1A gene family example.
PMID 15752422 · PMC555591 · BMC evolutionary biology · 2005 · 8 claims · 8 setups
A previously unidentified fourth human SULT1A gene (SULT1A4) exists on chromosome 16 and is transcriptionally active
-
Full-text index only
MILANO--custom annotation of microarray results using automatic literature searches.
PMID 15661078 · PMC547913 · BMC bioinformatics · 2005 · 7 claims · 4 setups
MILANO annotates microarray gene lists by counting literature co-occurrences of each gene with user-defined secondary terms
-
Full-text index only
SNP-RFLPing: restriction enzyme mining for SNPs in genomes.
PMID 16503968 · PMC1386656 · BMC genomics · 2006 · 8 claims · 2 setups
SNP-RFLPing accepts three flexible input types (dbSNP rs#/ss# IDs, HUGO gene name/Entrez gene ID, or free-form SNP-in-sequence including IUPAC or [dNTP1/dNTP2] formats) for human, rat, and mouse genomes
-
Has reproduction · 74
MicroPIPE: validating an end-to-end workflow for high-quality complete bacterial genome construction.
PMID 34172000 · PMC8235852 · BMC genomics · 2021 · 8 claims · 8 setups
MicroPIPE, an end-to-end Nextflow/Singularity-based pipeline built from systematically validated tool choices, produces high-quality complete bacterial genome assemblies without manual intervention.
-
Has reproduction · 58
A comparative study of techniques for differential expression analysis on RNA-Seq data.
PMID 25119138 · PMC4132098 · PloS one · 2014 · 8 claims · 8 setups
edgeR performs slightly better than DESeq and Cuffdiff2 in terms of the ability to uncover true positives.