Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 49
EDGE COVID-19: a web platform to generate submission-ready genomes from SARS-CoV-2 sequencing efforts.
PMID 35561186 · PMC9113274 · Bioinformatics (Oxford, England) · 2022 · 7 claims · 5 setups
EDGE COVID-19 (EC-19) is a web-based platform that automates QC, reference-based variant/consensus calling, lineage determination, and submission of SARS-CoV-2 genomes and metadata to GenBank, GISAID and INSDC for both Illumina and ONT data.
-
Has reproduction · 73
Genetic polyploid phasing from low-depth progeny samples.
PMID 35692633 · PMC9184567 · iScience · 2022 · 8 claims · 7 setups
WH-PPG phases polyploid parental samples by scoring informative variant pairs with a Bayesian log-likelihood model of progeny allele depths, clustering alleles by co-occurrence likelihood, and assigning clusters to haplotypes via interval scheduling
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Full-text index only
Optimizing data-driven excellence: Canada's approach to using pathogen test datasets for quality control, pipeline development and training initiatives.
PMID 41591806 · PMC12847982 · Microbial genomics · 2026 · 8 claims · 5 setups
Standardized SARS-CoV-2 test datasets (Illumina and Nanopore) were developed as benchmarks for validating sequencing/bioinformatics pipelines across Canadian public health labs
-
Has reproduction · 67
GEMmaker: process massive RNA-seq datasets on heterogeneous computational infrastructure.
PMID 35501696 · PMC9063052 · BMC bioinformatics · 2022 · 6 claims · 3 setups
GEMmaker, an nf-core compliant Nextflow workflow, can quantify gene expression from small to massive RNA-seq datasets while remaining reproducible via versioned containerized software.
-
Full-text index only
Rapid identification of microbial pathogens and antimicrobial resistance from bloodstream infections using long-read sequencing.
PMID 42274466 · PMC13256323 · Microbial genomics · 2026 · 8 claims · 8 setups
A novel ONT long-read sequencing laboratory and bioinformatic workflow rapidly identifies bacterial and fungal organisms and AMR determinants from positive blood cultures
-
Has reproduction · 86
LMAS: evaluating metagenomic short de novo assembly methods through defined communities.
PMID 36576131 · PMC9795473 · GigaScience · 2022 · 8 claims · 5 setups
LMAS (Last Metagenomic Assembler Standing) is a flexible, Nextflow-based, Docker-containerized automated workflow for benchmarking de novo metagenomic assemblers against defined mock communities, producing an interactive HTML report.