Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 58
A comparative study of techniques for differential expression analysis on RNA-Seq data.
PMID 25119138 · PMC4132098 · PloS one · 2014 · 8 claims · 8 setups
edgeR performs slightly better than DESeq and Cuffdiff2 in terms of the ability to uncover true positives.
-
Full-text index only
BFAST: an alignment tool for large scale genome resequencing.
PMID 19907642 · PMC2770639 · PloS one · 2009 · 7 claims · 4 setups
BFAST is a new algorithm and freely available software tool for aligning large-scale short-read sequencing data to large reference genomes with user-customizable speed and accuracy
-
Has reproduction · 69
TC-hunter: identification of the insertion site of a transgenic gene within the host genome.
PMID 35184734 · PMC8859905 · BMC genomics · 2022 · 7 claims · 4 setups
TC-hunter is an open-source Nextflow pipeline that identifies transgene insertion sites using chimeric reads and discordant read pairs from NGS data.
-
Full-text index only
Effect of read-mapping biases on detecting allele-specific expression from RNA-sequencing data.
PMID 19808877 · PMC2788925 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 6 setups
Reads mapped to the reference genome show a significant bias toward the reference allele at heterozygous SNPs
-
Has reproduction · 99
Evaluation of taxonomic classification and profiling methods for long-read shotgun metagenomic sequencing datasets.
PMID 36513983 · PMC9749362 · BMC bioinformatics · 2022 · 8 claims · 7 setups
Long-read classifiers generally performed best among the 11 methods tested
-
Has reproduction · 100
Intra-Host Co-Existing Strains of SARS-CoV-2 Reference Genome Uncovered by Exhaustive Computational Search.
PMID 37243151 · PMC10224212 · Viruses · 2023 · 8 claims · 7 setups
An exhaustive-search workflow can recover intra-host co-existing SARS-CoV-2 strains from the reference-genome read set (SRR11092062) that de Bruijn-graph assemblers discard.
-
Full-text index only
BOAT: Basic Oligonucleotide Alignment Tool.
PMID 19958483 · PMC2788372 · BMC genomics · 2009 · 7 claims · 3 setups
BOAT can accurately and efficiently map sequencing reads to a reference genome while handling several substitutions and indels simultaneously
-
Full-text index only
Rise of the machines.
PMID 18670625 · PMC2467494 · PLoS genetics · 2008 · 8 claims · 4 setups
New short-read sequencing platforms (Illumina Genome Analyzer, 454 FLX, ABI SOLiD) enable rapid, scalable whole-genome resequencing that was previously restricted to dedicated sequencing centers using Sanger methods.
-
Has reproduction · 50
MEDUSA: A Pipeline for Sensitive Taxonomic Classification and Flexible Functional Annotation of Metagenomic Shotgun Sequences.
PMID 35330728 · PMC8940201 · Frontiers in genetics · 2022 · 6 claims · 6 setups
MEDUSA is an automated, Conda-installable and Snakemake-managed pipeline performing preprocessing, assembly, alignment, taxonomic classification, and functional annotation on shotgun data.
-
Has reproduction · 95
transXpress: a Snakemake pipeline for streamlined de novo transcriptome assembly and annotation.
PMID 37016291 · PMC10074830 · BMC bioinformatics · 2023 · 6 claims · 7 setups
transXpress is a Snakemake pipeline that streamlines de novo transcriptome assembly, quantification, and annotation for non-model organisms
-
Has reproduction · 90
An improved assembly of the pearl millet reference genome using Oxford Nanopore long reads and optical mapping.
PMID 36891809 · PMC10151396 · G3 (Bethesda, Md.) · 2023 · 8 claims · 8 setups
Combining ONT long reads with Bionano optical maps produced a substantially more complete and contiguous pearl millet Tift 23D2B1-P1-P5 assembly than the prior short-read assembly.
-
Full-text index only
A comparison of random sequence reads versus 16S rDNA sequences for estimating the biodiversity of a metagenomic library.
PMID 18682527 · PMC2532719 · Nucleic acids research · 2008 · 8 claims · 7 setups
Biodiversity observed by RSR analysis is consistent with that obtained by 16S rDNA analysis
-
Has reproduction · 80
Chromosome-Scale Assembly of the Complete Genome Sequence of Leishmania (Mundinia) enriettii, Isolate CUR178, Strain LV763.
PMID 34498918 · PMC8428246 · Microbiology resource announcements · 2021 · 7 claims · 7 setups
Complete chromosome-scale genome sequence of Leishmania (Mundinia) enriettii isolate CUR178, strain LV763 was assembled using combined short-read and long-read sequencing
-
Has reproduction · 99
A haplotype-resolved genome assembly of the bocaccio rockfish, Sebastes paucispinis.
PMID 40323688 · PMC12584591 · The Journal of heredity · 2025 · 6 claims · 8 setups
This paper presents the first de novo, haplotype-resolved reference-quality genome assembly of Sebastes paucispinis (bocaccio rockfish).
-
Has reproduction · 80
SLDMS: A Tool for Calculating the Overlapping Regions of Sequences.
PMID 35046988 · PMC8761809 · Frontiers in plant science · 2021 · 8 claims · 5 setups
SLDMS is a novel method for computing overlapping regions of sequencing reads using suffix array (SA), longest common prefix (LCP) array, document array (DA), and a monotonic stack.
-
Has reproduction · 75
Sequencing of human genomes with nanopore technology.
PMID 31015479 · PMC6478738 · Nature communications · 2019 · 8 claims · 7 setups
A novel single-sample, reference panel-free, read-based phasing algorithm built on the STITCH model improves nanopore SNV calling from modest baseline levels.
-
Has reproduction · 91
Whole genome and transcriptome maps of the entirely black native Korean chicken breed Yeonsan Ogye.
PMID 30010758 · PMC6065499 · GigaScience · 2018 · 6 claims · 7 setups
A hybrid de novo assembly combining high-depth Illumina short reads (376.6X) and low-depth PacBio long reads (9.7X) produced the YO draft genome Ogye_1.1 with contig and scaffold NG50 of 362.3 Kbp and 16.8 Mbp.
-
Has reproduction · 71
Hyb: a bioinformatics pipeline for the analysis of CLASH (crosslinking, ligation and sequencing of hybrids) data.
PMID 24211736 · PMC3969109 · Methods (San Diego, Calif.) · 2014 · 8 claims · 6 setups
The 'hyb' pipeline detects, calls, folds and annotates chimeric reads from CLASH high-throughput sequencing data.
-
Has reproduction · 85
An extensive evaluation of read trimming effects on Illumina NGS data analysis.
PMID 24376861 · PMC3871669 · PloS one · 2013 · 8 claims · 8 setups
Read trimming increases the quality and reliability of downstream NGS analyses (RNA-Seq mapping, SNP identification, genome assembly) while reducing execution time and computational resources.
-
Full-text index only
An analysis of the feasibility of short read sequencing.
PMID 16275781 · PMC1278949 · Nucleic acids research · 2005 · 8 claims · 8 setups
Re-sequencing and de novo sequencing of the majority of a bacterial genome is possible with read lengths of 20-30 nt.