Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 71
RNAmountAlign: Efficient software for local, global, semiglobal pairwise and multiple RNA sequence/structure alignment.
PMID 31978147 · PMC6980424 · PloS one · 2020 · 8 claims · 6 setups
RNAmountAlign is the first RNA sequence/structure pairwise alignment algorithm based on incremental ensemble mountain distance, running in O(n^3) time and O(n^2) space for two sequences of length n.
-
Full-text index only
RotaC: a web-based tool for the complete genome classification of group A rotaviruses.
PMID 19930627 · PMC2785824 · BMC microbiology · 2009 · 7 claims · 4 setups
RotaC is a freely available web-based tool for complete genome classification of group A rotaviruses across all 11 gene segments.
-
Has reproduction · 75
geneshot: gene-level metagenomics identifies genome islands associated with immunotherapy response.
PMID 33952321 · PMC8097837 · Genome biology · 2021 · 8 claims · 4 setups
geneshot is a gene-level metagenomic bioinformatics tool that clusters de novo assembled protein-coding genes into co-abundant gene groups (CAGs) to reduce dimensionality and generate testable hypotheses from WGS microbiome data
-
Has reproduction · 38
RNA-Seq transcriptome profiling of upland cotton (Gossypium hirsutum L.) root tissue under water-deficit stress.
PMID 24324815 · PMC3855774 · PloS one · 2013 · 8 claims · 8 setups
A total of 1,530 transcripts were differentially expressed between well-watered and water-deficit stressed field-grown upland cotton root tissues (913 up-regulated, 617 down-regulated).
-
Has reproduction · 100
A Bioinformatics Workflow to Identify eccDNA Using ECCFP From Long-Read Nanopore Sequencing Data.
PMID 41924242 · PMC13037781 · Bio-protocol · 2026 · 7 claims · 5 setups
ECCFP significantly improves eccDNA detection sensitivity, accuracy, and runtime efficiency compared to other pipelines
-
Has reproduction · 80
Progressive transformation of the HIV-1 reservoir cell profile over two decades of antiviral therapy.
PMID 36596305 · PMC9839361 · Cell host & microbe · 2023 · 8 claims · 7 setups
After long-term ART, intact HIV-1 proviruses are predominantly integrated in heterochromatin locations, most prominently centromeric satellite/micro-satellite DNA.
-
Full-text index only
Versatile and open software for comparing large genomes.
PMID 14759262 · PMC395750 · Genome biology · 2004 · 8 claims · 8 setups
MUMmer 3.0 efficiently handles comparisons of large eukaryotic genomes at varying evolutionary distances
-
Has reproduction · 86
LMAS: evaluating metagenomic short de novo assembly methods through defined communities.
PMID 36576131 · PMC9795473 · GigaScience · 2022 · 8 claims · 5 setups
LMAS (Last Metagenomic Assembler Standing) is a flexible, Nextflow-based, Docker-containerized automated workflow for benchmarking de novo metagenomic assemblers against defined mock communities, producing an interactive HTML report.
-
Has reproduction · 50
Polymorphism identification and improved genome annotation of Brassica rapa through Deep RNA sequencing.
PMID 25122667 · PMC4232532 · G3 (Bethesda, Md.) · 2014 · 8 claims · 8 setups
330,995 SNPs were identified in transcribed regions between B. rapa genotypes R500 and IMB211, at an average frequency of one SNP per 200 bases.
-
Full-text index only
NCBI Reference Sequences: current status, policy and new initiatives.
PMID 18927115 · PMC2686572 · Nucleic acids research · 2009 · 7 claims · 5 setups
RefSeq is a curated, non-redundant, explicitly linked database of nucleotide and protein sequences spanning genomes, transcripts and proteins across prokaryotes, eukaryotes and viruses
-
Full-text index only
BFAST: an alignment tool for large scale genome resequencing.
PMID 19907642 · PMC2770639 · PloS one · 2009 · 7 claims · 4 setups
BFAST is a new algorithm and freely available software tool for aligning large-scale short-read sequencing data to large reference genomes with user-customizable speed and accuracy
-
Full-text index only
Phylogenomic approaches to common problems encountered in the analysis of low copy repeats: the sulfotransferase 1A gene family example.
PMID 15752422 · PMC555591 · BMC evolutionary biology · 2005 · 8 claims · 8 setups
A previously unidentified fourth human SULT1A gene (SULT1A4) exists on chromosome 16 and is transcriptionally active
-
Full-text index only
Designating eukaryotic orthology via processed transcription units.
PMID 18445630 · PMC2425467 · Nucleic acids research · 2008 · 8 claims · 5 setups
Existing ortholog databases discard/ignore alternative splicing via all-against-all protein comparisons, causing ambiguous ortholog calls and misclassification of AS isoforms as in-paralogs
-
Has reproduction · 86
Assessing Bos taurus introgression in the UOA Bos indicus assembly.
PMID 34922445 · PMC8684283 · Genetics, selection, evolution : GSE · 2021 · 7 claims · 6 setups
Aligning divergent (cross-subspecies) sequence data detects substantially more SNVs than aligning to a same-subspecies reference, indicating reference/assembly bias in variant calling.
-
Has reproduction · 65
SPEAQeasy: a scalable pipeline for expression analysis and quantification for R/bioconductor-powered RNA-seq analyses.
PMID 33932985 · PMC8088074 · BMC bioinformatics · 2021 · 8 claims · 5 setups
SPEAQeasy is a portable, easy-to-install, Nextflow-powered RNA-seq processing pipeline that lowers the computational entry barrier for biologists/clinicians
-
Has reproduction · 82
Whole-genome analysis of a multidrug-resistant Klebsiella michiganensis environmental isolate from an orthopedic ward in Mwanza, Tanzania reveals IncF-family plasmid replicon signatures associated with resistance determinants.
PMID 41957580 · PMC13173886 · BMC genomics · 2026 · 6 claims · 8 setups
Genome-based taxonomy (GTDB-Tk and ANI) reclassified the isolate A55848, phenotypically identified as K. oxytoca, as Klebsiella michiganensis
-
Has reproduction · 85
An extensive evaluation of read trimming effects on Illumina NGS data analysis.
PMID 24376861 · PMC3871669 · PloS one · 2013 · 8 claims · 8 setups
Read trimming increases the quality and reliability of downstream NGS analyses (RNA-Seq mapping, SNP identification, genome assembly) while reducing execution time and computational resources.
-
Has reproduction · 45
Identifying and classifying trait linked polymorphisms in non-reference species by walking coloured de bruijn graphs.
PMID 23536903 · PMC3607606 · PloS one · 2013 · 8 claims · 9 setups
Bubbleparse detects sequence variants directly from NGS reads without a reference genome, using the coloured de Bruijn graph implementation of Cortex plus a new depth-first bubble-finding module.
-
Full-text index only
Searching for SNPs with cloud computing.
PMID 19930550 · PMC3091327 · Genome biology · 2009 · 8 claims · 4 setups
Crossbow combines the Bowtie short-read aligner and SOAPsnp SNP caller into a seamless, automatic Hadoop/MapReduce pipeline for whole-genome resequencing analysis
-
Full-text index only
Slider--maximum use of probability information for alignment of short sequence reads and SNP detection.
PMID 18974170 · PMC2638935 · Bioinformatics (Oxford, England) · 2009 · 7 claims · 3 setups
Slider aligns reads using all bases above a probability threshold (baseMinPrb) from prb files, generating all possible read sequences above a read probability threshold (read_0_MinPrb), rather than only the most probable sequence