Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 89
HTSQualC is a flexible and one-step quality control software for high-throughput sequencing data analysis.
PMID 34548573 · PMC8455540 · Scientific reports · 2021 · 8 claims · 5 setups
HTSQualC is a standalone, one-step QC software that performs filtering and trimming of raw HTS data in a single run
-
Full-text index only
Identifying repeat domains in large genomes.
PMID 16507140 · PMC1431705 · Genome biology · 2006 · 7 claims · 5 setups
A repeat domain graph, built using a modified A-Bruijn graph framework, decomposes a repeat library into shared repeat domains and reveals the mosaic structure of repeat families.
-
Full-text index only
Ensembl 2006.
PMID 16381931 · PMC1347495 · Nucleic acids research · 2006 · 8 claims · 5 setups
Ensembl now provides annotation for 19 genomes, up from 4 the previous year, including new mammalian (Rhesus macaque, Opossum), chordate (Ciona intestinalis), and yeast genomes.
-
Has reproduction · 76
nf-core/circrna: a portable workflow for the quantification, miRNA target prediction and differential expression analysis of circular RNAs.
PMID 36694127 · PMC9875403 · BMC bioinformatics · 2023 · 8 claims · 4 setups
Existing circRNA workflows are limited: none delineate circRNA-miRNA interactions and only one performs differential expression analysis, requiring users to supplement missing analysis types with in-house expertise
-
Has reproduction · 83
Hobbes: optimized gram-based methods for efficient read alignment.
PMID 22199254 · PMC3315303 · Nucleic acids research · 2012 · 8 claims · 4 setups
Hobbes, a gram-based short-read mapper supporting Hamming and edit distance, is faster than all other read-mapping programs tested while maintaining high mapping quality.
-
Full-text index only
DAVID Knowledgebase: a gene-centered database integrating heterogeneous gene annotation resources to facilitate high-throughput gene functional analysis.
PMID 17980028 · PMC2186358 · BMC bioinformatics · 2007 · 7 claims · 3 setups
The DAVID Gene Concept, a single-linkage algorithm, merges gene clusters from Entrez Gene, UniRef100, and PIR-NREF100 that share protein IDs and species into unified DAVID gene clusters, improving cross-referencing between NCBI and UniProt systems
-
Full-text index only
Grammar-based distance in progressive multiple sequence alignment.
PMID 18616828 · PMC2478692 · BMC bioinformatics · 2008 · 7 claims · 3 setups
A grammar-based (LZ complexity) distance metric can be used to determine the order in which sequences are progressively pairwise aligned
-
Has reproduction · 50
MEDUSA: A Pipeline for Sensitive Taxonomic Classification and Flexible Functional Annotation of Metagenomic Shotgun Sequences.
PMID 35330728 · PMC8940201 · Frontiers in genetics · 2022 · 6 claims · 6 setups
MEDUSA is an automated, Conda-installable and Snakemake-managed pipeline performing preprocessing, assembly, alignment, taxonomic classification, and functional annotation on shotgun data.
-
Has reproduction · 67
binny: an automated binning algorithm to recover high-quality genomes from complex metagenomic datasets.
PMID 36239393 · PMC9677464 · Briefings in bioinformatics · 2022 · 8 claims · 8 setups
binny outperforms or is highly competitive with commonly used and state-of-the-art binning methods (MetaBAT2, MaxBin2, CONCOCT, VAMB, SemiBin, MetaDecoder)
-
Full-text index only
NCBI Reference Sequence (RefSeq): a curated non-redundant sequence database of genomes, transcripts and proteins.
PMID 15608248 · PMC539979 · Nucleic acids research · 2005 · 7 claims · 5 setups
RefSeq provides a curated, non-redundant, explicitly linked collection of genomic, transcript and protein sequences spanning prokaryotes, eukaryotes and viruses.
-
Has reproduction · 74
MicroPIPE: validating an end-to-end workflow for high-quality complete bacterial genome construction.
PMID 34172000 · PMC8235852 · BMC genomics · 2021 · 8 claims · 8 setups
MicroPIPE, an end-to-end Nextflow/Singularity-based pipeline built from systematically validated tool choices, produces high-quality complete bacterial genome assemblies without manual intervention.
-
Has reproduction · 67
SnakeMAGs: a simple, efficient, flexible and scalable workflow to reconstruct prokaryotic genomes from metagenomes.
PMID 36875992 · PMC9978240 · F1000Research · 2022 · 7 claims · 8 setups
SnakeMAGs is a simple, efficient, flexible and scalable Snakemake workflow that processes Illumina reads from raw data to MAG classification and relative abundance estimation
-
Full-text index only
SeqDoC: rapid SNP and mutation detection by direct comparison of DNA sequence chromatograms.
PMID 15927052 · PMC1156871 · BMC bioinformatics · 2005 · 8 claims · 6 setups
SeqDoC generates a subtracted difference trace between a reference and test chromatogram that highlights single base changes
-
Full-text index only
COMUS: Clinician-Oriented locus-specific MUtation detection and deposition System.
PMID 19958500 · PMC2788389 · BMC genomics · 2009 · 8 claims · 6 setups
COMUS is a bioinformatics system for detecting and depositing new mutations from patient DNA with a clinician-friendly interface
-
Full-text index only
In silico meets in vivo.
PMID 18304380 · PMC2374716 · Genome biology · 2008 · 8 claims · 8 setups
About 10% of positions in multiple sequence alignments of the human genome with other vertebrate genomes are likely incorrect.
-
Has reproduction · 58
HGA: de novo genome assembly method for bacterial genomes using high coverage short sequencing reads.
PMID 26945881 · PMC4779561 · BMC genomics · 2016 · 8 claims · 7 setups
HGA leads to significant improvement in assembly quality (N50 and corrected N50) for all 7 evaluated GAGE-B bacterial datasets using most of the 8 evaluated assemblers
-
Full-text index only
DiagHunter and GenoPix2D: programs for genomic comparisons, large-scale homology discovery and visualization.
PMID 14519203 · PMC328457 · Genome biology · 2003 · 7 claims · 5 setups
DiagHunter identifies large-scale synteny blocks within or between genomes efficiently despite background noise and genomic discontinuities, without performing sequence alignment
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.
-
Full-text index only
Analysis of the prostate cancer cell line LNCaP transcriptome using a sequencing-by-synthesis approach.
PMID 17010196 · PMC1592491 · BMC genomics · 2006 · 8 claims · 7 setups
High-throughput 454 sequencing-by-synthesis of LNCaP cDNA can profile transcript abundance across the transcriptome
-
Has reproduction · 86
Plasmid transmission dynamics and evolution of partner quality in a natural population of Rhizobium leguminosarum.
PMID 41212030 · PMC12691615 · mBio · 2025 · 8 claims · 8 setups
Of the four most frequent plasmid types, types II and III have more stable size, larger core genomes, and track the chromosomal phylogeny (more vertical transmission), while types I and IV (pSym) vary in size and gene content with phylogenies consistent with frequent horizontal transmission.