Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 89
Evaluating sequence data quality from the Swift Accel-Amplicon CFTR Panel.
PMID 31913291 · PMC6949293 · Scientific data · 2020 · 6 claims · 7 setups
The Accel-Amplicon CFTR panel generates sequencing data with high coverage depth and near 100% on-target reads.
-
Has reproduction · 85
An extensive evaluation of read trimming effects on Illumina NGS data analysis.
PMID 24376861 · PMC3871669 · PloS one · 2013 · 8 claims · 8 setups
Read trimming increases the quality and reliability of downstream NGS analyses (RNA-Seq mapping, SNP identification, genome assembly) while reducing execution time and computational resources.
-
Full-text index only
In silico meets in vivo.
PMID 18304380 · PMC2374716 · Genome biology · 2008 · 8 claims · 8 setups
About 10% of positions in multiple sequence alignments of the human genome with other vertebrate genomes are likely incorrect.
-
Has reproduction · 66
RiboTaxa: combined approaches for rRNA genes taxonomic resolution down to the species level from metagenomics data revealing novelties.
PMID 36159175 · PMC9492272 · NAR genomics and bioinformatics · 2022 · 8 claims · 6 setups
RiboTaxa, combining BBTools, FastQC, SortMeRNA, MetaRib, EMIRGE, VSEARCH, BBMap and QIIME 2's Sklearn classifier, was built as a pipeline for SSU rRNA-based taxonomic profiling of metagenomics data.
-
Has reproduction · 95
transXpress: a Snakemake pipeline for streamlined de novo transcriptome assembly and annotation.
PMID 37016291 · PMC10074830 · BMC bioinformatics · 2023 · 6 claims · 7 setups
transXpress is a Snakemake pipeline that streamlines de novo transcriptome assembly, quantification, and annotation for non-model organisms
-
Full-text index only
Integrating alternative splicing detection into gene prediction.
PMID 15705189 · PMC550657 · BMC bioinformatics · 2005 · 8 claims · 4 setups
An integrative intrinsic/extrinsic method was implemented in the gene finder EuGÈNE (as EuGÈNE-M) to detect AS evidence from aligned transcripts and generate alternative optimal gene predictions consistent with each detected AS event.
-
Full-text index only
NCBI Reference Sequences: current status, policy and new initiatives.
PMID 18927115 · PMC2686572 · Nucleic acids research · 2009 · 7 claims · 5 setups
RefSeq is a curated, non-redundant, explicitly linked database of nucleotide and protein sequences spanning genomes, transcripts and proteins across prokaryotes, eukaryotes and viruses
-
Has reproduction · 86
Plasmid transmission dynamics and evolution of partner quality in a natural population of Rhizobium leguminosarum.
PMID 41212030 · PMC12691615 · mBio · 2025 · 8 claims · 8 setups
Of the four most frequent plasmid types, types II and III have more stable size, larger core genomes, and track the chromosomal phylogeny (more vertical transmission), while types I and IV (pSym) vary in size and gene content with phylogenies consistent with frequent horizontal transmission.
-
Has reproduction · 91
Chromosome-level genome assembly of agar-producing red seaweed Gracilaria vermiculophylla.
PMID 41629338 · PMC12966425 · Scientific data · 2026 · 8 claims · 8 setups
A chromosome-level genome assembly of G. vermiculophylla was generated by combining DNBSeq short reads, Nanopore long reads, and Hi-C data.
-
Has reproduction · 100
A Bioinformatics Workflow to Identify eccDNA Using ECCFP From Long-Read Nanopore Sequencing Data.
PMID 41924242 · PMC13037781 · Bio-protocol · 2026 · 7 claims · 5 setups
ECCFP significantly improves eccDNA detection sensitivity, accuracy, and runtime efficiency compared to other pipelines
-
Full-text index only
A third approach to gene prediction suggests thousands of additional human transcribed regions.
PMID 16543943 · PMC1391917 · PLoS computational biology · 2006 · 8 claims · 7 setups
A third basic concept for gene prediction exists, based on detecting strand-specific 'transcription footprints' (mutational and selectional biases) rather than gene structure or sequence similarity.
-
Has reproduction · 87
A target enrichment method for gathering phylogenetic information from hundreds of loci: An example from the Compositae.
PMID 25202605 · PMC4103609 · Applications in plant sciences · 2014 · 8 claims · 8 setups
A custom sequence capture probe set (9678 baits targeting 1061 orthologous genes) was designed to enrich COS loci across the Compositae.
-
Full-text index only
Single-molecule sequencing of an individual human genome.
PMID 19668243 · PMC4117198 · Nature biotechnology · 2009 · 8 claims · 7 setups
Single-molecule sequencing without cloning, amplification or ligation can sequence an individual human genome on one instrument by a single operator in four runs
-
Full-text index only
Analyses of deep mammalian sequence alignments and constraint predictions for 1% of the human genome.
PMID 17567995 · PMC1891336 · Genome research · 2007 · 7 claims · 3 setups
Four different alignment methods show large-scale consistency but substantial differences in small-scale rearrangements, sensitivity, and specificity.
-
Has reproduction · 43
TransFlow: a Snakemake workflow for transmission analysis of Mycobacterium tuberculosis whole-genome sequencing data.
PMID 36469333 · PMC9825751 · Bioinformatics (Oxford, England) · 2023 · 8 claims · 8 setups
TransFlow is a Snakemake- and Conda-based workflow that combines state-of-the-art tools into a single, fast, scalable pipeline for MTBC WGS-based transmission analysis.
-
Has reproduction · 84
Genome of the endangered Guatemalan Beaded Lizard, Heloderma charlesbogerti, reveals evolutionary relationships of squamates and declines in effective population sizes.
PMID 36226801 · PMC9713440 · G3 (Bethesda, Md.) · 2022 · 8 claims · 7 setups
The assembled draft genome of H. charlesbogerti totals 2.31 Gb, similar in size to related species
-
Full-text index only
Processing and population genetic analysis of multigenic datasets with ProSeq3 software.
PMID 19797407 · PMC2778335 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 7 setups
ProSeq3 is a program with a graphic user interface that simplifies preparation and basic population genetic analysis of multigenic DNA polymorphism datasets
-
Full-text index only
BIPASS: BioInformatics Pipeline Alternative Splicing Services.
PMID 17584795 · PMC1933140 · Nucleic acids research · 2007 · 8 claims · 4 setups
BIPASS offers two complementary services for alternative splicing (AS) research: BIPAS-SpliceDB, a queryable pre-computed AS data warehouse, and BIPAS-Align&Splice, an online pipeline for user-submitted sequences.
-
Has reproduction · 45
Accurate sequence variant genotyping in cattle using variation-aware genome graphs.
PMID 31092189 · PMC6521551 · Genetics, selection, evolution : GSE · 2019 · 8 claims · 7 setups
Graphtyper outperformed GATK and SAMtools in genotype concordance, non-reference sensitivity, and non-reference discrepancy compared to microarray genotypes
-
Has reproduction · 57
Regulatory Noncoding Small RNAs Are Diverse and Abundant in an Extremophilic Microbial Community.
PMID 32019831 · PMC7002113 · mSystems · 2020 · 8 claims · 7 setups
Hundreds of intergenic (itsRNAs) and antisense (asRNAs) sRNAs are diverse and abundant in the halite endolithic microbial community, with 1,538 total ncRNAs discovered across Archaea and Bacteria.