Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 79
Enriched domain detector: a program for detection of wide genomic enrichment domains robust against local variations.
PMID 24782521 · PMC4066758 · Nucleic acids research · 2014 · 8 claims · 5 setups
EDD is a new algorithm that detects broad (megabase-size) enrichment domains from ChIP-seq data of widely distributed chromatin proteins such as A- and B-type lamins.
-
Has reproduction · 80
SLDMS: A Tool for Calculating the Overlapping Regions of Sequences.
PMID 35046988 · PMC8761809 · Frontiers in plant science · 2021 · 8 claims · 5 setups
SLDMS is a novel method for computing overlapping regions of sequencing reads using suffix array (SA), longest common prefix (LCP) array, document array (DA), and a monotonic stack.
-
Has reproduction · 86
LMAS: evaluating metagenomic short de novo assembly methods through defined communities.
PMID 36576131 · PMC9795473 · GigaScience · 2022 · 8 claims · 5 setups
LMAS (Last Metagenomic Assembler Standing) is a flexible, Nextflow-based, Docker-containerized automated workflow for benchmarking de novo metagenomic assemblers against defined mock communities, producing an interactive HTML report.
-
Has reproduction · 96
Calibration-free NGS quantitation of mutations below 0.01% VAF.
PMID 34675197 · PMC8531361 · Nature communications · 2021 · 8 claims · 6 setups
QBDA (Quantitative Blocker Displacement Amplification) integrates UMI molecular barcoding with BDA variant enrichment to enable calibration-free VAF quantitation
-
Has reproduction · 71
Hyb: a bioinformatics pipeline for the analysis of CLASH (crosslinking, ligation and sequencing of hybrids) data.
PMID 24211736 · PMC3969109 · Methods (San Diego, Calif.) · 2014 · 8 claims · 6 setups
The 'hyb' pipeline detects, calls, folds and annotates chimeric reads from CLASH high-throughput sequencing data.
-
Has reproduction · 42
KAGE: fast alignment-free graph-based genotyping of SNPs and short indels.
PMID 36195962 · PMC9531401 · Genome biology · 2022 · 7 claims · 7 setups
KAGE combines population-based kmer count modeling with single-variant prior adjustment into an alignment-free genotyper that matches the accuracy of the best existing alignment-free genotypers while being an order of magnitude faster.
-
Full-text index only
Primase-based whole genome amplification.
PMID 18559358 · PMC2490742 · Nucleic acids research · 2008 · 8 claims · 6 setups
A primase-based Whole Genome Amplification (pWGA) method was developed using T7 gp4 primase to synthesize primers on-template, removing the requirement for synthetic primers
-
Has reproduction · 76
nf-core/circrna: a portable workflow for the quantification, miRNA target prediction and differential expression analysis of circular RNAs.
PMID 36694127 · PMC9875403 · BMC bioinformatics · 2023 · 8 claims · 4 setups
Existing circRNA workflows are limited: none delineate circRNA-miRNA interactions and only one performs differential expression analysis, requiring users to supplement missing analysis types with in-house expertise
-
Has reproduction · 73
Vespucci: a system for building annotated databases of nascent transcripts.
PMID 24304890 · PMC3936758 · Nucleic acids research · 2014 · 8 claims · 7 setups
Existing ChIP-seq and RNA-seq analysis platforms (e.g. Cufflinks, peak callers) are unsuited to GRO-seq because they assume spliced/exonic reads, uniform density and paired-end data, and cannot identify transcriptional units de novo across the whole genome.
-
Has reproduction · 68
Bayesian transcriptome assembly.
PMID 25367074 · PMC4397945 · Genome biology · 2014 · 8 claims · 8 setups
Bayesembler, a probabilistic transcriptome assembler built on a Bayesian model of the RNA sequencing process with Gibbs sampling over expressed candidates, abundances and read assignments, is introduced.
-
Has reproduction · 85
Chromosome-level genome assembly of Lilford's wall lizard, Podarcis lilfordi (Günther, 1874) from the Balearic Islands (Spain).
PMID 37137526 · PMC10214862 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2023 · 8 claims · 8 setups
First high-quality chromosome-level genome assembly and annotation of P. lilfordi, generated via a mixed sequencing strategy (10X linked reads, ONT long reads, Hi-C) plus RNAseq/Iso-Seq
-
Has reproduction · 80
DMN-seq enriches DNA hypomethylated regions for biomarker discovery using 5-methylcytosine glycosylase.
PMID 41673887 · PMC13097799 · Genome biology · 2026 · 8 claims · 9 setups
DMN-seq (DMN+) uses DME to nick DNA specifically at 5mC sites, enabling 5mC detection at single-base resolution via selective adaptor ligation
-
Has reproduction · 69
High-resolution transcriptome and genome-wide dynamics of RNA polymerase and NusA in Mycobacterium tuberculosis.
PMID 23222129 · PMC3553938 · Nucleic acids research · 2013 · 8 claims · 7 setups
NusA interacts with RNAP ubiquitously throughout the M. tuberculosis chromosome and its ChIP-seq profile mirrors RNAP distribution in both exponential and stationary phase, despite NusA not binding DNA directly.
-
Full-text index only
Analysis of the prostate cancer cell line LNCaP transcriptome using a sequencing-by-synthesis approach.
PMID 17010196 · PMC1592491 · BMC genomics · 2006 · 8 claims · 7 setups
High-throughput 454 sequencing-by-synthesis of LNCaP cDNA can profile transcript abundance across the transcriptome
-
Has reproduction · 88
nf-core/isoseq: simple gene and isoform annotation with PacBio Iso-Seq long-read sequencing.
PMID 36961337 · PMC10199315 · Bioinformatics (Oxford, England) · 2023 · 7 claims · 4 setups
nf-core/isoseq is a new automated Nextflow-based pipeline that processes raw Iso-Seq subreads through to genome annotation (BED format) without requiring transcriptome assembly.
-
Has reproduction · 80
TP53 engagement with the genome occurs in distinct local chromatin environments via pioneer factor activity.
PMID 25391375 · PMC4315292 · Genome research · 2015 · 8 claims · 8 setups
TP53 binding events fall into three distinct categories defined by the local chromatin environment: TSS (H3K4me3+), enhancer (H3K4me1+/H3K4me3-), and distal (H3K4me1-/H3K4me3-) peaks.
-
Has reproduction · 59
WASP: a versatile, web-accessible single cell RNA-Seq processing platform.
PMID 33736596 · PMC7977290 · BMC genomics · 2021 · 7 claims · 7 setups
WASP is a software platform for processing Drop-Seq-based scRNA-seq data generated with ddSEQ or 10x protocols, combining a Snakemake pre-processing pipeline with an R Shiny post-processing application.
-
Has reproduction · 100
poreCov-An Easy to Use, Fast, and Robust Workflow for SARS-CoV-2 Genome Reconstruction via Nanopore Sequencing.
PMID 34394197 · PMC8355734 · Frontiers in genetics · 2021 · 8 claims · 8 setups
poreCov is an easy-to-use, fast, and robust Nextflow-based workflow for reference-based SARS-CoV-2 genome reconstruction and lineage determination from nanopore sequencing data
-
Has reproduction · 74
ChIP-seq guidelines and practices of the ENCODE and modENCODE consortia.
PMID 22955991 · PMC3431496 · Genome research · 2012 · 8 claims · 8 setups
ENCODE/modENCODE define a set of working standards and guidelines for ChIP-seq covering antibody validation, experimental replication, sequencing depth, data/metadata reporting, and data quality assessment.
-
Full-text index only
MrHAMER yields highly accurate single molecule viral sequences enabling analysis of intra-host evolution.
PMID 33849057 · PMC8266615 · Nucleic acids research · 2021 · 8 claims · 7 setups
MrHAMER yields >1000s of viral genomes per sample at 99.9% accuracy