Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 69
High-resolution transcriptome and genome-wide dynamics of RNA polymerase and NusA in Mycobacterium tuberculosis.
PMID 23222129 · PMC3553938 · Nucleic acids research · 2013 · 8 claims · 7 setups
NusA interacts with RNAP ubiquitously throughout the M. tuberculosis chromosome and its ChIP-seq profile mirrors RNAP distribution in both exponential and stationary phase, despite NusA not binding DNA directly.
-
Has reproduction · 80
Bisulfite sequencing of chromatin immunoprecipitated DNA (BisChIP-seq) directly informs methylation status of histone-modified DNA.
PMID 22466171 · PMC3371705 · Genome research · 2012 · 8 claims · 8 setups
BisChIP-seq — bisulfite sequencing of chromatin immunoprecipitated DNA — enables direct genome-wide, base-resolution interrogation of DNA methylation on histone-modified DNA molecules
-
Has reproduction · 87
Ultra-deep sequencing data from a liquid biopsy proficiency study demonstrating analytic validity.
PMID 35418127 · PMC9008010 · Scientific data · 2022 · 6 claims · 5 setups
This dataset is the most comprehensive public-facing dataset of ultra-deep ctDNA sequencing data generated to date
-
Has reproduction · 80
SLDMS: A Tool for Calculating the Overlapping Regions of Sequences.
PMID 35046988 · PMC8761809 · Frontiers in plant science · 2021 · 8 claims · 5 setups
SLDMS is a novel method for computing overlapping regions of sequencing reads using suffix array (SA), longest common prefix (LCP) array, document array (DA), and a monotonic stack.
-
Has reproduction · 85
Chromosome-level genome assembly of Lilford's wall lizard, Podarcis lilfordi (Günther, 1874) from the Balearic Islands (Spain).
PMID 37137526 · PMC10214862 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2023 · 8 claims · 8 setups
First high-quality chromosome-level genome assembly and annotation of P. lilfordi, generated via a mixed sequencing strategy (10X linked reads, ONT long reads, Hi-C) plus RNAseq/Iso-Seq
-
Has reproduction · 96
Calibration-free NGS quantitation of mutations below 0.01% VAF.
PMID 34675197 · PMC8531361 · Nature communications · 2021 · 8 claims · 6 setups
QBDA (Quantitative Blocker Displacement Amplification) integrates UMI molecular barcoding with BDA variant enrichment to enable calibration-free VAF quantitation
-
Has reproduction · 73
Vespucci: a system for building annotated databases of nascent transcripts.
PMID 24304890 · PMC3936758 · Nucleic acids research · 2014 · 8 claims · 7 setups
Existing ChIP-seq and RNA-seq analysis platforms (e.g. Cufflinks, peak callers) are unsuited to GRO-seq because they assume spliced/exonic reads, uniform density and paired-end data, and cannot identify transcriptional units de novo across the whole genome.
-
Full-text index only
Analysis of the prostate cancer cell line LNCaP transcriptome using a sequencing-by-synthesis approach.
PMID 17010196 · PMC1592491 · BMC genomics · 2006 · 8 claims · 7 setups
High-throughput 454 sequencing-by-synthesis of LNCaP cDNA can profile transcript abundance across the transcriptome
-
Full-text index only
Sequencing the regulatory genome.
PMID 18598374 · PMC2481419 · Genome biology · 2008 · 8 claims · 8 setups
Nuclear-lamina-associated domains (LADs) define chromatin regions with distinct transcriptional characteristics (fewer, lower-expressed genes, low RNA Pol II occupancy, H3K27me3-enriched borders)
-
Full-text index only
High resolution melting analysis for rapid and sensitive EGFR and KRAS mutation detection in formalin fixed paraffin embedded biopsies.
PMID 18495026 · PMC2408599 · BMC cancer · 2008 · 8 claims · 4 setups
HRM correctly identified all 73 EGFR-mutation-positive FFPE samples previously found by sequencing, giving 100% sensitivity and 90% specificity
-
Has reproduction · 76
nf-core/circrna: a portable workflow for the quantification, miRNA target prediction and differential expression analysis of circular RNAs.
PMID 36694127 · PMC9875403 · BMC bioinformatics · 2023 · 8 claims · 4 setups
Existing circRNA workflows are limited: none delineate circRNA-miRNA interactions and only one performs differential expression analysis, requiring users to supplement missing analysis types with in-house expertise
-
Full-text index only
Decoding of superimposed traces produced by direct sequencing of heterozygous indels.
PMID 18654614 · PMC2429969 · PLoS computational biology · 2008 · 7 claims · 3 setups
A dynamic programming method (implemented as web app Indelligent) can decode superimposed allelic sequences from a single mixed trace, using only the observed string of ambiguous peak calls, without a reference sequence or reverse trace.
-
Has reproduction · 88
nf-core/isoseq: simple gene and isoform annotation with PacBio Iso-Seq long-read sequencing.
PMID 36961337 · PMC10199315 · Bioinformatics (Oxford, England) · 2023 · 7 claims · 4 setups
nf-core/isoseq is a new automated Nextflow-based pipeline that processes raw Iso-Seq subreads through to genome annotation (BED format) without requiring transcriptome assembly.
-
Has reproduction · 85
High performance imputation of structural and single nucleotide variants using low-coverage whole genome sequencing.
PMID 40155798 · PMC11951665 · Genetics, selection, evolution : GSE · 2025 · 7 claims · 6 setups
SNVs are imputed with high accuracy and recall across all tested WGS depths (1-4x), including in samples external to the reference panel.
-
Has reproduction · 100
poreCov-An Easy to Use, Fast, and Robust Workflow for SARS-CoV-2 Genome Reconstruction via Nanopore Sequencing.
PMID 34394197 · PMC8355734 · Frontiers in genetics · 2021 · 8 claims · 8 setups
poreCov is an easy-to-use, fast, and robust Nextflow-based workflow for reference-based SARS-CoV-2 genome reconstruction and lineage determination from nanopore sequencing data
-
Has reproduction · 43
Compression of structured high-throughput sequencing data.
PMID 24260313 · PMC3832420 · PloS one · 2013 · 8 claims · 7 setups
Leveraging an explicit data schema (separate field encoding, field modeling, template compression, domain modeling) enables stronger compression of HTS alignment data than general-purpose compression of serialized bytes.
-
Full-text index only
Primase-based whole genome amplification.
PMID 18559358 · PMC2490742 · Nucleic acids research · 2008 · 8 claims · 6 setups
A primase-based Whole Genome Amplification (pWGA) method was developed using T7 gp4 primase to synthesize primers on-template, removing the requirement for synthetic primers
-
Has reproduction · 71
Hyb: a bioinformatics pipeline for the analysis of CLASH (crosslinking, ligation and sequencing of hybrids) data.
PMID 24211736 · PMC3969109 · Methods (San Diego, Calif.) · 2014 · 8 claims · 6 setups
The 'hyb' pipeline detects, calls, folds and annotates chimeric reads from CLASH high-throughput sequencing data.
-
Has reproduction · 79
Enriched domain detector: a program for detection of wide genomic enrichment domains robust against local variations.
PMID 24782521 · PMC4066758 · Nucleic acids research · 2014 · 8 claims · 5 setups
EDD is a new algorithm that detects broad (megabase-size) enrichment domains from ChIP-seq data of widely distributed chromatin proteins such as A- and B-type lamins.
-
Has reproduction · 80
DMN-seq enriches DNA hypomethylated regions for biomarker discovery using 5-methylcytosine glycosylase.
PMID 41673887 · PMC13097799 · Genome biology · 2026 · 7 claims · 8 setups
DME-mediated nicking enables DMN-seq (DMN+) to detect 5mC at single-base resolution by ligating adaptors only to 5mC-containing fragments generated by DME excision