Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 100
Integrative transcriptome sequencing identifies trans-splicing events with important roles in human embryonic stem cell pluripotency.
PMID 24131564 · PMC3875859 · Genome research · 2014 · 8 claims · 8 setups
TSscan, a computational pipeline integrating long- and short-read transcriptome sequencing from multiple hESC lines, can detect trans-splicing while minimizing false positives from experimental artifacts and genetic rearrangements.
-
Has reproduction · 90
LoRA-TV: read depth profile-based clustering of tumor cells in single-cell sequencing.
PMID 38877886 · PMC11179121 · Briefings in bioinformatics · 2024 · 6 claims · 2 setups
LoRA-TV jointly processes read-depth profiles of all cells by stacking them into a matrix and applying low-rank approximation plus total-variation smoothing to capture shared genomic signatures for clustering.
-
Full-text index only
CholeraSeq: a comprehensive genomic pipeline for cholera surveillance and near real-time outbreak investigation.
PMID 41400832 · PMC12790814 · Bioinformatics (Oxford, England) · 2026 · 7 claims · 6 setups
CholeraSeq is an automated, V. cholerae-specific Nextflow pipeline that processes WGS outbreak data from raw reads/assemblies to high-quality SNPs and phylogenies in near-real-time.
-
Full-text index only
NOD-like receptor repertoire in the chromosome-level genome of the demosponge Dysidea avara (Schmidt, 1862).
PMID 41710890 · PMC12909245 · Frontiers in immunology · 2026 · 8 claims · 8 setups
Dysidea avara has a chromosome-level genome assembly of 575 Mb, N50 41 Mb, 162 scaffolds, and 15 chromosomes.
-
Full-text index only
A repeat expansion in GOLGA8A is a major risk factor for atypical frontotemporal lobar degeneration with ubiquitin-positive inclusions.
PMID 41820575 · PMC13083237 · Nature genetics · 2026 · 5 claims · 3 setups
A genome-wide association study identifies a major risk locus for aFTLD-U on chromosome 15q14, with lead SNP rs549846383 (P=5.85×10^-21, OR=26.7)
-
Has reproduction · 95
nf-rnaSeqCount: A Nextflow pipeline for obtaining raw read counts from RNA-seq data.
PMID 35574063 · PMC9097006 · South African computer journal = Suid-Afrikaanse rekenaartydskrif · 2021 · 7 claims · 5 setups
nf-rnaSeqCount is a portable, reproducible Nextflow pipeline that maps RNA-seq reads to a reference genome and quantifies gene abundance for differential expression analysis
-
Has reproduction · 90
A genome-wide association analysis identifies 16 novel susceptibility loci for carpal tunnel syndrome.
PMID 30833571 · PMC6399342 · Nature communications · 2019 · 6 claims · 8 setups
A GWAS of 12,312 CTS cases and 389,344 controls in UK Biobank identifies 16 novel genome-wide significant susceptibility loci for CTS
-
Has reproduction · 43
TransFlow: a Snakemake workflow for transmission analysis of Mycobacterium tuberculosis whole-genome sequencing data.
PMID 36469333 · PMC9825751 · Bioinformatics (Oxford, England) · 2023 · 8 claims · 8 setups
TransFlow is a Snakemake- and Conda-based workflow that combines state-of-the-art tools into a single, fast, scalable pipeline for MTBC WGS-based transmission analysis.
-
Full-text index only
Unraveling Cefiderocol Resistance in NDM- and OXA-48-like Co-Producing Klebsiella pneumoniae Isolates Through Integrated Genomic and Phenotypic Analysis.
PMID 42192735 · PMC13203471 · Antibiotics (Basel, Switzerland) · 2026 · 8 claims · 6 setups
K. pneumoniae isolates co-producing NDM and OXA-48-like carbapenemases are predominantly clonal, belonging to the high-risk ST147 lineage.
-
Has reproduction · 88
nf-core/isoseq: simple gene and isoform annotation with PacBio Iso-Seq long-read sequencing.
PMID 36961337 · PMC10199315 · Bioinformatics (Oxford, England) · 2023 · 7 claims · 4 setups
nf-core/isoseq is a new automated Nextflow-based pipeline that processes raw Iso-Seq subreads through to genome annotation (BED format) without requiring transcriptome assembly.
-
Has reproduction
sRNAbench and sRNAtoolbox 2019: intuitive fast small RNA profiling and differential expression.
PMID 31114926 · PMC6602500 · Nucleic acids research · 2019 · 8 claims · 1 setups
sRNAtoolbox 2019 update adds automatic processing of the most used small RNA library preparation protocols, including UMI-based protocols, to sRNAbench
-
Has reproduction · 100
A comprehensive framework for analysis of microRNA sequencing data in metastatic colorectal cancer.
PMID 35047825 · PMC8759566 · NAR cancer · 2022 · 8 claims · 7 setups
Five miRNAs (Mir-210_3p, Mir-191_5p, Mir-8-P1b_3p, Mir-1307_5p, Mir-155_5p) are up-regulated at multiple metastatic CRC sites compared to primary CRC.
-
Full-text index only
TOFU-MAaPO: fast, scalable and reproducible analysis of large metagenome sequence data from the Sequence Read Archive.
PMID 42277027 · PMC13260335 · Nature communications · 2026 · 8 claims · 5 setups
TOFU-MAaPO yields significantly more high-quality MAGs than metaFun, nf-core/mag, and ATLAS due to integration of multiple complementary binning tools with unified MAGScoT refinement
-
Has reproduction · 95
MetaMap: an atlas of metatranscriptomic reads in human disease-related RNA-seq data.
PMID 29901703 · PMC6025204 · GigaScience · 2018 · 8 claims · 7 setups
The MetaMap pipeline recapitulates known infection agents in bona fide dual RNA-seq validation studies (Salmonella, HPV, HSV, rhinovirus)
-
Has reproduction · 82
Landscape of allele-specific transcription factor binding in the human genome.
PMID 33980847 · PMC8115691 · Nature communications · 2021 · 8 claims · 6 setups
A novel statistical framework (ADASTRA) calls allele-specific TF binding from existing ChIP-Seq alignments by jointly correcting for background allelic dosage (BAD, from aneuploidy/CNVs) and reference mapping bias.
-
Has reproduction · 98
maxATAC: Genome-scale transcription-factor binding prediction from ATAC-seq with deep neural networks.
PMID 36719906 · PMC9917285 · PLoS computational biology · 2023 · 8 claims · 6 setups
maxATAC is a suite of deep neural network models enabling state-of-the-art, genome-scale TFBS prediction from ATAC-seq, with models for 127 human transcription factors
-
Has reproduction · 73
Vespucci: a system for building annotated databases of nascent transcripts.
PMID 24304890 · PMC3936758 · Nucleic acids research · 2014 · 8 claims · 7 setups
Existing ChIP-seq and RNA-seq analysis platforms (e.g. Cufflinks, peak callers) are unsuited to GRO-seq because they assume spliced/exonic reads, uniform density and paired-end data, and cannot identify transcriptional units de novo across the whole genome.
-
Full-text index only
umite: fast quantification of Smart-seq3 libraries with improved UMI retrieval.
PMID 41692984 · PMC12989134 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 6 setups
umite offers efficient mismatch-tolerant (fuzzy) UMI detection that boosts UMI retrieval by 5%-15% compared to standard position-based matching
-
Full-text index only
Ensembl 2008.
PMID 18000006 · PMC2238821 · Nucleic acids research · 2008 · 8 claims · 6 setups
The Ensembl regulatory build integrates multiple genome-wide functional genomics datasets to automatically annotate regulatory regions and assign putative functions across the genome.
-
Has reproduction · 79
Enriched domain detector: a program for detection of wide genomic enrichment domains robust against local variations.
PMID 24782521 · PMC4066758 · Nucleic acids research · 2014 · 8 claims · 5 setups
EDD is a new algorithm that detects broad (megabase-size) enrichment domains from ChIP-seq data of widely distributed chromatin proteins such as A- and B-type lamins.