Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Comparative Genomics Provide Insight Into the Evolution of European Aphanomyces euteiches Strains.
PMID 41832745 · PMC13044513 · Genome biology and evolution · 2026 · 8 claims · 8 setups
Genome-wide SNP data confirm three genetically distinct A. euteiches populations in Europe, with Italian strains forming a clearly separated group
-
Full-text index only
The genome sequence of the flat clown beetle, Hololepta plana (Sulzer, 1776) (Coleoptera: Histeridae).
PMID 42255359 · PMC13237540 · Wellcome open research · 2026 · 8 claims · 8 setups
A chromosome-level genome assembly was produced for Hololepta plana (icHolPlan1) using the Darwin Tree of Life pipeline
-
Has reproduction · 94
BaRTv2: a highly resolved barley reference transcriptome for accurate transcript-specific RNA-seq quantification.
PMID 35704392 · PMC9546494 · The Plant journal : for cell and molecular biology · 2022 · 8 claims · 6 setups
BaRTv2.18 is the most comprehensive and resolved reference transcriptome in barley to date, containing 39,434 genes and 148,260 transcripts
-
Full-text index only
A chromosome-level reference genome and pangenome for barn swallow population genomics.
PMID 36662619 · PMC10044405 · Cell reports · 2023 · 8 claims · 8 setups
A chromosome-level, karyotype-validated reference genome (bHirRus1) was assembled using the VGP pipeline combining PacBio CLR, 10x Linked-Reads, Bionano optical maps, and Hi-C data
-
Has reproduction · 78
annotate_my_genomes: an easy-to-use pipeline to improve genome annotation and uncover neglected genes by hybrid RNA sequencing.
PMID 36472574 · PMC9724561 · GigaScience · 2022 · 7 claims · 8 setups
annotate_my_genomes is an easy-to-use genome-guided pipeline that uses hybrid (PacBio+Illumina) assembled transcripts to distinguish coding genes from long non-coding RNAs and reconcile them with prior annotations.
-
Full-text index only
Hybrid sequencing reveals incompleteness of the H37Rv reference genome and highlights lineage-specific genomic divergence in Mycobacterium tuberculosis.
PMID 42224013 · PMC13225438 · Microbial genomics · 2026 · 8 claims · 6 setups
The H37Rv_ref reference genome, sequenced in 1998 with early technology, is incomplete relative to modern hybrid-sequenced assemblies
-
Full-text index only
Haplotype-resolved and near telomere-to-telomere assembly of the autotetraploid potato genome.
PMID 41634861 · PMC12955163 · Genome biology · 2026 · 8 claims · 8 setups
PHap is a new pipeline that enables haplotype-resolved, near-T2T assembly of autopolyploid genomes using only standard HiFi, ONT-UL, and Hi-C sequencing data
-
Full-text index only
Transcriptome assemblies for two drug-type cannabis chemotypes by long-read RNA sequencing.
PMID 41968651 · PMC13071343 · The plant genome · 2026 · 8 claims · 7 setups
PacBio Iso-Seq transcriptome assemblies were generated for two contrasting drug-type cannabis chemotypes (THC-dominant and CBD-dominant)
-
Has reproduction · 88
nf-core/isoseq: simple gene and isoform annotation with PacBio Iso-Seq long-read sequencing.
PMID 36961337 · PMC10199315 · Bioinformatics (Oxford, England) · 2023 · 7 claims · 4 setups
nf-core/isoseq is a new automated Nextflow-based pipeline that processes raw Iso-Seq subreads through to genome annotation (BED format) without requiring transcriptome assembly.
-
Has reproduction · 81
Chromosome-scale Elaeis guineensis and E. oleifera assemblies: comparative genomics of oil palm and other Arecaceae.
PMID 38918881 · PMC11373658 · G3 (Bethesda, Md.) · 2024 · 8 claims · 8 setups
Improved E. guineensis genome assembly achieved with substantially increased continuity and completeness compared to prior assemblies
-
Full-text index only
Highly contiguous chromosome-level assembly of the rock goby (Gobius paganellus) genome.
PMID 41611730 · PMC12957465 · Scientific data · 2026 · 8 claims · 8 setups
Chromosome-level genome assembly of Gobius paganellus spans 813 Mb with >99.9% of sequence anchored to 23 pseudochromosomes
-
Has reproduction · 93
Characterization of protein isoform diversity in human umbilical vein endothelial cells via long-read proteogenomics.
PMID 36457147 · PMC9721438 · RNA biology · 2022 · 8 claims · 7 setups
Long-read RNA-seq detected 53,863 transcript isoforms from 10,426 genes in HUVECs, of which 22,195 were novel
-
Full-text index only
Accessing medically relevant complex regions with a pangenome graph of 20 near-complete Japanese haplotypes.
PMID 42203797 · PMC13216315 · Nature communications · 2026 · 8 claims · 8 setups
Generated 20 near-complete haplotypes from 10 Japanese male individuals using PacBio HiFi, ONT ultra-long, and Omni-C reads, all with contig N50 exceeding 100 Mbp
-
Full-text index only
isoSeQL: comparing long-read isoforms across multiple datasets.
PMID 41452740 · PMC12790818 · Bioinformatics (Oxford, England) · 2026 · 8 claims · 4 setups
isoSeQL enables comparison of long-read isoform profiles across multiple datasets by consolidating SQANTI3-annotated samples into a unified SQLite database with consistent isoform IDs
-
Full-text index only
The chromosomal genome sequence of the sponge Phakellia ventilabrum (Linnaeus, 1767) and its associated microbial metagenome sequences.
PMID 41625985 · PMC12859430 · Wellcome open research · 2026 · 8 claims · 8 setups
The Phakellia ventilabrum genome assembly spans 211.92 Mb with 99.97-99.98% scaffolded into 25 chromosomal pseudomolecules
-
Full-text index only
FEDRANN: effective long-read overlap detection based on dimensionality reduction and approximate nearest neighbors.
PMID 42102720 · PMC13201080 · GigaScience · 2026 · 8 claims · 6 setups
A pipeline combining IDF transformation, sparse random projection (SRP), and NNDescent (the FEDRANN strategy) enables accurate overlap detection across diverse long-read datasets
-
Has reproduction · 29
MOSAIK: a hash-based algorithm for accurate next-generation sequencing short-read mapping.
PMID 24599324 · PMC3944147 · PloS one · 2014 · 8 claims · 8 setups
MOSAIK is the only aligner that consistently aligns reads from all major sequencing platforms (Illumina, AB SOLiD, Roche 454, Ion Torrent, Pacific Biosciences SMRT) using the same algorithmic approach.
-
Full-text index only
IFDlong: a model-based isoform and fusion detector for accurate annotation and quantification of long-read RNA-seq data.
PMID 41851882 · PMC13113378 · Genome biology · 2026 · 8 claims · 2 setups
IFDlong is the only tool that can discover fusion transcripts at isoform resolution
-
Has reproduction · 73
Genetic polyploid phasing from low-depth progeny samples.
PMID 35692633 · PMC9184567 · iScience · 2022 · 8 claims · 7 setups
WH-PPG phases polyploid parental samples by scoring informative variant pairs with a Bayesian log-likelihood model of progeny allele depths, clustering alleles by co-occurrence likelihood, and assigning clusters to haplotypes via interval scheduling
-
Has reproduction · 97
Determination of complete chromosomal haplotypes by bulk DNA sequencing.
PMID 33957932 · PMC8101039 · Genome biology · 2021 · 8 claims · 8 setups
A hierarchical computational strategy that first builds high-confidence local haplotype blocks from long-range/linked-read linkage and then concatenates them into whole-chromosome haplotypes using Hi-C contacts