Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 88
nf-core/isoseq: simple gene and isoform annotation with PacBio Iso-Seq long-read sequencing.
PMID 36961337 · PMC10199315 · Bioinformatics (Oxford, England) · 2023 · 7 claims · 4 setups
nf-core/isoseq is a new automated Nextflow-based pipeline that processes raw Iso-Seq subreads through to genome annotation (BED format) without requiring transcriptome assembly.
-
Has reproduction · 57
Analysis and comprehensive comparison of PacBio and nanopore-based RNA sequencing of the Arabidopsis transcriptome.
PMID 32536962 · PMC7291481 · Plant methods · 2020 · 8 claims · 8 setups
ONT Pc produces higher raw data quality (higher alignment rate, lower error rate) than ONT Dc, while PacBio generates the longest reads
-
Has reproduction · 89
A near complete genome for goat genetic and genomic research.
PMID 34507524 · PMC8434745 · Genetics, selection, evolution : GSE · 2021 · 8 claims · 8 setups
Saanen_v1 is a near-complete de novo goat genome assembly generated from 117x PacBio and 118x Hi-C data, including the first goat Y chromosome scaffold
-
Full-text index only
The fungal pathogen Rhizoctonia solani AG-8 has 2 nuclear haplotypes that differ in abundance.
PMID 41124349 · PMC12774589 · G3 (Bethesda, Md.) · 2026 · 8 claims · 8 setups
R. solani isolates AG8-1 and AG8-3 each possess 2 distinct nuclear haplotypes, each ~50 Mbp assembled into 16 chromosomes
-
Has reproduction · 84
The SARS-CoV-2 subgenome landscape and its novel regulatory features.
PMID 33713597 · PMC7927579 · Molecular cell · 2021 · 8 claims · 6 setups
Template switching in SARS-CoV-2 can occur bidirectionally, generating diverse subgenomes through successive template-switching events
-
Full-text index only
Chromatin state architecture governs transcription factor accessibility across plant genomes.
PMID 41570051 · PMC12867329 · PLoS genetics · 2026 · 8 claims · 8 setups
Chromatin states show a large degree of functional conservation between Arabidopsis thaliana and Marchantia polymorpha across more than 450 million years of land plant evolution
-
Full-text index only
A new chromosome-level genome assembly for western painted turtle Chrysemys picta bellii, a model for extreme physiological adaptations.
PMID 41792601 · PMC13077969 · BMC genomics · 2026 · 6 claims · 8 setups
A new haplotype-resolved, chromosome-level reference genome assembly (SLU_Cpb5.0) was generated for C. picta bellii using combined PacBio HiFi, 10x Genomics Chromium, Hi-C, and Bionano optical mapping data from a single individual.
-
Has reproduction · 78
annotate_my_genomes: an easy-to-use pipeline to improve genome annotation and uncover neglected genes by hybrid RNA sequencing.
PMID 36472574 · PMC9724561 · GigaScience · 2022 · 7 claims · 8 setups
annotate_my_genomes is an easy-to-use genome-guided pipeline that uses hybrid (PacBio+Illumina) assembled transcripts to distinguish coding genes from long non-coding RNAs and reconcile them with prior annotations.
-
Has reproduction · 94
BaRTv2: a highly resolved barley reference transcriptome for accurate transcript-specific RNA-seq quantification.
PMID 35704392 · PMC9546494 · The Plant journal : for cell and molecular biology · 2022 · 8 claims · 6 setups
BaRTv2.18 is the most comprehensive and resolved reference transcriptome in barley to date, containing 39,434 genes and 148,260 transcripts
-
Has reproduction · 79
Enhanced protein isoform characterization through long-read proteogenomics.
PMID 35241129 · PMC8892804 · Genome biology · 2022 · 6 claims · 4 setups
A long-read proteogenomics pipeline integrating PacBio long-read RNA-seq with MS-based proteomics enhances isoform-resolved protein characterization
-
Has reproduction · 93
Characterization of protein isoform diversity in human umbilical vein endothelial cells via long-read proteogenomics.
PMID 36457147 · PMC9721438 · RNA biology · 2022 · 8 claims · 7 setups
Long-read RNA-seq detected 53,863 transcript isoforms from 10,426 genes in HUVECs, of which 22,195 were novel
-
Full-text index only
Optimizing Single-Cell Long-Read Sequencing for Enhanced Isoform Detection in Pancreatic Islets.
PMID 41563441 · PMC13007207 · Diabetes · 2026 · 8 claims · 7 setups
5′ single-cell library preparation protocols outperform 3′ protocols for transcript identification and read length
-
Full-text index only
FracFixR: a compositional statistical framework for absolute proportion estimation between fractions in RNA sequencing data.
PMID 41264734 · PMC12866640 · Bioinformatics (Oxford, England) · 2026 · 7 claims · 5 setups
FracFixR reconstructs original fraction proportions by modeling the compositional relationship between whole and fractionated RNA samples using non-negative least squares (NNLS) regression on selected transcripts
-
Full-text index only
Variant-resolved prediction of context-specific isoform variation with a graph-based attention model.
PMID 41547351 · PMC13069856 · Cell genomics · 2026 · 8 claims · 8 setups
Otari, an attention-based graph neural network trained on long-read transcriptomes across 30 tissues/brain regions, predicts tissue-specific differential isoform abundance
-
Full-text index only
Long-read sequencing and proteomics reveal blood transcriptome and protein expression profiles in multiple primary lung cancers.
PMID 41761147 · PMC13059216 · BMC cancer · 2026 · 6 claims · 7 setups
MPC patients exhibit significantly increased blood transcript complexity, with higher numbers of DEGs, DETs, and DTU events than OPLC and HC groups
-
Has reproduction · 50
Time course profiling of host cell response to herpesvirus infection using nanopore and synthetic long-read transcriptome sequencing.
PMID 34244540 · PMC8270970 · Scientific reports · 2021 · 8 claims · 5 setups
BoHV-1 infection causes substantial up- and down-regulation of host gene networks, including antiviral response and viral transcription/translation-associated genes
-
Has reproduction · 100
Integrative transcriptome sequencing identifies trans-splicing events with important roles in human embryonic stem cell pluripotency.
PMID 24131564 · PMC3875859 · Genome research · 2014 · 8 claims · 8 setups
TSscan, a computational pipeline integrating long- and short-read transcriptome sequencing from multiple hESC lines, can detect trans-splicing while minimizing false positives from experimental artifacts and genetic rearrangements.
-
Full-text index only
Fully haplotyped genome assemblies of healthy individuals reveal variability in 5'ss strength and support by splicing regulatory proteins.
PMID 40191587 · PMC11970367 · NAR genomics and bioinformatics · 2025 · 8 claims · 5 setups
44 individuals' fully haplotyped diploid genome assemblies (88 haplotypes) from the 1000 Genomes Project were used to comprehensively assess homozygous and heterozygous sequence variations around and within 5'ss
-
Full-text index only
Single-cell transcriptomic analysis of plant quiescent center by third-generation sequencing reveals developmental trajectories.
PMID 41664205 · PMC12990480 · Genome biology · 2026 · 8 claims · 8 setups
Developed an improved protoplasting/hand-picking protocol enabling isolation of intact QC cells for single-cell long-read RNA sequencing (SCAN-seq)
-
Has reproduction · 98
Diminutive, degraded but dissimilar: Wolbachia genomes from filarial nematodes do not conform to a single paradigm.
PMID 33295865 · PMC8116671 · Microbial genomics · 2020 · 8 claims · 4 setups
wCtub and wDcau (863 988 bp and 863 427 bp) are the smallest Wolbachia genomes sequenced to date and are the first genomes representing supergroup J.