Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 88
Evaluating genome sequencing strategies: trio, singleton, and standard testing in rare disease diagnosis.
PMID 40963120 · PMC12445032 · Genome medicine · 2025 · 7 claims · 4 setups
Trio genome sequencing (tGS) achieves higher prospective diagnostic yield than standard-of-care (SoC) and singleton genome sequencing (sGS) even when performed by a newly trained team.
-
Has reproduction · 84
Deep transcriptomics reveals cell-specific isoforms of pan-neuronal genes.
PMID 40379625 · PMC12084633 · Nature communications · 2025 · 8 claims · 5 setups
Pan-neuronal genes (expressed in many/all neurons) harbor highly cell-specific splice variants/isoforms restricted to single or few neuron types.
-
Has reproduction · 91
Chromosome-level genome assembly of agar-producing red seaweed Gracilaria vermiculophylla.
PMID 41629338 · PMC12966425 · Scientific data · 2026 · 8 claims · 8 setups
Assembled the first chromosome-level genome of G. vermiculophylla: 77.5 Mb, 22 pseudochromosomes, contig N50 2.61 Mb, scaffold N50 3.16 Mb
-
Has reproduction · 100
Intra-Host Co-Existing Strains of SARS-CoV-2 Reference Genome Uncovered by Exhaustive Computational Search.
PMID 37243151 · PMC10224212 · Viruses · 2023 · 8 claims · 7 setups
An exhaustive-search workflow can recover intra-host co-existing SARS-CoV-2 strains from the reference-genome read set (SRR11092062) that de Bruijn-graph assemblers discard.
-
Full-text index only
Contributions from molecular/biochemical approaches in epidemiology to cancer risk assessment and prevention.
PMID 1486845 · PMC1519598 · Environmental health perspectives · 1992 · 8 claims · 8 setups
Genotoxicity of chemicals is a continuous, graded property (agent score) rather than a simple dichotomy of mutagenic vs. nonmutagenic chemicals
-
Full-text index only
Identification of SLC26A4 gene mutations in Iranian families with hereditary hearing impairment.
PMID 18813951 · PMC4428656 · European journal of pediatrics · 2009 · 7 claims · 6 setups
SLC26A4 mutations are the most prevalent cause of syndromic hereditary hearing loss (Pendred syndrome) in Iran
-
Full-text index only
Nucleotide-resolution analysis of structural variants using BreakSeq and a breakpoint library.
PMID 20037582 · PMC2951730 · Nature biotechnology · 2010 · 8 claims · 7 setups
A standardized, non-redundant library of 1,889 breakpoint-resolved SVs was assembled from eight published surveys
-
Full-text index only
Impact of short-read sequencing on the misassembly of a plant genome.
PMID 33530937 · PMC7852129 · BMC genomics · 2021 · 7 claims · 6 setups
Short-read tomato assembly has substantial high-coverage (0.6%, 5.1 Mb) and low-coverage (9.7%, 79.6 Mb) regions relative to background coverage
-
Has reproduction · 29
MOSAIK: a hash-based algorithm for accurate next-generation sequencing short-read mapping.
PMID 24599324 · PMC3944147 · PloS one · 2014 · 8 claims · 8 setups
MOSAIK is the only aligner that consistently aligns reads from all major sequencing platforms (Illumina, AB SOLiD, Roche 454, Ion Torrent, Pacific Biosciences SMRT) using the same algorithmic approach.
-
Full-text index only
A cost-effective and scalable barcoded library construction method for deep mutational scanning studies.
PMID 41671286 · PMC12923136 · PLoS biology · 2026 · 8 claims · 4 setups
A two-step cloning strategy (Gibson assembly followed by Golden Gate assembly) using degenerate oligo pools (oPools) with co-synthesized DNA barcodes enables construction of scalable DMS libraries for large genes.
-
Full-text index only
BOAT: Basic Oligonucleotide Alignment Tool.
PMID 19958483 · PMC2788372 · BMC genomics · 2009 · 7 claims · 3 setups
BOAT can accurately and efficiently map sequencing reads to a reference genome while handling several substitutions and indels simultaneously
-
Full-text index only
The most frequent short sequences in non-coding DNA.
PMID 19966278 · PMC2831315 · Nucleic acids research · 2010 · 8 claims · 2 setups
Short frequent sequences (9-14 bases) in non-coding DNA may play a role in maintaining chromosome structure and function
-
Has reproduction · 80
Chromosome-level genome of the long-tailed marine-living ornate spiny lobster, Panulirus ornatus.
PMID 38909031 · PMC11193758 · Scientific data · 2024 · 7 claims · 8 setups
A chromosome-level genome of P. ornatus was assembled spanning 2.65 Gb with contig N50 of 51.05 Mb, with 99.11% of sequence anchored to 73 chromosomes
-
Has reproduction · 57
Analysis and comprehensive comparison of PacBio and nanopore-based RNA sequencing of the Arabidopsis transcriptome.
PMID 32536962 · PMC7291481 · Plant methods · 2020 · 8 claims · 8 setups
ONT Pc produces higher raw data quality (higher alignment rate, lower error rate) than ONT Dc, while PacBio generates the longest reads
-
Has reproduction · 92
Acquisition and loss of CTX-M plasmids in Shigella species associated with MSM transmission in the UK.
PMID 34427554 · PMC8549364 · Microbial genomics · 2021 · 8 claims · 8 setups
bla_CTX-M-27 is located on IncFII pKSR100-like plasmids, flanked by IS26 and IS903B
-
Has reproduction · 73
Genetic polyploid phasing from low-depth progeny samples.
PMID 35692633 · PMC9184567 · iScience · 2022 · 8 claims · 7 setups
WH-PPG phases polyploid parental samples by scoring informative variant pairs with a Bayesian log-likelihood model of progeny allele depths, clustering alleles by co-occurrence likelihood, and assigning clusters to haplotypes via interval scheduling
-
Full-text index only
Selecting additional tag SNPs for tolerating missing data in genotyping.
PMID 16259642 · PMC1316880 · BMC bioinformatics · 2005 · 7 claims · 6 setups
There exists a subset of SNPs (robust tag SNPs) that can distinguish all distinct haplotypes even when up to m SNPs are missing
-
Full-text index only
DDBJ dealing with mass data produced by the second generation sequencer.
PMID 18927114 · PMC2686496 · Nucleic acids research · 2009 · 8 claims · 7 setups
DDBJ collected and released 2,368,110 entries (1,415,106,598 bases) of original DNA sequence data from July 2007 to June 2008.
-
Full-text index only
Rise of the machines.
PMID 18670625 · PMC2467494 · PLoS genetics · 2008 · 8 claims · 4 setups
New short-read sequencing platforms (Illumina Genome Analyzer, 454 FLX, ABI SOLiD) enable rapid, scalable whole-genome resequencing that was previously restricted to dedicated sequencing centers using Sanger methods.
-
Full-text index only
BFAST: an alignment tool for large scale genome resequencing.
PMID 19907642 · PMC2770639 · PloS one · 2009 · 7 claims · 4 setups
BFAST is a new algorithm and freely available software tool for aligning large-scale short-read sequencing data to large reference genomes with user-customizable speed and accuracy