Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 88
Evaluating genome sequencing strategies: trio, singleton, and standard testing in rare disease diagnosis.
PMID 40963120 · PMC12445032 · Genome medicine · 2025 · 7 claims · 4 setups
Trio genome sequencing (tGS) achieves higher prospective diagnostic yield than standard-of-care (SoC) and singleton genome sequencing (sGS) even when performed by a newly trained team.
-
Has reproduction · 84
Deep transcriptomics reveals cell-specific isoforms of pan-neuronal genes.
PMID 40379625 · PMC12084633 · Nature communications · 2025 · 8 claims · 5 setups
Pan-neuronal genes (expressed in many/all neurons) harbor highly cell-specific splice variants/isoforms restricted to single or few neuron types.
-
Has reproduction · 75
Identification of Key Differentially Expressed Genes in Arabidopsis thaliana Under Short- and Long-Term High Light Stress.
PMID 40869111 · PMC12386182 · International journal of molecular sciences · 2025 · 7 claims · 5 setups
Short- and long-term HL responses in Arabidopsis leaves are driven by distinct transcriptional programs, with duration of HL treatment as the primary factor separating transcriptomic clusters.
-
Has reproduction · 82
Whole-genome analysis of a multidrug-resistant Klebsiella michiganensis environmental isolate from an orthopedic ward in Mwanza, Tanzania reveals IncF-family plasmid replicon signatures associated with resistance determinants.
PMID 41957580 · PMC13173886 · BMC genomics · 2026 · 6 claims · 8 setups
Genome-based taxonomy (GTDB-Tk and ANI) reclassified the isolate A55848, phenotypically identified as K. oxytoca, as Klebsiella michiganensis
-
Has reproduction
Human Retrotransposons and Effective Computational Detection Methods for Next-Generation Sequencing Data.
PMID 36295018 · PMC9605557 · Life (Basel, Switzerland) · 2022 · 8 claims · 7 setups
Transposable elements make up nearly 45% of the human genome, vastly exceeding the ~1.5% that is protein-coding.
-
Has reproduction · 100
Intra-Host Co-Existing Strains of SARS-CoV-2 Reference Genome Uncovered by Exhaustive Computational Search.
PMID 37243151 · PMC10224212 · Viruses · 2023 · 8 claims · 7 setups
An exhaustive-search workflow can recover intra-host co-existing SARS-CoV-2 strains from the reference-genome read set (SRR11092062) that de Bruijn-graph assemblers discard.
-
Full-text index only
Contributions from molecular/biochemical approaches in epidemiology to cancer risk assessment and prevention.
PMID 1486845 · PMC1519598 · Environmental health perspectives · 1992 · 8 claims · 8 setups
Genotoxicity of chemicals is a continuous, graded property (agent score) rather than a simple dichotomy of mutagenic vs. nonmutagenic chemicals
-
Full-text index only
Identification of SLC26A4 gene mutations in Iranian families with hereditary hearing impairment.
PMID 18813951 · PMC4428656 · European journal of pediatrics · 2009 · 7 claims · 6 setups
SLC26A4 mutations are the most prevalent cause of syndromic hereditary hearing loss (Pendred syndrome) in Iran
-
Full-text index only
Nucleotide-resolution analysis of structural variants using BreakSeq and a breakpoint library.
PMID 20037582 · PMC2951730 · Nature biotechnology · 2010 · 8 claims · 7 setups
A standardized, non-redundant library of 1,889 breakpoint-resolved SVs was assembled from eight published surveys
-
Full-text index only
BOAT: Basic Oligonucleotide Alignment Tool.
PMID 19958483 · PMC2788372 · BMC genomics · 2009 · 7 claims · 3 setups
BOAT can accurately and efficiently map sequencing reads to a reference genome while handling several substitutions and indels simultaneously
-
Full-text index only
The most frequent short sequences in non-coding DNA.
PMID 19966278 · PMC2831315 · Nucleic acids research · 2010 · 8 claims · 2 setups
Short frequent sequences (9-14 bases) in non-coding DNA may play a role in maintaining chromosome structure and function
-
Has reproduction · 80
Chromosome-level genome of the long-tailed marine-living ornate spiny lobster, Panulirus ornatus.
PMID 38909031 · PMC11193758 · Scientific data · 2024 · 6 claims · 5 setups
A chromosome-level genome of P. ornatus spanning 2.65 Gb was assembled with a contig N50 of 51.05 Mb, anchoring 99.11% of sequences to 73 chromosomes.
-
Has reproduction · 57
Analysis and comprehensive comparison of PacBio and nanopore-based RNA sequencing of the Arabidopsis transcriptome.
PMID 32536962 · PMC7291481 · Plant methods · 2020 · 8 claims · 8 setups
ONT Pc produces higher raw data quality (higher alignment rate, lower error rate) than ONT Dc, while PacBio generates the longest reads
-
Has reproduction · 92
Acquisition and loss of CTX-M plasmids in Shigella species associated with MSM transmission in the UK.
PMID 34427554 · PMC8549364 · Microbial genomics · 2021 · 8 claims · 8 setups
bla_CTX-M-27 is located on IncFII pKSR100-like plasmids, flanked by IS26 and IS903B
-
Has reproduction · 73
Genetic polyploid phasing from low-depth progeny samples.
PMID 35692633 · PMC9184567 · iScience · 2022 · 8 claims · 7 setups
WH-PPG phases polyploid parental samples by scoring informative variant pairs with a Bayesian log-likelihood model of progeny allele depths, clustering alleles by co-occurrence likelihood, and assigning clusters to haplotypes via interval scheduling
-
Full-text index only
Selecting additional tag SNPs for tolerating missing data in genotyping.
PMID 16259642 · PMC1316880 · BMC bioinformatics · 2005 · 7 claims · 6 setups
There exists a subset of SNPs (robust tag SNPs) that can distinguish all distinct haplotypes even when up to m SNPs are missing
-
Full-text index only
DDBJ dealing with mass data produced by the second generation sequencer.
PMID 18927114 · PMC2686496 · Nucleic acids research · 2009 · 8 claims · 7 setups
DDBJ collected and released 2,368,110 entries (1,415,106,598 bases) of original DNA sequence data from July 2007 to June 2008.
-
Full-text index only
Rise of the machines.
PMID 18670625 · PMC2467494 · PLoS genetics · 2008 · 8 claims · 4 setups
New short-read sequencing platforms (Illumina Genome Analyzer, 454 FLX, ABI SOLiD) enable rapid, scalable whole-genome resequencing that was previously restricted to dedicated sequencing centers using Sanger methods.
-
Full-text index only
BFAST: an alignment tool for large scale genome resequencing.
PMID 19907642 · PMC2770639 · PloS one · 2009 · 7 claims · 4 setups
BFAST is a new algorithm and freely available software tool for aligning large-scale short-read sequencing data to large reference genomes with user-customizable speed and accuracy
-
Full-text index only
Slider--maximum use of probability information for alignment of short sequence reads and SNP detection.
PMID 18974170 · PMC2638935 · Bioinformatics (Oxford, England) · 2009 · 7 claims · 3 setups
Slider aligns reads using all bases above a probability threshold (baseMinPrb) from prb files, generating all possible read sequences above a read probability threshold (read_0_MinPrb), rather than only the most probable sequence