Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 93
Population genomics of the Wolbachia endosymbiont in Drosophila melanogaster.
PMID 23284297 · PMC3527207 · PLoS genetics · 2012 · 8 claims · 8 setups
Wolbachia infection status can be accurately predicted in silico from whole-genome shotgun sequence of individual host strains, showing 99% concordance with diagnostic PCR.
-
Has reproduction
Genome-wide signatures of convergent evolution in echolocating mammals.
PMID 24005325 · PMC3836225 · Nature · 2013 · 8 claims · 8 setups
Genome-wide convergent sequence evolution between echolocating lineages is not rare but widespread and continuously distributed, with signatures consistent with convergence in nearly 200 loci out of 2,326 examined.
-
Has reproduction · 58
The Li2 mutation results in reduced subgenome expression bias in elongating fibers of allotetraploid cotton (Gossypium hirsutum L.).
PMID 24598808 · PMC3944810 · PloS one · 2014 · 8 claims · 7 setups
The Li2 mutation significantly reduces subgenome (homeolog) expression bias in the elongating fiber transcriptome.
-
Has reproduction · 75
Identification of Key Differentially Expressed Genes in Arabidopsis thaliana Under Short- and Long-Term High Light Stress.
PMID 40869111 · PMC12386182 · International journal of molecular sciences · 2025 · 8 claims · 6 setups
Meta-analysis of 21 experiments covering 58 HL conditions yielded ~218,000 DEG instances corresponding to ~19,000 unique A. thaliana genes
-
Has reproduction · 42
KAGE: fast alignment-free graph-based genotyping of SNPs and short indels.
PMID 36195962 · PMC9531401 · Genome biology · 2022 · 7 claims · 7 setups
KAGE combines population-based kmer count modeling with single-variant prior adjustment into an alignment-free genotyper that matches the accuracy of the best existing alignment-free genotypers while being an order of magnitude faster.
-
Has reproduction · 83
Hobbes: optimized gram-based methods for efficient read alignment.
PMID 22199254 · PMC3315303 · Nucleic acids research · 2012 · 8 claims · 4 setups
Hobbes, a gram-based short-read mapper supporting Hamming and edit distance, is faster than all other read-mapping programs tested while maintaining high mapping quality.
-
Has reproduction · 78
A case study for large-scale human microbiome analysis using JCVI's metagenomics reports (METAREP).
PMID 22719821 · PMC3374610 · PloS one · 2012 · 8 claims · 7 setups
METAREP version 1.3.1 is an open-source, scalable tool for querying, browsing and comparing extremely large volumes of metagenomic annotations, with an extended data model, dynamic weighting, distributed searches and advanced clustering.
-
Full-text index only
A Korean family with Arg1448Cys mutation of SCN4A channel causing paramyotonia congenita: electrophysiologic, histopathologic, and molecular genetic studies.
PMID 12483017 · PMC3054970 · Journal of Korean medical science · 2002 · 7 claims · 5 setups
A missense mutation (Arg1448Cys, R1448C) in SCN4A causes paramyotonia congenita in this Korean family
-
Has reproduction · 73
Vespucci: a system for building annotated databases of nascent transcripts.
PMID 24304890 · PMC3936758 · Nucleic acids research · 2014 · 8 claims · 7 setups
Existing ChIP-seq and RNA-seq analysis platforms (e.g. Cufflinks, peak callers) are unsuited to GRO-seq because they assume spliced/exonic reads, uniform density and paired-end data, and cannot identify transcriptional units de novo across the whole genome.
-
Full-text index only
Cruciform extrusion propensity of human translocation-mediating palindromic AT-rich repeats.
PMID 17264116 · PMC1851657 · Nucleic acids research · 2007 · 8 claims · 4 setups
Cruciform extrusion propensity of PATRRs depends on both length and central symmetry of the repeat.
-
Full-text index only
The diploid genome sequence of an Asian individual.
PMID 18987735 · PMC2716080 · Nature · 2008 · 8 claims · 8 setups
First diploid genome sequence of an Asian (Han Chinese) individual generated using massively parallel Illumina sequencing
-
Full-text index only
High throughput sequencing and proteomics to identify immunogenic proteins of a new pathogen: the dirty genome approach.
PMID 20037647 · PMC2793016 · PloS one · 2009 · 7 claims · 7 setups
A dirty genome approach using unfinished, unclosed genome sequences combined with proteomics can rapidly identify immunogenic proteins useful for diagnostic tool development
-
Full-text index only
Application of genomics to toxicology research.
PMID 12634120 · PMC1241273 · Environmental health perspectives · 2002 · 8 claims · 3 setups
Toxic chemical exposures alter gene expression, producing a diagnostic transcriptional 'fingerprint' that can be matched against known toxicants to classify untested chemicals' toxic potential.
-
Has reproduction · 83
Macrel: antimicrobial peptide screening in genomes and metagenomes.
PMID 33384902 · PMC7751412 · PeerJ · 2020 · 8 claims · 8 setups
Macrel introduces a novel set of 22 peptide features (6 local, 16 global), including a new Free Energy Transition (FET) feature group, for AMP and hemolytic activity classification
-
Has reproduction · 73
Genomics of Environmental Salmonella: Engaging Students in the Microbiology and Bioinformatics of Foodborne Pathogens.
PMID 33967968 · PMC8100199 · Frontiers in microbiology · 2021 · 8 claims · 8 setups
An undergraduate CURE combining field sampling, wet-lab microbiology, and genomic bioinformatics can be used to isolate and characterize environmental S. enterica strains.
-
Has reproduction · 89
A near complete genome for goat genetic and genomic research.
PMID 34507524 · PMC8434745 · Genetics, selection, evolution : GSE · 2021 · 8 claims · 8 setups
Saanen_v1 is a near-complete de novo goat genome assembly generated from 117x PacBio and 118x Hi-C data, including the first goat Y chromosome scaffold
-
Has reproduction · 71
A crowdsourced set of curated structural variants for the human genome.
PMID 32559231 · PMC7329145 · PLoS computational biology · 2020 · 8 claims · 8 setups
1235 manually curated SVs were produced that can be used to evaluate SV callers or train machine learning models
-
Has reproduction · 76
Organelle Genomes and Transcriptomes of Nymphaea Reveal the Interplay between Intron Splicing and RNA Editing.
PMID 34576004 · PMC8466565 · International journal of molecular sciences · 2021 · 8 claims · 8 setups
Both cis- and trans-splicing group II introns in Nymphaea organelle genomes are spliced in random order, generating diverse co-existing intermediates rather than following a fixed splicing sequence.
-
Has reproduction · 44
Population differentiation and epidemic tracking of Bursaphelenchus xylophilus in China based on chromosome-level assembly and whole-genome sequencing data.
PMID 34839581 · PMC9300093 · Pest management science · 2022 · 6 claims · 8 setups
Generated the first chromosome-level genome assembly (AH1) of B. xylophilus using PacBio, Illumina, BioNano, and Hi-C data
-
Has reproduction · 78
annotate_my_genomes: an easy-to-use pipeline to improve genome annotation and uncover neglected genes by hybrid RNA sequencing.
PMID 36472574 · PMC9724561 · GigaScience · 2022 · 7 claims · 8 setups
annotate_my_genomes is an easy-to-use genome-guided pipeline that uses hybrid (PacBio+Illumina) assembled transcripts to distinguish coding genes from long non-coding RNAs and reconcile them with prior annotations.