Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
MultiPhyl: a high-throughput phylogenomics webserver using distributed computing.
PMID 17553837 · PMC1933173 · Nucleic acids research · 2007 · 8 claims · 8 setups
MultiPhyl is the first high-throughput distributed phylogenetics platform capable of using idle computational resources of many heterogeneous non-dedicated machines to form a phylogenetics supercomputer
-
Full-text index only
Importance sampling for the infinite sites model.
PMID 18976228 · PMC2832804 · Statistical applications in genetics and molecular biology · 2008 · 7 claims · 2 setups
A new importance sampling proposal distribution for the ISM, derived from a new result on exact sampling from a single segregating site, generally shows greater efficiency than the GT and SD proposals.
-
Full-text index only
Evidence of recombination in Hepatitis C Virus populations infecting a hemophiliac patient.
PMID 19922637 · PMC2784780 · Virology journal · 2009 · 7 claims · 6 setups
A new intragenotypic recombinant HCV strain (1b/1a), named H23, was detected in 1 of 10 hemophiliac patients studied
-
Full-text index only
Molecular genetic evidence for unifocal origin of advanced epithelial ovarian cancer and for minor clonal divergence.
PMID 7577492 · PMC2033953 · British journal of cancer · 1995 · 7 claims · 4 setups
LOH analysis has higher sensitivity than DNA flow cytometry for detecting unifocal origin of bilateral ovarian tumors
-
Full-text index only
Getting positive about selection.
PMID 12914654 · PMC193638 · Genome biology · 2003 · 8 claims · 4 setups
Purifying selection is the predominant form of molecular evolution, preserving fitness by eliminating deleterious mutations, while positive selection is rare but critical for adaptation.
-
Full-text index only
Codon usage comparison of novel genes in clinical isolates of Haemophilus influenzae.
PMID 15983137 · PMC1160521 · Nucleic acids research · 2005 · 8 claims · 4 setups
A codon usage similarity statistic (ε, based on squared/absolute differences of codon frequencies with an optimized amino acid usage factor) was developed to compare ORFs against a set of 80 reference genomes.
-
Full-text index only
Evolutionary sequence analysis of complete eukaryote genomes.
PMID 15762985 · PMC1274250 · BMC bioinformatics · 2005 · 8 claims · 6 setups
A conservative genome-comparison method (MIA) identifies panorthologs — strict single-copy 1:1 orthologs containing only species divergences, no paralogy — to minimize errors from gene duplication in evolutionary sequence analysis.
-
Full-text index only
A third approach to gene prediction suggests thousands of additional human transcribed regions.
PMID 16543943 · PMC1391917 · PLoS computational biology · 2006 · 8 claims · 7 setups
A third basic concept for gene prediction exists, based on detecting strand-specific 'transcription footprints' (mutational and selectional biases) rather than gene structure or sequence similarity.
-
Full-text index only
Global distribution of rubella virus genotypes.
PMID 14720390 · PMC3034328 · Emerging infectious diseases · 2003 · 8 claims · 6 setups
Phylogenetic analysis of 103 E1 gene sequences from 17 countries confirms at least two rubella virus genotypes, RGI and RGII
-
Full-text index only
A computational screen for type I polyketide synthases in metagenomics shotgun data.
PMID 18953415 · PMC2568958 · PloS one · 2008 · 8 claims · 6 setups
Combining HMM domain searches with maximum-likelihood phylogenetic trees can discriminate true PKS I sequences from evolutionarily related but functionally different enzymes (e.g., FAS I) in metagenomic data.
-
Full-text index only
Genome-wide survey for biologically functional pseudogenes.
PMID 16680195 · PMC1456316 · PLoS computational biology · 2006 · 8 claims · 6 setups
A subset of ancient, cross-species-conserved pseudogenes (30 of 1,453 candidate quartets) show evidence consistent with retained biological function
-
Has reproduction · 94
Large-Scale Phylogenomics of the Lactobacillus casei Group Highlights Taxonomic Inconsistencies and Reveals Novel Clade-Associated Features.
PMID 28845461 · PMC5566788 · mSystems · 2017 · 8 claims · 8 setups
The L. casei group resolves into three distinct clades (A, B, C) supported by phylogeny, GC content, ANI, and TETRA, and many strains are misclassified relative to their nearest type strain.
-
Has reproduction · 65
FusionQ: a novel approach for gene fusion detection and quantification from paired-end RNA-Seq.
PMID 23768108 · PMC3691734 · BMC bioinformatics · 2013 · 8 claims · 8 setups
FusionQ is a novel tool that detects gene fusions, constructs chimerical transcript structures, and estimates their abundances from paired-end RNA-Seq data.
-
Full-text index only
A model-based approach to selection of tag SNPs.
PMID 16776821 · PMC1525207 · BMC bioinformatics · 2006 · 7 claims · 5 setups
The Li and Stephens hidden Markov model outperforms other tested models (simple Markov, two-state HMM, HMM-4D, greedy GR-1/GR-2) in description code-length, tag set information content, and prediction of tagged SNPs.
-
Full-text index only
Genome-wide survey of allele-specific splicing in humans.
PMID 18518984 · PMC2427040 · BMC genomics · 2008 · 8 claims · 5 setups
A genome-wide computational scan identified 30,977 SNPs located within predicted splicing regulatory sequences (donor sites, acceptor sites, branch points, and ESEs)
-
Has reproduction · 57
Diapause vs. reproductive programs: transcriptional phenotypes in a keystone copepod.
PMID 33782539 · PMC8007741 · Communications biology · 2021 · 8 claims · 7 setups
t-SNE clustering of all-gene expression data groups field-collected (diapause program) samples into one cluster while early and late culture (reproductive program) samples separate into two distinct phenotypes
-
Has reproduction · 92
Evaluation of core genome and whole genome multilocus sequence typing schemes for Campylobacter jejuni and Campylobacter coli outbreak detection in the USA.
PMID 37133905 · PMC10272873 · Microbial genomics · 2023 · 8 claims · 8 setups
cgMLST, wgMLST and hqSNP WGS-based analysis methods clustered C. jejuni and C. coli isolates in concordance with epidemiological data.
-
Full-text index only
Eighth major clade for hepatitis delta virus.
PMID 17073101 · PMC3294742 · Emerging infectious diseases · 2006 · 7 claims · 7 setups
Three HDV isolates (dFr644, dFr2072, dFr2736) form a monophyletic group distinct from HDV-1 through HDV-7, constituting a new eighth major clade (HDV-8) of the Deltavirus genus.
-
Full-text index only
Clustering of phosphorylation site recognition motifs can be exploited to predict the targets of cyclin-dependent kinase.
PMID 17316440 · PMC1852407 · Genome biology · 2007 · 8 claims · 6 setups
CDK consensus motifs are frequently clustered (closely spaced) in known CDK substrate proteins rather than uniformly distributed
-
Full-text index only
Direct evidence of extensive diversity of HIV-1 in Kinshasa by 1960.
PMID 18833279 · PMC3682493 · Nature · 2008 · 7 claims · 8 setups
Recovered and characterized HIV-1 sequences (DRC60) from a 1960 Bouin's-fixed paraffin-embedded lymph node biopsy from Léopoldville, Belgian Congo