Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Exhaustive prediction of disease susceptibility to coding base changes in the human genome.
PMID 18793467 · PMC2537574 · BMC bioinformatics · 2008 · 8 claims · 7 setups
Inter-species conservation is the strongest single predictor of disease-associated coding mutations among the factors tested.
-
Full-text index only
Testing groups of genomic locations for enrichment in disease loci using linkage scan data: a method for hypothesis testing.
PMID 16848972 · PMC3525155 · Human genomics · 2006 · 8 claims · 2 setups
A method testing enrichment of a group of genomic locations for disease loci by comparing the average NPL score of the group to a null distribution from randomly drawn groups of equal size
-
Full-text index only
Search for genomic alterations in monozygotic twins discordant for cleft lip and/or palate.
PMID 19803774 · PMC2893889 · Twin research and human genetics : the official journal of the International Society for Twin Studies · 2009 · 7 claims · 5 setups
Postzygotic genomic alterations are not a common cause of monozygotic twin discordance for isolated cleft lip and/or palate.
-
Full-text index only
iMapper: a web application for the automated analysis and mapping of insertional mutagenesis sequence data against Ensembl genomes.
PMID 18974167 · PMC2639305 · Bioinformatics (Oxford, England) · 2008 · 6 claims · 3 setups
iMapper is a web application for automated analysis and mapping of insertional mutagenesis sequence data against vertebrate and invertebrate Ensembl genomes (human, mouse, rat, zebrafish, Drosophila, S. cerevisiae).
-
Full-text index only
Versatile and open software for comparing large genomes.
PMID 14759262 · PMC395750 · Genome biology · 2004 · 8 claims · 8 setups
MUMmer 3.0 efficiently handles comparisons of large eukaryotic genomes at varying evolutionary distances
-
Full-text index only
Information-based methods for predicting gene function from systematic gene knock-downs.
PMID 18959798 · PMC2596148 · BMC bioinformatics · 2008 · 8 claims · 4 setups
Information-based metrics, which incorporate a phenotype's genomic frequency, outperform non-information-based metrics for detecting gene-gene functional similarity from phenotypic knock-down profiles.
-
Full-text index only
Apolipoprotein e, alcohol consumption, and risk of ischemic stroke: the Framingham Heart Study revisited.
PMID 19717024 · PMC2743951 · Journal of stroke and cerebrovascular diseases : the official journal of National Stroke Association · 2009 · 8 claims · 6 setups
ApoE E4 allele does not significantly modify the association between alcohol consumption and risk of ischemic stroke in subjects <65 years
-
Full-text index only
The biological function of some human transcription factor binding motifs varies with position relative to the transcription start site.
PMID 18367472 · PMC2377430 · Nucleic acids research · 2008 · 8 claims · 5 setups
1226 eight-letter DNA words show statistically significant positional preferences relative to the TSS across 7914 human promoter regions
-
Full-text index only
A genome search for primary vesicoureteral reflux shows further evidence for genetic heterogeneity.
PMID 18197425 · PMC2259258 · Pediatric nephrology (Berlin, Germany) · 2008 · 8 claims · 7 setups
Genome-wide linkage analysis identifies several novel loci for primary VUR on chromosomes 1, 3, 4, and 22, supporting genetic heterogeneity.
-
Full-text index only
Assessing the utility of whole-genome amplified serum DNA for array-based high throughput genotyping.
PMID 20021669 · PMC2803178 · BMC genetics · 2009 · 8 claims · 8 setups
WGA DNA from archived serum samples produced low genotyping call rates and is unsuitable for high-resolution genotyping on SNP 6.0 arrays
-
Has reproduction · 89
DFAST and DAGA: web-based integrated genome annotation tools and resources.
PMID 27867804 · PMC5107635 · Bioscience of microbiota, food and health · 2016 · 8 claims · 7 setups
DFAST is a web-based bacterial genome annotation and DDBJ submission pipeline with integrated CheckM quality assessment and ANI taxonomic assessment.
-
Has reproduction · 71
Newborn sex-specific transcriptome signatures and gestational exposure to fine particles: findings from the ENVIRONAGE birth cohort.
PMID 28583124 · PMC5458481 · Environmental health : a global access science source · 2017 · 7 claims · 6 setups
Gestational PM2.5 exposure is associated with sex-specific gene expression changes in newborn cord blood, with major differences between boys and girls.
-
Full-text index only
Increased DNA microarray hybridization specificity using sscDNA targets.
PMID 15847692 · PMC1090574 · BMC genomics · 2005 · 7 claims · 5 setups
A single round of ribo-SPIA amplification produces sufficient sscDNA for microarray hybridization from as little as 5 ng of starting total RNA
-
Full-text index only
Assessing the gene space in draft genomes.
PMID 19042974 · PMC2615622 · Nucleic acids research · 2009 · 6 claims · 7 setups
The proportion of mapped CEGs in a draft genome assembly is a useful metric for describing gene space completeness, complementing N50 and x-fold coverage.
-
Full-text index only
Statistical challenges in preprocessing in microarray experiments in cancer.
PMID 18829474 · PMC3529914 · Clinical cancer research : an official journal of the American Association for Cancer Research · 2008 · 8 claims · 7 setups
Choice of pre-processing method materially changes which features are found significantly associated with survival in the Beer et al. lung cancer microarray dataset
-
Has reproduction · 95
Comparative Genome Analysis of 16SrXII-A 'Candidatus Phytoplasma solani' POT Transmitted by Hyalesthes obsoletus.
PMID 41597744 · PMC12843639 · Microorganisms · 2026 · 7 claims · 8 setups
The complete 832,614 bp circular chromosome of the H. obsoletus-transmissible 'Ca. P. solani' 16SrXII-A strain POT was assembled and functionally reconstructed.
-
Has reproduction · 85
PowerBacGWAS: a computational pipeline to perform power calculations for bacterial genome-wide association studies.
PMID 35338232 · PMC8956664 · Communications biology · 2022 · 8 claims · 8 setups
Two computational approaches (sub-sampling and phenotype-simulation) can be implemented to perform power calculations for bacterial GWAS using existing genome collections, packaged as the PowerBacGWAS pipeline
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
Tracing the origin of functional and conserved domains in the human proteome: implications for protein evolution at the modular level.
PMID 17090320 · PMC1654190 · BMC evolutionary biology · 2006 · 8 claims · 5 setups
HHpred (HMM-HMM comparison) detects remote homologs in the human proteome with higher sensitivity than hmmpfam (HMMER), giving 10% more functional domain coverage and 20% higher residue coverage against Pfam-A families.
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.