Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Simultaneous analysis of all SNPs in genome-wide and re-sequencing association studies.
PMID 18654633 · PMC2464715 · PLoS genetics · 2008 · 8 claims · 5 setups
A Bayesian-inspired penalised maximum likelihood stochastic search method can simultaneously analyse all SNPs (up to 500K) from a GWA study in a few hours on a desktop workstation
-
Full-text index only
A Hidden Markov Model to estimate population mixture and allelic copy-numbers in cancers using Affymetrix SNP arrays.
PMID 17996079 · PMC2206057 · BMC bioinformatics · 2007 · 8 claims · 7 setups
An HMM using paired germline genotype calls and tumour allelic SNP intensities can estimate allele-specific copy-numbers, distinguishing events like uniparental disomy from allelic imbalance.
-
Full-text index only
Direct inference of SNP heterozygosity rates and resolution of LOH detection.
PMID 18052545 · PMC2098867 · PLoS computational biology · 2007 · 6 claims · 7 setups
A large proportion of SNPs in dbSNP have high-variance HET rate estimates, limiting their reliability for LOH study design.
-
Full-text index only
Screening large-scale association study data: exploiting interactions using random forests.
PMID 15588316 · PMC545646 · BMC genetics · 2004 · 7 claims · 3 setups
Random forest importance measure significantly outperforms the Fisher Exact test as a screening tool when risk SNPs interact.
-
Full-text index only
Bayesian estimates of linkage disequilibrium.
PMID 17592642 · PMC1924864 · BMC genetics · 2007 · 8 claims · 3 setups
The MLE of D' is biased toward disequilibrium, with the bias particularly severe in small samples (<100 subjects) and rare alleles (MAF<0.05)
-
Full-text index only
Cubic exact solutions for the estimation of pairwise haplotype frequencies: implications for linkage disequilibrium analyses and a web tool 'CubeX'.
PMID 17980034 · PMC2180187 · BMC bioinformatics · 2007 · 6 claims · 4 setups
CubeX, a Python program/web tool, computes the exact algebraic (Cardan/Nickalls) solution(s) of Hill's cubic equation to estimate pairwise haplotype frequencies, D', r2 and chi-square for each solution
-
Full-text index only
Inferring human colonization history using a copying model.
PMID 18497854 · PMC2367454 · PLoS genetics · 2008 · 8 claims · 6 setups
A copying-model approach using SNP haplotype sharing can infer both the order of population founding and the donor populations contributing ancestry to each new population.
-
Full-text index only
SiDCoN: a tool to aid scoring of DNA copy number changes in SNP chip data.
PMID 17971856 · PMC2034603 · PloS one · 2007 · 8 claims · 3 setups
SiDCoN is a spreadsheet-based application that simulates Ballele and logR plots for all known types of DNA copy number change, with or without stromal contamination, for up to 5000 SNP data points and up to 3 combined aberrations
-
Full-text index only
Quantitative analysis of single nucleotide polymorphisms within copy number variation.
PMID 19093001 · PMC2600609 · PloS one · 2008 · 8 claims · 2 setups
Copy number variation is a major factor in HWE violation for SNPs with small minor allele frequency, large sample size, and 0-1% genotyping error rate
-
Full-text index only
Computational tradeoffs in multiplex PCR assay design for SNP genotyping.
PMID 16042802 · PMC1190169 · BMC genomics · 2005 · 7 claims · 6 setups
Achieving high-multiplexing/high-coverage multiplex PCR designs is subject to a computational phase transition as the SNP-pair compatibility probability crosses a critical threshold
-
Full-text index only
A simple and efficient algorithm for genome-wide homozygosity analysis in disease.
PMID 19756043 · PMC2758715 · Molecular systems biology · 2009 · 8 claims · 4 setups
A genome-wide AH analysis (GAHA) algorithm can identify disease-associated loci by comparing frequencies of homozygous segments between cases and controls using a z-statistic proportion test
-
Full-text index only
Visualization of shared genomic regions and meiotic recombination in high-density SNP data.
PMID 19696932 · PMC2725774 · PloS one · 2009 · 8 claims · 7 setups
SNPduo is a command-line (SNPduo++) and web-accessible tool that analyzes and visualizes relatedness between two individuals using identity by state (IBS) from SNP genotypes.
-
Full-text index only
Information-theoretic identification of predictive SNPs and supervised visualization of genome-wide association studies.
PMID 16899448 · PMC1557808 · Nucleic acids research · 2006 · 7 claims · 4 setups
3D VizStruct (DFT-based radial mapping + KLD as z-axis) can identify SNPs/polymorphic markers that are predictive of underlying biological class distinctions across diverse datasets
-
Full-text index only
Calibrating the performance of SNP arrays for whole-genome association studies.
PMID 18584036 · PMC2432039 · PLoS genetics · 2008 · 8 claims · 7 setups
Previous SNP array genetic coverage estimates are inflated due to SNP overfitting and sample overfitting, since they were evaluated on the same HapMap SNPs/individuals used to design the arrays.
-
Full-text index only
Iterative pruning PCA improves resolution of highly structured populations.
PMID 19930644 · PMC2790469 · BMC bioinformatics · 2009 · 7 claims · 7 setups
ipPCA is a novel algorithm that assigns individuals to subpopulations and infers the total number of subpopulations (K) present in genotypic data
-
Full-text index only
Effect of read-mapping biases on detecting allele-specific expression from RNA-sequencing data.
PMID 19808877 · PMC2788925 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 6 setups
Reads mapped to the reference genome show a significant bias toward the reference allele at heterozygous SNPs
-
Full-text index only
A method for detecting epistasis in genome-wide studies using case-control multi-locus association analysis.
PMID 18667089 · PMC2533022 · BMC genomics · 2008 · 7 claims · 2 setups
HFCC is a method/software for genome-wide epistasis detection using case-control multi-locus association analysis, combining a fast computing algorithm with flexibility to test a variety of epistatic models.
-
Full-text index only
Analysis of concordance of different haplotype block partitioning algorithms.
PMID 16356172 · PMC1343594 · BMC bioinformatics · 2005 · 7 claims · 7 setups
Each block partitioning algorithm infers blocks differing in number, size, and coverage under different SNP density and allele frequency conditions.
-
Full-text index only
HapMap-based study of the 17q21 ERBB2 amplicon in susceptibility to breast cancer.
PMID 17117180 · PMC2360759 · British journal of cancer · 2006 · 6 claims · 5 setups
Common genetic variation (tSNPs and haplotypes) across the 400-kb 17q21 ERBB2 amplicon is not associated with breast cancer risk in British women.
-
Full-text index only
A genome-wide approach to identify genetic loci with a signature of natural selection in the Irish population.
PMID 16904005 · PMC1779589 · Genome biology · 2006 · 8 claims · 7 setups
Eight SNPs with extreme European-branch locus-specific branch length (LSBL) were selected from a genome-wide FST dataset as candidates for selection in Europe.