Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
PBAT: a comprehensive software package for genome-wide association analysis of complex family-based studies.
PMID 15814068 · PMC3525120 · Human genomics · 2005 · 8 claims · 1 setups
PBAT provides comprehensive tools for family-based association analysis, including nuclear families with missing parental genotypes, extended pedigrees, SNP and haplotype analysis, quantitative/qualitative/multivariate/longitudinal traits and time-to-onset phenotypes
-
Full-text index only
Whole genome association mapping by incompatibilities and local perfect phylogenies.
PMID 17042942 · PMC1624851 · BMC bioinformatics · 2006 · 8 claims · 8 setups
Blossoc scores the perfect phylogenetic tree spanning the largest compatible region around each marker as a decision tree for case/control status to detect association
-
Full-text index only
Bayesian survival analysis in genetic association studies.
PMID 18617538 · PMC2530885 · Bioinformatics (Oxford, England) · 2008 · 7 claims · 5 setups
A novel Bayesian method (BETA-Surv) extends prior case-control haplotype-clustering work to censored survival outcomes by clustering haplotypes via gene tree/perfect phylogeny topology and relative mutation age.
-
Full-text index only
Screening large-scale association study data: exploiting interactions using random forests.
PMID 15588316 · PMC545646 · BMC genetics · 2004 · 7 claims · 3 setups
Random forest importance measure significantly outperforms the Fisher Exact test as a screening tool when risk SNPs interact.
-
Full-text index only
Simultaneous analysis of all SNPs in genome-wide and re-sequencing association studies.
PMID 18654633 · PMC2464715 · PLoS genetics · 2008 · 8 claims · 5 setups
A Bayesian-inspired penalised maximum likelihood stochastic search method can simultaneously analyse all SNPs (up to 500K) from a GWA study in a few hours on a desktop workstation
-
Full-text index only
PedGenie: an analysis approach for genetic association testing in extended pedigrees and genealogies of arbitrary size.
PMID 16620382 · PMC1459209 · BMC bioinformatics · 2006 · 7 claims · 3 setups
PedGenie is a valid, flexible statistical tool for genetic association analysis in pedigrees of arbitrary size and structure using Monte Carlo significance testing
-
Full-text index only
Application of two machine learning algorithms to genetic association studies in the presence of covariates.
PMID 19014573 · PMC2620353 · BMC genetics · 2008 · 8 claims · 3 setups
The relative performance of RF and MARS for detecting genotype-trait associations depends on both the strategy used to handle covariates and the true underlying model of association (e.g., confounding vs. mediation vs. interaction).
-
Full-text index only
Is replication the gold standard for validating genome-wide association findings?
PMID 19112512 · PMC2605260 · PloS one · 2008 · 8 claims · 4 setups
The probability of replicating a specific GWA-identified variant decreases as the number of independent GWA/replication studies increases, when individual study power is less than 100%.
-
Full-text index only
The androgen receptor CAG repeat polymorphism and modification of breast cancer risk in BRCA1 and BRCA2 mutation carriers.
PMID 15743497 · PMC1064126 · Breast cancer research : BCR · 2005 · 7 claims · 5 setups
The AR CAG repeat polymorphism does not modify breast cancer risk in BRCA1 mutation carriers
-
Full-text index only
Detecting purely epistatic multi-locus interactions by an omnibus permutation test on ensembles of two-locus analyses.
PMID 19761607 · PMC2759961 · BMC bioinformatics · 2009 · 8 claims · 5 setups
2LOmb performs an omnibus permutation test on ensembles of two-locus analyses via a four-step algorithm (two-locus analysis, permutation test, global p-value determination, progressive ensemble search)
-
Full-text index only
Association testing of novel type 2 diabetes risk alleles in the JAZF1, CDC123/CAMK1D, TSPAN8, THADA, ADAMTS9, and NOTCH2 loci with insulin release, insulin sensitivity, and obesity in a population-based sample of 4,516 glucose-tolerant middle-aged Danes.
PMID 18567820 · PMC2518507 · Diabetes · 2008 · 8 claims · 5 setups
CDC123/CAMK1D rs12779790 risk allele (homozygous) is associated with decreased insulinogenic index, corrected insulin response (CIR), and AUC-insulin/AUC-glucose ratio, indicating impaired insulin release
-
Has reproduction · 62
Gbdmr: identifying differentially methylated CpG regions in the human genome via generalized beta regressions.
PMID 38443825 · PMC10916021 · BMC bioinformatics · 2024 · 8 claims · 4 setups
gbdmr models DNA methylation levels of CpG sites using a generalized beta distribution instead of assuming normality as in linear-regression-based methods
-
Full-text index only
What can genome-wide association studies tell us about the genetics of common disease?
PMID 18454206 · PMC2323402 · PLoS genetics · 2008 · 8 claims · 4 setups
Apparent patterns of common, low-effect disease-associated alleles largely reflect statistical power of studies rather than the true underlying distribution of disease variants
-
Full-text index only
A method for detecting epistasis in genome-wide studies using case-control multi-locus association analysis.
PMID 18667089 · PMC2533022 · BMC genomics · 2008 · 7 claims · 2 setups
HFCC is a method/software for genome-wide epistasis detection using case-control multi-locus association analysis, combining a fast computing algorithm with flexibility to test a variety of epistatic models.
-
Full-text index only
PRESTO: rapid calculation of order statistic distributions and multiple-testing adjusted P-values via permutation for one and two-stage genetic association studies.
PMID 18620604 · PMC2483288 · BMC bioinformatics · 2008 · 8 claims · 4 setups
PRESTO is an order of magnitude faster than other existing permutation testing software for genetic association studies.
-
Full-text index only
Imputation-based analysis of association studies: candidate regions and quantitative traits.
PMID 17676998 · PMC1934390 · PLoS genetics · 2007 · 8 claims · 2 setups
Imputation-based Bayesian regression increases power to detect association compared with standard single-SNP tests, even when the causal variant is directly typed
-
Full-text index only
Calibrating the performance of SNP arrays for whole-genome association studies.
PMID 18584036 · PMC2432039 · PLoS genetics · 2008 · 8 claims · 7 setups
Previous SNP array genetic coverage estimates are inflated due to SNP overfitting and sample overfitting, since they were evaluated on the same HapMap SNPs/individuals used to design the arrays.
-
Full-text index only
A hierarchical and modular approach to the discovery of robust associations in genome-wide association studies from pooled DNA samples.
PMID 18194558 · PMC2248205 · BMC genetics · 2008 · 8 claims · 5 setups
A hierarchical/modular approach integrating quality control, LD, physical distance, and gene ontology identifies authentic associations among those found by statistical tests in pooled DNA GWAS
-
Has reproduction · 91
A reference profile-free deconvolution method to infer cancer cell-intrinsic subtypes and tumor-type-specific stromal profiles.
PMID 32111252 · PMC7049190 · Genome medicine · 2020 · 8 claims · 8 setups
DeClust is a reference profile-free deconvolution method that simultaneously deconvolves bulk tumor expression into cancer, immune, and stromal compartments and clusters samples into cancer cell-intrinsic molecular subtypes, outputting subtype-specific reference profiles for the cohort rather than for individuals.
-
Has reproduction · 67
HArmonized single-cell RNA-seq Cell type Assisted Deconvolution (HASCAD).
PMID 37907883 · PMC10619225 · BMC medical genomics · 2023 · 6 claims · 5 setups
HASCAD, a DNN-based cell composition deconvolution model, predicts the fractions of up to 15 immune cell types from bulk RNA-seq.