Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A note on generalized Genome Scan Meta-Analysis statistics.
PMID 15717930 · PMC551600 · BMC bioinformatics · 2005 · 7 claims · 3 setups
An Edgeworth series approximation to the null distribution of the weighted GSMA statistic provides a more accurate representation than the normal approximation, especially in the tails
-
Full-text index only
A Hidden Markov Model to estimate population mixture and allelic copy-numbers in cancers using Affymetrix SNP arrays.
PMID 17996079 · PMC2206057 · BMC bioinformatics · 2007 · 8 claims · 7 setups
An HMM using paired germline genotype calls and tumour allelic SNP intensities can estimate allele-specific copy-numbers, distinguishing events like uniparental disomy from allelic imbalance.
-
Full-text index only
Testing whether genetic variation explains correlation of quantitative measures of gene expression, and application to genetic network analysis.
PMID 18444230 · PMC2729096 · Statistics in medicine · 2008 · 8 claims · 3 setups
A statistical test (delta method and Steiger-Browne optimal linear composites) is developed to test equality of the marginal correlation and the partial correlation of two gene expression traits conditional on a set of covariates.
-
Has reproduction · 10
RADAR: differential analysis of MeRIP-seq data with a random effect model.
PMID 31870409 · PMC6927177 · Genome biology · 2019 · 8 claims · 6 setups
RADAR is a novel analytical tool for differential methylation analysis of MeRIP-seq data combining gene-level INPUT normalization with a Poisson random effect model.
-
Has reproduction · 75
Revealing the critical state and identifying individualized dynamic network biomarker for type 2 diabetes through advanced analysis methods on individual basis.
PMID 39890881 · PMC11785715 · Scientific reports · 2025 · 8 claims · 5 setups
sJSD, NIG, and TNFE methods can detect critical states/tipping points before disease deterioration using only a single sample
-
Full-text index only
On the analysis of glycomics mass spectrometry data via the regularized area under the ROC curve.
PMID 18076765 · PMC2211327 · BMC bioinformatics · 2007 · 8 claims · 4 setups
The TGDR-AUC algorithm regularizes the empirical AUC by replacing the non-differentiable 0-1 loss with a smooth sigmoid surrogate function and applies constrained threshold gradient descent regularization
-
Full-text index only
A statistical model to identify differentially expressed proteins in 2D PAGE gels.
PMID 19763172 · PMC2734266 · PLoS computational biology · 2009 · 7 claims · 5 setups
A mixture likelihood model incorporating both detected and non-detected proteins has higher statistical power to detect differential expression than standard approaches like the Student's t-test.
-
Full-text index only
SiDCoN: a tool to aid scoring of DNA copy number changes in SNP chip data.
PMID 17971856 · PMC2034603 · PloS one · 2007 · 8 claims · 3 setups
SiDCoN is a spreadsheet-based application that simulates Ballele and logR plots for all known types of DNA copy number change, with or without stromal contamination, for up to 5000 SNP data points and up to 3 combined aberrations
-
Full-text index only
Optimality driven nearest centroid classification from genomic data.
PMID 17912341 · PMC1991588 · PloS one · 2007 · 7 claims · 5 setups
A theoretical result determines the subset of features of a given size that minimizes the misclassification rate for a nearest-centroid (LDA) classifier, based on equation (4).
-
Full-text index only
Accuracy of predicting the genetic risk of disease using a genome-wide approach.
PMID 18852893 · PMC2561058 · PloS one · 2008 · 8 claims · 4 setups
Deterministic equations can predict the accuracy (r_gĝ) of genome-wide genetic risk/value prediction for continuous, dichotomous, and case-control study designs.
-
Full-text index only
Design and analysis issues in genome-wide somatic mutation studies of cancer.
PMID 18692126 · PMC2820387 · Genomics · 2009 · 6 claims · 4 setups
Two-stage (discovery + validation) sequencing designs efficiently allocate resources and can produce highly informative candidate driver gene lists even with relatively small sample sizes.
-
Full-text index only
Large-scale copy number variants (CNVs): distribution in normal subjects and FISH/real-time qPCR analysis.
PMID 17565693 · PMC1920519 · BMC genomics · 2007 · 8 claims · 4 setups
42 different CNVs were detected in 27 phenotypically normal individuals using 1 Mb resolution BAC array-CGH
-
Full-text index only
A simple and efficient algorithm for genome-wide homozygosity analysis in disease.
PMID 19756043 · PMC2758715 · Molecular systems biology · 2009 · 8 claims · 4 setups
A genome-wide AH analysis (GAHA) algorithm can identify disease-associated loci by comparing frequencies of homozygous segments between cases and controls using a z-statistic proportion test
-
Full-text index only
A method for detecting epistasis in genome-wide studies using case-control multi-locus association analysis.
PMID 18667089 · PMC2533022 · BMC genomics · 2008 · 7 claims · 2 setups
HFCC is a method/software for genome-wide epistasis detection using case-control multi-locus association analysis, combining a fast computing algorithm with flexibility to test a variety of epistatic models.
-
Full-text index only
QuantiSNP: an Objective Bayes Hidden-Markov Model to detect and accurately map copy number variation using SNP genotyping data.
PMID 17341461 · PMC1874617 · Nucleic acids research · 2007 · 8 claims · 7 setups
QuantiSNP (OB-HMM) provides probabilistic quantification of copy number states and significantly improves accuracy of segmental aneuploidy identification and breakpoint mapping relative to existing tools (BeadStudio/Illumina)
-
Full-text index only
A non-parametric meta-analysis approach for combining independent microarray datasets: application using two microarray datasets pertaining to chronic allograft nephropathy.
PMID 18302764 · PMC2276496 · BMC genomics · 2008 · 8 claims · 6 setups
A novel non-parametric meta-analysis approach for combining independent microarray datasets is presented, requiring no distributional assumptions and being logically intuitive.
-
Full-text index only
Hit selection with false discovery rate control in genome-scale RNAi screens.
PMID 18628291 · PMC2504311 · Nucleic acids research · 2008 · 8 claims · 3 setups
A Bayesian FDR-controlling methodology for hit selection in genome-scale RNAi HTS is proposed, using a direct posterior probability approach analogous to Newton et al.
-
Has reproduction · 100
Intratumoral heterogeneity in microsatellite instability status at single-cell resolution.
PMID 41767255 · PMC12936829 · iScience · 2026 · 8 claims · 8 setups
MSI status can be heterogeneous at the single-cell level within a tumor, challenging its use as a binary biomarker
-
Full-text index only
Meta-analysis of inter-species liver co-expression networks elucidates traits associated with common human diseases.
PMID 20019805 · PMC2787626 · PLoS computational biology · 2009 · 8 claims · 8 setups
A novel semi-parametric meta-analysis method (based on a gene-centric Glass's d effect size) outperforms existing parametric and non-parametric meta-analysis methods at identifying functionally coherent gene pairs across species.
-
Has reproduction · 88
pwrEWAS: a user-friendly tool for comprehensive power estimation for epigenome wide association studies (EWAS).
PMID 31035919 · PMC6489300 · BMC bioinformatics · 2019 · 7 claims · 8 setups
pwrEWAS is a user-friendly tool (R package and Shiny web interface) for comprehensive power estimation in two-group EWAS using Illumina HumanMethylation BeadChip technology.