Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 75
Revealing the critical state and identifying individualized dynamic network biomarker for type 2 diabetes through advanced analysis methods on individual basis.
PMID 39890881 · PMC11785715 · Scientific reports · 2025 · 8 claims · 5 setups
sJSD, NIG, and TNFE methods can detect critical states/tipping points before disease deterioration using only a single sample
-
Has reproduction · 84
An accurate method for identifying recent recombinants from unaligned sequences.
PMID 35025988 · PMC8963311 · Bioinformatics (Oxford, England) · 2022 · 8 claims · 4 setups
A novel algorithm combining the JHMM (Zilversmit et al. 2013) mosaic representation with a distance-based triple comparison can identify recombinant sequences and their parents from unaligned, gene-length sequences without a reference panel.
-
Full-text index only
Analysis of concordance of different haplotype block partitioning algorithms.
PMID 16356172 · PMC1343594 · BMC bioinformatics · 2005 · 7 claims · 7 setups
Each block partitioning algorithm infers blocks differing in number, size, and coverage under different SNP density and allele frequency conditions.
-
Full-text index only
Modeling the amplification dynamics of human Alu retrotransposons.
PMID 16201008 · PMC1239904 · PLoS computational biology · 2005 · 8 claims · 4 setups
Combining sequence diversity (π) and insertion polymorphism level (IPL) statistics can statistically exclude implausible Alu amplification scenarios and narrow the range of plausible ones for individual subfamilies.
-
Full-text index only
Evaluation of six methods for estimating synonymous and nonsynonymous substitution rates.
PMID 17127215 · PMC5054070 · Genomics, proteomics & bioinformatics · 2006 · 8 claims · 4 setups
Incorporating more sequence evolution features (transition/transversion bias, nucleotide/codon frequency bias) into Ka/Ks estimation methods yields more accurate and reliable estimates.
-
Full-text index only
Computing Ka and Ks with a consideration of unequal transitional substitutions.
PMID 16740169 · PMC1552089 · BMC evolutionary biology · 2006 · 7 claims · 7 setups
MYN, a modified version of the Yang-Nielsen (YN) algorithm based on the Tamura-Nei Model, allows unequal transitional substitution rates between purines (κR) and pyrimidines (κY) plus codon frequency bias
-
Full-text index only
Direct maximum parsimony phylogeny reconstruction from genotype data.
PMID 18053244 · PMC2222657 · BMC bioinformatics · 2007 · 6 claims · 4 setups
The paper presents the first practical method for computing maximum parsimony phylogenies directly from genotype data, using integer linear programming.
-
Full-text index only
Inferring human colonization history using a copying model.
PMID 18497854 · PMC2367454 · PLoS genetics · 2008 · 8 claims · 6 setups
A copying-model approach using SNP haplotype sharing can infer both the order of population founding and the donor populations contributing ancestry to each new population.
-
Full-text index only
Testing whether genetic variation explains correlation of quantitative measures of gene expression, and application to genetic network analysis.
PMID 18444230 · PMC2729096 · Statistics in medicine · 2008 · 8 claims · 3 setups
A statistical test (delta method and Steiger-Browne optimal linear composites) is developed to test equality of the marginal correlation and the partial correlation of two gene expression traits conditional on a set of covariates.
-
Full-text index only
Bayesian estimates of linkage disequilibrium.
PMID 17592642 · PMC1924864 · BMC genetics · 2007 · 8 claims · 3 setups
The MLE of D' is biased toward disequilibrium, with the bias particularly severe in small samples (<100 subjects) and rare alleles (MAF<0.05)
-
Full-text index only
Importance sampling for the infinite sites model.
PMID 18976228 · PMC2832804 · Statistical applications in genetics and molecular biology · 2008 · 7 claims · 2 setups
A new importance sampling proposal distribution for the ISM, derived from a new result on exact sampling from a single segregating site, generally shows greater efficiency than the GT and SD proposals.
-
Has reproduction · 94
Deep learning from phylogenies to uncover the epidemiological dynamics of outbreaks.
PMID 35794110 · PMC9258765 · Nature communications · 2022 · 8 claims · 5 setups
Deep learning (FFNN-SS and CNN-CBLV) enables accurate and fast likelihood-free estimation of epidemiological parameters and model selection from phylogenies
-
Full-text index only
Computational tradeoffs in multiplex PCR assay design for SNP genotyping.
PMID 16042802 · PMC1190169 · BMC genomics · 2005 · 7 claims · 6 setups
Achieving high-multiplexing/high-coverage multiplex PCR designs is subject to a computational phase transition as the SNP-pair compatibility probability crosses a critical threshold
-
Full-text index only
Evolutionary distance estimation and fidelity of pair wise sequence alignment.
PMID 15840174 · PMC1087827 · BMC bioinformatics · 2005 · 8 claims · 8 setups
Evolutionary distance estimation is relatively unaffected by alignment error as long as 50% or more of homologous sites remain identical between sequences
-
Full-text index only
On the analysis of glycomics mass spectrometry data via the regularized area under the ROC curve.
PMID 18076765 · PMC2211327 · BMC bioinformatics · 2007 · 8 claims · 4 setups
The TGDR-AUC algorithm regularizes the empirical AUC by replacing the non-differentiable 0-1 loss with a smooth sigmoid surrogate function and applies constrained threshold gradient descent regularization
-
Full-text index only
Accuracy of predicting the genetic risk of disease using a genome-wide approach.
PMID 18852893 · PMC2561058 · PloS one · 2008 · 8 claims · 4 setups
Deterministic equations can predict the accuracy (r_gĝ) of genome-wide genetic risk/value prediction for continuous, dichotomous, and case-control study designs.
-
Full-text index only
Stability analysis of mixtures of mutagenetic trees.
PMID 18366778 · PMC2335279 · BMC bioinformatics · 2008 · 7 claims · 5 setups
Mutagenetic trees mixture models capture multiple alternative pathways of ordered accumulation of genetic events (e.g., HIV resistance mutations, cancer chromosomal aberrations).
-
Has reproduction · 100
nf-core/mag: a best-practice pipeline for metagenome hybrid assembly and binning.
PMID 35118380 · PMC8808542 · NAR genomics and bioinformatics · 2022 · 8 claims · 7 setups
nf-core/mag is a Nextflow/nf-core pipeline for hybrid metagenome assembly, binning and taxonomic classification of MAGs.
-
Full-text index only
A statistical change point model approach for the detection of DNA copy number variations in array CGH data.
PMID 19875853 · PMC4154476 · IEEE/ACM transactions on computational biology and bioinformatics · 2009 · 7 claims · 4 setups
A novel mean and variance change point model (MVCM) is proposed to detect CNVs/breakpoints in aCGH data.
-
Has reproduction · 62
Equivalent change enrichment analysis: assessing equivalent and inverse change in biological pathways between diverse experiments.
PMID 32093613 · PMC7041296 · BMC genomics · 2020 · 7 claims · 3 setups
The Equivalent Change Index (ECI), a gene-level statistic ranging from -1 to 1, quantifies whether a gene was changed to the same (1) or completely opposite (-1) degree across two experiments relative to their controls.