Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
A statistical approach designed for finding mathematically defined repeats in shotgun data and determining the length distribution of clone-inserts.
PMID 15626332 · PMC5172250 · Genomics, proteomics & bioinformatics · 2003 · 8 claims · 6 setups
Repeats of different copy number have distinct probabilities of appearance in shotgun data, which can be modeled statistically to define recognition thresholds (MDRs) at different shotgun coverages.
-
Full-text index only
A note on generalized Genome Scan Meta-Analysis statistics.
PMID 15717930 · PMC551600 · BMC bioinformatics · 2005 · 7 claims · 3 setups
An Edgeworth series approximation to the null distribution of the weighted GSMA statistic provides a more accurate representation than the normal approximation, especially in the tails
-
Has reproduction · 67
Generative and integrative modeling for transcriptomics with formalin fixed paraffin embedded material.
PMID 41029822 · PMC12486589 · Journal of translational medicine · 2025 · 8 claims · 5 setups
fRNA-seq transcript counts are best fit by the negative binomial distribution, with little evidence supporting zero-inflated extensions
-
Full-text index only
Insights into the coupling of duplication events and macroevolution from an age profile of animal transmembrane gene families.
PMID 16895434 · PMC1534073 · PLoS computational biology · 2006 · 8 claims · 7 setups
The density of transmembrane gene duplicates positively correlates with the estimated maximum number of cell types of common ancestors
-
Full-text index only
What can genome-wide association studies tell us about the genetics of common disease?
PMID 18454206 · PMC2323402 · PLoS genetics · 2008 · 8 claims · 4 setups
Apparent patterns of common, low-effect disease-associated alleles largely reflect statistical power of studies rather than the true underlying distribution of disease variants
-
Has reproduction · 85
A mechanistic model captures the emergence and implications of non-genetic heterogeneity and reversible drug resistance in ER+ breast cancer cells.
PMID 34316714 · PMC8271219 · NAR cancer · 2021 · 7 claims · 8 setups
EMT and tamoxifen-resistance (TamR) regulatory axes can drive one another, enabling non-genetic heterogeneity via six co-existing phenotypes (ES, ER, HS, HR, MS, MR)
-
Full-text index only
Genome-wide scans for loci under selection in humans.
PMID 16004726 · PMC3525256 · Human genomics · 2005 · 8 claims · 4 setups
Natural selection and population demographic history both distort patterns of genetic variation relative to the standard neutral model, so single-locus tests cannot unambiguously distinguish selection from demography.
-
Full-text index only
Direct inference of SNP heterozygosity rates and resolution of LOH detection.
PMID 18052545 · PMC2098867 · PLoS computational biology · 2007 · 6 claims · 7 setups
A large proportion of SNPs in dbSNP have high-variance HET rate estimates, limiting their reliability for LOH study design.
-
Full-text index only
A universal mechanism ties genotype to phenotype in trinucleotide diseases.
PMID 18039028 · PMC2082501 · PLoS computational biology · 2007 · 8 claims · 5 setups
A universal mechanism of somatic, length-dependent trinucleotide repeat expansion toward a disease-specific pathological threshold explains genotype-phenotype correlations common to trinucleotide diseases
-
Full-text index only
Accuracy of predicting the genetic risk of disease using a genome-wide approach.
PMID 18852893 · PMC2561058 · PloS one · 2008 · 8 claims · 4 setups
Deterministic equations can predict the accuracy (r_gĝ) of genome-wide genetic risk/value prediction for continuous, dichotomous, and case-control study designs.
-
Full-text index only
Modeling ChIP sequencing in silico with applications.
PMID 18725927 · PMC2507756 · PLoS computational biology · 2008 · 8 claims · 4 setups
Observed ChIP-seq tag counts follow an initial power-law distribution followed by a long right tail.
-
Full-text index only
Spontaneous symmetry breaking in genome evolution.
PMID 18367477 · PMC2377439 · Nucleic acids research · 2008 · 6 claims · 3 setups
Exon size distributions in sequenced genomes follow a lognormal pattern typical of a random Kolmogoroff fractioning process
-
Full-text index only
Threshold-dominated regulation hides genetic variation in gene expression networks.
PMID 18062810 · PMC2238762 · BMC systems biology · 2007 · 8 claims · 2 setups
Threshold robustness (insensitivity of a singular/regulating variable's equilibrium value to parameter perturbations, except threshold changes) increases with increasing response function steepness and is present even under Michaelis-Menten conditions, not just in the step-function limit.
-
Has reproduction · 100
Topological approximate Bayesian computation for parameter inference of an angiogenesis model.
PMID 35191485 · PMC9048691 · Bioinformatics (Oxford, England) · 2022 · 7 claims · 3 setups
TDA summary statistics can be combined with ABC to infer parameters (ρ, χ) of the Anderson–Chaplain angiogenesis model
-
Full-text index only
A computational study of off-target effects of RNA interference.
PMID 15800213 · PMC1072799 · Nucleic acids research · 2005 · 8 claims · 5 setups
The chance of RNAi off-target effects is considerable, ranging from 5% to 80% depending on organism and parameters, when using exact sequence identity between siRNA and transcripts.
-
Full-text index only
Duplication count distributions in DNA sequences.
PMID 19256873 · PMC3121164 · Physical review. E, Statistical, nonlinear, and soft matter physics · 2008 · 8 claims · 8 setups
Duplication count distributions N(c) for complex 40-mers show power-law-like decay for c roughly 3 to 50 (or higher) across human, C. elegans, A. thaliana, and D. melanogaster genomes.
-
Full-text index only
The whole alignment and nothing but the alignment: the problem of spurious alignment flanks.
PMID 18796526 · PMC2566872 · Nucleic acids research · 2008 · 8 claims · 4 setups
Some common scoring schemes tend to overextend alignments, generating spurious alignment flanks up to hundreds of bp/amino acids in length
-
Full-text index only
A statistical model to identify differentially expressed proteins in 2D PAGE gels.
PMID 19763172 · PMC2734266 · PLoS computational biology · 2009 · 7 claims · 5 setups
A mixture likelihood model incorporating both detected and non-detected proteins has higher statistical power to detect differential expression than standard approaches like the Student's t-test.
-
Full-text index only
Calculating expected DNA remnants from ancient founding events in human population genetics.
PMID 18928554 · PMC2588638 · BMC genetics · 2008 · 8 claims · 3 setups
Genetic parameters (native/migrant population size, mutation rate, generations since admixture) strongly determine the final frequency of migrant alleles detectable today.
-
Full-text index only
Benchmarking tools for the alignment of functional noncoding DNA.
PMID 14736341 · PMC344529 · BMC bioinformatics · 2004 · 8 claims · 4 setups
Global alignment tools (Avid, ClustalW, Lagan, Needle, DiAlign-G) typically have higher sensitivity over entire noncoding sequences and within constrained blocks than local tools