Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The origins of lactase persistence in Europe.
PMID 19714206 · PMC2722739 · PLoS computational biology · 2009 · 8 claims · 5 setups
The −13,910*T allele first underwent selection among dairying farmers around 7,500 years ago in a region between the central Balkans and central Europe, possibly linked to the Linearbandkeramik culture.
-
Full-text index only
High genetic variability of HIV-1 in female sex workers from Argentina.
PMID 17697319 · PMC1971708 · Retrovirology · 2007 · 7 claims · 6 setups
HIV-1 genetic diversity among Argentine FSWs is extensive, with BF recombinants predominating over subtypes B and C
-
Full-text index only
Protein ranking by semi-supervised network propagation.
PMID 16723003 · PMC1810311 · BMC bioinformatics · 2006 · 8 claims · 5 setups
RankProp, a diffusion-based network propagation algorithm on a PSI-BLAST-derived protein similarity network, significantly outperforms local search methods (BLAST/PSI-BLAST) at detecting remote homologs.
-
Full-text index only
A new method for 2D gel spot alignment: application to the analysis of large sample sets in clinical proteomics.
PMID 18957120 · PMC2628390 · BMC bioinformatics · 2008 · 8 claims · 2 setups
Sili2DGel represents recursive gel matching results as a weighted undirected graph and identifies SAP by finding cliques and pseudocliques (dense subgraphs) after edge-weight filtering, strength-metric-based graph reduction, and cluster refinement.
-
Full-text index only
Gene prediction in eukaryotes with a generalized hidden Markov model that uses hints from external sources.
PMID 16469098 · PMC1409804 · BMC bioinformatics · 2006 · 7 claims · 3 setups
AUGUSTUS+ extends the AUGUSTUS GHMM by combining intrinsic sequence information with extrinsic hints via an extended emission alphabet, so the GHMM jointly models the DNA sequence, gene structure, and hint collection.
-
Has reproduction · 94
Topological signatures in regulatory network enable phenotypic heterogeneity in small cell lung cancer.
PMID 33729159 · PMC8012062 · eLife · 2021 · 7 claims · 6 setups
Discrete (Boolean/Ising) and continuous (RACIPE) simulations of the SCLC regulatory network yield similar multistable phenotypic distributions, with four dominant steady states (X1-X4) that map onto experimentally observed SCLC molecular subtypes.
-
Has reproduction · 92
An integrative proteomics method identifies a regulator of translation during stem cell maintenance and differentiation.
PMID 34772928 · PMC8590018 · Nature communications · 2021 · 8 claims · 5 setups
PISA-Express is a method that simultaneously measures protein expression and thermal stability (solubility) changes using only two samples per replicate per cell type
-
Full-text index only
What can genome-wide association studies tell us about the genetics of common disease?
PMID 18454206 · PMC2323402 · PLoS genetics · 2008 · 8 claims · 4 setups
Apparent patterns of common, low-effect disease-associated alleles largely reflect statistical power of studies rather than the true underlying distribution of disease variants
-
Full-text index only
Antigenic diversity, transmission mechanisms, and the evolution of pathogens.
PMID 19847288 · PMC2759524 · PLoS computational biology · 2009 · 8 claims · 3 setups
Three distinct infection types (A, B, C) emerge as maxima in the pathogen fitness landscape, each with characteristic within-host dynamics, contact network structure, and transmission mode
-
Has reproduction · 49
oPOSSUM-3: advanced analysis of regulatory motif over-representation across genes or ChIP-Seq datasets.
PMID 22973536 · PMC3429929 · G3 (Bethesda, Md.) · 2012 · 8 claims · 6 setups
oPOSSUM-3 is a web-accessible system that identifies over-represented TFBS and TFBS families in DNA sequences of co-expressed genes or in sequences from high-throughput methods such as ChIP-Seq.
-
Full-text index only
Inconsistencies in Neanderthal genomic DNA sequences.
PMID 17937503 · PMC2014787 · PLoS genetics · 2007 · 8 claims · 6 setups
The Noonan et al. and Green et al. Neanderthal nuclear DNA datasets yield mutually inconsistent estimates of population split time and Neanderthal admixture proportion when analyzed with the same method
-
Has reproduction · 83
Accurate prediction of metagenome-assembled genome completeness by MAGISTA, a random forest model built on alignment-free intra-bin statistics.
PMID 35248155 · PMC8898458 · Environmental microbiome · 2022 · 7 claims · 7 setups
MAGISTA, a random forest model built on alignment-free intra-bin distance-distribution statistics, can estimate MAG completeness and purity without relying on reference marker genes.