Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The personal genome project.
PMID 16729065 · PMC1681452 · Molecular systems biology · 2005 · 8 claims · 1 setups
A Personal Genome Project (PGP) should be established as the natural successor to the Human Genome Project, providing integrated genome and phenome data for genetically diverse subjects.
-
Full-text index only
Searching for interpretable rules for disease mutations: a simulated annealing bump hunting strategy.
PMID 16984653 · PMC1618409 · BMC bioinformatics · 2006 · 8 claims · 6 setups
The proposed feature set outperforms existing published feature sets for predicting effects of amino acid substitutions
-
Full-text index only
Application of genomics to toxicology research.
PMID 12634120 · PMC1241273 · Environmental health perspectives · 2002 · 8 claims · 3 setups
Toxic chemical exposures alter gene expression, producing a diagnostic transcriptional 'fingerprint' that can be matched against known toxicants to classify untested chemicals' toxic potential.
-
Has reproduction · 89
Spatial information matters: are traditional imputation methods effective for spatial transcriptomics data?
PMID 41627342 · PMC12862982 · Briefings in bioinformatics · 2026 · 7 claims · 3 setups
No single existing SOTA imputation method consistently performs well across newer SRT platforms/datasets
-
Full-text index only
A statistical approach for array CGH data analysis.
PMID 15705208 · PMC549559 · BMC bioinformatics · 2005 · 8 claims · 4 setups
Existing model-selection criteria (AIC, BIC, and prior ad hoc penalties) are not well adapted to estimating the number of segments in array CGH data
-
Full-text index only
The complete genome sequence of Vibrio cholerae: a tale of two chromosomes and of two lifestyles.
PMID 11178241 · PMC138858 · Genome biology · 2000 · 8 claims · 4 setups
The V. cholerae O1 (El Tor) genome consists of two chromosomes with asymmetrically distributed gene functions
-
Full-text index only
The impact of low-cost, genome-wide resequencing on association studies.
PMID 16004723 · PMC3525269 · Human genomics · 2005 · 8 claims · 4 setups
The HapMap project provides more than 500,000 SNPs across three major human populations, enabling genome-wide association studies via linkage disequilibrium tagging.
-
Full-text index only
Metabolic syndrome: from epidemiology to systems biology.
PMID 18852695 · PMC2829312 · Nature reviews. Genetics · 2008 · 8 claims · 8 setups
MetSyn component traits (obesity, insulin resistance, dyslipidaemia, hypertension) exhibit causal interactions and common etiologies rather than being independent conditions
-
Full-text index only
A simple and robust method for connecting small-molecule drugs using gene-expression signatures.
PMID 18518950 · PMC2464610 · BMC bioinformatics · 2008 · 8 claims · 4 setups
A new method for building reference gene-expression profiles and scoring/testing connections improves on the original Connectivity Map by enabling statistical significance testing of connections.
-
Full-text index only
Incorporating genetics and genomics in risk assessment for inhaled manganese: from data to policy.
PMID 19646473 · PMC2765692 · Neurotoxicology · 2009 · 8 claims · 7 setups
Inhaled manganese bypasses normal homeostatic regulation and can accumulate in the brain, unlike dietary manganese which is readily excreted
-
Full-text index only
Comparative genomics of cyclin-dependent kinases suggest co-evolution of the RNAP II C-terminal domain and CTD-directed CDKs.
PMID 15380029 · PMC521075 · BMC genomics · 2004 · 8 claims · 6 setups
Cell-cycle related CDKs (orthologs of CDK1-6) are present in all sampled eukaryotic organisms, including the most ancestral protists.
-
Full-text index only
Optimality driven nearest centroid classification from genomic data.
PMID 17912341 · PMC1991588 · PloS one · 2007 · 7 claims · 5 setups
A theoretical result determines the subset of features of a given size that minimizes the misclassification rate for a nearest-centroid (LDA) classifier, based on equation (4).
-
Has reproduction · 75
PHA4GE quality control contextual data tags: standardized annotations for sharing public health sequence datasets with known quality issues to facilitate testing and training.
PMID 38860884 · PMC11261899 · Microbial genomics · 2024 · 7 claims · 3 setups
PHA4GE developed a set of standardized contextual data tags (five fields plus controlled-vocabulary terms) for annotating pathogen sequence datasets with known quality issues.
-
Full-text index only
A general definition and nomenclature for alternative splicing events.
PMID 18688268 · PMC2467475 · PLoS computational biology · 2008 · 6 claims · 4 setups
Existing AS nomenclatures (Malko et al.'s 5-letter strings, Nagasaki et al.'s bit matrices, and the ASD/ATD/AEdb system) are redundant, ambiguous, or incapable of representing complex or large splicing variations.