Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Species-specific protein sequence and fold optimizations.
PMID 12487631 · PMC139977 · BMC bioinformatics · 2002 · 7 claims · 7 setups
Environmental niche is a significant factor explaining variability in amino acid composition across 100 complete genomes
-
Has reproduction · 68
Improved precision of epigenetic clock estimates across tissues and its implication for biological ageing.
PMID 31443728 · PMC6708158 · Genome medicine · 2019 · 8 claims · 6 setups
The proportion of variance in chronological age explained by all DNA methylation probes is close to 1, so a near-perfect age predictor is in principle achievable with sufficient training data.
-
Full-text index only
Adaptively inferring human transcriptional subnetworks.
PMID 16760900 · PMC1681499 · Molecular systems biology · 2006 · 8 claims · 7 setups
A multivariate linear spline (MARS-based) model correlating PWM binding scores with log expression ratios can identify active cis-motif combinations in mammalian promoters without requiring gene clustering.
-
Has reproduction · 74
An open RNA-Seq data analysis pipeline tutorial with an example of reprocessing data from a recent Zika virus study.
PMID 27583132 · PMC4972086 · F1000Research · 2016 · 6 claims · 6 setups
An open-source, reproducible RNA-seq pipeline delivered as an IPython notebook and Docker image can process raw RNA-seq data into interactive PCA/HC plots, enrichment results, and small-molecule predictions with minimal setup overhead
-
Has reproduction · 88
Human methylome variation across Infinium 450K data on the Gene Expression Omnibus.
PMID 33937763 · PMC8061458 · NAR genomics and bioinformatics · 2021 · 8 claims · 6 setups
Approximately two-thirds of compiled HM450K samples are from blood, one-quarter from brain, and roughly one-third from cancer patients.
-
Full-text index only
Clear detection of ADIPOQ locus as the major gene for plasma adiponectin: results of genome-wide association analyses including 4659 European individuals.
PMID 20018283 · PMC2845297 · Atherosclerosis · 2010 · 8 claims · 7 setups
The ADIPOQ locus is the only major gene locus for plasma adiponectin, reaching genome-wide significance in combined and sex-stratified analyses
-
Has reproduction · 50
Exploiting convergent phenotypes to derive a pan-cancer cisplatin response gene expression signature.
PMID 37076665 · PMC10115855 · NPJ precision oncology · 2023 · 8 claims · 8 setups
A convergent-phenotype-based seed gene/co-expression method can extract consensus gene expression signatures predictive of response to chemotherapeutic drugs in the GDSC database
-
Has reproduction · 90
A Decentralized Kidney Transplant Biopsy Classifier for Transplant Rejection Developed Using Genes of the Banff-Human Organ Transplant Panel.
PMID 35619722 · PMC9128066 · Frontiers in immunology · 2022 · 6 claims · 6 setups
A random forest model trained solely on B-HOT panel genes (B-HOT Model) accurately classifies kidney transplant biopsies as NR, ABMR, or TCMR.
-
Full-text index only
Are the so-called low penetrance breast cancer genes, ATM, BRIP1, PALB2 and CHEK2, high risk for women with strong family histories?
PMID 18557994 · PMC2481495 · Breast cancer research : BCR · 2008 · 8 claims · 8 setups
Mutation frequencies in ATM, BRIP1, PALB2 and CHEK2 are many times higher in women with strong breast cancer family history (familial cases) than in population controls.