Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 94
Manually curated transcriptomics data collection for toxicogenomic assessment of engineered nanomaterials.
PMID 33558569 · PMC7870661 · Scientific data · 2021 · 7 claims · 7 setups
A unified collection of 101 manually curated and homogenized transcriptomics datasets covering human, mouse, and rat ENM exposures in vitro and in vivo was compiled.
-
Full-text index only
The Proteomics Identifications database: 2010 update.
PMID 19906717 · PMC2808904 · Nucleic acids research · 2010 · 8 claims · 6 setups
PRIDE has become one of the main repositories for MS-based proteomics data, with substantial growth in data holdings over the last two years.
-
Has reproduction · 71
Sustainable data analysis with Snakemake.
PMID 34035898 · PMC8114187 · F1000Research · 2021 · 8 claims · 4 setups
Reproducibility alone is insufficient for sustainable data analysis; transparency and adaptability are equally important additional properties.
-
Has reproduction · 74
Wide-Open: Accelerating public data release by automating detection of overdue datasets.
PMID 28594819 · PMC5464523 · PLoS biology · 2017 · 7 claims · 5 setups
Wide-Open is a general text-mining approach that automatically detects overdue datasets by scanning PubMed articles for dataset accession identifiers and querying repositories to determine if the datasets remain private.
-
Has reproduction · 96
Calibration-free NGS quantitation of mutations below 0.01% VAF.
PMID 34675197 · PMC8531361 · Nature communications · 2021 · 8 claims · 6 setups
QBDA (Quantitative Blocker Displacement Amplification) integrates UMI molecular barcoding with BDA variant enrichment to enable calibration-free VAF quantitation
-
Full-text index only
Effects of DNA mass on multiple displacement whole genome amplification and genotyping performance.
PMID 16168060 · PMC1249558 · BMC biotechnology · 2005 · 8 claims · 6 setups
Increased gDNA input into the MDA WGA reaction increases the proportion of double-stranded and human-specific PCR-amplifiable wgaDNA and improves genotyping performance.
-
Has reproduction · 99
The systematic assessment of completeness of public metadata accompanying omics studies in the Gene Expression Omnibus data repository.
PMID 40926267 · PMC12421755 · Genome biology · 2025 · 8 claims · 3 setups
Over 25% of critical metadata are omitted, with only 74.8% of relevant phenotypes available in publications or public repositories.
-
Has reproduction · 95
The archives are half-empty: an assessment of the availability of microbial community sequencing data.
PMID 32859925 · PMC7455719 · Communications biology · 2020 · 8 claims · 5 setups
More than half of surveyed amplicon sequencing studies were affected by lack of data deposition, improper file formatting, or inconsistent labeling that impede reuse.
-
Has reproduction · 50
Workflow sharing with automated metadata validation and test execution to improve the reusability of published workflows.
PMID 36810800 · PMC9944229 · GigaScience · 2022 · 8 claims · 5 setups
Yevis is a system that builds a workflow registry which automatically validates and tests workflows prior to publication, ensuring they are 'reusable with confidence'.
-
Full-text index only
Genomics and the prevention and control of common chronic diseases: emerging priorities for public health action.
PMID 15888216 · PMC1327699 · Preventing chronic disease · 2005 · 8 claims · 6 setups
Family history is the most consistent risk factor for almost all human diseases across the lifespan.
-
Full-text index only
Exploiting noise in array CGH data to improve detection of DNA copy number change.
PMID 17272296 · PMC1994778 · Nucleic acids research · 2007 · 7 claims · 4 setups
When aberrations are present, noise in BAC, 19k oligo, and 385k oligo array-CGH data is highly non-Gaussian and shows long-range spatial correlations.
-
Full-text index only
Next-generation high-density self-assembling functional protein arrays.
PMID 18469824 · PMC3070491 · Nature methods · 2008 · 8 claims · 7 setups
A next-generation NAPPA method produces high-density protein microarrays displaying over 1500 unique proteins with >90% expression success
-
Full-text index only
Functional copy-number alterations in cancer.
PMID 18784837 · PMC2527508 · PloS one · 2008 · 8 claims · 3 setups
RAE is a comprehensive computational framework that robustly maps chromosomal alterations in tumor samples and statistically assesses their functional importance in cancer.
-
Has reproduction · 92
Analytical code sharing practices in biomedical research.
PMID 38983240 · PMC11232620 · PeerJ. Computer science · 2024 · 8 claims · 4 setups
Nearly half (49.9%) of 453 examined biomedical manuscripts failed to share the analytical code used to generate their results
-
Has reproduction
Using prototyping to choose a bioinformatics workflow management system.
PMID 33630841 · PMC7906312 · PLoS computational biology · 2021 · 8 claims · 6 setups
Prototyping a subset of a project's actual workflow in candidate software offers a low-cost (time and effort) way to make a more informed software selection than relying solely on reviews and recommendations.
-
Full-text index only
Animal models of gene-nutrient interactions.
PMID 19037208 · PMC2703433 · Obesity (Silver Spring, Md.) · 2008 · 8 claims · 5 setups
Mice and rats are well-suited models for human food selection because they share food preferences with humans and are supported by extensive genetic tools (sequenced genome, inbred strains, gene targeting, Cre-lox).
-
Full-text index only
The RYR2-encoded ryanodine receptor/calcium release channel in patients diagnosed previously with either catecholaminergic polymorphic ventricular tachycardia or genotype negative, exercise-induced long QT syndrome: a comprehensive open reading frame mutational analysis.
PMID 19926015 · PMC2880864 · Journal of the American College of Cardiology · 2009 · 8 claims · 6 setups
Comprehensive open-reading-frame RYR2 mutational analysis reveals possible CPVT1 mutations located outside the three canonical hot-spot domains
-
Has reproduction · 67
GEMmaker: process massive RNA-seq datasets on heterogeneous computational infrastructure.
PMID 35501696 · PMC9063052 · BMC bioinformatics · 2022 · 6 claims · 3 setups
GEMmaker, an nf-core compliant Nextflow workflow, can quantify gene expression from small to massive RNA-seq datasets while remaining reproducible via versioned containerized software.