Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 87
A target enrichment method for gathering phylogenetic information from hundreds of loci: An example from the Compositae.
PMID 25202605 · PMC4103609 · Applications in plant sciences · 2014 · 8 claims · 8 setups
A custom sequence capture probe set (9678 baits targeting 1061 orthologous genes) was designed to enrich COS loci across the Compositae.
-
Full-text index only
Automatic discovery of cross-family sequence features associated with protein function.
PMID 16409628 · PMC1395344 · BMC bioinformatics · 2006 · 8 claims · 6 setups
A self-supervised data mining approach can find relationships between sequence features and functional annotations without preconceived functional categories.
-
Has reproduction · 95
Increased prevalence of hybrid epithelial/mesenchymal state and enhanced phenotypic heterogeneity in basal breast cancer.
PMID 38974967 · PMC11225361 · iScience · 2024 · 7 claims · 7 setups
Luminal breast cancer gene expression signature is closely/positively associated with an epithelial signature
-
Full-text index only
Inferring combinatorial regulation of transcription in silico.
PMID 15647509 · PMC546154 · Nucleic acids research · 2005 · 8 claims · 5 setups
Combining Cluster-Buster (TFBS cluster prediction) with GOSSIP (rigorous GO enrichment statistics with multiple-testing/FDR correction) predicts biological functions controlled by combinatorial transcription factor action, without prior knowledge of factor targets
-
Full-text index only
Assessment of algorithms for high throughput detection of genomic copy number variation in oligonucleotide microarray data.
PMID 17910767 · PMC2148068 · BMC bioinformatics · 2007 · 8 claims · 4 setups
Different CNV analysis software packages produce highly variable numbers and types of candidate CNVs from the same data
-
Full-text index only
Detection of venous thromboembolism by proteomic serum biomarkers.
PMID 17579716 · PMC1891085 · PloS one · 2007 · 5 claims · 8 setups
A neural network-based classifier built from direct MALDI-TOF MS serum protein expression profiles can diagnose VTE with sensitivity/specificity that exceeds D-dimer assays
-
Has reproduction · 74
An open RNA-Seq data analysis pipeline tutorial with an example of reprocessing data from a recent Zika virus study.
PMID 27583132 · PMC4972086 · F1000Research · 2016 · 6 claims · 6 setups
An open-source, reproducible RNA-seq pipeline delivered as an IPython notebook and Docker image can process raw RNA-seq data into interactive PCA/HC plots, enrichment results, and small-molecule predictions with minimal setup overhead
-
Has reproduction · 69
COVID-19 vaccination atlas using an integrative systems vaccinology approach.
PMID 40456760 · PMC12130191 · NPJ vaccines · 2025 · 8 claims · 6 setups
mRNA vaccines induce transient but strong immune responses after booster doses
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
Proteomic analysis of stage I primary lung adenocarcinoma aimed at individualisation of postoperative therapy.
PMID 18212748 · PMC2243141 · British journal of cancer · 2008 · 5 claims · 6 setups
LC-MS/MS proteomic analysis of stage I lung adenocarcinoma specimens identified myosin IIA and vimentin as candidate biomarker proteins with signal intensities that differed significantly among patient outcome groups
-
Full-text index only
Validation of pooled genotyping on the Affymetrix 500 k and SNP6.0 genotyping platforms using the polynomial-based probe-specific correction.
PMID 20003400 · PMC2806376 · BMC genetics · 2009 · 7 claims · 4 setups
Pooled genotyping on the Affymetrix 500k platform using PPC yields highly accurate allele frequency estimates (correlation 0.988) comparable to or better than the 10k/100k platforms.
-
Full-text index only
Functional features of gene expression profiles differentiating gastrointestinal stromal tumours according to KIT mutations and expression.
PMID 19943934 · PMC2794290 · BMC cancer · 2009 · 8 claims · 5 setups
Hundreds of genes differentiate GISTs according to KIT versus PDGFRA mutation and expression status, despite no discriminative profile for clinical/pathological parameters.
-
Has reproduction · 85
A mechanistic model captures the emergence and implications of non-genetic heterogeneity and reversible drug resistance in ER+ breast cancer cells.
PMID 34316714 · PMC8271219 · NAR cancer · 2021 · 7 claims · 8 setups
EMT and tamoxifen-resistance (TamR) regulatory axes can drive one another, enabling non-genetic heterogeneity via six co-existing phenotypes (ES, ER, HS, HR, MS, MR)
-
Full-text index only
Global distribution of rubella virus genotypes.
PMID 14720390 · PMC3034328 · Emerging infectious diseases · 2003 · 8 claims · 6 setups
Phylogenetic analysis of 103 E1 gene sequences from 17 countries confirms at least two rubella virus genotypes, RGI and RGII
-
Full-text index only
HapMap-based study of the 17q21 ERBB2 amplicon in susceptibility to breast cancer.
PMID 17117180 · PMC2360759 · British journal of cancer · 2006 · 6 claims · 5 setups
Common genetic variation (tSNPs and haplotypes) across the 400-kb 17q21 ERBB2 amplicon is not associated with breast cancer risk in British women.
-
Has reproduction · 42
The electrostatic profile of consecutive Cβ atoms applied to protein structure quality assessment.
PMID 25506420 · PMC4257144 · F1000Research · 2013 · 8 claims · 8 setups
The EPD between Cβ atoms of consecutive residues provides unique signatures of amino acid pair types and can discriminate native from decoy protein structures.
-
Full-text index only
Protein ranking by semi-supervised network propagation.
PMID 16723003 · PMC1810311 · BMC bioinformatics · 2006 · 8 claims · 5 setups
RankProp, a diffusion-based network propagation algorithm on a PSI-BLAST-derived protein similarity network, significantly outperforms local search methods (BLAST/PSI-BLAST) at detecting remote homologs.
-
Full-text index only
Combinatorial Mismatch Scan (CMS) for loci associated with dementia in the Amish.
PMID 16515697 · PMC1448207 · BMC medical genetics · 2006 · 8 claims · 7 setups
CMS compares IBS allele/genotype sharing between distantly related (beyond grandparental) affected and unaffected individuals from founder populations to detect disease loci while reducing confounding from population stratification and genetic heterogeneity.
-
Full-text index only
Adaptive history of single copy genes highly expressed in the term human placenta.
PMID 18848617 · PMC2759754 · Genomics · 2009 · 8 claims · 6 setups
222 single copy eutherian genes are highly expressed (>=3x median) in the term human placenta