Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Similarities and differences in genome-wide expression data of six organisms.
PMID 14737187 · PMC300882 · PLoS biology · 2004 · 8 claims · 8 setups
Coexpression of functionally related genes is frequently conserved across evolutionarily distant organisms
-
Has reproduction · 71
Parsimonious Gene Correlation Network Analysis (PGCNA): a tool to define modular gene co-expression for refined molecular stratification in cancer.
PMID 30993001 · PMC6459838 · NPJ systems biology and applications · 2019 · 8 claims · 7 setups
Retaining only the top ~3 most correlated edges per gene (EPG3) combined with FastUnfold clustering (termed PGCNA) produces gene co-expression modules with significantly better separation and enrichment of known biology than using all edges or other clustering methods.
-
Full-text index only
A procedure for the detection of linkage with high density SNP arrays in a large pedigree with colorectal cancer.
PMID 17222328 · PMC1784097 · BMC cancer · 2007 · 7 claims · 8 setups
A workflow combining Alohomora, Mega2, MENDEL, SNPLINK and SimWalk2 enables linkage analysis with high-density SNP arrays in large pedigrees (>35-40 bits) that exceed the capacity of single existing programs
-
Full-text index only
Extraction of human kinase mutations from literature, databases and genotyping studies.
PMID 19758464 · PMC2745582 · BMC bioinformatics · 2009 · 7 claims · 6 setups
A literature mining pipeline combining MutationFinder, false-positive filtering, and SVM-based classification can extract and disambiguate single-point mutation mentions from abstracts and full text
-
Full-text index only
How to find soluble proteins: a comprehensive analysis of alpha/beta hydrolases for recombinant expression in E. coli.
PMID 15804363 · PMC1079826 · BMC genomics · 2005 · 7 claims · 7 setups
Predicted solubility in E. coli (via CV-CV') depends on hydrolase size, phylogenetic origin, homologous family, and superfamily
-
Full-text index only
Predicting survival outcomes using subsets of significant genes in prognostic marker studies with microarrays.
PMID 16549007 · PMC1544357 · BMC bioinformatics · 2006 · 7 claims · 2 setups
A methodology combining Cox proportional hazards models with a compound covariate, cross-validated log partial likelihood (ACVL) for predictive accuracy, and permutation-based significance testing can identify an optimal subset of significant genes for survival prediction
-
Full-text index only
Microarray analysis: genome-scale hypothesis scanning.
PMID 14551912 · PMC212694 · PLoS biology · 2003 · 8 claims · 5 setups
Microarrays can be used to both test and generate hypotheses, not merely to fish for candidate genes.
-
Full-text index only
Molecular tumor profiling: translating genomic insights into clinical advances.
PMID 15287965 · PMC507868 · Genome biology · 2004 · 8 claims · 8 setups
Gene-expression profiling can distinguish BRCA1- and BRCA2-linked breast tumors from sporadic breast tumors with similar hormone-receptor status
-
Full-text index only
Relative impact of nucleotide and copy number variation on gene expression phenotypes.
PMID 17289997 · PMC2665772 · Science (New York, N.Y.) · 2007 · 8 claims · 5 setups
SNPs and CNVs capture largely non-overlapping signals of genetic variation affecting gene expression
-
Full-text index only
IDEAL-Q, an automated tool for label-free quantitation analysis using an efficient peptide alignment approach and spectral data validation.
PMID 19752006 · PMC2808259 · Molecular & cellular proteomics : MCP · 2010 · 6 claims · 5 setups
IDEAL-Q predicts the elution time of peptides unidentified in a given LC-MS/MS run (but identified in others) using a computation-efficient linear regression plus fragmental refining function, avoiding costly whole-dataset pattern recognition
-
Has reproduction · 85
ScLRTC: imputation for single-cell RNA-seq data via low-rank tensor completion.
PMID 34844559 · PMC8628418 · BMC genomics · 2021 · 8 claims · 8 setups
scLRTC imputes dropout entries closest to the original expression values on simulated datasets, outperforming other state-of-the-art methods by SSE and PCC.
-
Full-text index only
Expression genomics in breast cancer research: microarrays at the crossroads of biology and medicine.
PMID 17397520 · PMC1868923 · Breast cancer research : BCR · 2007 · 8 claims · 8 setups
Genome-wide expression microarray studies reveal transcriptional networks/signatures that explain breast cancer biological and clinical heterogeneity