Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Functional nsSNPs from carcinogenesis-related genes expressed in breast tissue: potential breast cancer risk alleles and their distribution across human populations.
PMID 16595073 · PMC3500178 · Human genomics · 2006 · 7 claims · 5 setups
A bioinformatics strategy cross-referencing carcinogenesis-related gene lists with breast-tissue expression data can identify candidate breast cancer risk nsSNPs.
-
Full-text index only
Sequence and structure signatures of cancer mutation hotspots in protein kinases.
PMID 19834613 · PMC2759519 · PloS one · 2009 · 8 claims · 6 setups
Developed CKMD (Composite Kinase Mutation Database), an integrated bioinformatics resource mapping genetic variation in protein kinase genes to sequence, structural, and functional data
-
Full-text index only
CLEAN: CLustering Enrichment ANalysis.
PMID 19640299 · PMC2734555 · BMC bioinformatics · 2009 · 8 claims · 4 setups
The gene-specific CLEAN score improves reproducibility of cluster analysis conclusions across independent datasets compared to the traditional cluster-wide score (cwCLEAN).
-
Has reproduction · 74
Identification and validation of a metabolic-related gene risk model predicting the prognosis of lung, colon, and breast cancers.
PMID 39779736 · PMC11711664 · Scientific reports · 2025 · 8 claims · 8 setups
540 DEGs overlap across breast, colorectal, and lung cancers out of 11,384 DEGs analyzed
-
Full-text index only
Patterns of evolutionary constraints on genes in humans.
PMID 18840274 · PMC2587479 · BMC evolutionary biology · 2008 · 7 claims · 6 setups
BaseDiver, a novel framework integrating GERP score and derived allele frequency (DAF) at nonsynonymous coding SNPs, can classify GO functional categories by patterns of evolutionary constraint
-
Full-text index only
Genetic variation in an individual human exome.
PMID 18704161 · PMC2493042 · PLoS genetics · 2008 · 8 claims · 7 setups
The ~12,500 nonsilent coding variants in the HuRef exome can be reduced ~8-fold to a set of ~1,600 variants most likely to affect protein function.
-
Has reproduction · 75
geneshot: gene-level metagenomics identifies genome islands associated with immunotherapy response.
PMID 33952321 · PMC8097837 · Genome biology · 2021 · 8 claims · 4 setups
geneshot is a gene-level metagenomic bioinformatics tool that clusters de novo assembled protein-coding genes into co-abundant gene groups (CAGs) to reduce dimensionality and generate testable hypotheses from WGS microbiome data
-
Full-text index only
A DNA microarray survey of gene expression in normal human tissues.
PMID 15774023 · PMC1088941 · Genome biology · 2005 · 6 claims · 6 setups
Unsupervised hierarchical clustering of gene expression groups normal tissue samples largely according to anatomic location, cellular composition, or physiologic function.
-
Full-text index only
Reconstruction of human protein interolog network using evolutionary conserved network.
PMID 17493278 · PMC1885812 · BMC bioinformatics · 2007 · 8 claims · 7 setups
A relative conservation score derived from maximal quasi-cliques in protein interaction networks, combined with other interaction features, can score and rank predicted human interologs for confidence.
-
Full-text index only
Mutations associated with HNPCC predisposition -- Update of ICG-HNPCC/INSiGHT mutation database.
PMID 15528792 · PMC3839397 · Disease markers · 2004 · 8 claims · 4 setups
The ICG-HNPCC/INSiGHT mutation database has grown from 126 predisposing mutations (1997) to 448 mutations occurring in 748 families (2003 update)
-
Has reproduction · 87
De Novo Transcriptome Meta-Assembly of the Mixotrophic Freshwater Microalga Euglena gracilis.
PMID 34072576 · PMC8227486 · Genes · 2021 · 6 claims · 8 setups
A consensus transcriptome assembled by combining reads from five independent studies is the most complete E. gracilis transcriptome released to date, outperforming the two previously available transcriptomes (GEFR01 and GDJR01).