Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Towards precise classification of cancers based on robust gene functional expression profiles.
PMID 15774002 · PMC1274255 · BMC bioinformatics · 2005 · 6 claims · 7 setups
Functional expression profiles (FEPs) achieve comparable or better classification performance than conventional gene expression profiles (GEPs) across four public microarray datasets
-
Has reproduction · 66
Integrative bioinformatics and artificial intelligence analyses of transcriptomics data identified genes associated with major depressive disorders including NRG1.
PMID 37583471 · PMC10423927 · Neurobiology of stress · 2023 · 7 claims · 5 setups
Differentially expressed genes in MDD patients are enriched in immune response, inflammatory response, neurodegeneration, and cerebellar atrophy pathways.
-
Has reproduction · 86
Molecular Classification Models for Triple Negative Breast Cancer Subtype Using Machine Learning.
PMID 34575658 · PMC8472680 · Journal of personalized medicine · 2021 · 6 claims · 4 setups
A training gene set of 719 unique upregulated DEGs (subtype-specific) can be used to build ML models that classify TNBC into BLIA, BLIS, MES, and LAR subtypes.
-
Has reproduction · 60
TRAPID 2.0: a web application for taxonomic and functional analysis of de novo transcriptomes.
PMID 34197621 · PMC8464036 · Nucleic acids research · 2021 · 8 claims · 8 setups
TRAPID 2.0 is a web application performing global characterization of de novo transcriptomes via structural, functional, and taxonomic annotation in an initial processing phase, followed by an exploratory phase of downstream analyses.
-
Has reproduction · 75
Why an integrated view of gene expression studies on hematopoiesis in mouse aging is better than the sum of their parts.
PMID 38627103 · PMC11586588 · FEBS letters · 2024 · 7 claims · 4 setups
Combining differentially expressed (DE) gene lists from multiple publications into a unified 'aging list' (AL) with citation counts, and deriving a shorter high-confidence 'aging signature' (AS, genes cited in >3 publications, ~200 genes), produces a more reliable and consistent picture of hematopoietic aging than any single study.
-
Has reproduction · 77
Metabolic reprogramming and prognostic insights in molecular landscapes driven by glycolysis in ovarian cancer.
PMID 40707588 · PMC12290113 · Scientific reports · 2025 · 7 claims · 8 setups
457 differentially expressed GRGs were identified between OC and normal ovarian tissue, of which 30 were significantly associated with prognosis
-
Full-text index only
Applications for protein sequence-function evolution data: mRNA/protein expression analysis and coding SNP scoring tools.
PMID 16912992 · PMC1538848 · Nucleic acids research · 2006 · 7 claims · 8 setups
PANTHER HMMs built from family/subfamily multiple sequence alignments can classify novel protein sequences into functional groups based on statistically significant HMM match scores
-
Full-text index only
Computational approaches for predicting the biological effect of p53 missense mutations: a comparison of three sequence analysis based methods.
PMID 16522644 · PMC1390679 · Nucleic acids research · 2006 · 7 claims · 6 setups
Align-GVGD predicts loss of transactivation activity with high specificity (~88%) but lower sensitivity (67.9-71.2%) for neutral mutants
-
Full-text index only
Variations in the transcriptome of Alzheimer's disease reveal molecular networks involved in cardiovascular diseases.
PMID 18842138 · PMC2760875 · Genome biology · 2008 · 8 claims · 6 setups
AD-related genes (APOE, A2M, PON2, MAP4) and CVD-associated genes (COMT, CBS, WNK1) congregate in a single co-expression module, linking AD and CVD at the transcriptional level
-
Full-text index only
Annotation and analysis of 10,000 expressed sequence tags from developing mouse eye and adult retina.
PMID 14519200 · PMC328454 · Genome biology · 2003 · 8 claims · 5 setups
Annotation of 8,633 high-quality non-mitochondrial/non-ribosomal ESTs shows 57% represent known genes and 43% are unknown or novel, with M15E having the highest proportion of novel ESTs
-
Has reproduction · 53
Combining evidence of preferential gene-tissue relationships from multiple sources.
PMID 23950964 · PMC3741196 · PloS one · 2013 · 8 claims · 8 setups
A high-level integration approach combining three methods across four human microarray datasets, merged by consensus voting and a rule-based inner/total score, predicts preferentially expressed genes while reducing method- and study-specific bias.
-
Full-text index only
Functional nsSNPs from carcinogenesis-related genes expressed in breast tissue: potential breast cancer risk alleles and their distribution across human populations.
PMID 16595073 · PMC3500178 · Human genomics · 2006 · 7 claims · 5 setups
A bioinformatics strategy cross-referencing carcinogenesis-related gene lists with breast-tissue expression data can identify candidate breast cancer risk nsSNPs.
-
Full-text index only
Up regulation in gene expression of chromatin remodelling factors in cervical intraepithelial neoplasia.
PMID 18248679 · PMC2277413 · BMC genomics · 2008 · 8 claims · 4 setups
Chromatin remodelling-associated genes SMARCC1, NCOR1, MRFAP1 and MORF4L2 are upregulated during progression of cervical intraepithelial neoplasia
-
Has reproduction · 59
Application of Machine Learning in Predicting Hepatic Metastasis or Primary Site in Gastroenteropancreatic Neuroendocrine Tumors.
PMID 37887568 · PMC10605255 · Current oncology (Toronto, Ont.) · 2023 · 8 claims · 7 setups
Multi-gene random forest models classify primary tumor vs. liver metastasis samples with 100% accuracy in training/test cohorts and >90% accuracy in an independent validation cohort
-
Has reproduction · 59
Refining breast cancer biomarker discovery and drug targeting through an advanced data-driven approach.
PMID 38253993 · PMC10810249 · BMC bioinformatics · 2024 · 8 claims · 8 setups
The BGWO_SA_Ens algorithm (hybrid BGWO + simulated annealing with an ensemble classifier objective function) selects predictive breast cancer biomarker genes with high classification performance
-
Has reproduction · 53
Blood RNA signature RISK4LEP predicts leprosy years before clinical onset.
PMID 34090257 · PMC8182229 · EBioMedicine · 2021 · 7 claims · 5 setups
A 4-gene blood RNA signature (RISK4LEP: MT-ND2, REX1BD, TPGS1, UBC) predicts leprosy development 4–61 months before clinical diagnosis
-
Full-text index only
SGCEdb: a flexible database and web interface integrating experimental results and analysis for structural genomics focusing on Caenorhabditis elegans.
PMID 16381914 · PMC1347399 · Nucleic acids research · 2006 · 8 claims · 8 setups
SGCEdb is a flexible, reusable database and web interface for reporting and analyzing structural genomics experiment results, focused on C. elegans
-
Has reproduction · 80
Colorectal Cancer Prediction Based on Weighted Gene Co-Expression Network Analysis and Variational Auto-Encoder.
PMID 32825264 · PMC7563725 · Biomolecules · 2020 · 6 claims · 7 setups
Combining WGCNA-derived hub genes with a VAE-derived 10-dimensional representation as features for an SVM classifier achieves high accuracy (0.9692) and AUC (0.9981) for colorectal cancer prediction.
-
Full-text index only
Structural organization and interactions of transmembrane domains in tetraspanin proteins.
PMID 15985154 · PMC1190194 · BMC structural biology · 2005 · 8 claims · 5 setups
TM1, TM2 and TM3 of human tetraspanins display a distinct heptad repeat motif (abcdefg)n, while TM4 lacks this motif.
-
Full-text index only
CGH-Profiler: data mining based on genomic aberration profiles.
PMID 16042799 · PMC1183191 · BMC bioinformatics · 2005 · 8 claims · 3 setups
CGH-Profiler circumvents ISCN nomenclature by importing CGH data from different vendor systems and converting it into a table format suitable for statistical analysis.