Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The role of positive selection in determining the molecular cause of species differences in disease.
PMID 18837980 · PMC2576240 · BMC evolutionary biology · 2008 · 8 claims · 6 setups
Genes predicted to be under positive selection during human evolution are implicated in diseases (epithelial cancers, schizophrenia, autoimmune diseases, Alzheimer's disease) that differ in prevalence and symptomatology between humans and other mammals
-
Has reproduction · 72
Prediction of prognostic signatures in triple-negative breast cancer based on the differential expression analysis via NanoString nCounter immune panel.
PMID 33138797 · PMC7607642 · BMC cancer · 2020 · 8 claims · 8 setups
edgeR-based DEG selection is more appropriate for feature selection than Elastic Net when sample sizes are small.
-
Full-text index only
Analysis of protein sequence and interaction data for candidate disease gene prediction.
PMID 17020920 · PMC1636487 · Nucleic acids research · 2006 · 8 claims · 7 setups
Combining CPS and CMP using known disease genes as input achieves sensitivity 0.52 and specificity 0.97, reducing candidate lists 13-fold
-
Full-text index only
The dystrobrevin-binding protein 1 gene: features and networks.
PMID 18663367 · PMC2859304 · Molecular psychiatry · 2009 · 8 claims · 6 setups
DTNBP1 gene structure, protein-coding sequence, and dysbindin domain are conserved across 13 vertebrate species, while noncoding sequence is diverse.
-
Full-text index only
Systems integration of biodefense omics data for analysis of pathogen-host interactions and identification of potential targets.
PMID 19779614 · PMC2745575 · PloS one · 2009 · 8 claims · 8 setups
A protein-centric data integration approach (Master Protein Directory) enables integration and mining of heterogeneous pathogen-host omics data across multiple research centers
-
Full-text index only
Towards the identification of essential genes using targeted genome sequencing and comparative analysis.
PMID 17052348 · PMC1624830 · BMC genomics · 2006 · 8 claims · 8 setups
Phyletic retention (ortholog presence across organisms) is the single most predictive feature of gene essentiality in both E. coli and S. cerevisiae.
-
Full-text index only
Prioritization of candidate cancer genes--an aid to oncogenomic studies.
PMID 18710882 · PMC2566894 · Nucleic acids research · 2008 · 8 claims · 8 setups
Computational classifiers using combinations of protein conservation, gene structure, protein domains, protein interactions, and regulatory data can distinguish known cancer genes (CD/CR) from unlabelled human genes
-
Has reproduction · 78
Machine learning and free energy clustering reveal PAH protein binding linked to AD risk.
PMID 41953002 · PMC13053772 · iScience · 2026 · 8 claims · 8 setups
PARP1, PTPN1, and ITGA4 are core PAH protein targets identified via PPI network analysis and XGBoost feature selection from AD-associated DEGs
-
Has reproduction · 92
A network-guided protocol to discover susceptibility genes in genome-wide association studies using stability selection.
PMID 36609152 · PMC9850185 · STAR protocols · 2023 · 5 claims · 5 setups
The protocol identifies genes that are both statistically associated with a phenotype and functionally interconnected in a biological network
-
Full-text index only
A surrogate-based approach for post-genomic partner identification.
PMID 11602024 · PMC57814 · BMC biotechnology · 2001 · 8 claims · 5 setups
Peptide surrogates derived from random phage display libraries contain amino acid sequence information that identifies the natural biological partner of the panned target via database searching.
-
Full-text index only
Proceedings of the First International Conference on Phylogenomics. March 15-19, 2006. Quebec, Canada.
PMID 17288567 · PMC1796603 · BMC evolutionary biology · 2007 · 8 claims · 8 setups
Gene tree parsimony applied to EST data with widespread gene duplication can infer an organismal phylogeny in excellent agreement with the expected angiosperm phylogeny.
-
Has reproduction · 59
Refining breast cancer biomarker discovery and drug targeting through an advanced data-driven approach.
PMID 38253993 · PMC10810249 · BMC bioinformatics · 2024 · 8 claims · 8 setups
The BGWO_SA_Ens algorithm (hybrid BGWO + simulated annealing with an ensemble classifier objective function) selects predictive breast cancer biomarker genes with high classification performance
-
Full-text index only
The DAVID Gene Functional Classification Tool: a novel biological module-centric algorithm to functionally analyze large gene lists.
PMID 17784955 · PMC2375021 · Genome biology · 2007 · 8 claims · 6 setups
Gene-gene functional similarity can be measured using kappa statistics applied to a binary gene-annotation-term matrix built from 14 annotation categories.
-
Full-text index only
Exploration of effective biomarkers for venous thrombosis embolism in Behçet's disease based on comprehensive bioinformatics analysis.
PMID 38987624 · PMC11236978 · Scientific reports · 2024 · 6 claims · 8 setups
Four hub genes (E2F1, GATA3, HDAC5, MSH2) serve as diagnostic biomarkers for VTE in BD with high accuracy (AUC 0.816)
-
Full-text index only
Revealing potential interfering genes between abdominal aortic aneurysm and periodontitis through machine learning and bioinformatics analysis.
PMID 40857330 · PMC12380325 · PloS one · 2025 · 6 claims · 8 setups
90 interacting genes were identified at the intersection of AAA DEGs and periodontitis WGCNA module genes
-
Has reproduction · 100
Shared and unique phosphoproteomics responses in skeletal muscle from exercise models and in hyperammonemic myotubes.
PMID 36345342 · PMC9636548 · iScience · 2022 · 8 claims · 7 setups
Comparative phosphoproteomics of hyperammonemic myotubes and exercise-model muscle identifies shared enriched pathways: PKA, calcium signaling, MAPK signaling, and protein homeostasis.
-
Has reproduction
Genome-wide signatures of convergent evolution in echolocating mammals.
PMID 24005325 · PMC3836225 · Nature · 2013 · 8 claims · 8 setups
Genome-wide convergent sequence evolution between echolocating lineages is not rare but widespread and continuously distributed, with signatures consistent with convergence in nearly 200 loci out of 2,326 examined.
-
Has reproduction · 80
Specific signature biomarkers highlight the potential mechanisms of circulating neutrophils in aneurysmal subarachnoid hemorrhage.
PMID 36438795 · PMC9685413 · Frontiers in pharmacology · 2022 · 8 claims · 8 setups
A neutrophil-related co-expression gene module (blue module) is significantly associated with aSAH occurrence
-
Has reproduction · 58
Identification of common genetic characteristics of rheumatoid arthritis and major depressive disorder by bioinformatics analysis and machine learning.
PMID 37415981 · PMC10320004 · Frontiers in immunology · 2023 · 7 claims · 8 setups
EAF1, SDCBP and RNF19B are common genetic characteristics (hub genes) shared by RA and MDD
-
Has reproduction · 83
Integrative transcriptomic and machine learning framework reveals candidate genes and potential mechanisms of aflatoxin B1 exposure in breast cancer.
PMID 41688730 · PMC12982753 · Scientific reports · 2026 · 7 claims · 8 setups
170 unique human AFB1 targets were identified by merging ChEMBL, SwissTargetPrediction, and PharmMapper predictions