Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Single-cell and spatial profiling highlights TB-induced myofibroblasts as drivers of lung pathology.
PMID 41489684 · PMC12767585 · The Journal of experimental medicine · 2026 · 8 claims · 8 setups
MMP1+CXCL5+ fibroblasts and SPP1+ macrophages are linked to TB disease and TB lung granuloma and reveal targetable cellular cross talk underlying TB immunopathology
-
Full-text index only
Integration of text- and data-mining using ontologies successfully selects disease gene candidates.
PMID 15767279 · PMC1065256 · Nucleic acids research · 2005 · 7 claims · 6 setups
Integrating eVOC anatomical ontology-based text-mining of PubMed abstracts with data-mining of gene expression annotation successfully selects and prioritizes candidate disease genes
-
Full-text index only
nsSNPAnalyzer: identifying disease-associated nonsynonymous single nucleotide polymorphisms.
PMID 15980516 · PMC1160133 · Nucleic acids research · 2005 · 6 claims · 4 setups
nsSNPAnalyzer is a web server that predicts whether a query nsSNP is disease-associated or functionally neutral using a Random Forest classifier combining structural and evolutionary information
-
Full-text index only
Searching for interpretable rules for disease mutations: a simulated annealing bump hunting strategy.
PMID 16984653 · PMC1618409 · BMC bioinformatics · 2006 · 8 claims · 6 setups
The proposed feature set outperforms existing published feature sets for predicting effects of amino acid substitutions
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
Analysis of protein sequence and interaction data for candidate disease gene prediction.
PMID 17020920 · PMC1636487 · Nucleic acids research · 2006 · 8 claims · 7 setups
Combining CPS and CMP using known disease genes as input achieves sensitivity 0.52 and specificity 0.97, reducing candidate lists 13-fold
-
Full-text index only
Biomarker discovery in neurodegenerative diseases: a proteomic approach.
PMID 18938247 · PMC2939006 · Neurobiology of disease · 2009 · 6 claims · 8 setups
Proteomic profiling of CSF and plasma can identify candidate protein biomarkers that distinguish AD patients from controls with high sensitivity and specificity
-
Full-text index only
Validation of reverse phase protein array for practical screening of potential biomarkers in serum and plasma: accurate detection of CA19-9 levels in pancreatic cancer.
PMID 18615426 · PMC2992687 · Proteomics · 2008 · 8 claims · 6 setups
RPPA-measured CA19-9 levels correlate strongly with ELISA-measured levels in the same patient samples
-
Full-text index only
Identification of deleterious non-synonymous single nucleotide polymorphisms using sequence-derived information.
PMID 18588693 · PMC2446391 · BMC bioinformatics · 2008 · 8 claims · 5 setups
A decision tree built on 10 selected sequence-derived features classifies SAPs as Disease or Polymorphism with 82.6% accuracy and 0.607 MCC in cross-validation.
-
Has reproduction · 95
MetaMap: an atlas of metatranscriptomic reads in human disease-related RNA-seq data.
PMID 29901703 · PMC6025204 · GigaScience · 2018 · 8 claims · 7 setups
The MetaMap pipeline recapitulates known infection agents in bona fide dual RNA-seq validation studies (Salmonella, HPV, HSV, rhinovirus)
-
Has reproduction · 100
Gene signature discovery and systematic validation across diverse clinical cohorts for TB prognosis and response to treatment.
PMID 37471455 · PMC10393163 · PLoS computational biology · 2023 · 8 claims · 7 setups
A network-based meta-analysis of 27 discovery cohorts identified a common 45-gene signature specific to active TB disease across studies.
-
Has reproduction · 37
RNA Editing Alterations Define Disease Manifestations in the Progression of Experimental Autoimmune Encephalomyelitis (EAE).
PMID 36429012 · PMC9688714 · Cells · 2022 · 7 claims · 7 setups
RNA-editing events mediated by APOBEC and ADAR deaminases are significantly reduced throughout the course of EAE disease progression
-
Has reproduction · 83
Discovery and validation of molecular patterns and immune characteristics in the peripheral blood of ischemic stroke patients.
PMID 38650649 · PMC11034498 · PeerJ · 2024 · 8 claims · 8 setups
188 differentially expressed genes (DEGs) between IS and control blood samples were identified and enriched in immune-related biological pathways
-
Has reproduction · 95
nf-rnaSeqCount: A Nextflow pipeline for obtaining raw read counts from RNA-seq data.
PMID 35574063 · PMC9097006 · South African computer journal = Suid-Afrikaanse rekenaartydskrif · 2021 · 7 claims · 5 setups
nf-rnaSeqCount is a portable, reproducible Nextflow pipeline that maps RNA-seq reads to a reference genome and quantifies gene abundance for differential expression analysis
-
Full-text index only
Diagnostic proteomics: serum proteomic patterns for the detection of early stage cancers.
PMID 15258335 · PMC3851082 · Disease markers · 2003 · 8 claims · 8 setups
Proteomic pattern analysis of serum mass spectra, without identifying the underlying proteins, can distinguish cancer patients from healthy controls with high sensitivity and specificity.
-
Full-text index only
Proteomics as a tool for biomarker discovery.
PMID 18057524 · PMC3851415 · Disease markers · 2007 · 8 claims · 7 setups
A useful clinical biomarker must be easily attainable, have adequate sensitivity, have adequate specificity, and lead to patient benefit through intervention
-
Full-text index only
Identification of candidate disease genes by integrating Gene Ontologies and protein-interaction networks: case study of primary immunodeficiencies.
PMID 19073697 · PMC2632920 · Nucleic acids research · 2009 · 8 claims · 5 setups
Combining high protein-interaction network scores with significant PID-related GO terms identifies novel PID candidate genes
-
Full-text index only
Genome-wide prioritization of disease genes and identification of disease-disease associations from an integrated human functional linkage network.
PMID 19728866 · PMC2768980 · Genome biology · 2009 · 6 claims · 6 setups
Integrating 16 genomic features (32 sub-features) via a naïve Bayes classifier produces a genome-scale FLN of 21,657 human genes and 22,388,609 weighted links that outperforms any individual data source for inferring functional linkages.
-
Full-text index only
Exploration of effective biomarkers for venous thrombosis embolism in Behçet's disease based on comprehensive bioinformatics analysis.
PMID 38987624 · PMC11236978 · Scientific reports · 2024 · 6 claims · 8 setups
Four hub genes (E2F1, GATA3, HDAC5, MSH2) serve as diagnostic biomarkers for VTE in BD with high accuracy (AUC 0.816)
-
Full-text index only
InSite: a computational method for identifying protein-protein interaction binding sites on a proteome-wide scale.
PMID 17868464 · PMC2375030 · Genome biology · 2007 · 8 claims · 8 setups
InSite predicts protein-pair-specific binding motifs ('Motif M on protein A binds to protein B') by integrating heterogeneous PPI and motif-motif interaction evidence within a Bayesian network trained by EM