Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Has reproduction · 77
Representing and querying disease networks using graph databases.
PMID 27462371 · PMC4960687 · BioData mining · 2016 · 7 claims · 8 setups
Graph databases are well suited for representing biological information that is highly connected, semi-structured, and unpredictable.
-
Has reproduction · 62
Metatranscriptomics of the human oral microbiome during health and disease.
PMID 24692635 · PMC3977359 · mBio · 2014 · 8 claims · 8 setups
Disease-associated periodontal communities display conserved community-level metabolic gene expression profiles between patients, whereas the metabolic gene expression of individual species is highly variable between patients.
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
Identification of deleterious non-synonymous single nucleotide polymorphisms using sequence-derived information.
PMID 18588693 · PMC2446391 · BMC bioinformatics · 2008 · 8 claims · 5 setups
A decision tree built on 10 selected sequence-derived features classifies SAPs as Disease or Polymorphism with 82.6% accuracy and 0.607 MCC in cross-validation.
-
Full-text index only
Mapping proteins to disease terminologies: from UniProt to MeSH.
PMID 18460185 · PMC2367626 · BMC bioinformatics · 2008 · 8 claims · 7 setups
Developed a three-step procedure (disease name extraction, exact matching, partial/similarity-based matching) to map UniProtKB/Swiss-Prot disease names to MeSH terms
-
Full-text index only
Putting science over supposition in the arena of personalized genomics.
PMID 18665132 · PMC2531214 · Nature genetics · 2008 · 6 claims · 3 setups
There is a rapidly widening gap between gene-disease association discovery and research into the public health/clinical utility of that information.
-
Has reproduction · 78
A single-cell compendium of human cerebrospinal fluid identifies disease-associated immune cell populations.
PMID 39744938 · PMC11684814 · The Journal of clinical investigation · 2025 · 8 claims · 4 setups
Integration of public and newly generated scRNA-seq datasets yields a compendium of 139 subjects (193 samples, 403,973 immune cells) spanning CSF and blood across healthy controls and multiple neurologic diseases.
-
Full-text index only
Towards precise classification of cancers based on robust gene functional expression profiles.
PMID 15774002 · PMC1274255 · BMC bioinformatics · 2005 · 6 claims · 7 setups
Functional expression profiles (FEPs) achieve comparable or better classification performance than conventional gene expression profiles (GEPs) across four public microarray datasets
-
Full-text index only
The FAS gene, brain volume, and disease progression in Alzheimer's disease.
PMID 19766542 · PMC3100774 · Alzheimer's & dementia : the journal of the Alzheimer's Association · 2010 · 8 claims · 4 setups
The minor (T) allele of rs1468063 in FAS is significantly associated with faster AD progression after permutation-based multiple-testing correction
-
Has reproduction · 71
Gene Set Enrichment Analysis Reveals Individual Variability in Host Responses in Tuberculosis Patients.
PMID 34421903 · PMC8375662 · Frontiers in immunology · 2021 · 8 claims · 8 setups
TB patients show substantial individual variability in the intensity of hallmark IFN responses, as well as in complement system, metabolic, and other pathway responses.
-
Full-text index only
Construction of a nasopharyngeal carcinoma 2D/MS repository with Open Source XML database--Xindice.
PMID 16403238 · PMC1351203 · BMC bioinformatics · 2006 · 8 claims · 4 setups
No NPC proteome database existed prior to this work despite availability of other cancer proteome databases
-
Full-text index only
SePaCS--a web-based application for classification of seroreactivity profiles.
PMID 17478503 · PMC1933220 · Nucleic acids research · 2007 · 8 claims · 4 setups
SePaCS is a freely available web-based tool that trains and applies multiple classification methods (4 Naive Bayes variants, SVM with RBF kernel, LDA, DLDA) to seroreactivity profiles and outputs results as a summary table plus a detailed PDF report
-
Full-text index only
Mutation analysis in the long isoform of USH2A in American patients with Usher Syndrome type II.
PMID 19881469 · PMC4511341 · Journal of human genetics · 2009 · 8 claims · 6 setups
Screening all 72 exons of USH2A (long isoform) identifies significantly more mutations than screening only the short-isoform exons 1-21
-
Has reproduction · 94
Hierarchical cell-type identifier accurately distinguishes immune-cell subtypes enabling precise profiling of tissue microenvironment with single-cell RNA-sequencing.
PMID 36681937 · PMC10025442 · Briefings in bioinformatics · 2023 · 8 claims · 8 setups
HiCAT is a hierarchical, marker-based cell-type identifier that uses gene set analysis (GSA) scoring with markers structured in a three-level taxonomy tree (major-type, minor-type, subset)
-
Has reproduction · 100
Identification of a PRDM1-regulated T cell network to regulate atherosclerotic plaque inflammation.
PMID 41039608 · PMC12490039 · Genome medicine · 2025 · 6 claims · 7 setups
A distinct gene co-expression module with a prominent T cell signature is enriched in unstable plaques and distinguishes high-risk from low-risk lesions.
-
Full-text index only
Optimal step length EM algorithm (OSLEM) for the estimation of haplotype frequency and its application in lipoprotein lipase genotyping.
PMID 12529185 · PMC149347 · BMC bioinformatics · 2003 · 5 claims · 4 setups
OSLEM (Optimal Step Length EM), which approximates an optimal step length via a fixed-point search (D_N = D_{N-1} + λ(D_preN - D_{N-1})), runs about twice as fast as standard EM while producing the same haplotype frequency estimates.
-
Full-text index only
COMP mutation screening as an aid for the clinical diagnosis and counselling of patients with a suspected diagnosis of pseudoachondroplasia or multiple epiphyseal dysplasia.
PMID 15756302 · PMC2673054 · European journal of human genetics : EJHG · 2005 · 8 claims · 4 setups
COMP mutations were identified in 78% of families referred with PSACH
-
Full-text index only
A genome-wide siRNA screen reveals diverse cellular processes and pathways that mediate genome stability.
PMID 19647519 · PMC2772893 · Molecular cell · 2009 · 8 claims · 6 setups
A genome-wide siRNA screen in HeLa cells using γH2AX as a readout identifies genes whose knockdown elevates DNA damage/genome instability
-
Full-text index only
The molecular landscape of ASPM mutations in primary microcephaly.
PMID 19028728 · PMC2658750 · Journal of medical genetics · 2009 · 8 claims · 7 setups
ASPM mutations are the most common cause of MCPH
-
Full-text index only
Integrated proteomic analysis of human cancer cells and plasma from tumor bearing mice for ovarian cancer biomarker discovery.
PMID 19936259 · PMC2775948 · PloS one · 2009 · 8 claims · 8 setups
Integrated proteomic analysis of a cancer mouse model and human cancer cell populations provides an effective approach to identify potential circulating protein biomarkers.