Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Copy number variants and common disorders: filling the gaps and exploring complexity in genome-wide association studies.
PMID 17953491 · PMC2039766 · PLoS genetics · 2007 · 8 claims · 5 setups
CNVs are not easily tagged by SNPs and often fall in genomic regions poorly covered by whole-genome SNP arrays or not genotyped by HapMap, so current GWASs have largely missed their contribution to complex disorders.
-
Full-text index only
Human disease classification in the postgenomic era: a complex systems approach to human pathobiology.
PMID 17625512 · PMC1948102 · Molecular systems biology · 2007 · 8 claims · 5 setups
Current syndromic disease classification lacks specificity despite historically serving clinicians well
-
Full-text index only
Gene-disease relationship discovery based on model-driven data integration and database view definition.
PMID 19042916 · PMC2639000 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 4 setups
Explicit gene–disease relationships can be formulated as candidate gene definitions (e.g., co-localization, dysregulation, functional similarity) that may include intermediary orthologous or interacting genes
-
Full-text index only
Short tandem repeats in human exons: a target for disease mutations.
PMID 18789129 · PMC2543027 · BMC genomics · 2008 · 8 claims · 6 setups
STRs are present in exons of 92% of known human genes, unlike longer tandem repeats which are rare in exons
-
Full-text index only
Mapping proteins to disease terminologies: from UniProt to MeSH.
PMID 18460185 · PMC2367626 · BMC bioinformatics · 2008 · 8 claims · 7 setups
Developed a three-step procedure (disease name extraction, exact matching, partial/similarity-based matching) to map UniProtKB/Swiss-Prot disease names to MeSH terms
-
Full-text index only
Genome-wide prioritization of disease genes and identification of disease-disease associations from an integrated human functional linkage network.
PMID 19728866 · PMC2768980 · Genome biology · 2009 · 6 claims · 6 setups
Integrating 16 genomic features (32 sub-features) via a naïve Bayes classifier produces a genome-scale FLN of 21,657 human genes and 22,388,609 weighted links that outperforms any individual data source for inferring functional linkages.
-
Full-text index only
A repeat expansion in GOLGA8A is a major risk factor for atypical frontotemporal lobar degeneration with ubiquitin-positive inclusions.
PMID 41820575 · PMC13083237 · Nature genetics · 2026 · 5 claims · 3 setups
A genome-wide association study identifies a major risk locus for aFTLD-U on chromosome 15q14, with lead SNP rs549846383 (P=5.85×10^-21, OR=26.7)
-
Has reproduction
Dissection of multiple sclerosis genetics identifies B and CD4+ T cells as driver cell subsets.
PMID 35672799 · PMC9175345 · Genome biology · 2022 · 8 claims · 7 setups
CD4+ T cells and B cells independently and significantly contribute to MS GWAS heritability enrichment, distinct from a merely shared immune regulatory landscape.
-
Full-text index only
MACSIMS: multiple alignment of complete sequences information management system.
PMID 16792820 · PMC1539025 · BMC bioinformatics · 2006 · 8 claims · 5 setups
MACSIMS is a multiple alignment-based information management system combining knowledge-based database mining with ab initio sequence predictions
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
F-SNP: computationally predicted functional SNPs for disease association studies.
PMID 17986460 · PMC2238878 · Nucleic acids research · 2008 · 6 claims · 8 setups
F-SNP is a database integrating functional effect predictions for SNPs from 16 bioinformatics tools/databases across four categories: splicing, transcription, translation, and post-translation
-
Full-text index only
hORFeome v3.1: a resource of human open reading frames representing over 10,000 human genes.
PMID 17207965 · PMC4647941 · Genomics · 2007 · 8 claims · 7 setups
hORFeome v3.1 is a resource of 12,212 cloned human ORFs representing 10,214 genes, a 51% expansion over hORFeome v1.1
-
Full-text index only
Systems genetics of alcoholism.
PMID 23584748 · PMC3860445 · Alcohol research & health : the journal of the National Institute on Alcohol Abuse and Alcoholism · 2008 · 8 claims · 8 setups
Alcoholism is a multifactorial disease driven by interacting genetic, social, and environmental factors, with genetics accounting for 50-60% of risk.
-
Has reproduction · 26
Integrating transcriptomic datasets across neurological disease identifies unique myeloid subpopulations driving disease-specific signatures.
PMID 36527260 · PMC10952672 · Glia · 2023 · 6 claims · 3 setups
The bulk microglial and monocyte transcriptomic program is highly contingent on the disease environment, challenging the notion of a universal microglial disease signature
-
Full-text index only
Inherited disorder phenotypes: controlled annotation and statistical analysis for knowledge mining from gene lists.
PMID 16351744 · PMC1866390 · BMC bioinformatics · 2005 · 5 claims · 3 setups
OMIM Clinical Synopsis free-text phenotype and location names can be normalized and hierarchically structured into a controlled vocabulary suitable for computational analysis
-
Full-text index only
Given the complexity of the human genome, can 'personalised medicine' or 'individualised drug therapy' ever be achieved?
PMID 19706359 · PMC3525196 · Human genomics · 2009 · 7 claims · 3 setups
The human genome is far too complex, given current understanding, for personalised medicine or individualised drug therapy to be realised in the near term
-
Full-text index only
Size matters: just how big is BIG?: Quantifying realistic sample size requirements for human genome epidemiology.
PMID 18676414 · PMC2639365 · International journal of epidemiology · 2009 · 7 claims · 2 setups
Conventional power calculations for case-control studies disregard analytic complexity (e.g. clinical assessment errors, unmeasured aetiological determinants) and can seriously underestimate true sample size requirements
-
Full-text index only
The role of X-chromosome inactivation in female predisposition to autoimmunity.
PMID 11056674 · PMC17816 · Arthritis research · 2000 · 6 claims · 2 setups
Skewed X-chromosome inactivation in the thymus could lead to inadequate thymic deletion of T cells reactive to X-linked polymorphic self-antigens, predisposing to autoimmunity
-
Full-text index only
The UCSC Genome Browser Database: 2008 update.
PMID 18086701 · PMC2238835 · Nucleic acids research · 2008 · 8 claims · 8 setups
The UCSC Genome Browser Database (GBD) provides integrated sequence and annotation data for a large collection of vertebrate and model organism genomes.