Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
The UCSC Genome Browser Database: 2008 update.
PMID 18086701 · PMC2238835 · Nucleic acids research · 2008 · 8 claims · 8 setups
The UCSC Genome Browser Database (GBD) provides integrated sequence and annotation data for a large collection of vertebrate and model organism genomes.
-
Full-text index only
Prognostic significance of p53 and ras gene abnormalities in lung adenocarcinoma patients with stage I disease after curative resection.
PMID 7852188 · PMC5919395 · Japanese journal of cancer research : Gann · 1994 · 6 claims · 5 setups
Presence of p53 gene abnormalities is associated with shorter disease-free survival in stage I lung adenocarcinoma
-
Full-text index only
Analysis of protein sequence and interaction data for candidate disease gene prediction.
PMID 17020920 · PMC1636487 · Nucleic acids research · 2006 · 8 claims · 7 setups
Combining CPS and CMP using known disease genes as input achieves sensitivity 0.52 and specificity 0.97, reducing candidate lists 13-fold
-
Full-text index only
Integrated weighted gene co-expression network analysis with an application to chronic fatigue syndrome.
PMID 18986552 · PMC2625353 · BMC systems biology · 2008 · 8 claims · 6 setups
Integrated WGCNA (IWGCNA), which adds genetic marker-based causality testing to standard WGCNA, can identify a disease-related module and its causal drivers
-
Full-text index only
Variations in the transcriptome of Alzheimer's disease reveal molecular networks involved in cardiovascular diseases.
PMID 18842138 · PMC2760875 · Genome biology · 2008 · 8 claims · 6 setups
AD-related genes (APOE, A2M, PON2, MAP4) and CVD-associated genes (COMT, CBS, WNK1) congregate in a single co-expression module, linking AD and CVD at the transcriptional level
-
Full-text index only
Disease-aging network reveals significant roles of aging genes in connecting genetic diseases.
PMID 19779549 · PMC2739292 · PLoS computational biology · 2009 · 8 claims · 8 setups
Human disease genes are much closer to aging genes in the PPI network than expected by chance
-
Full-text index only
L2L: a simple tool for discovering the hidden significance in microarray expression data.
PMID 16168088 · PMC1242216 · Genome biology · 2005 · 8 claims · 4 setups
L2L systematically compares a user's differentially expressed gene list against a database of published differentially expressed gene lists to find statistically significant overlaps and generate hypotheses about shared mechanisms
-
Full-text index only
What can genome-wide association studies tell us about the genetics of common disease?
PMID 18454206 · PMC2323402 · PLoS genetics · 2008 · 8 claims · 4 setups
Apparent patterns of common, low-effect disease-associated alleles largely reflect statistical power of studies rather than the true underlying distribution of disease variants
-
Full-text index only
Phenotypic categorization of genetic skin diseases reveals new relations between phenotypes, genes and pathways.
PMID 19744994 · PMC2773259 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 5 setups
560 genetic skin diseases can be decomposed into 71 elementary phenotypic features (42 dermatologic, 29 systemic) that combine to represent each disease as a point in a multidimensional phenotype space
-
Full-text index only
An analysis of human microRNA and disease associations.
PMID 18923704 · PMC2559869 · PloS one · 2008 · 8 claims · 8 setups
MicroRNAs tend to show similar dysfunctional evidence (both up- or both down-regulated) for diseases within the same disease cluster, and different dysfunctional evidence between different disease clusters.
-
Full-text index only
Localized-statistical quantification of human serum proteome associated with type 2 diabetes.
PMID 18795103 · PMC2529402 · PloS one · 2008 · 8 claims · 5 setups
Developed LSPAD (localized statistics of protein abundance distribution) to calculate statistical significance of protein-abundance bias between two serum cohorts using a local Fisher's exact test window
-
Full-text index only
A non-parametric meta-analysis approach for combining independent microarray datasets: application using two microarray datasets pertaining to chronic allograft nephropathy.
PMID 18302764 · PMC2276496 · BMC genomics · 2008 · 8 claims · 6 setups
A novel non-parametric meta-analysis approach for combining independent microarray datasets is presented, requiring no distributional assumptions and being logically intuitive.
-
Full-text index only
Identification of gene interactions associated with disease from gene expression data using synergy networks.
PMID 18234101 · PMC2258206 · BMC systems biology · 2008 · 8 claims · 4 setups
Synergy of a gene pair with respect to disease, defined as I(G1,G2;C) - [I(G1;C)+I(G2;C)], identifies gene pairs that interact cooperatively with respect to a phenotype rather than independently.
-
Full-text index only
Testing groups of genomic locations for enrichment in disease loci using linkage scan data: a method for hypothesis testing.
PMID 16848972 · PMC3525155 · Human genomics · 2006 · 8 claims · 2 setups
A method testing enrichment of a group of genomic locations for disease loci by comparing the average NPL score of the group to a null distribution from randomly drawn groups of equal size
-
Full-text index only
A simple and efficient algorithm for genome-wide homozygosity analysis in disease.
PMID 19756043 · PMC2758715 · Molecular systems biology · 2009 · 8 claims · 4 setups
A genome-wide AH analysis (GAHA) algorithm can identify disease-associated loci by comparing frequencies of homozygous segments between cases and controls using a z-statistic proportion test
-
Full-text index only
Screening large-scale association study data: exploiting interactions using random forests.
PMID 15588316 · PMC545646 · BMC genetics · 2004 · 7 claims · 3 setups
Random forest importance measure significantly outperforms the Fisher Exact test as a screening tool when risk SNPs interact.
-
Full-text index only
Size matters: just how big is BIG?: Quantifying realistic sample size requirements for human genome epidemiology.
PMID 18676414 · PMC2639365 · International journal of epidemiology · 2009 · 7 claims · 2 setups
Conventional power calculations for case-control studies disregard analytic complexity (e.g. clinical assessment errors, unmeasured aetiological determinants) and can seriously underestimate true sample size requirements
-
Full-text index only
An online database for brain disease research.
PMID 16594998 · PMC1489945 · BMC genomics · 2006 · 7 claims · 5 setups
SMRIDB is a comprehensive web-based database integrating gene expression data and clinical metadata to aid understanding of the genetic effects of brain disease (bipolar disorder, schizophrenia, depression)
-
Full-text index only
Molecular characterization of Campylobacter jejuni clones: a basis for epidemiologic investigation.
PMID 12194772 · PMC2732546 · Emerging infectious diseases · 2002 · 8 claims · 5 setups
Clonal complex, as defined by MLST, is an epidemiologically relevant unit for long- and short-term investigation of C. jejuni epidemiology.
-
Full-text index only
Satellog: a database for the identification and prioritization of satellite repeats in disease association studies.
PMID 15949044 · PMC1181805 · BMC bioinformatics · 2005 · 7 claims · 6 setups
Satellog is a database cataloging all pure 1-16 unit satellite repeats in the human genome with supplementary polymorphism, gene-location, and expression data for prioritizing repeats in disease-association studies.