Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Speeding disease gene discovery by sequence based candidate prioritization.
PMID 15766383 · PMC1274252 · BMC bioinformatics · 2005 · 7 claims · 8 setups
Disease genes (OMIM) differ significantly from non-disease genes in sequence-based features including gene/cDNA/protein size, exon number, homolog conservation, secretion signal, 3' UTR length, CpG islands, and distance to nearest gene.
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
DG-CST (Disease Gene Conserved Sequence Tags), a database of human-mouse conserved elements associated to disease genes.
PMID 15608249 · PMC539965 · Nucleic acids research · 2005 · 5 claims · 8 setups
Comparative human-mouse genome analysis identifies conserved sequence tags (CSTs, >=70% identity over >=100bp) that frequently correspond to non-coding elements with putative regulatory or structural roles
-
Full-text index only
Differences in the evolutionary history of disease genes affected by dominant or recessive mutations.
PMID 16817963 · PMC1534034 · BMC genomics · 2006 · 8 claims · 8 setups
Dominant disease genes are more conserved at the protein level (mouse orthologues) than recessive disease genes.
-
Full-text index only
Genome bioinformatic analysis of nonsynonymous SNPs.
PMID 17708757 · PMC1978506 · BMC bioinformatics · 2007 · 8 claims · 8 setups
Structure- and sequence-based prediction tools can generally distinguish disease-causing mutations from neutral ones
-
Full-text index only
The role of positive selection in determining the molecular cause of species differences in disease.
PMID 18837980 · PMC2576240 · BMC evolutionary biology · 2008 · 8 claims · 6 setups
Genes predicted to be under positive selection during human evolution are implicated in diseases (epithelial cancers, schizophrenia, autoimmune diseases, Alzheimer's disease) that differ in prevalence and symptomatology between humans and other mammals
-
Full-text index only
Correlation of microsynteny conservation and disease gene distribution in mammalian genomes.
PMID 19909546 · PMC2779822 · BMC genomics · 2009 · 7 claims · 8 setups
Density of mouse orthologs of human disease genes correlates with regions of conserved microsynteny in the mouse genome
-
Full-text index only
Application of OMICS technologies in occupational and environmental health research; current status and projections.
PMID 19933307 · PMC2910417 · Occupational and environmental medicine · 2010 · 8 claims · 6 setups
Five OMICS technologies are well established: genotyping, transcriptomics, epigenomics, proteomics, and metabolomics
-
Full-text index only
A chromosomally integrated bacteriophage in invasive meningococci.
PMID 15967821 · PMC2212043 · The Journal of experimental medicine · 2005 · 8 claims · 7 setups
An 8-kb genetic island (NMA1792-NMA1799) is specifically associated with hyperinvasive meningococcal clonal complexes
-
Full-text index only
The use of whole genome amplification to study chromosomal changes in prostate cancer: insights into genome-wide signature of preneoplasia associated with cancer progression.
PMID 16573809 · PMC1450280 · BMC genomics · 2006 · 7 claims · 8 setups
MDA-amplified DNA does not introduce major distortion of copy number imbalance assignments compared to unamplified DNA in control CGH experiments.
-
Full-text index only
A genome-wide screen for copy number alterations in Aicardi syndrome.
PMID 19760649 · PMC3640635 · American journal of medical genetics. Part A · 2009 · 7 claims · 4 setups
Aicardi syndrome is thought to result from heterozygous defects in an essential X-linked gene, or from a sex-limited autosomal gene defect, due to its occurrence almost exclusively in females and in 47,XXY males.
-
Has reproduction · 81
SEMdag: Fast learning of Directed Acyclic Graphs via node or layer ordering.
PMID 39775401 · PMC11709272 · PloS one · 2025 · 8 claims · 5 setups
SEMdag() is a two-step order-based algorithm for fast learning of high-dimensional linear SEMs, using knowledge-based (KB) or data-driven bottom-up (BU) node/layer ordering followed by penalized (L1) DAG estimation
-
Full-text index only
Functional coverage of the human genome by existing structures, structural genomics targets, and homology models.
PMID 16118666 · PMC1188274 · PLoS computational biology · 2005 · 8 claims · 5 setups
Existing PDB structures provide single-domain coverage for 37% of functional classes in the human genome and complete (whole-protein) structure coverage for 25%.
-
Full-text index only
Disease-specific proteins from rheumatoid arthritis patients.
PMID 16778393 · PMC2729955 · Journal of Korean medical science · 2006 · 8 claims · 6 setups
Fibronectin, semaphorin 7A precursor, growth factor receptor-bound protein 7 (GRB7), and immunoglobulin µ chain specifically associate with antibodies purified from RA synovial fluid
-
Full-text index only
Copy number variants and common disorders: filling the gaps and exploring complexity in genome-wide association studies.
PMID 17953491 · PMC2039766 · PLoS genetics · 2007 · 8 claims · 5 setups
CNVs are not easily tagged by SNPs and often fall in genomic regions poorly covered by whole-genome SNP arrays or not genotyped by HapMap, so current GWASs have largely missed their contribution to complex disorders.
-
Full-text index only
Gene Prospector: an evidence gateway for evaluating potential susceptibility genes and interacting risk factors for human diseases.
PMID 19063745 · PMC2613935 · BMC bioinformatics · 2008 · 8 claims · 5 setups
Gene Prospector is a Web-based application that selects and prioritizes potential disease-related genes using a curated, updated literature database of genetic association studies
-
Full-text index only
Exhaustive prediction of disease susceptibility to coding base changes in the human genome.
PMID 18793467 · PMC2537574 · BMC bioinformatics · 2008 · 8 claims · 7 setups
Inter-species conservation is the strongest single predictor of disease-associated coding mutations among the factors tested.
-
Full-text index only
Predicting phenotype and emerging strains among Chlamydia trachomatis infections.
PMID 19788805 · PMC2819883 · Emerging infectious diseases · 2009 · 8 claims · 7 setups
A 7-locus MLST scheme selected from conserved housekeeping genes shared across 4 Chlamydiaceae species (7 genomes) can genotype diverse C. trachomatis reference and clinical isolates.
-
Full-text index only
Variation in conserved non-coding sequences on chromosome 5q and susceptibility to asthma and atopy.
PMID 16336695 · PMC1325232 · Respiratory research · 2005 · 6 claims · 8 setups
There is overall little sequence variation in the conserved non-coding elements (CNEs) on 5q31, including none detected in CNE-B/CNS-1
-
Full-text index only
Bias of selection on human copy-number variants.
PMID 16482228 · PMC1366494 · PLoS genetics · 2006 · 8 claims · 8 setups
Human CNVs are significantly overrepresented near telomeres and centromeres and enriched in simple tandem repeats relative to the genome as a whole