Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Towards the identification of essential genes using targeted genome sequencing and comparative analysis.
PMID 17052348 · PMC1624830 · BMC genomics · 2006 · 8 claims · 8 setups
Phyletic retention (ortholog presence across organisms) is the single most predictive feature of gene essentiality in both E. coli and S. cerevisiae.
-
Has reproduction · 76
Bayesian prediction of microbial oxygen requirement.
PMID 26913185 · PMC4743139 · F1000Research · 2013 · 7 claims · 8 setups
A naive Bayesian classifier based on presence/absence of class-associated Pfam-A domains can distinguish three oxygen requirement classes (aerobe, anaerobe, facultative anaerobe) from genome sequence, unlike prior studies that only made pairwise distinctions.
-
Full-text index only
Analysis of nucleotide diversity of NAT2 coding region reveals homogeneity across Native American populations and high intra-population diversity.
PMID 16847467 · PMC3099416 · The pharmacogenomics journal · 2007 · 8 claims · 6 setups
NAT2 variants are homogeneously distributed across native populations of the American continent
-
Full-text index only
Identification of serum biomarkers for colon cancer by proteomic analysis.
PMID 16755300 · PMC2361335 · British journal of cancer · 2006 · 8 claims · 8 setups
Complement C3a des-arg, α1-antitrypsin and transferrin were identified as serum proteins with diagnostic potential for CRC.
-
Full-text index only
An evaluation of the performance of tag SNPs derived from HapMap in a Caucasian population.
PMID 16532062 · PMC1391920 · PLoS genetics · 2006 · 8 claims · 5 setups
CEU HapMap-derived tSNPs capture most of the genetic variation observed in the Estonian (EGP) population sample
-
Full-text index only
Application of machine learning in SNP discovery.
PMID 16398931 · PMC1955739 · BMC bioinformatics · 2006 · 8 claims · 6 setups
PolyBayes produces high false-positive SNP predictions even with stringent parameters
-
Full-text index only
Systems biology approach for mapping the response of human urothelial cells to infection by Enterococcus faecalis.
PMID 18047719 · PMC2099488 · BMC bioinformatics · 2007 · 8 claims · 5 setups
Deconvoluting gene expression variance into technical (Gaussian, ~6.5% relative SD) and biological components identifies hypervariable (HV) genes that reflect true biological response to infection without requiring replicates
-
Full-text index only
Mutation screening and haplotype analysis of the rhodopsin gene locus in Japanese patients with retinitis pigmentosa.
PMID 17653048 · PMC2776539 · Molecular vision · 2007 · 8 claims · 4 setups
No RP patient among 68 Japanese subjects carried a RHO mutation causing an amino acid substitution
-
Has reproduction · 30
Minimal metabolic pathway structure is consistent with associated biomolecular interactions.
PMID 24987116 · PMC4299494 · Molecular systems biology · 2014 · 8 claims · 8 setups
MinSpan, a mixed-integer linear optimization algorithm, computes the shortest, linearly independent pathways (sparsest basis of the null space of the stoichiometric matrix S) for genome-scale metabolic networks, which convex approaches (extreme pathways, elementary flux modes) cannot do at genome scale.
-
Full-text index only
Integrated proteomic and transcriptomic profiling of mouse lung development and Nmyc target genes.
PMID 17486137 · PMC2673710 · Molecular systems biology · 2007 · 8 claims · 7 setups
Global MudPIT-based proteomic profiling across six mouse lung developmental time points (E13.5–P56) identifies thousands of proteins and captures developmental/cell-biological expression patterns.
-
Full-text index only
FatiGO +: a functional profiling tool for genomic data. Integration of functional annotation, regulatory motifs and interaction data with microarray experiments.
PMID 17478504 · PMC1933151 · Nucleic acids research · 2007 · 8 claims · 8 setups
FatiGO+ is a web-based tool for functional profiling of genome-scale experiments that integrates functional annotation, regulatory motifs and interaction data
-
Full-text index only
Peptide bioinformatics: peptide classification using peptide machines.
PMID 19065810 · PMC7122642 · Methods in molecular biology (Clifton, N.J.) · 2008 · 8 claims · 4 setups
The bio-basis function, which converts peptides into numerical vectors using nongapped pairwise homology alignment scores against indicator peptides, can statistically quantify peptide similarity for classification.
-
Full-text index only
Integrative microRNA and proteomic approaches identify novel osteoarthritis genes and their collaborative metabolic and inflammatory networks.
PMID 19011694 · PMC2582945 · PloS one · 2008 · 8 claims · 8 setups
A 16-microRNA signature (9 up, 7 down) distinguishes osteoarthritic from normal cartilage.
-
Full-text index only
Information-based methods for predicting gene function from systematic gene knock-downs.
PMID 18959798 · PMC2596148 · BMC bioinformatics · 2008 · 8 claims · 4 setups
Information-based metrics, which incorporate a phenotype's genomic frequency, outperform non-information-based metrics for detecting gene-gene functional similarity from phenotypic knock-down profiles.
-
Full-text index only
Genetical genomics: spotlight on QTL hotspots.
PMID 18949031 · PMC2563687 · PLoS genetics · 2008 · 8 claims · 4 setups
Distant eQTL hotspots are rare and difficult to reliably verify across published genetical genomics studies
-
Full-text index only
Integrated analysis of genetic and proteomic data identifies biomarkers associated with adverse events following smallpox vaccination.
PMID 18923431 · PMC2692715 · Genes and immunity · 2009 · 7 claims · 6 setups
A two-stage strategy (Random Forest filtering followed by decision tree modeling) can integrate categorical genetic and continuous proteomic data to identify biomarkers of AE risk
-
Full-text index only
Variations in the transcriptome of Alzheimer's disease reveal molecular networks involved in cardiovascular diseases.
PMID 18842138 · PMC2760875 · Genome biology · 2008 · 8 claims · 6 setups
AD-related genes (APOE, A2M, PON2, MAP4) and CVD-associated genes (COMT, CBS, WNK1) congregate in a single co-expression module, linking AD and CVD at the transcriptional level
-
Full-text index only
The role of positive selection in determining the molecular cause of species differences in disease.
PMID 18837980 · PMC2576240 · BMC evolutionary biology · 2008 · 8 claims · 6 setups
Genes predicted to be under positive selection during human evolution are implicated in diseases (epithelial cancers, schizophrenia, autoimmune diseases, Alzheimer's disease) that differ in prevalence and symptomatology between humans and other mammals
-
Full-text index only
Comparative Toxicogenomics Database: a knowledgebase and discovery tool for chemical-gene-disease networks.
PMID 18782832 · PMC2686584 · Nucleic acids research · 2009 · 8 claims · 5 setups
CTD is a manually curated knowledgebase that integrates chemical-gene interactions, chemical-disease relationships, and gene-disease relationships into a chemical-gene-disease triad
-
Full-text index only
Prioritization of candidate cancer genes--an aid to oncogenomic studies.
PMID 18710882 · PMC2566894 · Nucleic acids research · 2008 · 8 claims · 8 setups
Computational classifiers using combinations of protein conservation, gene structure, protein domains, protein interactions, and regulatory data can distinguish known cancer genes (CD/CR) from unlabelled human genes