Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Gene-disease relationship discovery based on model-driven data integration and database view definition.
PMID 19042916 · PMC2639000 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 4 setups
Explicit gene–disease relationships can be formulated as candidate gene definitions (e.g., co-localization, dysregulation, functional similarity) that may include intermediary orthologous or interacting genes
-
Full-text index only
CapsID: a web-based tool for developing parsimonious sets of CAPS molecular markers for genotyping.
PMID 16686952 · PMC1471797 · BMC genetics · 2006 · 7 claims · 1 setups
CapsID identifies snip-SNPs (SNPs that alter restriction endonuclease recognition sites) within reference sequence alignments and designs PCR primers around them
-
Full-text index only
Ontological Discovery Environment: a system for integrating gene-phenotype associations.
PMID 19733230 · PMC2783409 · Genomics · 2009 · 8 claims · 8 setups
ODE is a web-based system for storing, sharing, retrieving and analyzing phenotype-centered genomic data sets across species and experimental systems
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
Prioritization of candidate cancer genes--an aid to oncogenomic studies.
PMID 18710882 · PMC2566894 · Nucleic acids research · 2008 · 8 claims · 8 setups
Computational classifiers using combinations of protein conservation, gene structure, protein domains, protein interactions, and regulatory data can distinguish known cancer genes (CD/CR) from unlabelled human genes
-
Full-text index only
Prediction of candidate primary immunodeficiency disease genes using a support vector machine learning approach.
PMID 19801557 · PMC2780952 · DNA research : an international journal for rapid publication of reports on genes and genomes · 2009 · 6 claims · 3 setups
An SVM trained on 69 binary features of known PID genes can accurately classify PID vs non-PID genes and predict novel candidate PID genes
-
Full-text index only
Functional analysis of human hematopoietic stem cell gene expression using zebrafish.
PMID 16089502 · PMC1166352 · PLoS biology · 2005 · 8 claims · 8 setups
277 unique transcripts are differentially expressed between Rho lo and Rho hi HSC-enriched/depleted populations, conserved across both umbilical cord blood and bone marrow
-
Full-text index only
A computational screen for type I polyketide synthases in metagenomics shotgun data.
PMID 18953415 · PMC2568958 · PloS one · 2008 · 8 claims · 6 setups
Combining HMM domain searches with maximum-likelihood phylogenetic trees can discriminate true PKS I sequences from evolutionarily related but functionally different enzymes (e.g., FAS I) in metagenomic data.
-
Full-text index only
Identification of candidate prostate cancer genes through comparative expression-profiling of seminal vesicle.
PMID 18500686 · PMC2516917 · The Prostate · 2008 · 8 claims · 5 setups
Identified 32 genes (38 cDNAs) with an expression pattern of highest levels in seminal vesicle, lower in normal prostate, and lowest in prostate cancer
-
Full-text index only
Adaptations to climate in candidate genes for common metabolic disorders.
PMID 18282109 · PMC2242814 · PLoS genetics · 2008 · 8 claims · 7 setups
A network-based bioinformatics approach (Molecular Triangulation) was used to select 82 candidate genes belonging to the core subnetwork of metabolic syndrome phenotypes.
-
Full-text index only
Linking disease-associated genes to regulatory networks via promoter organization.
PMID 15701758 · PMC549397 · Nucleic acids research · 2005 · 8 claims · 7 setups
Pairs of TFBSs conserved both vertically (orthologous genes) and horizontally (co-regulated genes) can serve as seeds to build promoter models representing potential co-regulation networks
-
Has reproduction · 83
Hobbes: optimized gram-based methods for efficient read alignment.
PMID 22199254 · PMC3315303 · Nucleic acids research · 2012 · 8 claims · 4 setups
Hobbes, a gram-based short-read mapper supporting Hamming and edit distance, is faster than all other read-mapping programs tested while maintaining high mapping quality.
-
Full-text index only
Function2Gene: a gene selection tool to increase the power of genetic association studies by utilizing public databases and expert knowledge.
PMID 18631403 · PMC2500032 · BMC bioinformatics · 2008 · 6 claims · 5 setups
Function2Gene is a set of Perl programs that queries public databases (NCBI, GeneCards, Harvester, with Uniprot/Ensembl also supported) using expert-selected keywords to rank genes by prior probability of disease association.
-
Full-text index only
Steps toward broad-spectrum therapeutics: discovering virulence-associated genes present in diverse human pathogens.
PMID 19874620 · PMC2774872 · BMC genomics · 2009 · 8 claims · 8 setups
Phylogenetic profiling of protein clusters across pathogen and non-pathogen genomes can identify candidate generic virulence factors
-
Has reproduction · 58
Identification of common genetic characteristics of rheumatoid arthritis and major depressive disorder by bioinformatics analysis and machine learning.
PMID 37415981 · PMC10320004 · Frontiers in immunology · 2023 · 7 claims · 8 setups
EAF1, SDCBP and RNF19B are common genetic characteristics (hub genes) shared by RA and MDD
-
Full-text index only
Phylogenetic profiling of the Arabidopsis thaliana proteome: what proteins distinguish plants from other organisms?
PMID 15287975 · PMC507878 · Genome biology · 2004 · 8 claims · 6 setups
3,848 Arabidopsis proteins were identified as likely plant-specific based on phylogenetic profiling and EST confirmation in multiple plant species
-
Full-text index only
Ab initio identification of human microRNAs based on structure motifs.
PMID 18088431 · PMC2238772 · BMC bioinformatics · 2007 · 8 claims · 7 setups
MiRPred predicts miRNA precursors ab initio using only predicted secondary structure motifs, ignoring nucleotide sequence
-
Full-text index only
COSMIC (the Catalogue of Somatic Mutations in Cancer): a resource to investigate acquired mutations in human cancer.
PMID 19906727 · PMC2808858 · Nucleic acids research · 2010 · 8 claims · 6 setups
COSMIC is the largest public resource for information on somatically acquired mutations in human cancer, freely available without restriction
-
Full-text index only
Getting positive about selection.
PMID 12914654 · PMC193638 · Genome biology · 2003 · 8 claims · 4 setups
Purifying selection is the predominant form of molecular evolution, preserving fitness by eliminating deleterious mutations, while positive selection is rare but critical for adaptation.