Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Integration of text- and data-mining using ontologies successfully selects disease gene candidates.
PMID 15767279 · PMC1065256 · Nucleic acids research · 2005 · 7 claims · 6 setups
Integrating eVOC anatomical ontology-based text-mining of PubMed abstracts with data-mining of gene expression annotation successfully selects and prioritizes candidate disease genes
-
Full-text index only
Web-based resources for comparative genomics.
PMID 16197736 · PMC3525128 · Human genomics · 2005 · 8 claims · 8 setups
Comparative genomics is an indispensable tool for identifying functional genome elements and exploring evolutionary genome dynamics
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Full-text index only
A re-annotation pipeline for Illumina BeadArrays: improving the interpretation of gene expression data.
PMID 19923232 · PMC2817484 · Nucleic acids research · 2010 · 8 claims · 7 setups
A Perl-based pipeline that BLASTs/BLATs Illumina probe sequences against genomes and transcript databases (RefSeq, UCSC Known Genes, UniGene/GenBank, Ensembl) can classify probes by quality grade (Perfect/Good/Bad/No match) and is applicable across 8 BeadArray platforms and other array types
-
Full-text index only
INTERFEROME: the database of interferon regulated genes.
PMID 18996892 · PMC2686605 · Nucleic acids research · 2009 · 8 claims · 6 setups
INTERFEROME is an open-access database integrating IRG expression data with annotation, orthologue sequences from 37 species, tissue expression, and gene regulatory (TFBS) information
-
Full-text index only
BABELOMICS: a systems biology perspective in the functional annotation of genome-scale experiments.
PMID 16845052 · PMC1538844 · Nucleic acids research · 2006 · 8 claims · 8 setups
Babelomics is presented as an updated, complete suite of web tools for functional analysis of genome-scale experiments with new and improved modules
-
Full-text index only
Functional annotation and identification of candidate disease genes by computational analysis of normal tissue gene expression data.
PMID 18560577 · PMC2409962 · PloS one · 2008 · 7 claims · 5 setups
Ranked Coexpression Groups (RCG) built from k=6 nearest coexpressed genes, combined with a majority-rule functional characterization, integrate multiple datasets/coexpression measures to generate high-confidence functional annotation predictions
-
Full-text index only
Recent additions and improvements to the Onto-Tools.
PMID 15980579 · PMC1160233 · Nucleic acids research · 2005 · 7 claims · 3 setups
The Onto-Tools back-end database was redesigned around the Entrez Gene data model after NCBI phased out LocusLink in February 2005.
-
Full-text index only
BioGPS: an extensible and customizable portal for querying and organizing gene annotation resources.
PMID 19919682 · PMC3091323 · Genome biology · 2009 · 8 claims · 4 setups
BioGPS aggregates distributed, third-party gene annotation resources into a single customizable portal for human, mouse, and rat genes.
-
Full-text index only
Functional nsSNPs from carcinogenesis-related genes expressed in breast tissue: potential breast cancer risk alleles and their distribution across human populations.
PMID 16595073 · PMC3500178 · Human genomics · 2006 · 7 claims · 5 setups
A bioinformatics strategy cross-referencing carcinogenesis-related gene lists with breast-tissue expression data can identify candidate breast cancer risk nsSNPs.
-
Has reproduction · 75
FEM: mining biological meaning from cell level in single-cell RNA sequencing data.
PMID 34909283 · PMC8641482 · PeerJ · 2021 · 7 claims · 5 setups
The FEM algorithm converts each cell's gene expression matrix (GEM) into a functional expression matrix by applying Fisher's exact test enrichment per cell and per gene set, then encoding adjusted p-values as information content.
-
Full-text index only
Genomic approaches to the genetics of alcoholism.
PMID 12875046 · PMC6683845 · Alcohol research & health : the journal of the National Institute on Alcohol Abuse and Alcoholism · 2002 · 8 claims · 4 setups
Alcoholism is a complex disease that develops from a combination of numerous genetic and environmental factors, unlike single-gene disorders such as cystic fibrosis or Huntington's disease.
-
Full-text index only
Molecular cloning, genomic characterization and over-expression of a novel gene, XRRA1, identified from human colorectal cancer cell HCT116Clone2_XRR and macaque testis.
PMID 12908878 · PMC194569 · BMC genomics · 2003 · 8 claims · 7 setups
XRRA1 is a novel gene down-regulated ~2-fold in XR-resistant HCT116 Clone2_XRR relative to HCT116 Clone10, identified via cDNA microarray
-
Full-text index only
Comprehensive annotation of bidirectional promoters identifies co-regulation among breast and ovarian cancer genes.
PMID 17447839 · PMC1853124 · PLoS computational biology · 2007 · 8 claims · 8 setups
A new algorithm using spliced ESTs (cross-validated against Known Genes and GenBank mRNA) comprehensively maps bidirectional promoters in the human genome
-
Full-text index only
The mammalian phenotype ontology: enabling robust annotation and comparative analysis.
PMID 20052305 · PMC2801442 · Wiley interdisciplinary reviews. Systems biology and medicine · 2009 · 8 claims · 6 setups
The Mammalian Phenotype (MP) Ontology enables classification and organization of phenotypic data for mouse and other mammalian species in a computationally useful, standardized manner.
-
Full-text index only
The complete genome sequence of Vibrio cholerae: a tale of two chromosomes and of two lifestyles.
PMID 11178241 · PMC138858 · Genome biology · 2000 · 8 claims · 4 setups
The V. cholerae O1 (El Tor) genome consists of two chromosomes with asymmetrically distributed gene functions
-
Has reproduction · 30
First step toward gene expression data integration: transcriptomic data acquisition with COMMAND>_.
PMID 30691411 · PMC6348648 · BMC bioinformatics · 2019 · 7 claims · 3 setups
COMMAND>_ is a flexible multi-user web application that searches, downloads, parses, re-annotates, and imports gene expression experiments into a coherent data model.
-
Full-text index only
Improvements to cardiovascular gene ontology.
PMID 19046747 · PMC2706316 · Atherosclerosis · 2009 · 8 claims · 8 setups
Gene Ontology (GO) provides a controlled vocabulary that links current functional knowledge of genes to high-throughput genomic and proteomic datasets, aiding data interpretation.
-
Has reproduction · 74
Discovery of a novel filamentous prophage in the genome of the Mimosa pudica microsymbiont Cupriavidus taiwanensis STM 6018.
PMID 36925474 · PMC10011098 · Frontiers in microbiology · 2023 · 8 claims · 8 setups
The STM 6018 genome contains two prophages: a complete Mu-like capsular phage and a filamentous phage that integrates into a putative dif site.
-
Has reproduction
Using random walks to identify cancer-associated modules in expression data.
PMID 24128261 · PMC4015830 · BioData mining · 2013 · 8 claims · 8 setups
Walktrap-GM, a random-walk community detection algorithm adapted with stopping criteria (maximum modularity, maximum size, maximum module score), identifies modules significantly enriched with cancer genes in expression-weighted interaction networks.