Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
KEGG for linking genomes to life and the environment.
PMID 18077471 · PMC2238879 · Nucleic acids research · 2008 · 8 claims · 4 setups
KEGG provides a reference knowledge base for linking genomes to life via PATHWAY mapping and to the environment via BRITE mapping.
-
Full-text index only
pSTIING: a 'systems' approach towards integrating signalling pathways, interaction and transcriptional regulatory networks in inflammation and cancer.
PMID 16381926 · PMC1347407 · Nucleic acids research · 2006 · 8 claims · 3 setups
pSTIING is a publicly accessible web-based knowledgebase integrating protein-protein, protein-lipid, protein-small molecule interactions, transcriptional regulatory associations, ligand-receptor-cell type information, and signal transduction modules, with a focus on inflammation, cell migration and cancer.
-
Full-text index only
Quadratic regression analysis for gene discovery and pattern recognition for non-cyclic short time-course microarray experiments.
PMID 15850479 · PMC1127068 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A step-down quadratic regression method (fitting quadratic, then linear, then null models per gene) identifies differentially expressed genes and classifies them into 9 temporal expression patterns using continuous time information.
-
Full-text index only
Identification of the proliferation/differentiation switch in the cellular network of multicellular organisms.
PMID 17166053 · PMC1664705 · PLoS computational biology · 2006 · 8 claims · 8 setups
Integrating interactome and transcriptome data reveals a pair of transcriptionally anticorrelated network modules (P and D) each comprising hundreds of genes, present across individuals and species.
-
Full-text index only
Unravelling the hidden heterogeneities of diffuse large B-cell lymphoma based on coupled two-way clustering.
PMID 17888167 · PMC2082044 · BMC genomics · 2007 · 8 claims · 6 setups
A proposed coupled two-way clustering (CTWC/SPC) method combined with a GO-based functional concept consistency score can identify compact, robust gene subsets that define clinically meaningful DLBCL subtypes
-
Full-text index only
Genomic expression during human myelopoiesis.
PMID 17683550 · PMC2045681 · BMC genomics · 2007 · 8 claims · 5 setups
An integrated myelopoiesis expression dataset of 9,425 genes, each mapped to a unique genomic position, was generated from 24 microarray experiments across 8 myeloid cell types.
-
Full-text index only
Motif discovery in promoters of genes co-localized and co-expressed during myeloid cells differentiation.
PMID 19059999 · PMC2632922 · Nucleic acids research · 2009 · 6 claims · 8 setups
A novel multi-step computational method (built on approximate pattern enumeration, binomial over-representation scoring with FDR correction, and k-medoids clustering) can identify over-represented motifs in a selected set of promoters relative to a background promoter set.
-
Full-text index only
Functional annotation and identification of candidate disease genes by computational analysis of normal tissue gene expression data.
PMID 18560577 · PMC2409962 · PloS one · 2008 · 7 claims · 5 setups
Ranked Coexpression Groups (RCG) built from k=6 nearest coexpressed genes, combined with a majority-rule functional characterization, integrate multiple datasets/coexpression measures to generate high-confidence functional annotation predictions
-
Full-text index only
Transcription network construction for large-scale microarray datasets using a high-performance computing approach.
PMID 18366618 · PMC2386070 · BMC genomics · 2008 · 8 claims · 7 setups
RMT removes the random noise component of the gene expression correlation matrix by testing its eigenvalue statistics against a null hypothesis derived from a truly random correlation matrix
-
Full-text index only
In silico discovery of transcription regulatory elements in Plasmodium falciparum.
PMID 18257930 · PMC2268928 · BMC genomics · 2008 · 7 claims · 8 setups
GEMS, using hypergeometric scoring and PWM parameter optimization, reliably identifies high-confidence cis-regulatory elements in the AT-rich, repeat-rich P. falciparum genome
-
Has reproduction · 30
First step toward gene expression data integration: transcriptomic data acquisition with COMMAND>_.
PMID 30691411 · PMC6348648 · BMC bioinformatics · 2019 · 7 claims · 3 setups
COMMAND>_ is a flexible multi-user web application that searches, downloads, parses, re-annotates, and imports gene expression experiments into a coherent data model.
-
Has reproduction · 71
Artificial intelligence-guided discovery of gastric cancer continuum.
PMID 36692601 · PMC9871434 · Gastric cancer : official journal of the International Gastric Cancer Association and the Japanese Gastric Cancer Association · 2023 · 8 claims · 8 setups
A Boolean implication network built from GSE66229 yields a GC-BoNE gene signature (Boolean paths C#11-2-4-14 and C#7-13-14) that classifies tumor vs normal/adjacent-normal gastric samples
-
Has reproduction · 77
Representing and querying disease networks using graph databases.
PMID 27462371 · PMC4960687 · BioData mining · 2016 · 7 claims · 8 setups
Graph databases are well suited for representing biological information that is highly connected, semi-structured, and unpredictable.
-
Full-text index only
Phylogenetic profiling of the Arabidopsis thaliana proteome: what proteins distinguish plants from other organisms?
PMID 15287975 · PMC507878 · Genome biology · 2004 · 8 claims · 6 setups
3,848 Arabidopsis proteins were identified as likely plant-specific based on phylogenetic profiling and EST confirmation in multiple plant species
-
Full-text index only
The DAVID Gene Functional Classification Tool: a novel biological module-centric algorithm to functionally analyze large gene lists.
PMID 17784955 · PMC2375021 · Genome biology · 2007 · 8 claims · 6 setups
Gene-gene functional similarity can be measured using kappa statistics applied to a binary gene-annotation-term matrix built from 14 annotation categories.
-
Full-text index only
GOLD.db: genomics of lipid-associated disorders database.
PMID 15588328 · PMC544894 · BMC genomics · 2004 · 8 claims · 4 setups
GOLD.db integrates annotated pathways, gene/protein reference information, and curated gene expression datasets for lipid-associated disorders research
-
Full-text index only
Toxicogenomics: an emerging discipline.
PMID 12460812 · PMC1241126 · Environmental health perspectives · 2002 · 8 claims · 6 setups
Toxicogenomics applies genomic tools (microarrays, proteomics, metabolomics) to characterize how cells and organisms respond to chemical/drug exposures.
-
Full-text index only
VectorBase: a home for invertebrate vectors of human pathogens.
PMID 17145709 · PMC1751530 · Nucleic acids research · 2007 · 8 claims · 5 setups
VectorBase is a web-accessible data repository for information about invertebrate vectors of human pathogens
-
Full-text index only
Identification and characterization of a novel mammalian Mg2+ transporter with channel-like properties.
PMID 15804357 · PMC1129089 · BMC genomics · 2005 · 8 claims · 6 setups
MagT1 is a novel mammalian Mg2+ transporter with channel-like properties, showing no amino acid sequence identity to other known transporters
-
Full-text index only
PlasmoDraft: a database of Plasmodium falciparum gene function predictions based on postgenomic data.
PMID 18925948 · PMC2605471 · BMC bioinformatics · 2008 · 8 claims · 4 setups
Gonna, a supervised k-nearest-neighbor Guilt-By-Association predictor, proposes GO annotations for a gene based on similarity of its transcriptome, proteome, or interactome profile to genes already annotated by GeneDB