Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
DAVID Knowledgebase: a gene-centered database integrating heterogeneous gene annotation resources to facilitate high-throughput gene functional analysis.
PMID 17980028 · PMC2186358 · BMC bioinformatics · 2007 · 7 claims · 3 setups
The DAVID Gene Concept, a single-linkage algorithm, merges gene clusters from Entrez Gene, UniRef100, and PIR-NREF100 that share protein IDs and species into unified DAVID gene clusters, improving cross-referencing between NCBI and UniProt systems
-
Full-text index only
Gene-disease relationship discovery based on model-driven data integration and database view definition.
PMID 19042916 · PMC2639000 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 4 setups
Explicit gene–disease relationships can be formulated as candidate gene definitions (e.g., co-localization, dysregulation, functional similarity) that may include intermediary orthologous or interacting genes
-
Full-text index only
SVC: structured visualization of evolutionary sequence conservation.
PMID 15991338 · PMC1160265 · Nucleic acids research · 2005 · 7 claims · 5 setups
SVC aligns protein-coding sequences of orthologous gene pairs and maps them back onto their encoding exons/introns to generate a scaffold of conserved gene structure.
-
Full-text index only
POCUS: mining genomic sequence annotation to predict disease genes.
PMID 14611661 · PMC329128 · Genome biology · 2003 · 8 claims · 6 setups
Genes predisposing to the same disease tend to share functional annotation IDs (GO/InterPro) more than expected by chance
-
Full-text index only
Integrating alternative splicing detection into gene prediction.
PMID 15705189 · PMC550657 · BMC bioinformatics · 2005 · 8 claims · 4 setups
An integrative intrinsic/extrinsic method was implemented in the gene finder EuGÈNE (as EuGÈNE-M) to detect AS evidence from aligned transcripts and generate alternative optimal gene predictions consistent with each detected AS event.
-
Full-text index only
Comparative Toxicogenomics Database: a knowledgebase and discovery tool for chemical-gene-disease networks.
PMID 18782832 · PMC2686584 · Nucleic acids research · 2009 · 8 claims · 5 setups
CTD is a manually curated knowledgebase that integrates chemical-gene interactions, chemical-disease relationships, and gene-disease relationships into a chemical-gene-disease triad
-
Has reproduction · 85
Single-Cell Differential Network Analysis with Sparse Bayesian Factor Models.
PMID 35186014 · PMC8855158 · Frontiers in genetics · 2021 · 8 claims · 2 setups
A hierarchical Bayesian factor model using treatment-dependent latent factor loadings can construct gene co-expression networks from scRNA-seq data and identify differences in network structure between two (or more) biological conditions.
-
Full-text index only
Integration of text- and data-mining using ontologies successfully selects disease gene candidates.
PMID 15767279 · PMC1065256 · Nucleic acids research · 2005 · 7 claims · 6 setups
Integrating eVOC anatomical ontology-based text-mining of PubMed abstracts with data-mining of gene expression annotation successfully selects and prioritizes candidate disease genes
-
Full-text index only
The MAPPER database: a multi-genome catalog of putative transcription factor binding sites.
PMID 15608292 · PMC540057 · Nucleic acids research · 2005 · 8 claims · 6 setups
Built a library of 1134 HMM models (359 matrix-derived, 718 factor-derived, 57 JASPAR-derived), corresponding to 863 distinct TF names, from TRANSFAC and JASPAR binding site data
-
Has reproduction · 85
Exploring microproteins from various model organisms using the mip-mining database.
PMID 37919660 · PMC10623795 · BMC genomics · 2023 · 5 claims · 4 setups
Mip-mining is a database of 336 curated RNA-seq datasets from 8626 samples across nine species, built specifically to explore microprotein functions under stress and disease conditions
-
Has reproduction
D2H2: diabetes data and hypothesis hub.
PMID 38107655 · PMC10723036 · Bioinformatics advances · 2023 · 6 claims · 5 setups
D2H2 is a web portal hosting hundreds of curated, uniformly reprocessed diabetes-relevant transcriptomics datasets from GEO with per-study visualization, differential expression, and single-gene queries.
-
Full-text index only
Network of Cancer Genes: a web resource to analyze duplicability, orthology and network properties of cancer genes.
PMID 19906700 · PMC2808873 · Nucleic acids research · 2010 · 7 claims · 4 setups
NCG is a web database integrating duplicability, orthology, evolutionary appearance, and network topology data for 736 human cancer genes
-
Full-text index only
New genes, new dilemmas: FTLD genetics and its implications for families.
PMID 18166610 · PMC10846215 · American journal of Alzheimer's disease and other dementias · 2007 · 8 claims · 8 setups
MAPT and PGRN mutations account for the largest number of familial FTLD cases and differ fundamentally in disease mechanism