Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Gene Prospector: an evidence gateway for evaluating potential susceptibility genes and interacting risk factors for human diseases.
PMID 19063745 · PMC2613935 · BMC bioinformatics · 2008 · 8 claims · 5 setups
Gene Prospector is a Web-based application that selects and prioritizes potential disease-related genes using a curated, updated literature database of genetic association studies
-
Full-text index only
Finding disease candidate genes by liquid association.
PMID 17915034 · PMC2246280 · Genome biology · 2007 · 7 claims · 6 setups
LA can detect functionally associated genes that are not directly co-expressed by identifying a mediating gene Z whose expression level changes the correlation between X and Y.
-
Full-text index only
g:Profiler--a web-based toolset for functional profiling of gene lists from large-scale experiments.
PMID 17478515 · PMC1933153 · Nucleic acids research · 2007 · 8 claims · 5 setups
g:Profiler integrates four modules (g:Profiler core, g:Convert, g:Orth, g:Sorter) into a single cross-linked web tool for gene list analysis
-
Full-text index only
GeneKeyDB: a lightweight, gene-centric, relational database to support data mining environments.
PMID 15790402 · PMC1274265 · BMC bioinformatics · 2005 · 8 claims · 6 setups
GeneKeyDB is a lightweight, gene-centric relational database that supports data mining and integration with computational analysis tools.
-
Full-text index only
RotaC: a web-based tool for the complete genome classification of group A rotaviruses.
PMID 19930627 · PMC2785824 · BMC microbiology · 2009 · 7 claims · 4 setups
RotaC is a freely available web-based tool for complete genome classification of group A rotaviruses across all 11 gene segments.
-
Full-text index only
Gene-disease relationship discovery based on model-driven data integration and database view definition.
PMID 19042916 · PMC2639000 · Bioinformatics (Oxford, England) · 2009 · 8 claims · 4 setups
Explicit gene–disease relationships can be formulated as candidate gene definitions (e.g., co-localization, dysregulation, functional similarity) that may include intermediary orthologous or interacting genes
-
Full-text index only
Cataloging coding sequence variations in human genome databases.
PMID 18974781 · PMC2570488 · PloS one · 2008 · 8 claims · 7 setups
A significant proportion of CVs overlap between HGMD and dbSNP (4.36% of HGMD CVs registered in dbSNP; 8.11% of dbSNP CVs registered in HGMD), warranting caution when interpreting phenotypic relevance of concurrent CVs.
-
Full-text index only
Genome annotation errors in pathway databases due to semantic ambiguity in partial EC numbers.
PMID 16034025 · PMC1179732 · Nucleic acids research · 2005 · 7 claims · 4 setups
Partial EC numbers are semantically ambiguous, and databases that assign a gene to all reactions sharing the same partial EC number make a faulty inference, causing systematic misannotation.
-
Full-text index only
FGF: a web tool for Fishing Gene Family in a whole genome database.
PMID 17584790 · PMC1933194 · Nucleic acids research · 2007 · 6 claims · 3 setups
FGF efficiently searches and identifies gene families in whole-genome databases and outputs visual phylogenetic trees annotated with gene structure, chromosome position, duplication fate, and selective pressure (Ka/Ks)
-
Has reproduction · 55
Identification and verification of diagnostic biomarkers in recurrent pregnancy loss via machine learning algorithm and WGCNA.
PMID 37691920 · PMC10485775 · Frontiers in immunology · 2023 · 8 claims · 8 setups
352 DEGs (198 up-regulated, 154 down-regulated) were identified between RPL and control endometrial samples
-
Has reproduction · 68
Cell-type annotation with accurate unseen cell-type identification using multiple references.
PMID 37379341 · PMC10335708 · PLoS computational biology · 2023 · 8 claims · 4 setups
mtANN integrates multiple reference datasets and eight gene selection methods via ensemble learning (multiple deep classification models + majority voting) to improve cell-type annotation accuracy
-
Has reproduction · 94
Systematic assessment of pathway databases, based on a diverse collection of user-submitted experiments.
PMID 36088548 · PMC9487593 · Briefings in bioinformatics · 2022 · 8 claims · 6 setups
Well-established, hierarchically organized pathway annotation systems (e.g. GO, Reactome, KEGG) yield the best overall enrichment performance despite covering much of the human genome only in general terms.
-
Has reproduction · 78
A case study for large-scale human microbiome analysis using JCVI's metagenomics reports (METAREP).
PMID 22719821 · PMC3374610 · PloS one · 2012 · 8 claims · 7 setups
METAREP version 1.3.1 is an open-source, scalable tool for querying, browsing and comparing extremely large volumes of metagenomic annotations, with an extended data model, dynamic weighting, distributed searches and advanced clustering.
-
Full-text index only
An integrated database-pipeline system for studying single nucleotide polymorphisms and diseases.
PMID 19091018 · PMC2638159 · BMC bioinformatics · 2008 · 6 claims · 5 setups
Existing SNP/disease databases are fragmented; no combined resource widely supports gene-, SNP-, and disease-related information together
-
Full-text index only
Systems integration of biodefense omics data for analysis of pathogen-host interactions and identification of potential targets.
PMID 19779614 · PMC2745575 · PloS one · 2009 · 8 claims · 8 setups
A protein-centric data integration approach (Master Protein Directory) enables integration and mining of heterogeneous pathogen-host omics data across multiple research centers
-
Has reproduction · 60
TRAPID 2.0: a web application for taxonomic and functional analysis of de novo transcriptomes.
PMID 34197621 · PMC8464036 · Nucleic acids research · 2021 · 8 claims · 8 setups
TRAPID 2.0 is a web application performing global characterization of de novo transcriptomes via structural, functional, and taxonomic annotation in an initial processing phase, followed by an exploratory phase of downstream analyses.
-
Has reproduction · 95
Determining virus-host interactions and glycerol metabolism profiles in geographically diverse solar salterns with metagenomics.
PMID 28097058 · PMC5228507 · PeerJ · 2017 · 8 claims · 8 setups
Similar virus-host interactions and glycerol metabolism gene associations (notably dihydroxyacetone kinase with Haloquadratum/Halorubrum) exist across geographically diverse solar salterns
-
Full-text index only
TRED: a Transcriptional Regulatory Element Database and a platform for in silico gene regulation studies.
PMID 15608156 · PMC539958 · Nucleic acids research · 2005 · 8 claims · 5 setups
TRED is a database collecting both cis-regulatory elements (promoters) and trans-regulatory elements (transcription factor binding/regulation data) with linked access.
-
Full-text index only
SpliceMiner: a high-throughput database implementation of the NCBI Evidence Viewer for microarray splice variant analysis.
PMID 17338820 · PMC1839109 · BMC bioinformatics · 2007 · 6 claims · 4 setups
EVDB is a comprehensive, non-redundant relational database of known human splice variants built from NCBI Entrez Gene and Evidence Viewer data
-
Full-text index only
The UCSC Genome Browser Database: update 2006.
PMID 16381938 · PMC1347506 · Nucleic acids research · 2006 · 8 claims · 8 setups
The UCSC Genome Browser Database (GBD) provides integrated sequence and annotation data, with web tools (Genome Browser, Table Browser, Proteome Browser, Gene Sorter, BLAT, In Silico PCR) for visualizing and querying genomes of about a dozen vertebrate species and several model organisms.