Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Protein co-evolution, co-adaptation and interactions.
PMID 18818697 · PMC2556093 · The EMBO journal · 2008 · 8 claims · 6 setups
The mirrortree method predicts protein-protein interactions by detecting pairs of protein families with similar phylogenetic trees (quantified as Pearson correlation of sequence similarity matrices).
-
Full-text index only
DiagHunter and GenoPix2D: programs for genomic comparisons, large-scale homology discovery and visualization.
PMID 14519203 · PMC328457 · Genome biology · 2003 · 7 claims · 5 setups
DiagHunter identifies large-scale synteny blocks within or between genomes efficiently despite background noise and genomic discontinuities, without performing sequence alignment
-
Full-text index only
Evolution of insect proteomes: insights into synapse organization and synaptic vesicle life cycle.
PMID 18257909 · PMC2374702 · Genome biology · 2008 · 8 claims · 3 setups
Compiled a list of 120 core presynaptic gene prototypes (PS120) and catalogued their conservation across insect proteomes.
-
Has reproduction · 77
Representing and querying disease networks using graph databases.
PMID 27462371 · PMC4960687 · BioData mining · 2016 · 7 claims · 8 setups
Graph databases are well suited for representing biological information because it is typically highly connected, semi-structured and unpredictable, unlike relational databases which require rigid schemas.
-
Full-text index only
The DAVID Gene Functional Classification Tool: a novel biological module-centric algorithm to functionally analyze large gene lists.
PMID 17784955 · PMC2375021 · Genome biology · 2007 · 8 claims · 6 setups
Gene-gene functional similarity can be measured using kappa statistics applied to a binary gene-annotation-term matrix built from 14 annotation categories.
-
Full-text index only
Analysis of protein sequence and interaction data for candidate disease gene prediction.
PMID 17020920 · PMC1636487 · Nucleic acids research · 2006 · 8 claims · 7 setups
Combining CPS and CMP using known disease genes as input achieves sensitivity 0.52 and specificity 0.97, reducing candidate lists 13-fold
-
Full-text index only
targetTB: a target identification pipeline for Mycobacterium tuberculosis through an interactome, reactome and genome-scale structural analysis.
PMID 19099550 · PMC2651862 · BMC systems biology · 2008 · 8 claims · 8 setups
A comprehensive in silico target identification pipeline (targetTB) integrating interactome, reactome, essentiality, sequence and structural analyses can identify high-confidence drug targets for Mtb
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
Biocomputing enters its adolescence.
PMID 15960815 · PMC1175967 · Genome biology · 2005 · 8 claims · 8 setups
A 'match augmentation' algorithm efficiently matches structural motifs by prioritizing functionally significant residues, enabling function prediction between evolutionarily unrelated proteins
-
Full-text index only
What makes species unique? The contribution of proteins with obscure features.
PMID 16859532 · PMC1779552 · Genome biology · 2006 · 7 claims · 8 setups
POFs constitute 18-38% (average 26%) of a typical eukaryotic proteome
-
Full-text index only
Identification of the proliferation/differentiation switch in the cellular network of multicellular organisms.
PMID 17166053 · PMC1664705 · PLoS computational biology · 2006 · 8 claims · 8 setups
Integrating interactome and transcriptome data reveals a pair of transcriptionally anticorrelated network modules (P and D) each comprising hundreds of genes, present across individuals and species.
-
Full-text index only
Human disease classification in the postgenomic era: a complex systems approach to human pathobiology.
PMID 17625512 · PMC1948102 · Molecular systems biology · 2007 · 8 claims · 5 setups
Current syndromic disease classification lacks specificity despite historically serving clinicians well
-
Has reproduction · 76
GeneSetCart: assembling, augmenting, combining, visualizing, and analyzing gene sets.
PMID 40208796 · PMC11984350 · GigaScience · 2025 · 8 claims · 8 setups
GeneSetCart is a web-based platform that lets users assemble, augment, combine, visualize, and analyze gene sets from multiple sources in one place
-
Has reproduction · 94
Systematic assessment of pathway databases, based on a diverse collection of user-submitted experiments.
PMID 36088548 · PMC9487593 · Briefings in bioinformatics · 2022 · 8 claims · 6 setups
Well-established, hierarchically organized pathway annotation systems (e.g. GO, Reactome, KEGG) yield the best overall enrichment performance despite covering much of the human genome only in general terms.
-
Full-text index only
Benchmarking ortholog identification methods using functional genomics data.
PMID 16613613 · PMC1557999 · Genome biology · 2006 · 8 claims · 7 setups
InParanoid is the best overall ortholog identification method for identifying functionally equivalent proteins when sensitivity and selectivity are combined into an overall score.
-
Full-text index only
Update of the G2D tool for prioritization of gene candidates to inherited diseases.
PMID 17478516 · PMC1933178 · Nucleic acids research · 2007 · 8 claims · 4 setups
G2D is a web server that prioritizes candidate genes for inherited diseases using three distinct algorithms based on different input information.
-
Full-text index only
PlasmoDraft: a database of Plasmodium falciparum gene function predictions based on postgenomic data.
PMID 18925948 · PMC2605471 · BMC bioinformatics · 2008 · 8 claims · 4 setups
Gonna, a supervised k-nearest-neighbor Guilt-By-Association predictor, proposes GO annotations for a gene based on similarity of its transcriptome, proteome, or interactome profile to genes already annotated by GeneDB
-
Full-text index only
Network properties of complex human disease genes identified through genome-wide association studies.
PMID 19956617 · PMC2779513 · PloS one · 2009 · 7 claims · 6 setups
Complex disease genes are significantly less central (lower degree/closeness, higher eccentricity) in the human interactome than essential and monogenic disease genes, occupying an intermediate niche between monogenic disease genes and non-disease genes
-
Full-text index only
Discovery and hypothesis generation through bioinformatics.
PMID 16522224 · PMC1431734 · Genome biology · 2006 · 8 claims · 8 setups
Bioinformatics should be used as a tool for discovery and hypothesis generation, not merely to manage biological data
-
Has reproduction · 30
Minimal metabolic pathway structure is consistent with associated biomolecular interactions.
PMID 24987116 · PMC4299494 · Molecular systems biology · 2014 · 8 claims · 8 setups
MinSpan, a mixed-integer linear optimization algorithm, computes the shortest, linearly independent pathways (sparsest basis of the null space of the stoichiometric matrix S) for genome-scale metabolic networks, which convex approaches (extreme pathways, elementary flux modes) cannot do at genome scale.