Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Large-scale trends in the evolution of gene structures within 11 animal genomes.
PMID 16518452 · PMC1386723 · PLoS computational biology · 2006 · 8 claims · 5 setups
Change in intron–exon gene structure is gradual, clock-like, and largely independent of coding-sequence (protein) evolution
-
Full-text index only
The repertoire of G protein-coupled receptors in the sea squirt Ciona intestinalis.
PMID 18452600 · PMC2396169 · BMC evolutionary biology · 2008 · 8 claims · 5 setups
169 gene products in the Ciona genome were identified as putative GPCRs
-
Has reproduction · 77
Representing and querying disease networks using graph databases.
PMID 27462371 · PMC4960687 · BioData mining · 2016 · 7 claims · 8 setups
Graph databases are well suited for representing biological information because it is typically highly connected, semi-structured and unpredictable, unlike relational databases which require rigid schemas.
-
Full-text index only
Phylogenetic profiling of the Arabidopsis thaliana proteome: what proteins distinguish plants from other organisms?
PMID 15287975 · PMC507878 · Genome biology · 2004 · 8 claims · 6 setups
3,848 Arabidopsis proteins were identified as likely plant-specific based on phylogenetic profiling and EST confirmation in multiple plant species
-
Full-text index only
DiagHunter and GenoPix2D: programs for genomic comparisons, large-scale homology discovery and visualization.
PMID 14519203 · PMC328457 · Genome biology · 2003 · 7 claims · 5 setups
DiagHunter identifies large-scale synteny blocks within or between genomes efficiently despite background noise and genomic discontinuities, without performing sequence alignment
-
Full-text index only
Exploring the immunome: A brave new world for human vaccine development.
PMID 20009527 · PMC2919815 · Human vaccines · 2009 · 7 claims · 7 setups
Screening the Mtb proteome in silico for epitopes ('fishing for antigens using epitopes as bait') revealed a remarkable diversity of human immune responses to Mtb proteins without an ascribed function, suggesting human immune response to Mtb is omnivorous rather than focused on single immunodominant proteins.
-
Full-text index only
InParanoid 6: eukaryotic ortholog clusters with inparalogs.
PMID 18055500 · PMC2238924 · Nucleic acids research · 2008 · 8 claims · 3 setups
InParanoid 6 is an updated eukaryotic ortholog database covering 35 species (34 eukaryotes plus E. coli as outgroup), providing pairwise ortholog clusters with inparalogs for all species pairs.
-
Full-text index only
Functional coverage of the human genome by existing structures, structural genomics targets, and homology models.
PMID 16118666 · PMC1188274 · PLoS computational biology · 2005 · 8 claims · 5 setups
Existing PDB structures provide single-domain coverage for 37% of functional classes in the human genome and complete (whole-protein) structure coverage for 25%.
-
Full-text index only
Analysis of protein sequence and interaction data for candidate disease gene prediction.
PMID 17020920 · PMC1636487 · Nucleic acids research · 2006 · 8 claims · 7 setups
Combining CPS and CMP using known disease genes as input achieves sensitivity 0.52 and specificity 0.97, reducing candidate lists 13-fold
-
Full-text index only
BioDrugScreen: a computational drug design resource for ranking molecules docked to the human proteome.
PMID 19923229 · PMC2808957 · Nucleic acids research · 2010 · 6 claims · 5 setups
BioDrugScreen is a web resource providing pre-docked and pre-scored receptor-ligand complexes for ranking molecules against human proteome targets
-
Full-text index only
Comparative genomics.
PMID 14624258 · PMC261895 · PLoS biology · 2003 · 8 claims · 7 setups
Conserved DNA between species tends to encode shared functional features, while divergent DNA underlies species differences
-
Full-text index only
targetTB: a target identification pipeline for Mycobacterium tuberculosis through an interactome, reactome and genome-scale structural analysis.
PMID 19099550 · PMC2651862 · BMC systems biology · 2008 · 8 claims · 8 setups
A comprehensive in silico target identification pipeline (targetTB) integrating interactome, reactome, essentiality, sequence and structural analyses can identify high-confidence drug targets for Mtb
-
Full-text index only
Comparative analysis reveals signatures of differentiation amid genomic polymorphism in Lake Malawi cichlids.
PMID 18616806 · PMC2530870 · Genome biology · 2008 · 8 claims · 8 setups
Lake Malawi cichlids are phenotypically and behaviorally diverse but appear genetically like a single subdivided population rather than distinct species
-
Full-text index only
Coiled-coil protein composition of 22 proteomes--differences and common themes in subcellular infrastructure and traffic control.
PMID 16288662 · PMC1322226 · BMC evolutionary biology · 2005 · 7 claims · 5 setups
Proteins with extended coiled-coil domains (>250 amino acids) are largely absent from bacterial genomes but present in archaea and eukaryotes.
-
Full-text index only
TPRpred: a tool for prediction of TPR-, PPR- and SEL1-like repeats from protein sequences.
PMID 17199898 · PMC1774580 · BMC bioinformatics · 2007 · 7 claims · 8 setups
TPRpred detects divergent/remote-homolog TPR repeat units that existing resources (Pfam, SMART, REP) fail to detect
-
Full-text index only
Inparanoid: a comprehensive database of eukaryotic orthologs.
PMID 15608241 · PMC540061 · Nucleic acids research · 2005 · 8 claims · 4 setups
The Inparanoid algorithm identifies true ortholog clusters by seeding on reciprocal best-matching pairs, gathering inparalogs (post-speciation duplicates) while excluding outparalogs (pre-speciation duplicates)
-
Full-text index only
PlasmoDraft: a database of Plasmodium falciparum gene function predictions based on postgenomic data.
PMID 18925948 · PMC2605471 · BMC bioinformatics · 2008 · 8 claims · 4 setups
Gonna, a supervised k-nearest-neighbor Guilt-By-Association predictor, proposes GO annotations for a gene based on similarity of its transcriptome, proteome, or interactome profile to genes already annotated by GeneDB
-
Has reproduction · 50
Genetic parallels in biomineralization of the calcareous sponge Sycon ciliatum and stony corals.
PMID 40922549 · PMC12419799 · eLife · 2025 · 8 claims · 8 setups
829 genes are overexpressed in the oscular region of increased calcite spicule formation in S. ciliatum
-
Full-text index only
Identification of mitochondrial disease genes through integrative analysis of multiple datasets.
PMID 18930150 · PMC2774125 · Methods (San Diego, Calif.) · 2008 · 8 claims · 8 setups
Data integration of multiple functional genomics datasets effectively predicts mitochondrial gene function and prioritizes candidate mitochondrial disease genes.
-
Has reproduction · 58
Revised annotations, sex-biased expression, and lineage-specific genes in the Drosophila melanogaster group.
PMID 25273863 · PMC4267930 · G3 (Bethesda, Md.) · 2014 · 8 claims · 6 setups
Revised RNA-seq-based gene models for D. ananassae, D. yakuba, and D. simulans include UTRs, empirically verified intron-exon boundaries, and previously unannotated novel exons, improving on r1.3 comparative-genomics annotations that lack UTRs.