Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Biocomputing enters its adolescence.
PMID 15960815 · PMC1175967 · Genome biology · 2005 · 8 claims · 8 setups
A 'match augmentation' algorithm efficiently matches structural motifs by prioritizing functionally significant residues, enabling function prediction between evolutionarily unrelated proteins
-
Full-text index only
targetTB: a target identification pipeline for Mycobacterium tuberculosis through an interactome, reactome and genome-scale structural analysis.
PMID 19099550 · PMC2651862 · BMC systems biology · 2008 · 8 claims · 8 setups
A comprehensive in silico target identification pipeline (targetTB) integrating interactome, reactome, essentiality, sequence and structural analyses can identify high-confidence drug targets for Mtb
-
Full-text index only
Predicting deleterious nsSNPs: an analysis of sequence and structural attributes.
PMID 16630345 · PMC1489951 · BMC bioinformatics · 2006 · 8 claims · 7 setups
Sequence conservation (PSIC score difference) at the nsSNP position is the single most useful attribute for predicting deleterious vs neutral status.
-
Full-text index only
Structural genomics and drug discovery for infectious diseases.
PMID 19860716 · PMC2789569 · Infectious disorders drug targets · 2009 · 7 claims · 4 setups
CSGID applies high-throughput X-ray crystallography structural genomics to NIAID category A-C pathogen proteins to enable structure-aided drug discovery, with a goal of 400 protein/protein-ligand structures
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
Assembling a jigsaw puzzle with 20,000 parts.
PMID 12801408 · PMC193613 · Genome biology · 2003 · 8 claims · 8 setups
Re-routing the intracellular interaction domains of receptor tyrosine kinases can redirect their signaling output, e.g. converting a growth signal into an apoptosis signal.
-
Full-text index only
Interaction profile-based protein classification of death domain.
PMID 15189571 · PMC459208 · BMC bioinformatics · 2004 · 7 claims · 6 setups
An SVM-based classifier using Residue Pair Interaction Profiles (RPIPs) can classify death domain superfamily members into subfamilies with 89% average cross-validation accuracy
-
Full-text index only
Crystal structure of the HSV-1 Fc receptor bound to Fc reveals a mechanism for antibody bipolar bridging.
PMID 16646632 · PMC1450327 · PLoS biology · 2006 · 8 claims · 5 setups
The C-terminal domain of the gE ectodomain (CgE) is the minimal Fc-binding domain of gE-gI
-
Full-text index only
InSite: a computational method for identifying protein-protein interaction binding sites on a proteome-wide scale.
PMID 17868464 · PMC2375030 · Genome biology · 2007 · 8 claims · 8 setups
InSite predicts protein-pair-specific binding motifs ('Motif M on protein A binds to protein B') by integrating heterogeneous PPI and motif-motif interaction evidence within a Bayesian network trained by EM
-
Has reproduction · 81
Identification of Proteins Deregulated by Platinum-Based Chemotherapy as Novel Biomarkers and Therapeutic Targets in Non-Small Cell Lung Cancer.
PMID 33777753 · PMC7991912 · Frontiers in oncology · 2021 · 7 claims · 8 setups
Cisplatin exposure induces significant deregulation of protein expression networks in NSCLC cells
-
Full-text index only
Filtering high-throughput protein-protein interaction data using a combination of genomic features.
PMID 15833142 · PMC1127019 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A combination of three genomic features (interacting Pfam domains, GO annotations, sequence homology) using naive Bayesian networks predicts true protein-protein interactions with high sensitivity and good specificity.
-
Full-text index only
Global sequencing of proteolytic cleavage sites in apoptosis by specific labeling of protein N termini.
PMID 18722006 · PMC2566540 · Cell · 2008 · 7 claims · 8 setups
A subtiligase-based N-terminal biotinylation and enrichment method enables global identification and sequencing of protease cleavage sites in complex mixtures
-
Full-text index only
'Genome design' model and multicellular complexity: golden middle.
PMID 17062620 · PMC1635334 · Nucleic acids research · 2006 · 8 claims · 8 setups
Intermediately expressed human genes are the longest genes genome-wide, in both coding and intronic sequence, longer than housekeeping or tissue-specific genes.
-
Full-text index only
Discovery and hypothesis generation through bioinformatics.
PMID 16522224 · PMC1431734 · Genome biology · 2006 · 8 claims · 8 setups
Bioinformatics should be used as a tool for discovery and hypothesis generation, not merely to manage biological data
-
Full-text index only
Phenotypic variation meets systems biology.
PMID 19664197 · PMC2745761 · Genome biology · 2009 · 8 claims · 8 setups
Cellular differentiation states are constrained by complex networks with substantial positive and negative regulation, challenging the concept of single 'master regulators'
-
Has reproduction · 67
CDKL1 variants affecting ciliary formation predispose to thoracic aortic aneurysm and dissection.
PMID 41056017 · PMC12646653 · The Journal of clinical investigation · 2025 · 8 claims · 8 setups
Heterozygous CDKL1 missense variants (Cys143Arg, Ser206Leu, Thr135Met) were identified in 6 patients from 3 families with TAAD spectrum disorders
-
Has reproduction · 87
Variants in LRRC7 lead to intellectual disability, autism, aggression and abnormal eating behaviors.
PMID 39256359 · PMC11387733 · Nature communications · 2024 · 8 claims · 7 setups
Heterozygous missense or loss-of-function variants in LRRC7 cause a dominant neurodevelopmental disorder in 33 identified individuals