Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Columba: an integrated database of proteins, structures, and annotations.
PMID 15801979 · PMC1087474 · BMC bioinformatics · 2005 · 8 claims · 6 setups
COLUMBA physically integrates data from twelve protein structure-related databases (PDB, KEGG, Swiss-Prot, CATH, SCOP, Gene Ontology, ENZYME, etc.) into a single PostgreSQL data warehouse.
-
Full-text index only
Prediction by graph theoretic measures of structural effects in proteins arising from non-synonymous single nucleotide polymorphisms.
PMID 18654622 · PMC2447880 · PLoS computational biology · 2008 · 8 claims · 5 setups
Bongo identifies mutations causing local and global structural effects with a remarkably low false positive rate
-
Has reproduction · 87
Enhanced Generalizability of RNA Secondary Structure Prediction via Convolutional Block Attention Network and Ensemble Learning.
PMID 40871599 · PMC12388828 · Molecules (Basel, Switzerland) · 2025 · 8 claims · 8 setups
TrioFold integrates base-pairing clues from thermodynamic- and DL-based methods via ensemble learning and a convolutional block attention mechanism to enhance RSS prediction generalizability.
-
Full-text index only
Random amino acid mutations and protein misfolding lead to Shannon limit in sequence-structure communication.
PMID 18769673 · PMC2518838 · PloS one · 2008 · 8 claims · 6 setups
The protein sequence-structure map behaves as a noisy digital communication channel whose capacity C exceeds the transmission rate R for native structures, satisfying Shannon's noisy channel theorem
-
Full-text index only
Oligomeric protein structure networks: insights into protein-protein interactions.
PMID 16336694 · PMC1326230 · BMC bioinformatics · 2005 · 8 claims · 6 setups
Interface amino acid clusters identified at Imin=6% correlate well with residues losing accessible surface area (δASA) upon oligomerization
-
Full-text index only
Protein under-wrapping causes dosage sensitivity and decreases gene duplicability.
PMID 18208334 · PMC2211539 · PLoS genetics · 2008 · 7 claims · 6 setups
Protein under-wrapping extent is negatively correlated with gene duplicability (family size) across six organisms (E. coli, yeast, worm, fly, human, thale cress)
-
Full-text index only
The European Bioinformatics Institute's data resources: towards systems biology.
PMID 15608238 · PMC539980 · Nucleic acids research · 2005 · 8 claims · 5 setups
Since 2003 the EBI has launched new databases covering protein-protein interactions (IntAct), pathways (Reactome) and small molecules (ChEBI)
-
Full-text index only
Natural history of S-adenosylmethionine-binding proteins.
PMID 16225687 · PMC1282579 · BMC structural biology · 2005 · 8 claims · 6 setups
The last universal common ancestor (LUCA) of cellular life had between 10 and 20 SAM-binding proteins from at least 5 fold classes
-
Full-text index only
DAVID Knowledgebase: a gene-centered database integrating heterogeneous gene annotation resources to facilitate high-throughput gene functional analysis.
PMID 17980028 · PMC2186358 · BMC bioinformatics · 2007 · 7 claims · 3 setups
The DAVID Gene Concept, a single-linkage algorithm, merges gene clusters from Entrez Gene, UniRef100, and PIR-NREF100 that share protein IDs and species into unified DAVID gene clusters, improving cross-referencing between NCBI and UniProt systems
-
Has reproduction · 67
Evaluating native-like structures of RNA-protein complexes through the deep learning method.
PMID 36828844 · PMC9958188 · Nature communications · 2023 · 8 claims · 7 setups
DRPScore identifies native-like RNA-protein structures with higher success rates than ITScore-PR, DARS-RNP, and 3dRPC across bound and unbound testing sets.
-
Full-text index only
A catalog of human cDNA expression clones and its application to structural genomics.
PMID 15345055 · PMC522878 · Genome biology · 2004 · 8 claims · 7 setups
A high-throughput screening approach can identify human cDNA clones from the hEx1 library that express soluble protein in E. coli
-
Full-text index only
Comparative sequence analysis of leucine-rich repeats (LRRs) within vertebrate toll-like receptors.
PMID 17517123 · PMC1899181 · BMC genomics · 2007 · 8 claims · 4 setups
A new method combining known LRR structures, multiple sequence alignment, and secondary structure prediction identifies and aligns LRRs in TLRs more accurately than PFAM/InterPro/SMART
-
Has reproduction · 100
Structure of a mitochondrial ribosome with fragmented rRNA in complex with membrane-targeting elements.
PMID 36253367 · PMC9576764 · Nature communications · 2022 · 8 claims · 4 setups
The P. magna mitoribosome contains rRNA split into 13 fragments (LSU1-8, SSU1-4, mt-5S)
-
Full-text index only
Intrinsic structural disorder confers cellular viability on oncogenic fusion proteins.
PMID 19888473 · PMC2768585 · PLoS computational biology · 2009 · 8 claims · 5 setups
Translocation-related human proteins are significantly enriched in intrinsic structural disorder compared to all human proteins
-
Full-text index only
InSite: a computational method for identifying protein-protein interaction binding sites on a proteome-wide scale.
PMID 17868464 · PMC2375030 · Genome biology · 2007 · 8 claims · 8 setups
InSite predicts protein-pair-specific binding motifs ('Motif M on protein A binds to protein B') by integrating heterogeneous PPI and motif-motif interaction evidence within a Bayesian network trained by EM
-
Full-text index only
Sequence occurrence and structural uniqueness of a G-quadruplex in the human c-kit promoter.
PMID 17720713 · PMC2034477 · Nucleic acids research · 2007 · 8 claims · 4 setups
The native 22-nt c-kit87 sequence occurs only once in the entire human genome.
-
Full-text index only
High-throughput crystallography for structural genomics.
PMID 19765976 · PMC2764548 · Current opinion in structural biology · 2009 · 8 claims · 8 setups
SG programs use genomic sequence data to select structurally novel protein targets, avoiding proteins with known structural homologues
-
Full-text index only
BioAfrica's HIV-1 proteomics resource: combining protein data with bioinformatics tools.
PMID 15757512 · PMC555852 · Retrovirology · 2005 · 8 claims · 3 setups
BioAfrica's HIV-1 Proteomics Resource integrates protein structure, gene expression, post-translational modification, functional activity and protein-macromolecule interaction data with bioinformatics tools in a single website.
-
Has reproduction · 71
Protein structure quality assessment based on the distance profiles of consecutive backbone Cα atoms.
PMID 24555103 · PMC3892923 · F1000Research · 2013 · 8 claims · 8 setups
The distance between consecutive backbone Cα atoms in high-quality structures is normally distributed with mean 3.8 Å and standard deviation 0.04 Å, justifying a reference state in which all consecutive Cα atoms are 3.8 Å apart.
-
Full-text index only
Ethnic differences and functional analysis of MET mutations in lung cancer.
PMID 19723643 · PMC2767337 · Clinical cancer research : an official journal of the American Association for Cancer Research · 2009 · 8 claims · 7 setups
MET mutations identified in lung tumors are predominantly germline rather than somatic