Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Genomic structure and expression of Jmjd6 and evolutionary analysis in the context of related JmjC domain containing proteins.
PMID 18564434 · PMC2453528 · BMC genomics · 2008 · 8 claims · 6 setups
Jmjd6 has been misleadingly annotated as a transmembrane receptor for engulfment of apoptotic cells; recent evidence contradicts this transmembrane receptor function
-
Full-text index only
Protein function assignment through mining cross-species protein-protein interactions.
PMID 18253506 · PMC2216687 · PloS one · 2008 · 8 claims · 6 setups
CSIDOP predicts protein molecular function with 95.42% accuracy using 2,972 GO functional categories in H. sapiens
-
Full-text index only
DBD--taxonomically broad transcription factor predictions: new content and functionality.
PMID 18073188 · PMC2238844 · Nucleic acids research · 2008 · 8 claims · 3 setups
DBD is a database of predicted sequence-specific DNA-binding transcription factors covering over 700 publicly available proteomes, up from 150 in the initial version.
-
Full-text index only
Filtering high-throughput protein-protein interaction data using a combination of genomic features.
PMID 15833142 · PMC1127019 · BMC bioinformatics · 2005 · 8 claims · 8 setups
A combination of three genomic features (interacting Pfam domains, GO annotations, sequence homology) using naive Bayesian networks predicts true protein-protein interactions with high sensitivity and good specificity.
-
Full-text index only
Protein coding potential of retroviruses and other transposable elements in vertebrate genomes.
PMID 15716312 · PMC549403 · Nucleic acids research · 2005 · 8 claims · 5 setups
About 1000 genes across four vertebrate gene sets analyzed contain at least one RETRA marker protein domain
-
Full-text index only
Organization of physical interactomes as uncovered by network schemas.
PMID 18949022 · PMC2561054 · PLoS computational biology · 2008 · 7 claims · 5 setups
A computational procedure can systematically identify 'emergent' network schemas that are both recurrent and over-represented relative to randomized networks preserving lower-order subschema distributions
-
Full-text index only
Comparative phosphoproteomics reveals evolutionary and functional conservation of phosphorylation across eukaryotes.
PMID 18828897 · PMC2760871 · Genome biology · 2008 · 8 claims · 8 setups
The overlap between phosphoproteomes of six eukaryotes (human, mouse, fly, yeast, plant, zebrafish) is significantly greater than expected by chance.
-
Full-text index only
Pseudofam: the pseudogene families database.
PMID 18957444 · PMC2686518 · Nucleic acids research · 2009 · 8 claims · 7 setups
Pseudofam is an online database of pseudogene families built by mapping pseudogenes to Pfam protein families, providing query tools, statistics, and sequence alignments
-
Full-text index only
The genome of the simian and human malaria parasite Plasmodium knowlesi.
PMID 18843368 · PMC2656934 · Nature · 2008 · 8 claims · 7 setups
The P. knowlesi (H strain) nuclear genome was sequenced and assembled: 23.5 Mb across 14 chromosomes with 5,188 predicted protein-encoding genes.
-
Full-text index only
InParanoid 7: new algorithms and tools for eukaryotic orthology analysis.
PMID 19892828 · PMC2808972 · Nucleic acids research · 2010 · 8 claims · 7 setups
InParanoid 7 expands the database by an order of magnitude to 100 species, 1.3 million proteins, and 42.7 million pairwise ortholog groups.
-
Full-text index only
MEROPS: the peptidase database.
PMID 19892822 · PMC2808883 · Nucleic acids research · 2010 · 8 claims · 5 setups
MEROPS is a manually curated hierarchical classification of peptidases and protein inhibitors organized into protein species, families, and clans based on sequence and structural homology.
-
Full-text index only
The EH1 motif in metazoan transcription factors.
PMID 16309560 · PMC1310626 · BMC genomics · 2005 · 8 claims · 5 setups
There is a statistically significant association between EH1hox motif HMM score and transcription factor function across human, Drosophila and C. elegans proteomes.
-
Full-text index only
Protein length in eukaryotic and prokaryotic proteomes.
PMID 15951512 · PMC1150220 · Nucleic acids research · 2005 · 7 claims · 5 setups
Eukaryotic proteins are significantly longer than prokaryotic proteins across virtually all functional categories and the majority of protein families
-
Has reproduction · 92
Telomere-to-telomere reference genome for Panax ginseng highlights the evolution of saponin biosynthesis.
PMID 38883331 · PMC11179851 · Horticulture research · 2024 · 8 claims · 8 setups
A telomere-to-telomere reference genome of P. ginseng was assembled (3.45 Gb, 24 chromosomes, 77266 protein-coding genes)
-
Full-text index only
NovelFam3000--uncharacterized human protein domains conserved across model organisms.
PMID 16533400 · PMC1440326 · BMC genomics · 2006 · 8 claims · 7 setups
NovelFam3000 is an online data centre unifying bioinformatics resource links, news, comments, and user-submitted experimental data (including a Gene Characterization Index) for ~3000 uncharacterized Pfam-B/DUF domain families conserved across worm, fly, and human