Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
An SVD-based comparison of nine whole eukaryotic genomes supports a coelomate rather than ecdysozoan lineage.
PMID 15606920 · PMC544558 · BMC bioinformatics · 2004 · 8 claims · 7 setups
SVD-based analysis of tetrapeptide frequency vectors can compare whole eukaryotic proteomes without pre-defining orthologs or aligning homologous sites
-
Full-text index only
Inventory and analysis of the protein subunits of the ribonucleases P and MRP provides further evidence of homology between the yeast and human enzymes.
PMID 16998185 · PMC1636426 · Nucleic acids research · 2006 · 8 claims · 6 setups
Fungal Pop8 is evolutionarily related to the Rpp14/Pop5 protein family, suggesting Pop8 is the fungal orthologue of Rpp14
-
Full-text index only
Automatic discovery of cross-family sequence features associated with protein function.
PMID 16409628 · PMC1395344 · BMC bioinformatics · 2006 · 8 claims · 6 setups
A self-supervised data mining approach can find relationships between sequence features and functional annotations without preconceived functional categories.
-
Full-text index only
From endosymbiont to host-controlled organelle: the hijacking of mitochondrial protein synthesis and metabolism.
PMID 17983265 · PMC2062474 · PLoS computational biology · 2007 · 8 claims · 7 setups
There has been a large turnover of the mitochondrial proteome during evolution: cell envelope synthesis proteins virtually disappeared, and replication, transcription, cell division, transport, regulation, and signal transduction proteins were replaced by eukaryotic proteins
-
Full-text index only
The human phylome.
PMID 17567924 · PMC2394744 · Genome biology · 2007 · 6 claims · 5 setups
Reconstruction of the human phylome: evolutionary trees for all human proteins and their homologs among 39 fully sequenced eukaryotic genomes, using a pipeline combining alignment trimming, NJ, ML (PhyML) and Bayesian (MrBayes) methods.
-
Full-text index only
Gene loss rate: a probabilistic measure for the conservation of eukaryotic genes.
PMID 17158152 · PMC1802574 · Nucleic acids research · 2007 · 8 claims · 8 setups
GLR is a novel maximum-likelihood measure of gene loss rate that probabilistically weighs all possible ancestral phyletic patterns rather than relying on a single parsimonious reconstruction.
-
Full-text index only
Genome-wide in silico identification and analysis of cis natural antisense transcripts (cis-NATs) in ten species.
PMID 16849434 · PMC1524920 · Nucleic acids research · 2006 · 8 claims · 7 setups
A fast integrative in silico pipeline combining UniGene mRNA/EST mapping to GoldenPath genomes with CDS, poly(A) signal, poly(A) tail and splicing site evidence can reliably identify cis-NATs genome-wide across multiple species
-
Full-text index only
Comparison of characteristics and function of translation termination signals between and within prokaryotic and eukaryotic organisms.
PMID 16614446 · PMC1435984 · Nucleic acids research · 2006 · 8 claims · 5 setups
A core termination signal of 4 nt (stop codon plus the following nucleotide) is preferred across most prokaryotic and eukaryotic genomes
-
Full-text index only
BLASTO: a tool for searching orthologous groups.
PMID 17483516 · PMC1933156 · Nucleic acids research · 2007 · 7 claims · 2 setups
BLASTO treats each orthologous group as a unit and outputs a ranked list of orthologous groups instead of single sequences
-
Full-text index only
SelenoDB 1.0 : a database of selenoprotein genes, proteins and SECIS elements.
PMID 18174224 · PMC2238826 · Nucleic acids research · 2008 · 6 claims · 5 setups
Standard genome annotation pipelines misannotate selenoprotein genes because they rely on UGA as a universal stop codon, failing to recognize its dual role as the selenocysteine-recoding codon.
-
Full-text index only
Comparative genomics of cyclin-dependent kinases suggest co-evolution of the RNAP II C-terminal domain and CTD-directed CDKs.
PMID 15380029 · PMC521075 · BMC genomics · 2004 · 8 claims · 6 setups
Cell-cycle related CDKs (orthologs of CDK1-6) are present in all sampled eukaryotic organisms, including the most ancestral protists.
-
Full-text index only
SECIS elements in the coding regions of selenoprotein transcripts are functional in higher eukaryotes.
PMID 17169995 · PMC1802603 · Nucleic acids research · 2007 · 8 claims · 5 setups
SECIS elements located within coding regions of selenoprotein mRNAs support functional Sec insertion in mammalian cells
-
Has reproduction · 98
Sequence-based pangenomic core detection.
PMID 35663029 · PMC9160775 · iScience · 2022 · 7 claims · 3 setups
Sequence-based pangenomic core detection can be performed directly on unannotated genome sequences using a colored de Bruijn graph, avoiding bias from error-prone gene annotations
-
Full-text index only
The TIGR Gene Indices: clustering and assembling EST and known genes and integration with eukaryotic genomes.
PMID 15608288 · PMC540018 · Nucleic acids research · 2005 · 8 claims · 8 setups
The TIGR Gene Indices (TGI) are a collection of 77 species-specific databases that cluster and assemble EST and known gene sequences into tentative consensus (TC) sequences to identify and characterize expressed transcripts.
-
Full-text index only
Dyneins across eukaryotes: a comparative genomic analysis.
PMID 17897317 · PMC2239267 · Traffic (Copenhagen, Denmark) · 2007 · 8 claims · 6 setups
Phylogenetic inference identified nine DHC families (two cytoplasmic, seven axonemal) and six IC families (one cytoplasmic)
-
Full-text index only
MODBASE, a database of annotated comparative protein structure models and associated resources.
PMID 18948282 · PMC2686492 · Nucleic acids research · 2009 · 8 claims · 8 setups
MODBASE contains 5,152,695 reliable comparative protein structure models for 1,593,209 unique protein sequences.
-
Full-text index only
Evolutionary origins of human apoptosis and genome-stability gene networks.
PMID 18832373 · PMC2577361 · Nucleic acids research · 2008 · 8 claims · 8 setups
The entanglement of DNA repair, chromosome stability and apoptosis gene networks appears with the caspase gene family and the antiapoptotic gene BCL2.
-
Full-text index only
Inverse symmetry in complete genomes and whole-genome inverse duplication.
PMID 19898631 · PMC2771390 · PloS one · 2009 · 8 claims · 5 setups
Reverse and complement symmetries are essentially absent in genomic sequences at all scales.
-
Has reproduction · 58
Mucospheres produced by a mixotrophic protist impact ocean carbon cycling.
PMID 35288549 · PMC8921327 · Nature communications · 2022 · 8 claims · 8 setups
P. cf. balticum produces carbon-rich 'mucospheres' that attract, capture and immobilise a wide range of prokaryotic and eukaryotic prey
-
Has reproduction · 78
Long-read nanopore shotgun metagenomic DNA sequencing for river biodiversity, wildlife, pollution, and environmental health monitoring.
PMID 42038409 · PMC13107125 · NAR genomics and bioinformatics · 2026 · 7 claims · 7 setups
Long-read shotgun metagenomic sequencing of eDNA can simultaneously detect and quantify organismal DNA from viruses to mammals in a single assay