Experiments
Searchable full-text extractions: founding hypothesis, core claims, experimental setups, key results and statistics — pulled out of each paper as structure. Search a cell line, an assay or an entity (e.g. HUH7) and find every paper that worked with it. This corpus stands on its own: most entries carry no reproduction assessment (yet).
-
Full-text index only
Bioinformatic mapping of AlkB homology domains in viruses.
PMID 15627404 · PMC544882 · BMC genomics · 2005 · 8 claims · 8 setups
AlkB-like domains are found in at least 22 different single-stranded RNA positive-strand plant viruses, mainly within a subgroup of the Flexiviridae family.
-
Full-text index only
Protein ranking by semi-supervised network propagation.
PMID 16723003 · PMC1810311 · BMC bioinformatics · 2006 · 8 claims · 5 setups
RankProp, a diffusion-based network propagation algorithm on a PSI-BLAST-derived protein similarity network, significantly outperforms local search methods (BLAST/PSI-BLAST) at detecting remote homologs.
-
Full-text index only
Coiled-coil protein composition of 22 proteomes--differences and common themes in subcellular infrastructure and traffic control.
PMID 16288662 · PMC1322226 · BMC evolutionary biology · 2005 · 7 claims · 5 setups
Proteins with extended coiled-coil domains (>250 amino acids) are largely absent from bacterial genomes but present in archaea and eukaryotes.
-
Full-text index only
Molecular evolution of Cide family proteins: novel domain formation in early vertebrates and the subsequent divergence.
PMID 18500987 · PMC2426694 · BMC evolutionary biology · 2008 · 8 claims · 5 setups
Sequences homologous to the CIDE-N domain/NCD show a wide phylogenetic distribution, from hydra and sea anemone to mammals, while true Cide proteins are restricted to vertebrates.
-
Full-text index only
The repertoire of G protein-coupled receptors in the sea squirt Ciona intestinalis.
PMID 18452600 · PMC2396169 · BMC evolutionary biology · 2008 · 8 claims · 5 setups
169 gene products in the Ciona genome were identified as putative GPCRs
-
Full-text index only
Evolution of insect proteomes: insights into synapse organization and synaptic vesicle life cycle.
PMID 18257909 · PMC2374702 · Genome biology · 2008 · 8 claims · 3 setups
Compiled a list of 120 core presynaptic gene prototypes (PS120) and catalogued their conservation across insect proteomes.
-
Full-text index only
Protein co-evolution, co-adaptation and interactions.
PMID 18818697 · PMC2556093 · The EMBO journal · 2008 · 8 claims · 6 setups
The mirrortree method predicts protein-protein interactions by detecting pairs of protein families with similar phylogenetic trees (quantified as Pearson correlation of sequence similarity matrices).
-
Full-text index only
Identifying repeat domains in large genomes.
PMID 16507140 · PMC1431705 · Genome biology · 2006 · 7 claims · 5 setups
A repeat domain graph, built using a modified A-Bruijn graph framework, decomposes a repeat library into shared repeat domains and reveals the mosaic structure of repeat families.
-
Full-text index only
Origin and diversification of the basic helix-loop-helix gene family in metazoans: insights from comparative genomics.
PMID 17335570 · PMC1828162 · BMC evolutionary biology · 2007 · 8 claims · 4 setups
An initial diversification of bHLHs occurred in the pre-Cambrian, prior to metazoan cladogenesis
-
Full-text index only
The Vertebrate Genome Annotation (Vega) database.
PMID 15608237 · PMC540089 · Nucleic acids research · 2005 · 8 claims · 8 setups
Vega is a community database for browsing manual annotation of finished vertebrate genome sequences, based on an extended Ensembl-style schema.
-
Has reproduction · 60
TRAPID 2.0: a web application for taxonomic and functional analysis of de novo transcriptomes.
PMID 34197621 · PMC8464036 · Nucleic acids research · 2021 · 8 claims · 8 setups
TRAPID 2.0 is a web application performing global characterization of de novo transcriptomes via structural, functional, and taxonomic annotation in an initial processing phase, followed by an exploratory phase of downstream analyses.
-
Full-text index only
TB database: an integrated platform for tuberculosis research.
PMID 18835847 · PMC2686437 · Nucleic acids research · 2009 · 8 claims · 8 setups
TBDB is an integrated database providing access to TB genomic data and resources relevant to discovery/development of TB drugs, vaccines and biomarkers.
-
Has reproduction · 67
Satellitome Analysis and Transposable Elements Comparison in Geographically Distant Populations of Spodoptera frugiperda.
PMID 35455012 · PMC9026859 · Life (Basel, Switzerland) · 2022 · 8 claims · 5 setups
Most transposable elements are commonly shared across all eight geographically distant S. frugiperda samples, except Maverick and PIF/Harbinger elements which show divergent repeat copies
-
Full-text index only
Protein coding potential of retroviruses and other transposable elements in vertebrate genomes.
PMID 15716312 · PMC549403 · Nucleic acids research · 2005 · 8 claims · 5 setups
About 1000 genes across four vertebrate gene sets analyzed contain at least one RETRA marker protein domain
-
Full-text index only
Systematic analysis of human kinase genes: a large number of genes and alternative splicing events result in functional and structural diversity.
PMID 16351747 · PMC1866387 · BMC bioinformatics · 2005 · 8 claims · 7 setups
Systematic in silico search identified 5 novel human kinase genes (on chromosomes 1, 11, 13, 15, 16) and 1 pseudogene (chromosome X) absent from KinBase
-
Full-text index only
Analysis of protein sequence and interaction data for candidate disease gene prediction.
PMID 17020920 · PMC1636487 · Nucleic acids research · 2006 · 8 claims · 7 setups
Combining CPS and CMP using known disease genes as input achieves sensitivity 0.52 and specificity 0.97, reducing candidate lists 13-fold
-
Has reproduction · 55
Genome-Wide Survey and Development of the First Microsatellite Markers Database (AnCorDB) in Anemone coronaria L.
PMID 35328546 · PMC8949970 · International journal of molecular sciences · 2022 · 8 claims · 8 setups
Generated the first draft genome assembly of A. coronaria by Illumina sequencing a haploid androgenetic plant
-
Full-text index only
Computational disease gene identification: a concert of methods prioritizes type 2 diabetes and obesity candidate genes.
PMID 16757574 · PMC1475747 · Nucleic acids research · 2006 · 6 claims · 8 setups
Applying seven independent computational disease-gene prioritization methods in concert to 9556 positional candidate genes identifies a prioritized set of likely T2D and obesity candidate genes
-
Has reproduction · 95
Identification and Characterization of Small Noncoding RNAs in Genome Sequences of the Edible Fungus Pleurotus ostreatus.
PMID 27703969 · PMC5040776 · BioMed research international · 2016 · 7 claims · 8 setups
Genome-scale identification detected 254 small noncoding RNAs (snRNAs, snoRNAs, tRNAs, miRNAs, and other Rfam-classified sncRNAs) in the P. ostreatus CCEF00389 genome assembly